-
CatLIP: CLIP-level Visual Recognition Accuracy with 2.7x Faster Pre-training on Web-scale Image-Text Data
Paper • 2404.15653 • Published • 28 -
MoDE: CLIP Data Experts via Clustering
Paper • 2404.16030 • Published • 13 -
MoRA: High-Rank Updating for Parameter-Efficient Fine-Tuning
Paper • 2405.12130 • Published • 49 -
Reducing Transformer Key-Value Cache Size with Cross-Layer Attention
Paper • 2405.12981 • Published • 33
Collections
Discover the best community collections!
Collections including paper arxiv:2505.07812
-
Continuous Diffusion Model for Language Modeling
Paper • 2502.11564 • Published • 53 -
EasyControl: Adding Efficient and Flexible Control for Diffusion Transformer
Paper • 2503.07027 • Published • 30 -
Efficient Generative Model Training via Embedded Representation Warmup
Paper • 2504.10188 • Published • 12 -
Improving Editability in Image Generation with Layer-wise Memory
Paper • 2505.01079 • Published • 29
-
CatLIP: CLIP-level Visual Recognition Accuracy with 2.7x Faster Pre-training on Web-scale Image-Text Data
Paper • 2404.15653 • Published • 28 -
MoDE: CLIP Data Experts via Clustering
Paper • 2404.16030 • Published • 13 -
MoRA: High-Rank Updating for Parameter-Efficient Fine-Tuning
Paper • 2405.12130 • Published • 49 -
Reducing Transformer Key-Value Cache Size with Cross-Layer Attention
Paper • 2405.12981 • Published • 33
-
Efficient Generative Modeling with Residual Vector Quantization-Based Tokens
Paper • 2412.10208 • Published • 19 -
Normalizing Flows are Capable Generative Models
Paper • 2412.06329 • Published • 10 -
A Noise is Worth Diffusion Guidance
Paper • 2412.03895 • Published • 29 -
Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
Paper • 2501.01423 • Published • 44
-
CatLIP: CLIP-level Visual Recognition Accuracy with 2.7x Faster Pre-training on Web-scale Image-Text Data
Paper • 2404.15653 • Published • 28 -
MoDE: CLIP Data Experts via Clustering
Paper • 2404.16030 • Published • 13 -
MoRA: High-Rank Updating for Parameter-Efficient Fine-Tuning
Paper • 2405.12130 • Published • 49 -
Reducing Transformer Key-Value Cache Size with Cross-Layer Attention
Paper • 2405.12981 • Published • 33
-
Continuous Diffusion Model for Language Modeling
Paper • 2502.11564 • Published • 53 -
EasyControl: Adding Efficient and Flexible Control for Diffusion Transformer
Paper • 2503.07027 • Published • 30 -
Efficient Generative Model Training via Embedded Representation Warmup
Paper • 2504.10188 • Published • 12 -
Improving Editability in Image Generation with Layer-wise Memory
Paper • 2505.01079 • Published • 29
-
Efficient Generative Modeling with Residual Vector Quantization-Based Tokens
Paper • 2412.10208 • Published • 19 -
Normalizing Flows are Capable Generative Models
Paper • 2412.06329 • Published • 10 -
A Noise is Worth Diffusion Guidance
Paper • 2412.03895 • Published • 29 -
Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
Paper • 2501.01423 • Published • 44
-
CatLIP: CLIP-level Visual Recognition Accuracy with 2.7x Faster Pre-training on Web-scale Image-Text Data
Paper • 2404.15653 • Published • 28 -
MoDE: CLIP Data Experts via Clustering
Paper • 2404.16030 • Published • 13 -
MoRA: High-Rank Updating for Parameter-Efficient Fine-Tuning
Paper • 2405.12130 • Published • 49 -
Reducing Transformer Key-Value Cache Size with Cross-Layer Attention
Paper • 2405.12981 • Published • 33