- 
	
	
	
Fine-Tuning Language Models from Human Preferences
Paper • 1909.08593 • Published • 3 - 
	
	
	
PromptCoT: Synthesizing Olympiad-level Problems for Mathematical Reasoning in Large Language Models
Paper • 2503.02324 • Published - 
	
	
	
How Difficulty-Aware Staged Reinforcement Learning Enhances LLMs' Reasoning Capabilities: A Preliminary Experimental Study
Paper • 2504.00829 • Published - 
	
	
	
GPG: A Simple and Strong Reinforcement Learning Baseline for Model Reasoning
Paper • 2504.02546 • Published • 2 
Collections
Discover the best community collections!
Collections including paper arxiv:2404.00987 
						
					
				- 
	
	
	
3D Congealing: 3D-Aware Image Alignment in the Wild
Paper • 2404.02125 • Published • 10 - 
	
	
	
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 23 - 
	
	
	
ST-LLM: Large Language Models Are Effective Temporal Learners
Paper • 2404.00308 • Published • 8 - 
	
	
	
InstantSplat: Unbounded Sparse-view Pose-free Gaussian Splatting in 40 Seconds
Paper • 2403.20309 • Published • 19 
- 
	
	
	
Condition-Aware Neural Network for Controlled Image Generation
Paper • 2404.01143 • Published • 13 - 
	
	
	
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 23 - 
	
	
	
Advancing LLM Reasoning Generalists with Preference Trees
Paper • 2404.02078 • Published • 46 - 
	
	
	
ChatGLM-Math: Improving Math Problem-Solving in Large Language Models with a Self-Critique Pipeline
Paper • 2404.02893 • Published • 22 
- 
	
	
	
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
Paper • 2404.02905 • Published • 74 - 
	
	
	
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
Paper • 2404.02733 • Published • 22 - 
	
	
	
Cross-Attention Makes Inference Cumbersome in Text-to-Image Diffusion Models
Paper • 2404.02747 • Published • 13 - 
	
	
	
Bigger is not Always Better: Scaling Properties of Latent Diffusion Models
Paper • 2404.01367 • Published • 22 
- 
	
	
	
Fine-Tuning Language Models from Human Preferences
Paper • 1909.08593 • Published • 3 - 
	
	
	
PromptCoT: Synthesizing Olympiad-level Problems for Mathematical Reasoning in Large Language Models
Paper • 2503.02324 • Published - 
	
	
	
How Difficulty-Aware Staged Reinforcement Learning Enhances LLMs' Reasoning Capabilities: A Preliminary Experimental Study
Paper • 2504.00829 • Published - 
	
	
	
GPG: A Simple and Strong Reinforcement Learning Baseline for Model Reasoning
Paper • 2504.02546 • Published • 2 
- 
	
	
	
3D Congealing: 3D-Aware Image Alignment in the Wild
Paper • 2404.02125 • Published • 10 - 
	
	
	
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 23 - 
	
	
	
ST-LLM: Large Language Models Are Effective Temporal Learners
Paper • 2404.00308 • Published • 8 - 
	
	
	
InstantSplat: Unbounded Sparse-view Pose-free Gaussian Splatting in 40 Seconds
Paper • 2403.20309 • Published • 19 
- 
	
	
	
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
Paper • 2404.02905 • Published • 74 - 
	
	
	
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
Paper • 2404.02733 • Published • 22 - 
	
	
	
Cross-Attention Makes Inference Cumbersome in Text-to-Image Diffusion Models
Paper • 2404.02747 • Published • 13 - 
	
	
	
Bigger is not Always Better: Scaling Properties of Latent Diffusion Models
Paper • 2404.01367 • Published • 22 
- 
	
	
	
Condition-Aware Neural Network for Controlled Image Generation
Paper • 2404.01143 • Published • 13 - 
	
	
	
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 23 - 
	
	
	
Advancing LLM Reasoning Generalists with Preference Trees
Paper • 2404.02078 • Published • 46 - 
	
	
	
ChatGLM-Math: Improving Math Problem-Solving in Large Language Models with a Self-Critique Pipeline
Paper • 2404.02893 • Published • 22