-
Fine-Tuning Language Models from Human Preferences
Paper • 1909.08593 • Published • 3 -
PromptCoT: Synthesizing Olympiad-level Problems for Mathematical Reasoning in Large Language Models
Paper • 2503.02324 • Published -
How Difficulty-Aware Staged Reinforcement Learning Enhances LLMs' Reasoning Capabilities: A Preliminary Experimental Study
Paper • 2504.00829 • Published -
GPG: A Simple and Strong Reinforcement Learning Baseline for Model Reasoning
Paper • 2504.02546 • Published • 1
Collections
Discover the best community collections!
Collections including paper arxiv:2404.00987
-
3D Congealing: 3D-Aware Image Alignment in the Wild
Paper • 2404.02125 • Published • 10 -
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 24 -
ST-LLM: Large Language Models Are Effective Temporal Learners
Paper • 2404.00308 • Published • 8 -
InstantSplat: Unbounded Sparse-view Pose-free Gaussian Splatting in 40 Seconds
Paper • 2403.20309 • Published • 19
-
Condition-Aware Neural Network for Controlled Image Generation
Paper • 2404.01143 • Published • 13 -
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 24 -
Advancing LLM Reasoning Generalists with Preference Trees
Paper • 2404.02078 • Published • 47 -
ChatGLM-Math: Improving Math Problem-Solving in Large Language Models with a Self-Critique Pipeline
Paper • 2404.02893 • Published • 23
-
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
Paper • 2404.02905 • Published • 74 -
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
Paper • 2404.02733 • Published • 23 -
Cross-Attention Makes Inference Cumbersome in Text-to-Image Diffusion Models
Paper • 2404.02747 • Published • 13 -
Bigger is not Always Better: Scaling Properties of Latent Diffusion Models
Paper • 2404.01367 • Published • 23
-
Fine-Tuning Language Models from Human Preferences
Paper • 1909.08593 • Published • 3 -
PromptCoT: Synthesizing Olympiad-level Problems for Mathematical Reasoning in Large Language Models
Paper • 2503.02324 • Published -
How Difficulty-Aware Staged Reinforcement Learning Enhances LLMs' Reasoning Capabilities: A Preliminary Experimental Study
Paper • 2504.00829 • Published -
GPG: A Simple and Strong Reinforcement Learning Baseline for Model Reasoning
Paper • 2504.02546 • Published • 1
-
3D Congealing: 3D-Aware Image Alignment in the Wild
Paper • 2404.02125 • Published • 10 -
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 24 -
ST-LLM: Large Language Models Are Effective Temporal Learners
Paper • 2404.00308 • Published • 8 -
InstantSplat: Unbounded Sparse-view Pose-free Gaussian Splatting in 40 Seconds
Paper • 2403.20309 • Published • 19
-
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
Paper • 2404.02905 • Published • 74 -
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
Paper • 2404.02733 • Published • 23 -
Cross-Attention Makes Inference Cumbersome in Text-to-Image Diffusion Models
Paper • 2404.02747 • Published • 13 -
Bigger is not Always Better: Scaling Properties of Latent Diffusion Models
Paper • 2404.01367 • Published • 23
-
Condition-Aware Neural Network for Controlled Image Generation
Paper • 2404.01143 • Published • 13 -
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
Paper • 2404.00987 • Published • 24 -
Advancing LLM Reasoning Generalists with Preference Trees
Paper • 2404.02078 • Published • 47 -
ChatGLM-Math: Improving Math Problem-Solving in Large Language Models with a Self-Critique Pipeline
Paper • 2404.02893 • Published • 23