-
Self-Rewarding Language Models
Paper • 2401.10020 • Published • 143 -
Orion-14B: Open-source Multilingual Large Language Models
Paper • 2401.12246 • Published • 11 -
MambaByte: Token-free Selective State Space Model
Paper • 2401.13660 • Published • 50 -
MM-LLMs: Recent Advances in MultiModal Large Language Models
Paper • 2401.13601 • Published • 44
Collections
Discover the best community collections!
Collections including paper arxiv:2409.17115
-
LinFusion: 1 GPU, 1 Minute, 16K Image
Paper • 2409.02097 • Published • 31 -
Phidias: A Generative Model for Creating 3D Content from Text, Image, and 3D Conditions with Reference-Augmented Diffusion
Paper • 2409.11406 • Published • 25 -
Diffusion Models Are Real-Time Game Engines
Paper • 2408.14837 • Published • 121 -
Segment Anything with Multiple Modalities
Paper • 2408.09085 • Published • 21
-
Programming Every Example: Lifting Pre-training Data Quality like Experts at Scale
Paper • 2409.17115 • Published • 59 -
gair-prox/FineWeb-pro
Viewer • Updated • 63.1M • 3.51k • 14 -
gair-prox/open-web-math-pro
Viewer • Updated • 2.58M • 2.87k • 8 -
gair-prox/RedPajama-pro
Viewer • Updated • 10.2M • 1.27k • 4
-
Qwen2.5-Coder Technical Report
Paper • 2409.12186 • Published • 125 -
Attention Heads of Large Language Models: A Survey
Paper • 2409.03752 • Published • 87 -
Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency
Paper • 2409.02634 • Published • 87 -
OmniGen: Unified Image Generation
Paper • 2409.11340 • Published • 106
-
Automated Design of Agentic Systems
Paper • 2408.08435 • Published • 38 -
On the limits of agency in agent-based models
Paper • 2409.10568 • Published • 12 -
On the Diagram of Thought
Paper • 2409.10038 • Published • 11 -
DSBench: How Far Are Data Science Agents to Becoming Data Science Experts?
Paper • 2409.07703 • Published • 66