Collections
Discover the best community collections!
Collections including paper arxiv:2603.28589
-
SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization
Paper • 2604.02268 • Published • 99 -
ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning
Paper • 2603.05863 • Published • 6 -
GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning
Paper • 2604.02721 • Published • 85 -
GLM-5: from Vibe Coding to Agentic Engineering
Paper • 2602.15763 • Published • 221
-
TOUCAN: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments
Paper • 2510.01179 • Published • 29 -
Towards a Medical AI Scientist
Paper • 2603.28589 • Published • 90 -
MinT: Managed Infrastructure for Training and Serving Millions of LLMs
Paper • 2605.13779 • Published • 225 -
Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling
Paper • 2605.13301 • Published • 165
-
BitNet: Scaling 1-bit Transformers for Large Language Models
Paper • 2310.11453 • Published • 108 -
Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection
Paper • 2310.11511 • Published • 79 -
In-Context Learning Creates Task Vectors
Paper • 2310.15916 • Published • 43 -
Matryoshka Diffusion Models
Paper • 2310.15111 • Published • 46
-
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain
Paper • 2509.26507 • Published • 557 -
mHC: Manifold-Constrained Hyper-Connections
Paper • 2512.24880 • Published • 336 -
NeoVerse: Enhancing 4D World Model with in-the-wild Monocular Videos
Paper • 2601.00393 • Published • 133 -
LTX-2: Efficient Joint Audio-Visual Foundation Model
Paper • 2601.03233 • Published • 196
-
SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization
Paper • 2604.02268 • Published • 99 -
ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning
Paper • 2603.05863 • Published • 6 -
GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning
Paper • 2604.02721 • Published • 85 -
GLM-5: from Vibe Coding to Agentic Engineering
Paper • 2602.15763 • Published • 221
-
TOUCAN: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments
Paper • 2510.01179 • Published • 29 -
Towards a Medical AI Scientist
Paper • 2603.28589 • Published • 90 -
MinT: Managed Infrastructure for Training and Serving Millions of LLMs
Paper • 2605.13779 • Published • 225 -
Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling
Paper • 2605.13301 • Published • 165
-
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain
Paper • 2509.26507 • Published • 557 -
mHC: Manifold-Constrained Hyper-Connections
Paper • 2512.24880 • Published • 336 -
NeoVerse: Enhancing 4D World Model with in-the-wild Monocular Videos
Paper • 2601.00393 • Published • 133 -
LTX-2: Efficient Joint Audio-Visual Foundation Model
Paper • 2601.03233 • Published • 196
-
BitNet: Scaling 1-bit Transformers for Large Language Models
Paper • 2310.11453 • Published • 108 -
Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection
Paper • 2310.11511 • Published • 79 -
In-Context Learning Creates Task Vectors
Paper • 2310.15916 • Published • 43 -
Matryoshka Diffusion Models
Paper • 2310.15111 • Published • 46