Invent a Dataset: Measuring dataset generation abilities with zero seed
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search
Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090
MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies
Muon is Scalable for LLM Training
Kimi K3: Open Frontier Intelligence
LLMRouterBench: A Massive Benchmark and Unified Framework for LLM Routing
Nemotron 3 Nano Open Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
SERA Soft-Verified Efficient Repository Agents
Nemotron 3 Super Open Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
Scaling Latent Reasoning via Looped Language Models
dots.ocr Multilingual Document Layout Parsing in a Single Vision-Language Model
2 OLMo 2 Furious
Qwen3 Technical Report