AutoResearch-RL: Perpetual Self-Evaluating Reinforcement Learning Agents for Autonomous Neural Architecture Discovery Paper • 2603.07300 • Published Mar 7 • 20
view article Article Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident +2 hlarcher, XciD, raphael-gl, chris-rannou • Jul 27 • 489
Group-in-Group Policy Optimization for LLM Agent Training Paper • 2505.10978 • Published May 16, 2025 • 24
GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation Identity Paper • 2607.00152 • Published Jun 30 • 9
When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling Paper • 2606.28661 • Published Jun 27 • 8
MolmoAct2 Models Collection Collection of the base models for MolmoAct2 • 6 items • Updated May 5 • 24
A General Theoretical Paradigm to Understand Learning from Human Preferences Paper • 2310.12036 • Published Oct 18, 2023 • 20
ProcessBench: Identifying Process Errors in Mathematical Reasoning Paper • 2412.06559 • Published Dec 9, 2024 • 87
Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document Collections Paper • 2603.12180 • Published Mar 12 • 65
NVIDIA Nemotron v3 Collection Open, Production-ready Enterprise Models • 33 items • Updated 25 days ago • 369
Scaling Open Discrete Audio Foundation Models with Interleaved Semantic, Acoustic, and Text Tokens Paper • 2602.16687 • Published Feb 18 • 5
Typhoon-S: Minimal Open Post-Training for Sovereign Large Language Models Paper • 2601.18129 • Published Jan 26 • 11
Proof or Bluff? Evaluating LLMs on 2025 USA Math Olympiad Paper • 2503.21934 • Published Mar 27, 2025 • 1
GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models Paper • 2508.06471 • Published Aug 8, 2025 • 214
DeepSearch: Overcome the Bottleneck of Reinforcement Learning with Verifiable Rewards via Monte Carlo Tree Search Paper • 2509.25454 • Published Sep 29, 2025 • 147
Typhoon ASR Real-time: FastConformer-Transducer for Thai Automatic Speech Recognition Paper • 2601.13044 • Published Jan 19 • 12
On the Robustness of Answer Formats in Medical Reasoning Models Paper • 2509.20866 • Published Sep 25, 2025 • 2
Typhoon OCR: Open Vision-Language Model For Thai Document Extraction Paper • 2601.14722 • Published Jan 21 • 15