LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks Paper • 2608.01964 • Published 30 days ago • 183
Progressive Agent Skill Generation via Reinforcement Learning Paper • 2608.01678 • Published 30 days ago • 60
UEmbed: Unified Sparse and Dense Multimodal Embeddings Paper • 2608.02583 • Published 30 days ago • 52
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs Paper • 2608.01755 • Published 30 days ago • 142
SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation Paper • 2608.02287 • Published 30 days ago • 31
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Paper • 2607.29613 • Published Jul 31 • 28
GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning Paper • 2608.02585 • Published 30 days ago • 24
Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures Paper • 2607.28802 • Published Jul 30 • 10
RecHarness: A Bandit-Routed Agentic Harness for Self-Evolving Recommender Systems Paper • 2607.29241 • Published Jul 31 • 11
SG-WAM: Self-Guided World Modeling in Geometry-Aware Policy Space Paper • 2608.01397 • Published about 1 month ago • 10