Data Scaling Laws in Imitation Learning for Robotic Manipulation Paper • 2410.18647 • Published Oct 24, 2024 • 8
DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset Paper • 2403.12945 • Published Mar 19, 2024 • 2
AutoSaddler: Automatic Harness Optimization with Durable Updates from Agent Execution Traces Paper • 2608.23041 • Published 4 days ago • 59
Qwythos v1 Collection A collection of our uncensored Qwen Claude Mythos fine tunes • 2 items • Updated Jul 11 • 36
KG-Agent: An Efficient Autonomous Agent Framework for Complex Reasoning over Knowledge Graph Paper • 2402.11163 • Published Feb 17, 2024 • 3
view article Article Introducing smolagents: simple agents that write actions in code. +1 m-ric, merve, thomwolf • Dec 31, 2024 • 1.21k
view article Article Harness, Scaffold, and the AI Agent Terms Worth Getting Right sergiopaniego, ariG23498 • May 25 • 143
view article Article DeepSeek-V4: a million-token context that agents can actually use burtenshaw • Apr 24 • 56
Communication is All You Need: Persuasion Dataset Construction via Multi-LLM Communication Paper • 2502.08896 • Published Feb 13, 2025 • 1
Eliciting and Analyzing Emergent Misalignment in State-of-the-Art Large Language Models Paper • 2508.04196 • Published Aug 6, 2025 • 2
Language of Persuasion and Misrepresentation in Business Communication: A Textual Detection Approach Paper • 2508.09935 • Published Aug 13, 2025 • 1
Natural Emergent Misalignment from Reward Hacking in Production RL Paper • 2511.18397 • Published Nov 23, 2025 • 2
LLM Can be a Dangerous Persuader: Empirical Study of Persuasion Safety in Large Language Models Paper • 2504.10430 • Published Apr 14, 2025 • 6