HuatuoGPT-3: RL-Only Domain Adaptation from Base Models Paper • 2610.05966 • Published 6 days ago • 39
VeriHarness: Scaling Agentic Verification for Long-Horizon Tasks Paper • 2610.00972 • Published 10 days ago • 60
PointWAM: 3D World Action Modeling for Dexterous Robotic Manipulation Paper • 2610.02840 • Published 9 days ago • 61
Coding Agents for Generalized Task and Motion Planning Problems Paper • 2609.30233 • Published 17 days ago • 28
TrackEverything: Long Horizon Dense Tracking via De-Duplicating 3D Scene Representations Paper • 2609.30222 • Published 17 days ago • 14
Disaggregated Quantization: Specializing LLM Prefill and Decode Paper • 2609.26333 • Published 19 days ago • 93
Depth-adaptive Inference of Looped Language Models via Continuous Depth Batching Paper • 2608.09444 • Published 16 days ago • 13
WorldCrafter: Consistent Video World Model with Implicit 3D-aware Memory Paper • 2609.24984 • Published 20 days ago • 158
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 20 days ago • 225