HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? Paper • 2609.01437 • Published 6 days ago • 240
HarnessEval-W: Agentifying the Evaluation of Visual Worlds Paper • 2608.16859 • Published 21 days ago • 340
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published 26 days ago • 291
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 28 days ago • 342
Recursive Synthesis for Long-Horizon Terminal Tasks Paper • 2608.05466 • Published about 1 month ago • 252
TARS: Timestep-Aware Data Scaling for 3D-Free Video Re-Shooting Paper • 2607.28261 • Published Jul 30 • 116
DistillAlign: Coordinating Mode Covering and Mode Seeking in Autoregressive Video Distillation Paper • 2607.26811 • Published Jul 29 • 93
dRAE: Representation Autoencoder with Hyper-Spherical Codes Paper • 2607.22148 • Published Jul 24 • 12
Where Did It Go Wrong? Process-Level Evaluation of Web Agents with Semantic State Tracking Paper • 2606.15673 • Published Apr 8 • 14