SKILLER: Language-Level Reinforcement Learning for Reusable Skill Extraction in Small Language Models Paper • 2608.10538 • Published 8 days ago • 14
Power law graph attention: exact generalization of scaled dot-product attention, empirical collapse at inference Paper • 2608.10288 • Published 9 days ago • 7
FocusMem: Factorizing Content, Readout, and Trust in Latent GUI Memory Paper • 2608.04530 • Published 14 days ago • 14
Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing Paper • 2608.02711 • Published 15 days ago • 89
QQWorld: Quantile-Quantile Matching for World Model Regularization Paper • 2607.28415 • Published 20 days ago • 30