view article Article Introducing Real World VoiceEQ: Measuring the human quality of voice AI +11 dayllon, aliceebaird, jeffbrooks, francamps, jpc, tlebryk02, jens-hume-ai, itsolyaossi, sharath25, hoon-hume, tig88, rashisht, tzirakis • 25 days ago • 30
view article Article TutorMoments: Do AI tutors know when to help and when to hold back? allenai • 1 day ago • 14
CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks Paper • 2608.06352 • Published 3 days ago • 16
PaDoc: Layout-Grounded Parallel Decoding for Document Parsing Paper • 2608.06146 • Published 3 days ago • 16
ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment Paper • 2608.05102 • Published 4 days ago • 63
EffectLearner: World-Aware Object-Effect Reasoning for Real-World Video Object Removal Paper • 2608.05565 • Published 3 days ago • 16
Learning from Failures: Retrieval-Centric CoT via Hard Negatives for Unified Multimodal Retrieval Paper • 2608.06060 • Published 3 days ago • 31
CAPEval: A Decoupled Caption Evaluation across Understanding and Generation Paper • 2608.02589 • Published 6 days ago • 25
Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging Paper • 2608.03316 • Published 5 days ago • 26
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models Paper • 2608.03812 • Published 5 days ago • 26
LLaDA MoE v2: Scaling Mixture-of-Experts Diffusion Language Models Paper • 2608.03457 • Published 5 days ago • 31
Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent Paper • 2608.03979 • Published 5 days ago • 49
Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing Paper • 2608.02711 • Published 6 days ago • 82
AURORA-LM: Autoencoding Unified Representation for Continuous-Latent Diffusion Language Modeling Paper • 2608.02602 • Published 6 days ago • 77
JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion Paper • 2608.03974 • Published 5 days ago • 87
ReToken: One Token to Improve Vision-Language Models for Visual Retrieval Paper • 2607.28627 • Published 10 days ago • 9