RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States Paper • 2608.02508 • Published 25 days ago • 13
ReferTrack: Referring Then Tracking for Embodied Visual Tracking Paper • 2607.20061 • Published Jul 22 • 53
nightmedia/Qwen3.6-27B-Architect-Polaris2-Fable-B-F451-Tess-1M-qx64-hi-mlx Image-Text-to-Text • 27B • Updated Jul 27 • 339 • 4
Bibby AI: An Editor-Native Agentic Platform for Academic Research, Writing, and Publishing Paper • 2607.05435 • Published Jul 3 • 6
OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation Paper • 2606.17628 • Published Jun 16 • 29
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models Paper • 2606.11324 • Published Jun 9 • 173
DecMem: Towards Minute-Long Consistent World Generation with Decoupled Memory Paper • 2605.31336 • Published May 29 • 13
Joint Training of Multi-Token Prediction in Reinforcement Learning via Optimal Coefficient Calibration Paper • 2605.28184 • Published May 27 • 6