Act with Intent: Distilling Behavior Intent for Vision-Language-Action Models Paper • 2608.23478 • Published 8 days ago • 24
LoopArena: Benchmarking Models as Runtime Controllers for Loop Engineering Paper • 2608.28281 • Published 4 days ago • 82
Training Agents to Evolve with Their Harness: TaoLive Digital Avatar Agent Technical Report Paper • 2608.15763 • Published 10 days ago • 47
VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning Paper • 2608.26105 • Published 6 days ago • 266
PILOT in the Loop: Live Self-Improvement for Long-Horizon Agents Paper • 2608.26530 • Published 5 days ago • 30
MA-VLA: Multi-Arm Vision-Language-Action Model for Collaboration and Compositional Generalization Paper • 2608.25864 • Published 6 days ago • 9
JIT-Agent: Scaling Harness Intelligence via Just-in-Time Harness Evolution Paper • 2608.25593 • Published 6 days ago • 111
GigaBrain-0.7: Scaling Embodied Foundation Models to Emergent Capabilities with a Three-System Architecture Paper • 2608.15875 • Published 16 days ago • 103
Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs Paper • 2608.20492 • Published 12 days ago • 110
TileMix: Tile-Centric Mixed-Precision Attention for LLM Inference Acceleration Paper • 2608.17336 • Published 14 days ago • 5
From Generation to Simulation: How Far Are World Models from Being True Simulators? Paper • 2608.23070 • Published 8 days ago • 4
Task-CoEvolve: Efficient Harness Optimization via Adaptive Validation Task Selection Paper • 2608.20169 • Published 8 days ago • 11
Towards a Densing Law for User Representation Learning at Billion-Scale Capacity Paper • 2608.23392 • Published 8 days ago • 28
TLive-Omni: An Omni-Modal Understanding Model for E-Commerce Live Streaming Paper • 2608.20958 • Published 11 days ago • 58