WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report Paper • 2608.24053 • Published 6 days ago • 67
Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs Paper • 2608.20492 • Published 11 days ago • 108
Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models Paper • 2608.25518 • Published 5 days ago • 175
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 5 days ago • 173
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 5 days ago • 173
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 5 days ago • 173
Spreadsheet-RL: Advancing Large Language Model Agents on Realistic Spreadsheet Tasks via Reinforcement Learning Paper • 2605.22642 • Published May 21 • 36
Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling Paper • 2604.28185 • Published Apr 30 • 92
Watch Before You Answer: Learning from Visually Grounded Post-Training Paper • 2604.05117 • Published Apr 6 • 36