PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation Paper • 2609.34759 • Published 4 days ago • 153
PAWBench: How Far Are We from Probabilistically Aligned World Modeling? Paper • 2608.27345 • Published Aug 27 • 77
PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation Paper • 2609.34759 • Published 4 days ago • 153
PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives Paper • 2608.13552 • Published Aug 13 • 47
PanoWorld: Towards Spatial Supersensing in 360$^\circ$ Panorama World Paper • 2605.13169 • Published May 13 • 20
GDRO: Group-level Reward Post-training Suitable for Diffusion Models Paper • 2601.02036 • Published Jan 5
OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions Paper • 2506.23361 • Published Jun 29, 2025 • 1
MemFlow: Flowing Adaptive Memory for Consistent and Efficient Long Video Narratives Paper • 2512.14699 • Published Dec 16, 2025 • 29
DiffDoctor: Diagnosing Image Diffusion Models Before Treating Paper • 2501.12382 • Published Jan 21, 2025
PanoWorld: Towards Spatial Supersensing in 360^circ Panorama World Paper • 2605.13169 • Published May 13 • 20
TGDPO: Harnessing Token-Level Reward Guidance for Enhancing Direct Preference Optimization Paper • 2506.14574 • Published Jun 17, 2025 • 1
Animate-X++: Universal Character Image Animation with Dynamic Backgrounds Paper • 2508.09454 • Published Aug 13, 2025
Factuality Matters: When Image Generation and Editing Meet Structured Visuals Paper • 2510.05091 • Published Oct 6, 2025 • 20
From Noisy Traces to Stable Gradients: Bias-Variance Optimized Preference Optimization for Aligning Large Reasoning Models Paper • 2510.05095 • Published Oct 6, 2025 • 1
Stratified GRPO: Handling Structural Heterogeneity in Reinforcement Learning of LLM Search Agents Paper • 2510.06214 • Published Oct 7, 2025 • 1
PhysMaster: Mastering Physical Representation for Video Generation via Reinforcement Learning Paper • 2510.13809 • Published Oct 15, 2025 • 38
PICABench: How Far Are We from Physically Realistic Image Editing? Paper • 2510.17681 • Published Oct 20, 2025 • 49
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Paper • 2506.22434 • Published Jun 27, 2025 • 10
ObjectMover: Generative Object Movement with Video Prior Paper • 2503.08037 • Published Mar 11, 2025 • 5