PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation Paper • 2609.34759 • Published 14 days ago • 153
PAWBench: How Far Are We from Probabilistically Aligned World Modeling? Paper • 2608.27345 • Published Aug 27 • 62
PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation Paper • 2609.34759 • Published 14 days ago • 153
PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation Paper • 2609.34759 • Published 14 days ago • 153
PAWBench: How Far Are We from Probabilistically Aligned World Modeling? Paper • 2608.27345 • Published Aug 27 • 62
PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives Paper • 2608.13552 • Published Aug 13 • 48
PanoWorld: Towards Spatial Supersensing in 360$^\circ$ Panorama World Paper • 2605.13169 • Published May 13 • 20
GDRO: Group-level Reward Post-training Suitable for Diffusion Models Paper • 2601.02036 • Published Jan 5
OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions Paper • 2506.23361 • Published Jun 29, 2025 • 1
MemFlow: Flowing Adaptive Memory for Consistent and Efficient Long Video Narratives Paper • 2512.14699 • Published Dec 16, 2025 • 29
DiffDoctor: Diagnosing Image Diffusion Models Before Treating Paper • 2501.12382 • Published Jan 21, 2025
PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives Paper • 2608.13552 • Published Aug 13 • 48
PanoWorld: Towards Spatial Supersensing in 360^circ Panorama World Paper • 2605.13169 • Published May 13 • 20
PanoWorld: Towards Spatial Supersensing in 360^circ Panorama World Paper • 2605.13169 • Published May 13 • 20
Tstars-Tryon 1.0: Robust and Realistic Virtual Try-On for Diverse Fashion Items Paper • 2604.19748 • Published Apr 21 • 205
ShowUI-π: Flow-based Generative Models as GUI Dexterous Hands Paper • 2512.24965 • Published Dec 31, 2025 • 43