OpenVisTool: An Open Recipe for Synthesizing Instructive Visual Tool-Use Trajectories Paper • 2608.08557 • Published 12 days ago • 2
Enjoy Your Talk: A Human-Centered Benchmark for Multi-Turn Dialogue with Decoupled User Simulation, Target Modeling, and Judging Paper • 2607.10428 • Published 28 days ago • 1
Omni-Perception Policy Optimization for Multimodal Emotion Reasoning Paper • 2606.25325 • Published Jun 24
MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy Paper • 2606.27652 • Published Jun 26
EgoPro-Bench: Benchmarking Personalized Proactive Interaction in Egocentric Video Streams Paper • 2605.07299 • Published May 8
ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories Paper • 2606.11520 • Published Jun 9
From Pixels to Words -- Towards Native One-Vision Models at Scale Paper • 2605.28820 • Published May 27 • 76
GeoSym127K: Scalable Symbolically-verifiable Synthesis for Multimodal Geometric Reasoning Paper • 2605.16371 • Published May 10
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture Paper • 2605.12500 • Published May 12 • 195
OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesis Paper • 2604.15093 • Published Apr 16 • 30
ReinDriveGen: Reinforcement Post-Training for Out-of-Distribution Driving Scene Generation Paper • 2604.01129 • Published Apr 1 • 8
EVA: Efficient Reinforcement Learning for End-to-End Video Agent Paper • 2603.22918 • Published Mar 24 • 44
Delving into the Devils of Bird's-eye-view Perception: A Review, Evaluation and Recipe Paper • 2209.05324 • Published Sep 12, 2022
ICA: Information-Aware Credit Assignment for Visually Grounded Long-Horizon Information-Seeking Agents Paper • 2602.10863 • Published Feb 11 • 10
SenseNova-MARS: Empowering Multimodal Agentic Reasoning and Search via Reinforcement Learning Paper • 2512.24330 • Published Dec 30, 2025 • 36
Scaling Spatial Intelligence with Multimodal Foundation Models Paper • 2511.13719 • Published Nov 17, 2025 • 50
Enhancing the Outcome Reward-based RL Training of MLLMs with Self-Consistency Sampling Paper • 2511.10648 • Published Nov 13, 2025 • 1