今日从 arXiv 订阅中筛选 10 篇论文。

⚡ RecurTrace: Adaptive Latent Reasoning with Loop-Time Memory

RecurTrace: Adaptive Latent Reasoning with Loop-Time Memory

⚡ WorldReward: Reward Modeling for Camera-Conditioned World Models

WorldReward: Reward Modeling for Camera-Conditioned World Models

⚡ SV-WAM: An Efficient Surround-View World-Action Model for End-to-End Autonomous Driving

SV-WAM: An Efficient Surround-View World-Action Model for End-to-End Autonomous Driving

⚡ Drive-HWM: Hierarchical World Models for Dynamic-Latent Guided Autonomous Driving

Drive-HWM: Hierarchical World Models for Dynamic-Latent Guided Autonomous Driving

⚡ Continuous Actions from Discrete Minds: Latent-Aligned Planning for End-to-End Autonomous Driving

Continuous Actions from Discrete Minds: Latent-Aligned Planning for End-to-End Autonomous Driving

⚡ Beyond Retrieval: Progressive Latent Memory Evolution for Streaming Video Understanding

Beyond Retrieval: Progressive Latent Memory Evolution for Streaming Video Understanding

⚡ Unfold The World: Factorize 4D Properties in Reinforcing Spatial Reasoning

Unfold The World: Factorize 4D Properties in Reinforcing Spatial Reasoning

⚡ Do Video Generators Track the World Across Segments? A Benchmark and Method for World-State Reasoning in Video Continuation

Do Video Generators Track the World Across Segments? A Benchmark and Method for World-State Reasoning in Video Continuation

⚡ The Shape of Time: Video-Token Contrast for Temporal Understanding in VideoLMs

The Shape of Time: Video-Token Contrast for Temporal Understanding in VideoLMs

⚡ Puffin-World: Scaling a Unified Multimodal Model with Native 3D World States

Puffin-World: Scaling a Unified Multimodal Model with Native 3D World States

自动生成于 2026-09-05 · 基于 arXiv Daily Digest