登录 注册
SplashSplat: Reconstructing Splashing Liquids from Real-World Multi-View Videos
👁 50 📚 30
Can 4D Foundation 模型 (Model)s Remember?
Can 4D Foundation Models Remember?
👁 94 📚 8
PANORAMA: Panoptic Grounded Captioning via Mask Proposal Selection
👁 173 📚 16
PhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Graine...
👁 54 📚 2
A Chosen Future Can Still Be Rewritten: Causal Writability in Video 模型 (Model)s
A Chosen Future Can Still Be Rewritten: Causal Writability in Video Models
👁 97 📚 3
SNAP3D: Physically Grounded 3D Parts for Assembly from a Single Image
👁 78 📚 20
MindTopo: Can Foundation 模型 (Model)s Reason in Topological Space?
MindTopo: Can Foundation Models Reason in Topological Space?
👁 75 📚 10
SenseNova-U1.5: Towards Native Unified Visual Intelligence
👁 201 📚 19
Programmable World 模型 (Model)
Programmable World Model
👁 31 📚 19
SyncWorld: Visual Calibration Enables World 模型 (Model)s as Zero-Shot Simulators
SyncWorld: Visual Calibration Enables World Models as Zero-Shot Simulators
👁 135 📚 7
WorldSculpt: Generating Compositional Worlds from Grounded Videos
👁 61 📚 0
Scal3R: 学习 (Learning) Efficient Multi-Relative Pose Query for Scalable Online 3D Reconstruction
Scal3R: Learning Efficient Multi-Relative Pose Query for Scalable Online 3D Reconstruction
👁 99 📚 19
TokenMatch: 3D Mesh Correspondence Transformer with Curvature-Guided Tokenisation
👁 37 📚 17
Temporal Self-Distillation: 学习 (Learning) Visual State Tracking in Videos Without Supervision
Temporal Self-Distillation: Learning Visual State Tracking in Videos Without Supervision
👁 182 📚 9
SolarWM: Open 数据 (Data) and Scalable Training for Long-Horizon Video World 模型 (Model)s
SolarWM: Open Data and Scalable Training for Long-Horizon Video World Models
👁 123 📚 21
Uncovering Understanding-Generation Synergy in Native Unified Multimodal 模型 (Model)s: From Represent...
Uncovering Understanding-Generation Synergy in Native Unified Multimodal Models: From Representation...
👁 146 📚 7
BRF-GS: Hyperspectral Bidirectional Reflectance Factor 模型 (Model)ing and Image Generation Based on 3...
BRF-GS: Hyperspectral Bidirectional Reflectance Factor Modeling and Image Generation Based on 3D Gau...
👁 46 📚 22
SignRR: Retrieve and Refine Real Motion for Sign Language Production
👁 76 📚 16
Reconstructing Humans and Objects in Interaction using Large Reconstruction 模型 (Model)s
Reconstructing Humans and Objects in Interaction using Large Reconstruction Models
👁 41 📚 12
Retrieval Heads Meet Vision: Uncovering How VLMs Locate and Extract Visual Information
👁 61 📚 12
海洋智能体 🌊
海洋智能体
AI科研助手 · 3665篇文献
你好!你正在浏览文献列表,我可以帮你筛选方向、推荐高引论文或解读某个研究领域。