登录 注册
Read It Back: Pretrained MLLMs Are Zero-Shot Reward 模型 (Model)s for Text-to-Image Generation
Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation
👁 187 📚 26
PanoWorld: Real-World Panoramic Generation
👁 196 📚 7
LongE2V: Long-Horizon Event-based Video Reconstruction, 预测 (Prediction), and Frame Interpolation wit...
LongE2V: Long-Horizon Event-based Video Reconstruction, Prediction, and Frame Interpolation with Vid...
👁 55 📚 23
ZipDepth: Bringing Lightweight Zero-Shot Monocular Depth Anywhere, on Any Device
👁 64 📚 13
Wat3R: Underwater 3D Geometry 学习 (Learning) without Annotations
Wat3R: Underwater 3D Geometry Learning without Annotations
👁 130 📚 21
Selective Timestep Weighting and Advantage-Based Replay for Sample-Efficient Diffusion RLHF
👁 38 📚 19
Lift3D-VLA: Lifting VLA 模型 (Model)s to 3D Geometry and Dynamics-Aware Manipulation
Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation
👁 86 📚 11
SynCity 3000: Bootstrapping Scene-Scale 3D Diffusion
👁 59 📚 26
PointDiT: Pixel-Space Diffusion for Monocular Geometry 估计 (Estimation)
PointDiT: Pixel-Space Diffusion for Monocular Geometry Estimation
👁 66 📚 16
Alignment Is All You Need For X-to-4D Generation
👁 163 📚 4
WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory
👁 83 📚 15
Lyophilized platelet-rich fibrin (PRF) promotes craniofacial bone regeneration through Runx2.
👁 16 📚 0
Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative 模型 (Model)s
Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models
👁 87 📚 19
FaceMoE: Mixture of Experts for Low-Resolution Face Recognition
👁 83 📚 18
Open-Vocabulary and Referring Segmentation for 3D Gaussians Using 2D Detectors
👁 77 📚 17
PerceptionRubrics: Calibrating Multimodal Evaluation to Human Perception
👁 206 📚 19
World Action 模型 (Model)s Enable Continual Imitation 学习 (Learning) with Recurrent Generative Replays
World Action Models Enable Continual Imitation Learning with Recurrent Generative Replays
👁 161 📚 12
Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consist...
👁 144 📚 11
学习 (Learning) Action Priors for Cross-embodiment Robot Manipulation
Learning Action Priors for Cross-embodiment Robot Manipulation
👁 186 📚 20
DiffusionBench: On Holistic Evaluation of Diffusion Transformers
👁 166 📚 11
海洋智能体 🌊
海洋智能体
AI科研助手 · 2983篇文献
你好!你正在浏览文献列表,我可以帮你筛选方向、推荐高引论文或解读某个研究领域。