Gespeichert in:
| Hauptverfasser: | Litman, Yehonathan, Liu, Shikun, Seyb, Dario, Milef, Nicholas, Zhou, Yang, Marshall, Carl, Tulsiani, Shubham, Leak, Caleb |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.15031 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LightSwitch: Multi-view Relighting with Material-guided Diffusion
von: Litman, Yehonathan, et al.
Veröffentlicht: (2025)
von: Litman, Yehonathan, et al.
Veröffentlicht: (2025)
MaterialFusion: Enhancing Inverse Rendering with Material Diffusion Priors
von: Litman, Yehonathan, et al.
Veröffentlicht: (2024)
von: Litman, Yehonathan, et al.
Veröffentlicht: (2024)
Towards Unstructured Unlabeled Optical Mocap: A Video Helps!
von: Milef, Nicholas, et al.
Veröffentlicht: (2024)
von: Milef, Nicholas, et al.
Veröffentlicht: (2024)
Real‐Time Neural Materials on Mobile VR
von: Zilin Xu, et al.
Veröffentlicht: (2026)
von: Zilin Xu, et al.
Veröffentlicht: (2026)
Sparse-view Pose Estimation and Reconstruction via Analysis by Generative Synthesis
von: Zhao, Qitao, et al.
Veröffentlicht: (2024)
von: Zhao, Qitao, et al.
Veröffentlicht: (2024)
SceneFactor: Factored Latent 3D Diffusion for Controllable 3D Scene Generation
von: Bokhovkin, Alexey, et al.
Veröffentlicht: (2024)
von: Bokhovkin, Alexey, et al.
Veröffentlicht: (2024)
REST3D: Reconstructing Physically Stable 3D Scenes from a Single Image
von: Ma, Xiaoxuan, et al.
Veröffentlicht: (2026)
von: Ma, Xiaoxuan, et al.
Veröffentlicht: (2026)
DriveCtrl: Conditioned Sim-to-Real Driving Video Generation
von: Zhao, Haonan, et al.
Veröffentlicht: (2026)
von: Zhao, Haonan, et al.
Veröffentlicht: (2026)
Dex4D: Task-Agnostic Point Track Policy for Sim-to-Real Dexterous Manipulation
von: Kuang, Yuxuan, et al.
Veröffentlicht: (2026)
von: Kuang, Yuxuan, et al.
Veröffentlicht: (2026)
Track2Act: Predicting Point Tracks from Internet Videos enables Generalizable Robot Manipulation
von: Bharadhwaj, Homanga, et al.
Veröffentlicht: (2024)
von: Bharadhwaj, Homanga, et al.
Veröffentlicht: (2024)
EgoEdit: Dataset, Real-Time Streaming Model, and Benchmark for Egocentric Video Editing
von: Li, Runjia, et al.
Veröffentlicht: (2025)
von: Li, Runjia, et al.
Veröffentlicht: (2025)
CameraCtrl: Enabling Camera Control for Text-to-Video Generation
von: He, Hao, et al.
Veröffentlicht: (2024)
von: He, Hao, et al.
Veröffentlicht: (2024)
PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation
von: Wang, Chen, et al.
Veröffentlicht: (2025)
von: Wang, Chen, et al.
Veröffentlicht: (2025)
DemoDiffusion: One-Shot Human Imitation using pre-trained Diffusion Policy
von: Park, Sungjae, et al.
Veröffentlicht: (2025)
von: Park, Sungjae, et al.
Veröffentlicht: (2025)
G-HOP: Generative Hand-Object Prior for Interaction Reconstruction and Grasp Synthesis
von: Ye, Yufei, et al.
Veröffentlicht: (2024)
von: Ye, Yufei, et al.
Veröffentlicht: (2024)
CRISP: Contact-Guided Real2Sim from Monocular Video with Planar Scene Primitives
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
MVD-Fusion: Single-view 3D via Depth-consistent Multi-view Generation
von: Hu, Hanzhe, et al.
Veröffentlicht: (2024)
von: Hu, Hanzhe, et al.
Veröffentlicht: (2024)
CtrlVDiff: Controllable Video Generation via Unified Multimodal Video Diffusion
von: Xi, Dianbing, et al.
Veröffentlicht: (2025)
von: Xi, Dianbing, et al.
Veröffentlicht: (2025)
DressRecon: Freeform 4D Human Reconstruction from Monocular Video
von: Tan, Jeff, et al.
Veröffentlicht: (2024)
von: Tan, Jeff, et al.
Veröffentlicht: (2024)
BlobCtrl: Taming Controllable Blob for Element-level Image Editing
von: Li, Yaowei, et al.
Veröffentlicht: (2025)
von: Li, Yaowei, et al.
Veröffentlicht: (2025)
MotionCtrl: A Unified and Flexible Motion Controller for Video Generation
von: Wang, Zhouxia, et al.
Veröffentlicht: (2023)
von: Wang, Zhouxia, et al.
Veröffentlicht: (2023)
LightCtrl: Training-free Controllable Video Relighting
von: Peng, Yizuo, et al.
Veröffentlicht: (2026)
von: Peng, Yizuo, et al.
Veröffentlicht: (2026)
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation
von: Bharadhwaj, Homanga, et al.
Veröffentlicht: (2024)
von: Bharadhwaj, Homanga, et al.
Veröffentlicht: (2024)
Ctrl-VI: Controllable Video Synthesis via Variational Inference
von: Duan, Haoyi, et al.
Veröffentlicht: (2025)
von: Duan, Haoyi, et al.
Veröffentlicht: (2025)
Flow3r: Factored Flow Prediction for Scalable Visual Geometry Learning
von: Cong, Zhongxiao, et al.
Veröffentlicht: (2026)
von: Cong, Zhongxiao, et al.
Veröffentlicht: (2026)
Diverse Score Distillation
von: Xu, Yanbo, et al.
Veröffentlicht: (2024)
von: Xu, Yanbo, et al.
Veröffentlicht: (2024)
Ctrl-V: Higher Fidelity Video Generation with Bounding-Box Controlled Object Motion
von: Luo, Ge Ya, et al.
Veröffentlicht: (2024)
von: Luo, Ge Ya, et al.
Veröffentlicht: (2024)
EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
DisProtEdit: Exploring Disentangled Representations for Multi-Attribute Protein Editing
von: Ku, Max, et al.
Veröffentlicht: (2025)
von: Ku, Max, et al.
Veröffentlicht: (2025)
Predicting 4D Hand Trajectory from Monocular Videos
von: Ye, Yufei, et al.
Veröffentlicht: (2025)
von: Ye, Yufei, et al.
Veröffentlicht: (2025)
ControlEdit: A MultiModal Local Clothing Image Editing Method
von: Cheng, Di, et al.
Veröffentlicht: (2024)
von: Cheng, Di, et al.
Veröffentlicht: (2024)
Edit As You Wish: Video Caption Editing with Multi-grained User Control
von: Yao, Linli, et al.
Veröffentlicht: (2023)
von: Yao, Linli, et al.
Veröffentlicht: (2023)
PPS-Ctrl: Controllable Sim-to-Real Translation for Colonoscopy Depth Estimation
von: Xiong, Xinqi, et al.
Veröffentlicht: (2025)
von: Xiong, Xinqi, et al.
Veröffentlicht: (2025)
DexCtrl: Towards Sim-to-Real Dexterity with Adaptive Controller Learning
von: Zhao, Shuqi, et al.
Veröffentlicht: (2025)
von: Zhao, Shuqi, et al.
Veröffentlicht: (2025)
EmoCtrl: Controllable Emotional Image Content Generation
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
List Decoding Expander-Based Codes up to Capacity in Near-Linear Time
von: Srivastava, Shashank, et al.
Veröffentlicht: (2025)
von: Srivastava, Shashank, et al.
Veröffentlicht: (2025)
ExpressEdit: Video Editing with Natural Language and Sketching
von: Tilekbay, Bekzat, et al.
Veröffentlicht: (2024)
von: Tilekbay, Bekzat, et al.
Veröffentlicht: (2024)
UniPhy: Learning a Unified Constitutive Model for Inverse Physics Simulation
von: Mittal, Himangi, et al.
Veröffentlicht: (2025)
von: Mittal, Himangi, et al.
Veröffentlicht: (2025)
Edit-Your-Motion: Space-Time Diffusion Decoupling Learning for Video Motion Editing
von: Zuo, Yi, et al.
Veröffentlicht: (2024)
von: Zuo, Yi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
LightSwitch: Multi-view Relighting with Material-guided Diffusion
von: Litman, Yehonathan, et al.
Veröffentlicht: (2025) -
MaterialFusion: Enhancing Inverse Rendering with Material Diffusion Priors
von: Litman, Yehonathan, et al.
Veröffentlicht: (2024) -
Towards Unstructured Unlabeled Optical Mocap: A Video Helps!
von: Milef, Nicholas, et al.
Veröffentlicht: (2024) -
Real‐Time Neural Materials on Mobile VR
von: Zilin Xu, et al.
Veröffentlicht: (2026) -
Sparse-view Pose Estimation and Reconstruction via Analysis by Generative Synthesis
von: Zhao, Qitao, et al.
Veröffentlicht: (2024)