JOG3R: Towards 3D-Consistent Video Generators
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Chun-Hao Paul, Mitra, Niloy, Jeong, Hyeonho, Yoon, Jae Shin, Ceylan, Duygu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
di: Jeong, Hyeonho, et al.
Pubblicazione: (2024)
di: Jeong, Hyeonho, et al.
Pubblicazione: (2024)
Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing
di: Lee, Dohun, et al.
Pubblicazione: (2026)
di: Lee, Dohun, et al.
Pubblicazione: (2026)
SuperGaussian: Repurposing Video Models for 3D Super Resolution
di: Shen, Yuan, et al.
Pubblicazione: (2024)
di: Shen, Yuan, et al.
Pubblicazione: (2024)
Improving Video Diffusion Transformer Training by Multi-Feature Fusion and Alignment from Self-Supervised Vision Encoders
di: Lee, Dohun, et al.
Pubblicazione: (2025)
di: Lee, Dohun, et al.
Pubblicazione: (2025)
BLiSS: Bootstrapped Linear Shape Space
di: Muralikrishnan, Sanjeev, et al.
Pubblicazione: (2023)
di: Muralikrishnan, Sanjeev, et al.
Pubblicazione: (2023)
VideoHandles: Editing 3D Object Compositions in Videos Using Video Generative Priors
di: Koo, Juil, et al.
Pubblicazione: (2025)
di: Koo, Juil, et al.
Pubblicazione: (2025)
EM-Vid: Training-Free Entity-Centric Memory for Efficient and Consistent Multi-Shot Video Generation
di: Vandersanden, Jente, et al.
Pubblicazione: (2026)
di: Vandersanden, Jente, et al.
Pubblicazione: (2026)
GANFusion: Feed-Forward Text-to-3D with Diffusion in GAN Space
di: Attaiki, Souhaib, et al.
Pubblicazione: (2024)
di: Attaiki, Souhaib, et al.
Pubblicazione: (2024)
WorldReel: 4D Video Generation with Consistent Geometry and Motion Modeling
di: Fang, Shaoheng, et al.
Pubblicazione: (2025)
di: Fang, Shaoheng, et al.
Pubblicazione: (2025)
Reangle-A-Video: 4D Video Generation as Video-to-Video Translation
di: Jeong, Hyeonho, et al.
Pubblicazione: (2025)
di: Jeong, Hyeonho, et al.
Pubblicazione: (2025)
LoST: Level of Semantics Tokenization for 3D Shapes
di: Dutt, Niladri Shekhar, et al.
Pubblicazione: (2026)
di: Dutt, Niladri Shekhar, et al.
Pubblicazione: (2026)
Boosting Camera Motion Control for Video Diffusion Transformers
di: Cheong, Soon Yau, et al.
Pubblicazione: (2024)
di: Cheong, Soon Yau, et al.
Pubblicazione: (2024)
LAMP: Language-Assisted Motion Planning for Controllable Video Generation
di: Kizil, Muhammed Burak, et al.
Pubblicazione: (2025)
di: Kizil, Muhammed Burak, et al.
Pubblicazione: (2025)
MonetGPT: Solving Puzzles Enhances MLLMs' Image Retouching Skills
di: Dutt, Niladri Shekhar, et al.
Pubblicazione: (2025)
di: Dutt, Niladri Shekhar, et al.
Pubblicazione: (2025)
Auteur: Language-Driven Cinematographic Framing for Human-Centric Video Generation
di: Kizil, Muhammed Burak, et al.
Pubblicazione: (2026)
di: Kizil, Muhammed Burak, et al.
Pubblicazione: (2026)
Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models
di: Jeong, Hyeonho, et al.
Pubblicazione: (2023)
di: Jeong, Hyeonho, et al.
Pubblicazione: (2023)
TrajectoryMover: Generative Movement of Object Trajectories in Videos
di: Chhatre, Kiran, et al.
Pubblicazione: (2026)
di: Chhatre, Kiran, et al.
Pubblicazione: (2026)
GeoFusionLRM: Geometry-Aware Self-Correction for Consistent 3D Reconstruction
di: Yildirim, Ahmet Burak, et al.
Pubblicazione: (2026)
di: Yildirim, Ahmet Burak, et al.
Pubblicazione: (2026)
Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models
di: Jeon, Wooseok, et al.
Pubblicazione: (2026)
di: Jeon, Wooseok, et al.
Pubblicazione: (2026)
Animal Avatars: Reconstructing Animatable 3D Animals from Casual Videos
di: Sabathier, Remy, et al.
Pubblicazione: (2024)
di: Sabathier, Remy, et al.
Pubblicazione: (2024)
3D Stylization via Large Reconstruction Model
di: Oztas, Ipek, et al.
Pubblicazione: (2025)
di: Oztas, Ipek, et al.
Pubblicazione: (2025)
SpaceTimePilot: Generative Rendering of Dynamic Scenes Across Space and Time
di: Huang, Zhening, et al.
Pubblicazione: (2025)
di: Huang, Zhening, et al.
Pubblicazione: (2025)
ActionMesh: Animated 3D Mesh Generation with Temporal 3D Diffusion
di: Sabathier, Remy, et al.
Pubblicazione: (2026)
di: Sabathier, Remy, et al.
Pubblicazione: (2026)
Improving Black-Box Generative Attacks via Generator Semantic Consistency
di: Jeong, Jongoh, et al.
Pubblicazione: (2025)
di: Jeong, Jongoh, et al.
Pubblicazione: (2025)
DreamMotion: Space-Time Self-Similar Score Distillation for Zero-Shot Video Editing
di: Jeong, Hyeonho, et al.
Pubblicazione: (2024)
di: Jeong, Hyeonho, et al.
Pubblicazione: (2024)
SAGE: Structure-Aware Generative Video Transitions between Diverse Clips
di: Kan, Mia, et al.
Pubblicazione: (2025)
di: Kan, Mia, et al.
Pubblicazione: (2025)
From Programs to Poses: Factored Real-World Scene Generation via Learned Program Libraries
di: Hsu, Joy, et al.
Pubblicazione: (2025)
di: Hsu, Joy, et al.
Pubblicazione: (2025)
Enhancing Temporal Consistency in Video Editing by Reconstructing Videos with 3D Gaussian Splatting
di: Shin, Inkyu, et al.
Pubblicazione: (2024)
di: Shin, Inkyu, et al.
Pubblicazione: (2024)
LIM: Large Interpolator Model for Dynamic Reconstruction
di: Sabathier, Remy, et al.
Pubblicazione: (2025)
di: Sabathier, Remy, et al.
Pubblicazione: (2025)
MatAtlas: Text-driven Consistent Geometry Texturing and Material Assignment
di: Ceylan, Duygu, et al.
Pubblicazione: (2024)
di: Ceylan, Duygu, et al.
Pubblicazione: (2024)
Spectral Motion Alignment for Video Motion Transfer using Diffusion Models
di: Park, Geon Yeong, et al.
Pubblicazione: (2024)
di: Park, Geon Yeong, et al.
Pubblicazione: (2024)
Extend3D: Town-Scale 3D Generation
di: Yoon, Seungwoo, et al.
Pubblicazione: (2026)
di: Yoon, Seungwoo, et al.
Pubblicazione: (2026)
VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation
di: Du, Hongyang, et al.
Pubblicazione: (2026)
di: Du, Hongyang, et al.
Pubblicazione: (2026)
MD-ProjTex: Texturing 3D Shapes with Multi-Diffusion Projection
di: Yildirim, Ahmet Burak, et al.
Pubblicazione: (2025)
di: Yildirim, Ahmet Burak, et al.
Pubblicazione: (2025)
Diffusion 3D Features (Diff3F): Decorating Untextured Shapes with Distilled Semantic Features
di: Dutt, Niladri Shekhar, et al.
Pubblicazione: (2023)
di: Dutt, Niladri Shekhar, et al.
Pubblicazione: (2023)
Leveraging VLM-Based Pipelines to Annotate 3D Objects
di: Kabra, Rishabh, et al.
Pubblicazione: (2023)
di: Kabra, Rishabh, et al.
Pubblicazione: (2023)
3D-Adapter: Geometry-Consistent Multi-View Diffusion for High-Quality 3D Generation
di: Chen, Hansheng, et al.
Pubblicazione: (2024)
di: Chen, Hansheng, et al.
Pubblicazione: (2024)
DISN: Deep Implicit Surface Network for High-quality Single-view 3D Reconstruction
di: Xu, Qiangeng, et al.
Pubblicazione: (2019)
di: Xu, Qiangeng, et al.
Pubblicazione: (2019)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
di: Cha, Junuk, et al.
Pubblicazione: (2024)
di: Cha, Junuk, et al.
Pubblicazione: (2024)
3DPhysVideo: Consistency-Guided Flow SDE for Video Generation via 3D Scene Reconstruction and Physical Simulation
di: Kim, Hwidong, et al.
Pubblicazione: (2026)
di: Kim, Hwidong, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
di: Jeong, Hyeonho, et al.
Pubblicazione: (2024) -
Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing
di: Lee, Dohun, et al.
Pubblicazione: (2026) -
SuperGaussian: Repurposing Video Models for 3D Super Resolution
di: Shen, Yuan, et al.
Pubblicazione: (2024) -
Improving Video Diffusion Transformer Training by Multi-Feature Fusion and Alignment from Self-Supervised Vision Encoders
di: Lee, Dohun, et al.
Pubblicazione: (2025) -
BLiSS: Bootstrapped Linear Shape Space
di: Muralikrishnan, Sanjeev, et al.
Pubblicazione: (2023)