Drag4D: Align Your Motion with Text-Driven 3D Scene Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Kang, Minjun, Shin, Inkyu, Lee, Taeyeop, Kweon, In So, Yoon, Kuk-Jin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis
by: Kang, Minjun, et al.
Published: (2026)
by: Kang, Minjun, et al.
Published: (2026)
Any6D: Model-free 6D Pose Estimation of Novel Objects
by: Lee, Taeyeop, et al.
Published: (2025)
by: Lee, Taeyeop, et al.
Published: (2025)
Stable Surface Regularization for Fast Few-Shot NeRF
by: Joung, Byeongin, et al.
Published: (2024)
by: Joung, Byeongin, et al.
Published: (2024)
Enhancing Temporal Consistency in Video Editing by Reconstructing Videos with 3D Gaussian Splatting
by: Shin, Inkyu, et al.
Published: (2024)
by: Shin, Inkyu, et al.
Published: (2024)
Event6D: Event-based Novel Object 6D Pose Tracking
by: Kang, Jae-Young, et al.
Published: (2026)
by: Kang, Jae-Young, et al.
Published: (2026)
DeLTa: Demonstration and Language-Guided Novel Transparent Object Manipulation
by: Lee, Taeyeop, et al.
Published: (2025)
by: Lee, Taeyeop, et al.
Published: (2025)
TALoS: Enhancing Semantic Scene Completion via Test-time Adaptation on the Line of Sight
by: Jang, Hyun-Kurl, et al.
Published: (2024)
by: Jang, Hyun-Kurl, et al.
Published: (2024)
MTMMC: A Large-Scale Real-World Multi-Modal Camera Tracking Benchmark
by: Woo, Sanghyun, et al.
Published: (2024)
by: Woo, Sanghyun, et al.
Published: (2024)
ControlDreamer: Blending Geometry and Style in Text-to-3D
by: Oh, Yeongtak, et al.
Published: (2023)
by: Oh, Yeongtak, et al.
Published: (2023)
Unleashing the Temporal Potential of Stereo Event Cameras for Continuous-Time 3D Object Detection
by: Kang, Jae-Young, et al.
Published: (2025)
by: Kang, Jae-Young, et al.
Published: (2025)
Finding Meaning in Points: Weakly Supervised Semantic Segmentation for Event Cameras
by: Cho, Hoonhee, et al.
Published: (2024)
by: Cho, Hoonhee, et al.
Published: (2024)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
by: Cha, Junuk, et al.
Published: (2024)
by: Cha, Junuk, et al.
Published: (2024)
Bootstrapping Video Semantic Segmentation Model via Distillation-assisted Test-Time Adaptation
by: Kim, Jihun, et al.
Published: (2026)
by: Kim, Jihun, et al.
Published: (2026)
From Sharp to Blur: Unsupervised Domain Adaptation for 2D Human Pose Estimation Under Extreme Motion Blur Using Event Cameras
by: Kim, Youngho, et al.
Published: (2025)
by: Kim, Youngho, et al.
Published: (2025)
Ev-3DOD: Pushing the Temporal Boundaries of 3D Object Detection with Event Cameras
by: Cho, Hoonhee, et al.
Published: (2025)
by: Cho, Hoonhee, et al.
Published: (2025)
DragScene: Interactive 3D Scene Editing with Single-view Drag Instructions
by: Gu, Chenghao, et al.
Published: (2024)
by: Gu, Chenghao, et al.
Published: (2024)
Motion4D: Learning 3D-Consistent Motion and Semantics for 4D Scene Understanding
by: Zhou, Haoran, et al.
Published: (2025)
by: Zhou, Haoran, et al.
Published: (2025)
MoDA: Leveraging Motion Priors from Videos for Advancing Unsupervised Domain Adaptation in Semantic Segmentation
by: Pan, Fei, et al.
Published: (2023)
by: Pan, Fei, et al.
Published: (2023)
FastScene: Text-Driven Fast 3D Indoor Scene Generation via Panoramic Gaussian Splatting
by: Ma, Yikun, et al.
Published: (2024)
by: Ma, Yikun, et al.
Published: (2024)
Generating Human Motion in 3D Scenes from Text Descriptions
by: Cen, Zhi, et al.
Published: (2024)
by: Cen, Zhi, et al.
Published: (2024)
DC-TTA: Divide-and-Conquer Framework for Test-Time Adaptation of Interactive Segmentation
by: Kim, Jihun, et al.
Published: (2025)
by: Kim, Jihun, et al.
Published: (2025)
3D-SceneDreamer: Text-Driven 3D-Consistent Scene Generation
by: Zhang, Frank, et al.
Published: (2024)
by: Zhang, Frank, et al.
Published: (2024)
Preference-Aligned LoRA Merging: Preserving Subspace Coverage and Addressing Directional Anisotropy
by: Jeong, Wooseong, et al.
Published: (2026)
by: Jeong, Wooseong, et al.
Published: (2026)
GMT: Enhancing Generalizable Neural Rendering via Geometry-Driven Multi-Reference Texture Transfer
by: Yoon, Youngho, et al.
Published: (2024)
by: Yoon, Youngho, et al.
Published: (2024)
Drag Your Gaussian: Effective Drag-Based Editing with Score Distillation for 3D Gaussian Splatting
by: Qu, Yansong, et al.
Published: (2025)
by: Qu, Yansong, et al.
Published: (2025)
FlowDrag: 3D-aware Drag-based Image Editing with Mesh-guided Deformation Vector Flow Fields
by: Koo, Gwanhyeong, et al.
Published: (2025)
by: Koo, Gwanhyeong, et al.
Published: (2025)
AnoStyler: Text-Driven Localized Anomaly Generation via Lightweight Style Transfer
by: So, Yulim, et al.
Published: (2025)
by: So, Yulim, et al.
Published: (2025)
Multi-agent Long-term 3D Human Pose Forecasting via Interaction-aware Trajectory Conditioning
by: Jeong, Jaewoo, et al.
Published: (2024)
by: Jeong, Jaewoo, et al.
Published: (2024)
PASTA: Part-Aware Sketch-to-3D Shape Generation with Text-Aligned Prior
by: Lee, Seunggwan, et al.
Published: (2025)
by: Lee, Seunggwan, et al.
Published: (2025)
VR-Drive: Viewpoint-Robust End-to-End Driving with Feed-Forward 3D Gaussian Splatting
by: Cho, Hoonhee, et al.
Published: (2025)
by: Cho, Hoonhee, et al.
Published: (2025)
Align Your Gaussians: Text-to-4D with Dynamic 3D Gaussians and Composed Diffusion Models
by: Ling, Huan, et al.
Published: (2023)
by: Ling, Huan, et al.
Published: (2023)
GPT-Connect: Interaction between Text-Driven Human Motion Generator and 3D Scenes in a Training-free Manner
by: Qu, Haoxuan, et al.
Published: (2024)
by: Qu, Haoxuan, et al.
Published: (2024)
AVOID: The Adverse Visual Conditions Dataset with Obstacles for Driving Scene Understanding
by: Jeong, Jongoh, et al.
Published: (2025)
by: Jeong, Jongoh, et al.
Published: (2025)
C-Drag: Chain-of-Thought Driven Motion Controller for Video Generation
by: Li, Yuhao, et al.
Published: (2025)
by: Li, Yuhao, et al.
Published: (2025)
DSERT-RoLL: Robust Multi-Modal Perception for Diverse Driving Conditions with Stereo Event-RGB-Thermal Cameras, 4D Radar, and Dual-LiDAR
by: Cho, Hoonhee, et al.
Published: (2026)
by: Cho, Hoonhee, et al.
Published: (2026)
DreamScene: 3D Gaussian-based Text-to-3D Scene Generation via Formation Pattern Sampling
by: Li, Haoran, et al.
Published: (2024)
by: Li, Haoran, et al.
Published: (2024)
Sitcom-Crafter: A Plot-Driven Human Motion Generation System in 3D Scenes
by: Chen, Jianqi, et al.
Published: (2024)
by: Chen, Jianqi, et al.
Published: (2024)
MvDrag3D: Drag-based Creative 3D Editing via Multi-view Generation-Reconstruction Priors
by: Chen, Honghua, et al.
Published: (2024)
by: Chen, Honghua, et al.
Published: (2024)
Synchronizing Task Behavior: Aligning Multiple Tasks during Test-Time Training
by: Jeong, Wooseong, et al.
Published: (2025)
by: Jeong, Wooseong, et al.
Published: (2025)
Temporal Event Stereo via Joint Learning with Stereoscopic Flow
by: Cho, Hoonhee, et al.
Published: (2024)
by: Cho, Hoonhee, et al.
Published: (2024)
Similar Items
-
GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis
by: Kang, Minjun, et al.
Published: (2026) -
Any6D: Model-free 6D Pose Estimation of Novel Objects
by: Lee, Taeyeop, et al.
Published: (2025) -
Stable Surface Regularization for Fast Few-Shot NeRF
by: Joung, Byeongin, et al.
Published: (2024) -
Enhancing Temporal Consistency in Video Editing by Reconstructing Videos with 3D Gaussian Splatting
by: Shin, Inkyu, et al.
Published: (2024) -
Event6D: Event-based Novel Object 6D Pose Tracking
by: Kang, Jae-Young, et al.
Published: (2026)