3DreamBooth: High-Fidelity 3D Subject-Driven Video Generation Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ko, Hyun-kyu, Park, Jihyeon, Kim, Younghyun, Park, Dongheok, Park, Eunbyung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sequence Matters: Harnessing Video Models in 3D Super-Resolution
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2024)
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2024)
Image-based Facial Rig Inversion
von: Yang, Tianxiang, et al.
Veröffentlicht: (2025)
von: Yang, Tianxiang, et al.
Veröffentlicht: (2025)
Objaverse++: Curated 3D Object Dataset with Quality Annotations
von: Lin, Chendi, et al.
Veröffentlicht: (2025)
von: Lin, Chendi, et al.
Veröffentlicht: (2025)
TSPE-GS: Probabilistic Depth Extraction for Semi-Transparent Surface Reconstruction via 3D Gaussian Splatting
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025)
VDPP: Video Depth Post-Processing for Speed and Scalability
von: Yoon, Daewon, et al.
Veröffentlicht: (2026)
von: Yoon, Daewon, et al.
Veröffentlicht: (2026)
Gaussian Splatting: 3D Reconstruction and Novel View Synthesis, a Review
von: Dalal, Anurag, et al.
Veröffentlicht: (2024)
von: Dalal, Anurag, et al.
Veröffentlicht: (2024)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
ROI-GS: Interest-based Local Quality 3D Gaussian Splatting
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
Beyond RNNs: Benchmarking Attention-Based Image Captioning Models
von: Yanambakkam, Hemanth Teja, et al.
Veröffentlicht: (2025)
von: Yanambakkam, Hemanth Teja, et al.
Veröffentlicht: (2025)
VisChainBench: A Benchmark for Multi-Turn, Multi-Image Visual Reasoning Beyond Language Priors
von: Lyu, Wenbo, et al.
Veröffentlicht: (2025)
von: Lyu, Wenbo, et al.
Veröffentlicht: (2025)
Smelly, dense, and spreaded: The Object Detection for Olfactory References (ODOR) dataset
von: Zinnen, Mathias, et al.
Veröffentlicht: (2025)
von: Zinnen, Mathias, et al.
Veröffentlicht: (2025)
Experimental Evaluation of Road-Crossing Decisions by Autonomous Wheelchairs against Environmental Factors
von: Corradini, Franca, et al.
Veröffentlicht: (2024)
von: Corradini, Franca, et al.
Veröffentlicht: (2024)
MSGS: Multispectral 3D Gaussian Splatting
von: Zheng, Iris, et al.
Veröffentlicht: (2026)
von: Zheng, Iris, et al.
Veröffentlicht: (2026)
SSD-GS: Scattering and Shadow Decomposition for Relightable 3D Gaussian Splatting
von: Zheng, Iris, et al.
Veröffentlicht: (2026)
von: Zheng, Iris, et al.
Veröffentlicht: (2026)
VLM-NCD:Novel Class Discovery with Vision-Based Large Language Models
von: Su, Yuetong, et al.
Veröffentlicht: (2025)
von: Su, Yuetong, et al.
Veröffentlicht: (2025)
From Gaze to Insight: Bridging Human Visual Attention and Vision Language Model Explanation for Weakly-Supervised Medical Image Segmentation
von: Chen, Jingkun, et al.
Veröffentlicht: (2025)
von: Chen, Jingkun, et al.
Veröffentlicht: (2025)
ROI-NeRFs: Hi-Fi Visualization of Objects of Interest within a Scene by NeRFs Composition
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
Semantic2Graph: Graph-based Multi-modal Feature Fusion for Action Segmentation in Videos
von: Zhang, Junbin, et al.
Veröffentlicht: (2022)
von: Zhang, Junbin, et al.
Veröffentlicht: (2022)
RCI: A Score for Evaluating Global and Local Reasoning in Multimodal Benchmarks
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
VideoMind: An Omni-Modal Video Dataset with Intent Grounding for Deep-Cognitive Video Understanding
von: Yang, Baoyao, et al.
Veröffentlicht: (2025)
von: Yang, Baoyao, et al.
Veröffentlicht: (2025)
FastGS: Training 3D Gaussian Splatting in 100 Seconds
von: Ren, Shiwei, et al.
Veröffentlicht: (2025)
von: Ren, Shiwei, et al.
Veröffentlicht: (2025)
SEGS-SLAM: Structure-enhanced 3D Gaussian Splatting SLAM with Appearance Embedding
von: Wen, Tianci, et al.
Veröffentlicht: (2025)
von: Wen, Tianci, et al.
Veröffentlicht: (2025)
OpenFusion++: An Open-vocabulary Real-time Scene Understanding System
von: Jin, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Jin, Xiaofeng, et al.
Veröffentlicht: (2025)
DSER: Spectral Epipolar Representation for Efficient Light Field Depth Estimation
von: Mohammad, Noor Islam S., et al.
Veröffentlicht: (2025)
von: Mohammad, Noor Islam S., et al.
Veröffentlicht: (2025)
Hierarchical Spatial Algorithms for High-Resolution Image Quantization and Feature Extraction
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025)
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025)
Predictive Modeling of Maritime Radar Data Using Transformer Architecture
von: Qesaraku, Bjorna, et al.
Veröffentlicht: (2025)
von: Qesaraku, Bjorna, et al.
Veröffentlicht: (2025)
TACIT: Transformation-Aware Capturing of Implicit Thought
von: Nobrega, Daniel
Veröffentlicht: (2026)
von: Nobrega, Daniel
Veröffentlicht: (2026)
Rethinking Visual Intelligence: Insights from Video Pretraining
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
Consistent Zero-shot 3D Texture Synthesis Using Geometry-aware Diffusion and Temporal Video Models
von: Kang, Donggoo, et al.
Veröffentlicht: (2025)
von: Kang, Donggoo, et al.
Veröffentlicht: (2025)
Dual-sensing driving detection model
von: K, Leon C. C., et al.
Veröffentlicht: (2025)
von: K, Leon C. C., et al.
Veröffentlicht: (2025)
Cooperative Perception: A Resource-Efficient Framework for Multi-Drone 3D Scene Reconstruction Using Federated Diffusion and NeRF
von: Pourmandi, Massoud
Veröffentlicht: (2025)
von: Pourmandi, Massoud
Veröffentlicht: (2025)
HEDGE: Hallucination Estimation via Dense Geometric Entropy for VQA with Vision-Language Models
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
Point, Detect, Count: Multi-Task Medical Image Understanding with Instruction-Tuned Vision-Language Models
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
SpectralGaussians: Semantic, spectral 3D Gaussian splatting for multi-spectral scene representation, visualization and analysis
von: Sinha, Saptarshi Neil, et al.
Veröffentlicht: (2024)
von: Sinha, Saptarshi Neil, et al.
Veröffentlicht: (2024)
Do Generative Metrics Predict YOLO Performance? An Evaluation Across Models, Augmentation Ratios, and Dataset Complexity
von: Marian, Vasile, et al.
Veröffentlicht: (2026)
von: Marian, Vasile, et al.
Veröffentlicht: (2026)
Short-Window Sliding Learning for Real-Time Violence Detection via LLM-based Auto-Labeling
von: Jung, Seoik, et al.
Veröffentlicht: (2025)
von: Jung, Seoik, et al.
Veröffentlicht: (2025)
IMKD: Intensity-Aware Multi-Level Knowledge Distillation for Camera-Radar Fusion
von: Mishra, Shashank, et al.
Veröffentlicht: (2025)
von: Mishra, Shashank, et al.
Veröffentlicht: (2025)
TACIT Benchmark: A Programmatic Visual Reasoning Benchmark for Generative and Discriminative Models
von: Medeiros, Daniel Nobrega
Veröffentlicht: (2026)
von: Medeiros, Daniel Nobrega
Veröffentlicht: (2026)
GraphiContact: Pose-aware Human-Scene Robust Contact Perception for Interactive Systems
von: Lin, Xiaojian, et al.
Veröffentlicht: (2026)
von: Lin, Xiaojian, et al.
Veröffentlicht: (2026)
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Sequence Matters: Harnessing Video Models in 3D Super-Resolution
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2024) -
Image-based Facial Rig Inversion
von: Yang, Tianxiang, et al.
Veröffentlicht: (2025) -
Objaverse++: Curated 3D Object Dataset with Quality Annotations
von: Lin, Chendi, et al.
Veröffentlicht: (2025) -
TSPE-GS: Probabilistic Depth Extraction for Semi-Transparent Surface Reconstruction via 3D Gaussian Splatting
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025) -
VDPP: Video Depth Post-Processing for Speed and Scalability
von: Yoon, Daewon, et al.
Veröffentlicht: (2026)