ORIGEN: Zero-Shot 3D Orientation Grounding in Text-to-Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Min, Yunhong, Choi, Daehyeon, Yeo, Kyeongmin, Lee, Jihyun, Sung, Minhyuk |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Psi-Sampler: Initial Particle Sampling for SMC-Based Inference-Time Reward Alignment in Score Models
di: Yoon, Taehoon, et al.
Pubblicazione: (2025)
di: Yoon, Taehoon, et al.
Pubblicazione: (2025)
MatLat: Material Latent Space for PBR Texture Generation
di: Yeo, Kyeongmin, et al.
Pubblicazione: (2025)
di: Yeo, Kyeongmin, et al.
Pubblicazione: (2025)
StochSync: Stochastic Diffusion Synchronization for Image Generation in Arbitrary Spaces
di: Yeo, Kyeongmin, et al.
Pubblicazione: (2025)
di: Yeo, Kyeongmin, et al.
Pubblicazione: (2025)
SyncTweedies: A General Generative Framework Based on Synchronized Diffusions
di: Kim, Jaihoon, et al.
Pubblicazione: (2024)
di: Kim, Jaihoon, et al.
Pubblicazione: (2024)
Neural Pose Representation Learning for Generating and Transferring Non-Rigid Object Poses
di: Yoo, Seungwoo, et al.
Pubblicazione: (2024)
di: Yoo, Seungwoo, et al.
Pubblicazione: (2024)
Toward Ambulatory Vision: Learning Visually-Grounded Active View Selection
di: Koo, Juil, et al.
Pubblicazione: (2025)
di: Koo, Juil, et al.
Pubblicazione: (2025)
Moment- and Power-Spectrum-Based Gaussianity Regularization for Text-to-Image Models
di: Hwang, Jisung, et al.
Pubblicazione: (2025)
di: Hwang, Jisung, et al.
Pubblicazione: (2025)
Grounding Descriptions in Images informs Zero-Shot Visual Recognition
di: Halbe, Shaunak, et al.
Pubblicazione: (2024)
di: Halbe, Shaunak, et al.
Pubblicazione: (2024)
Occupancy-Based Dual Contouring
di: Hwang, Jisung, et al.
Pubblicazione: (2024)
di: Hwang, Jisung, et al.
Pubblicazione: (2024)
Inference-Time Scaling for Flow Models via Stochastic Generation and Rollover Budget Forcing
di: Kim, Jaihoon, et al.
Pubblicazione: (2025)
di: Kim, Jaihoon, et al.
Pubblicazione: (2025)
ReGround: Improving Textual and Spatial Grounding at No Cost
di: Lee, Phillip Y., et al.
Pubblicazione: (2024)
di: Lee, Phillip Y., et al.
Pubblicazione: (2024)
Infusing Environmental Captions for Long-Form Video Language Grounding
di: Lee, Hyogun, et al.
Pubblicazione: (2024)
di: Lee, Hyogun, et al.
Pubblicazione: (2024)
MCL-AD: Multimodal Collaboration Learning for Zero-Shot 3D Anomaly Detection
di: Li, Gang, et al.
Pubblicazione: (2025)
di: Li, Gang, et al.
Pubblicazione: (2025)
Lipsum-FT: Robust Fine-Tuning of Zero-Shot Models Using Random Text Guidance
di: Nam, Giung, et al.
Pubblicazione: (2024)
di: Nam, Giung, et al.
Pubblicazione: (2024)
Zero-Shot Medical Phrase Grounding with Off-the-shelf Diffusion Models
di: Vilouras, Konstantinos, et al.
Pubblicazione: (2024)
di: Vilouras, Konstantinos, et al.
Pubblicazione: (2024)
Unleashing Text-to-Image Diffusion Prior for Zero-Shot Image Captioning
di: Luo, Jianjie, et al.
Pubblicazione: (2024)
di: Luo, Jianjie, et al.
Pubblicazione: (2024)
Improving Factuality of 3D Brain MRI Report Generation with Paired Image-domain Retrieval and Text-domain Augmentation
di: Lee, Junhyeok, et al.
Pubblicazione: (2024)
di: Lee, Junhyeok, et al.
Pubblicazione: (2024)
Image-Caption Encoding for Improving Zero-Shot Generalization
di: Yu, Eric Yang, et al.
Pubblicazione: (2024)
di: Yu, Eric Yang, et al.
Pubblicazione: (2024)
DiffuSAM: Diffusion Guided Zero-Shot Object Grounding for Remote Sensing Imagery
di: Sethi, Geet, et al.
Pubblicazione: (2026)
di: Sethi, Geet, et al.
Pubblicazione: (2026)
A Survey on Generative Modeling with Limited Data, Few Shots, and Zero Shot
di: Abdollahzadeh, Milad, et al.
Pubblicazione: (2023)
di: Abdollahzadeh, Milad, et al.
Pubblicazione: (2023)
T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis
di: Imam, Raza, et al.
Pubblicazione: (2025)
di: Imam, Raza, et al.
Pubblicazione: (2025)
CLIPoint3D: Language-Grounded Few-Shot Unsupervised 3D Point Cloud Domain Adaptation
di: Singha, Mainak, et al.
Pubblicazione: (2026)
di: Singha, Mainak, et al.
Pubblicazione: (2026)
Budgeted Online Continual Learning by Adaptive Layer Freezing and Frequency-based Sampling
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
InterHandGen: Two-Hand Interaction Generation via Cascaded Reverse Diffusion
di: Lee, Jihyun, et al.
Pubblicazione: (2024)
di: Lee, Jihyun, et al.
Pubblicazione: (2024)
GenOL: Generating Diverse Examples for Name-only Online Learning
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
D$^3$Fields: Dynamic 3D Descriptor Fields for Zero-Shot Generalizable Rearrangement
di: Wang, Yixuan, et al.
Pubblicazione: (2023)
di: Wang, Yixuan, et al.
Pubblicazione: (2023)
TSP3D: Text-guided Sparse Voxel Pruning for Efficient 3D Visual Grounding
di: Guo, Wenxuan, et al.
Pubblicazione: (2025)
di: Guo, Wenxuan, et al.
Pubblicazione: (2025)
Robust 3D Shape Reconstruction in Zero-Shot from a Single Image in the Wild
di: Cho, Junhyeong, et al.
Pubblicazione: (2024)
di: Cho, Junhyeong, et al.
Pubblicazione: (2024)
Symmetry-Robust 3D Orientation Estimation
di: Scarvelis, Christopher, et al.
Pubblicazione: (2024)
di: Scarvelis, Christopher, et al.
Pubblicazione: (2024)
Training-Free Zero-Shot Anomaly Detection in 3D Brain MRI with 2D Foundation Models
di: Le-Gia, Tai, et al.
Pubblicazione: (2026)
di: Le-Gia, Tai, et al.
Pubblicazione: (2026)
Towards Lifelong Few-Shot Customization of Text-to-Image Diffusion
di: Song, Nan, et al.
Pubblicazione: (2024)
di: Song, Nan, et al.
Pubblicazione: (2024)
Retrieval-Augmented Score Distillation for Text-to-3D Generation
di: Seo, Junyoung, et al.
Pubblicazione: (2024)
di: Seo, Junyoung, et al.
Pubblicazione: (2024)
Investigating the Effectiveness of Cross-Attention to Unlock Zero-Shot Editing of Text-to-Video Diffusion Models
di: Motamed, Saman, et al.
Pubblicazione: (2024)
di: Motamed, Saman, et al.
Pubblicazione: (2024)
ActCam: Zero-Shot Joint Camera and 3D Motion Control for Video Generation
di: Khalifi, Omar El, et al.
Pubblicazione: (2026)
di: Khalifi, Omar El, et al.
Pubblicazione: (2026)
PartSTAD: 2D-to-3D Part Segmentation Task Adaptation
di: Kim, Hyunjin, et al.
Pubblicazione: (2024)
di: Kim, Hyunjin, et al.
Pubblicazione: (2024)
Coordinate-Based Neural Representation Enabling Zero-Shot Learning for 3D Multiparametric Quantitative MRI
di: Lao, Guoyan, et al.
Pubblicazione: (2024)
di: Lao, Guoyan, et al.
Pubblicazione: (2024)
Evaluating Zero-Shot GPT-4V Performance on 3D Visual Question Answering Benchmarks
di: Singh, Simranjit, et al.
Pubblicazione: (2024)
di: Singh, Simranjit, et al.
Pubblicazione: (2024)
Text2Model: Text-based Model Induction for Zero-shot Image Classification
di: Amosy, Ohad, et al.
Pubblicazione: (2022)
di: Amosy, Ohad, et al.
Pubblicazione: (2022)
Diffusion Self-Distillation for Zero-Shot Customized Image Generation
di: Cai, Shengqu, et al.
Pubblicazione: (2024)
di: Cai, Shengqu, et al.
Pubblicazione: (2024)
Zero-Shot Video Restoration and Enhancement Using Pre-Trained Image Diffusion Model
di: Cao, Cong, et al.
Pubblicazione: (2024)
di: Cao, Cong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Psi-Sampler: Initial Particle Sampling for SMC-Based Inference-Time Reward Alignment in Score Models
di: Yoon, Taehoon, et al.
Pubblicazione: (2025) -
MatLat: Material Latent Space for PBR Texture Generation
di: Yeo, Kyeongmin, et al.
Pubblicazione: (2025) -
StochSync: Stochastic Diffusion Synchronization for Image Generation in Arbitrary Spaces
di: Yeo, Kyeongmin, et al.
Pubblicazione: (2025) -
SyncTweedies: A General Generative Framework Based on Synchronized Diffusions
di: Kim, Jaihoon, et al.
Pubblicazione: (2024) -
Neural Pose Representation Learning for Generating and Transferring Non-Rigid Object Poses
di: Yoo, Seungwoo, et al.
Pubblicazione: (2024)