Composing Parts for Expressive Object Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Rangwani, Harsh, Agarwal, Aishwarya, Kulkarni, Kuldeep, Babu, R. Venkatesh, Karanam, Srikrishna |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning from Limited and Imperfect Data
por: Rangwani, Harsh
Publicado: (2025)
por: Rangwani, Harsh
Publicado: (2025)
DeiT-LT Distillation Strikes Back for Vision Transformer Training on Long-Tailed Datasets
por: Rangwani, Harsh, et al.
Publicado: (2024)
por: Rangwani, Harsh, et al.
Publicado: (2024)
Selective Mixup Fine-Tuning for Optimizing Non-Decomposable Objectives
por: Ramasubramanian, Shrinivas, et al.
Publicado: (2024)
por: Ramasubramanian, Shrinivas, et al.
Publicado: (2024)
SafaRi:Adaptive Sequence Transformer for Weakly Supervised Referring Expression Segmentation
por: Nag, Sayan, et al.
Publicado: (2024)
por: Nag, Sayan, et al.
Publicado: (2024)
TIDE: Training Locally Interpretable Domain Generalization Models Enables Test-time Correction
por: Agarwal, Aishwarya, et al.
Publicado: (2024)
por: Agarwal, Aishwarya, et al.
Publicado: (2024)
Concept Regions Matter: Benchmarking CLIP with a New Cluster-Importance Approach
por: Agarwal, Aishwarya, et al.
Publicado: (2025)
por: Agarwal, Aishwarya, et al.
Publicado: (2025)
LiteEmbed: Adapting CLIP to Rare Classes
por: Agarwal, Aishwarya, et al.
Publicado: (2026)
por: Agarwal, Aishwarya, et al.
Publicado: (2026)
Enhanced Rooftop Solar Panel Detection by Efficiently Aggregating Local Features
por: Kurte, Kuldeep, et al.
Publicado: (2025)
por: Kurte, Kuldeep, et al.
Publicado: (2025)
Training-free Color-Style Disentanglement for Constrained Text-to-Image Synthesis
por: Agarwal, Aishwarya, et al.
Publicado: (2024)
por: Agarwal, Aishwarya, et al.
Publicado: (2024)
AlignIT: Enhancing Prompt Alignment in Customization of Text-to-Image Models
por: Agarwal, Aishwarya, et al.
Publicado: (2024)
por: Agarwal, Aishwarya, et al.
Publicado: (2024)
DynamicEval: Rethinking Evaluation for Dynamic Text-to-Video Synthesis
por: Babu, Nithin C., et al.
Publicado: (2025)
por: Babu, Nithin C., et al.
Publicado: (2025)
Composable Part-Based Manipulation
por: Liu, Weiyu, et al.
Publicado: (2024)
por: Liu, Weiyu, et al.
Publicado: (2024)
DreamComposer: Controllable 3D Object Generation via Multi-View Conditions
por: Yang, Yunhan, et al.
Publicado: (2023)
por: Yang, Yunhan, et al.
Publicado: (2023)
CoCoNO: Attention Contrast-and-Complete for Initial Noise Optimization in Text-to-Image Synthesis
por: Sundaram, Aravindan, et al.
Publicado: (2024)
por: Sundaram, Aravindan, et al.
Publicado: (2024)
CTRL-O: Language-Controllable Object-Centric Visual Representation Learning
por: Didolkar, Aniket, et al.
Publicado: (2025)
por: Didolkar, Aniket, et al.
Publicado: (2025)
Learning from Limited and Imperfect Data
por: Rangwani, Harsh
Publicado: (2024)
por: Rangwani, Harsh
Publicado: (2024)
Composing Pre-Trained Object-Centric Representations for Robotics From "What" and "Where" Foundation Models
por: Shi, Junyao, et al.
Publicado: (2024)
por: Shi, Junyao, et al.
Publicado: (2024)
How Diffusion Models Learn to Factorize and Compose
por: Liang, Qiyao, et al.
Publicado: (2024)
por: Liang, Qiyao, et al.
Publicado: (2024)
Towards High-Order Mean Flow Generative Models: Feasibility, Expressivity, and Provably Efficient Criteria
por: Cao, Yang, et al.
Publicado: (2025)
por: Cao, Yang, et al.
Publicado: (2025)
ComposableNav: Instruction-Following Navigation in Dynamic Environments via Composable Diffusion
por: Hu, Zichao, et al.
Publicado: (2025)
por: Hu, Zichao, et al.
Publicado: (2025)
Federated Domain Generalization with Latent Space Inversion
por: Palakkadavath, Ragja, et al.
Publicado: (2025)
por: Palakkadavath, Ragja, et al.
Publicado: (2025)
An Examination of the Robustness of Reference-Free Image Captioning Evaluation Metrics
por: Ahmadi, Saba, et al.
Publicado: (2023)
por: Ahmadi, Saba, et al.
Publicado: (2023)
Pix2Code: Learning to Compose Neural Visual Concepts as Programs
por: Wüst, Antonia, et al.
Publicado: (2024)
por: Wüst, Antonia, et al.
Publicado: (2024)
Learning 3D Texture-Aware Representations for Parsing Diverse Human Clothing and Body Parts
por: Chhatre, Kiran, et al.
Publicado: (2025)
por: Chhatre, Kiran, et al.
Publicado: (2025)
DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models
por: Kim, Sungnyun, et al.
Publicado: (2023)
por: Kim, Sungnyun, et al.
Publicado: (2023)
Bridging KAN and MLP: MJKAN, a Hybrid Architecture with Both Efficiency and Expressiveness
por: Joo, Hanseon, et al.
Publicado: (2025)
por: Joo, Hanseon, et al.
Publicado: (2025)
Part$^{2}$GS: Part-aware Modeling of Articulated Objects using 3D Gaussian Splatting
por: Yu, Tianjiao, et al.
Publicado: (2025)
por: Yu, Tianjiao, et al.
Publicado: (2025)
A Composable Multimodal Framework for cine CMR-Text-Driven Prediction of Heart Failure Outcomes
por: Chen, Jianzhou, et al.
Publicado: (2025)
por: Chen, Jianzhou, et al.
Publicado: (2025)
DreamPartGen: Semantically Grounded Part-Level 3D Generation via Collaborative Latent Denoising
por: Yu, Tianjiao, et al.
Publicado: (2026)
por: Yu, Tianjiao, et al.
Publicado: (2026)
Improving Automatic VQA Evaluation Using Large Language Models
por: Mañas, Oscar, et al.
Publicado: (2023)
por: Mañas, Oscar, et al.
Publicado: (2023)
X-Mark: Saliency-Guided Robust Dataset Ownership Verification for Medical Imaging
por: Kulkarni, Pranav, et al.
Publicado: (2026)
por: Kulkarni, Pranav, et al.
Publicado: (2026)
SPADE: Spatial Transcriptomics and Pathology Alignment Using a Mixture of Data Experts for an Expressive Latent Space
por: Redekop, Ekaterina, et al.
Publicado: (2025)
por: Redekop, Ekaterina, et al.
Publicado: (2025)
On Computational Limits of FlowAR Models: Expressivity and Efficiency
por: Cao, Yang, et al.
Publicado: (2025)
por: Cao, Yang, et al.
Publicado: (2025)
Real2Code: Reconstruct Articulated Objects via Code Generation
por: Mandi, Zhao, et al.
Publicado: (2024)
por: Mandi, Zhao, et al.
Publicado: (2024)
Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback
por: Furuta, Hiroki, et al.
Publicado: (2024)
por: Furuta, Hiroki, et al.
Publicado: (2024)
A Foundation Model for General Moving Object Segmentation in Medical Images
por: Yan, Zhongnuo, et al.
Publicado: (2023)
por: Yan, Zhongnuo, et al.
Publicado: (2023)
DQE-CIR: Distinctive Query Embeddings through Learnable Attribute Weights and Target Relative Negative Sampling in Composed Image Retrieval
por: Park, Geon, et al.
Publicado: (2026)
por: Park, Geon, et al.
Publicado: (2026)
Towards Virtual Clinical Trials of Radiology AI with Conditional Generative Modeling
por: Killeen, Benjamin D., et al.
Publicado: (2025)
por: Killeen, Benjamin D., et al.
Publicado: (2025)
Learning Part Knowledge to Facilitate Category Understanding for Fine-Grained Generalized Category Discovery
por: Wang, Enguang, et al.
Publicado: (2025)
por: Wang, Enguang, et al.
Publicado: (2025)
Towards Size-invariant Salient Object Detection: A Generic Evaluation and Optimization Approach
por: Bao, Shilong, et al.
Publicado: (2025)
por: Bao, Shilong, et al.
Publicado: (2025)
Ejemplares similares
-
Learning from Limited and Imperfect Data
por: Rangwani, Harsh
Publicado: (2025) -
DeiT-LT Distillation Strikes Back for Vision Transformer Training on Long-Tailed Datasets
por: Rangwani, Harsh, et al.
Publicado: (2024) -
Selective Mixup Fine-Tuning for Optimizing Non-Decomposable Objectives
por: Ramasubramanian, Shrinivas, et al.
Publicado: (2024) -
SafaRi:Adaptive Sequence Transformer for Weakly Supervised Referring Expression Segmentation
por: Nag, Sayan, et al.
Publicado: (2024) -
TIDE: Training Locally Interpretable Domain Generalization Models Enables Test-time Correction
por: Agarwal, Aishwarya, et al.
Publicado: (2024)