Unified Framework for Open-World Compositional Zero-shot Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Jayasekara, Hirunima, Pham, Khoi, Saini, Nirat, Shrivastava, Abhinav |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
All-in-One Conditioning for Text-to-Image Synthesis
by: Jayasekara, Hirunima, et al.
Published: (2026)
by: Jayasekara, Hirunima, et al.
Published: (2026)
InVi: Object Insertion In Videos Using Off-the-Shelf Diffusion Models
by: Saini, Nirat, et al.
Published: (2024)
by: Saini, Nirat, et al.
Published: (2024)
Composing Object Relations and Attributes for Image-Text Matching
by: Pham, Khoi, et al.
Published: (2024)
by: Pham, Khoi, et al.
Published: (2024)
Efficient Continuous Video Flow Model for Video Prediction
by: Shrivastava, Gaurav, et al.
Published: (2024)
by: Shrivastava, Gaurav, et al.
Published: (2024)
A Conditional Probability Framework for Compositional Zero-shot Learning
by: Wu, Peng, et al.
Published: (2025)
by: Wu, Peng, et al.
Published: (2025)
LP-OVOD: Open-Vocabulary Object Detection by Linear Probing
by: Pham, Chau, et al.
Published: (2023)
by: Pham, Chau, et al.
Published: (2023)
Trajectory-aligned Space-time Tokens for Few-shot Action Recognition
by: Kumar, Pulkit, et al.
Published: (2024)
by: Kumar, Pulkit, et al.
Published: (2024)
Learning Clustering-based Prototypes for Compositional Zero-shot Learning
by: Qu, Hongyu, et al.
Published: (2025)
by: Qu, Hongyu, et al.
Published: (2025)
Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction
by: Shrivastava, Gaurav, et al.
Published: (2024)
by: Shrivastava, Gaurav, et al.
Published: (2024)
Compositional Zero-shot Learning via Progressive Language-based Observations
by: Li, Lin, et al.
Published: (2023)
by: Li, Lin, et al.
Published: (2023)
Unified Language-driven Zero-shot Domain Adaptation
by: Yang, Senqiao, et al.
Published: (2024)
by: Yang, Senqiao, et al.
Published: (2024)
Improved Feature Generating Framework for Transductive Zero-shot Learning
by: Ye, Zihan, et al.
Published: (2024)
by: Ye, Zihan, et al.
Published: (2024)
Graph-guided Cross-composition Feature Disentanglement for Compositional Zero-shot Learning
by: Geng, Yuxia, et al.
Published: (2024)
by: Geng, Yuxia, et al.
Published: (2024)
No Hard Negatives Required: Concept Centric Learning Leads to Compositionality without Degrading Zero-shot Capabilities of Contrastive Models
by: Pham, Hai X., et al.
Published: (2026)
by: Pham, Hai X., et al.
Published: (2026)
V-VIPE: Variational View Invariant Pose Embedding
by: Levy, Mara, et al.
Published: (2024)
by: Levy, Mara, et al.
Published: (2024)
Video Decomposition Prior: A Methodology to Decompose Videos into Layers
by: Shrivastava, Gaurav, et al.
Published: (2024)
by: Shrivastava, Gaurav, et al.
Published: (2024)
Evolutionary Caching to Accelerate Your Off-the-Shelf Diffusion Model
by: Aggarwal, Anirud, et al.
Published: (2025)
by: Aggarwal, Anirud, et al.
Published: (2025)
TeCoNeRV: Leveraging Temporal Coherence for Compressible Neural Representations for Videos
by: Padmanabhan, Namitha, et al.
Published: (2026)
by: Padmanabhan, Namitha, et al.
Published: (2026)
Mitigating Hallucinations in Diffusion Models through Adaptive Attention Modulation
by: Oorloff, Trevine, et al.
Published: (2025)
by: Oorloff, Trevine, et al.
Published: (2025)
FlowOVD: Learning Generative Latent Flows for Zero-shot Open-vocabulary Detection
by: Wei, Yao, et al.
Published: (2026)
by: Wei, Yao, et al.
Published: (2026)
Attention Based Simple Primitives for Open World Compositional Zero-Shot Learning
by: Munir, Ans, et al.
Published: (2024)
by: Munir, Ans, et al.
Published: (2024)
Contextual Interaction via Primitive-based Adversarial Training For Compositional Zero-shot Learning
by: Li, Suyi, et al.
Published: (2024)
by: Li, Suyi, et al.
Published: (2024)
Open-RGBT: Open-vocabulary RGB-T Zero-shot Semantic Segmentation in Open-world Environments
by: Yu, Meng, et al.
Published: (2024)
by: Yu, Meng, et al.
Published: (2024)
Zero-shot Compositional Action Recognition with Neural Logic Constraints
by: Ye, Gefan, et al.
Published: (2025)
by: Ye, Gefan, et al.
Published: (2025)
ZeroComp: Zero-shot Object Compositing from Image Intrinsics via Diffusion
by: Zhang, Zitian, et al.
Published: (2024)
by: Zhang, Zitian, et al.
Published: (2024)
Latent-INR: A Flexible Framework for Implicit Representations of Videos with Discriminative Semantics
by: Maiya, Shishira R, et al.
Published: (2024)
by: Maiya, Shishira R, et al.
Published: (2024)
EAGLES: Efficient Accelerated 3D Gaussians with Lightweight EncodingS
by: Girish, Sharath, et al.
Published: (2023)
by: Girish, Sharath, et al.
Published: (2023)
Part-aware Unified Representation of Language and Skeleton for Zero-shot Action Recognition
by: Zhu, Anqi, et al.
Published: (2024)
by: Zhu, Anqi, et al.
Published: (2024)
Explanatory Instructions: Towards Unified Vision Tasks Understanding and Zero-shot Generalization
by: Shen, Yang, et al.
Published: (2024)
by: Shen, Yang, et al.
Published: (2024)
Prompting without Panic: Attribute-aware, Zero-shot, Test-Time Calibration
by: Hebbalaguppe, Ramya, et al.
Published: (2025)
by: Hebbalaguppe, Ramya, et al.
Published: (2025)
OpenKD: Opening Prompt Diversity for Zero- and Few-shot Keypoint Detection
by: Lu, Changsheng, et al.
Published: (2024)
by: Lu, Changsheng, et al.
Published: (2024)
Visual Programming for Zero-shot Open-Vocabulary 3D Visual Grounding
by: Yuan, Zhihao, et al.
Published: (2023)
by: Yuan, Zhihao, et al.
Published: (2023)
Stereo Anything: Unifying Zero-shot Stereo Matching with Large-Scale Mixed Data
by: Guo, Xianda, et al.
Published: (2024)
by: Guo, Xianda, et al.
Published: (2024)
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
by: Nguyen, Phuc D. A., et al.
Published: (2023)
by: Nguyen, Phuc D. A., et al.
Published: (2023)
LiFT: A Surprisingly Simple Lightweight Feature Transform for Dense ViT Descriptors
by: Suri, Saksham, et al.
Published: (2024)
by: Suri, Saksham, et al.
Published: (2024)
UPLiFT: Efficient Pixel-Dense Feature Upsampling with Local Attenders
by: Walmer, Matthew, et al.
Published: (2026)
by: Walmer, Matthew, et al.
Published: (2026)
Zero-shot World Models Are Developmentally Efficient Learners
by: Aw, Khai Loong, et al.
Published: (2026)
by: Aw, Khai Loong, et al.
Published: (2026)
Prioritized Semantic Learning for Zero-shot Instance Navigation
by: Sun, Xinyu, et al.
Published: (2024)
by: Sun, Xinyu, et al.
Published: (2024)
ZeroDiff++: Substantial Unseen Visual-semantic Correlation in Zero-shot Learning
by: Ye, Zihan, et al.
Published: (2026)
by: Ye, Zihan, et al.
Published: (2026)
Interpretable Zero-shot Learning with Infinite Class Concepts
by: Ye, Zihan, et al.
Published: (2025)
by: Ye, Zihan, et al.
Published: (2025)
Similar Items
-
All-in-One Conditioning for Text-to-Image Synthesis
by: Jayasekara, Hirunima, et al.
Published: (2026) -
InVi: Object Insertion In Videos Using Off-the-Shelf Diffusion Models
by: Saini, Nirat, et al.
Published: (2024) -
Composing Object Relations and Attributes for Image-Text Matching
by: Pham, Khoi, et al.
Published: (2024) -
Efficient Continuous Video Flow Model for Video Prediction
by: Shrivastava, Gaurav, et al.
Published: (2024) -
A Conditional Probability Framework for Compositional Zero-shot Learning
by: Wu, Peng, et al.
Published: (2025)