Saved in:
| Main Authors: | Lei, Jiaming, Li, Lin, Wang, Chunping, Xiao, Jun, Chen, Long |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2404.15785 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
$\text{H}^2$em: Learning Hierarchical Hyperbolic Embeddings for Compositional Zero-Shot Learning
by: Li, Lin, et al.
Published: (2025)
by: Li, Lin, et al.
Published: (2025)
SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
by: Li, Rong, et al.
Published: (2024)
by: Li, Rong, et al.
Published: (2024)
Compositional Zero-shot Learning via Progressive Language-based Observations
by: Li, Lin, et al.
Published: (2023)
by: Li, Lin, et al.
Published: (2023)
Solving Zero-Shot 3D Visual Grounding as Constraint Satisfaction Problems
by: Yuan, Qihao, et al.
Published: (2024)
by: Yuan, Qihao, et al.
Published: (2024)
Zero-shot Compositional Action Recognition with Neural Logic Constraints
by: Ye, Gefan, et al.
Published: (2025)
by: Ye, Gefan, et al.
Published: (2025)
Compositional Feature Augmentation for Unbiased Scene Graph Generation
by: Li, Lin, et al.
Published: (2023)
by: Li, Lin, et al.
Published: (2023)
Beyond Heuristic Prompting: A Concept-Guided Bayesian Framework for Zero-Shot Image Recognition
by: Liu, Hui, et al.
Published: (2026)
by: Liu, Hui, et al.
Published: (2026)
Language Model as Visual Explainer
by: Yang, Xingyi, et al.
Published: (2024)
by: Yang, Xingyi, et al.
Published: (2024)
Task-Specific Distance Correlation Matching for Few-Shot Action Recognition
by: Long, Fei, et al.
Published: (2025)
by: Long, Fei, et al.
Published: (2025)
Distributed Zero-Shot Learning for Visual Recognition
by: Chen, Zhi, et al.
Published: (2025)
by: Chen, Zhi, et al.
Published: (2025)
FlowComposer: Composable Flows for Compositional Zero-Shot Learning
by: He, Zhenqi, et al.
Published: (2026)
by: He, Zhenqi, et al.
Published: (2026)
$\text{Di}^2\text{Pose}$: Discrete Diffusion Model for Occluded 3D Human Pose Estimation
by: Wang, Weiquan, et al.
Published: (2024)
by: Wang, Weiquan, et al.
Published: (2024)
DA-Font: Few-Shot Font Generation via Dual-Attention Hybrid Integration
by: Chen, Weiran, et al.
Published: (2025)
by: Chen, Weiran, et al.
Published: (2025)
Grounding Descriptions in Images informs Zero-Shot Visual Recognition
by: Halbe, Shaunak, et al.
Published: (2024)
by: Halbe, Shaunak, et al.
Published: (2024)
C2C: Component-to-Composition Learning for Zero-Shot Compositional Action Recognition
by: Li, Rongchang, et al.
Published: (2024)
by: Li, Rongchang, et al.
Published: (2024)
Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing
by: Gao, Kaifeng, et al.
Published: (2024)
by: Gao, Kaifeng, et al.
Published: (2024)
Situat3DChange: Situated 3D Change Understanding Dataset for Multimodal Large Language Model
by: Liu, Ruiping, et al.
Published: (2025)
by: Liu, Ruiping, et al.
Published: (2025)
EZSR: Event-based Zero-Shot Recognition
by: Yang, Yan, et al.
Published: (2024)
by: Yang, Yan, et al.
Published: (2024)
Liberating Seen Classes: Boosting Few-Shot and Zero-Shot Text Classification via Anchor Generation and Classification Reframing
by: Liu, Han, et al.
Published: (2024)
by: Liu, Han, et al.
Published: (2024)
Zero-Shot Visual Grounding in 3D Gaussians via View Retrieval
by: Liao, Liwei, et al.
Published: (2025)
by: Liao, Liwei, et al.
Published: (2025)
Zero-Shot 3D Visual Grounding from Vision-Language Models
by: Li, Rong, et al.
Published: (2025)
by: Li, Rong, et al.
Published: (2025)
Zero-Shot Underwater Gesture Recognition
by: Sarma, Sandipan, et al.
Published: (2024)
by: Sarma, Sandipan, et al.
Published: (2024)
From Semantics, Scene to Instance-awareness: Distilling Foundation Model for Grounded Open-vocabulary Situation Recognition
by: Cai, Chen, et al.
Published: (2025)
by: Cai, Chen, et al.
Published: (2025)
Zero-P-to-3: Zero-Shot Partial-View Images to 3D Object
by: Lin, Yuxuan, et al.
Published: (2025)
by: Lin, Yuxuan, et al.
Published: (2025)
ZONE: Zero-Shot Instruction-Guided Local Editing
by: Li, Shanglin, et al.
Published: (2023)
by: Li, Shanglin, et al.
Published: (2023)
VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
by: Xu, Runsen, et al.
Published: (2024)
by: Xu, Runsen, et al.
Published: (2024)
Seeing Beyond the Scene: Enhancing Vision-Language Models with Interactional Reasoning
by: Liang, Dayong, et al.
Published: (2025)
by: Liang, Dayong, et al.
Published: (2025)
What Do You See? Enhancing Zero-Shot Image Classification with Multimodal Large Language Models
by: Abdelhamed, Abdelrahman, et al.
Published: (2024)
by: Abdelhamed, Abdelrahman, et al.
Published: (2024)
Zero-Shot Class Unlearning in CLIP with Synthetic Samples
by: Kravets, A., et al.
Published: (2024)
by: Kravets, A., et al.
Published: (2024)
Benchmarking Zero-Shot Recognition with Vision-Language Models: Challenges on Granularity and Specificity
by: Xu, Zhenlin, et al.
Published: (2023)
by: Xu, Zhenlin, et al.
Published: (2023)
Think, Act, Build: An Agentic Framework with Vision Language Models for Zero-Shot 3D Visual Grounding
by: Wang, Haibo, et al.
Published: (2026)
by: Wang, Haibo, et al.
Published: (2026)
What's in a Name? Beyond Class Indices for Image Recognition
by: Han, Kai, et al.
Published: (2023)
by: Han, Kai, et al.
Published: (2023)
Zero-Shot Synthetic-to-Real Handwritten Text Recognition via Task Analogies
by: Garrido-Munoz, Carlos, et al.
Published: (2026)
by: Garrido-Munoz, Carlos, et al.
Published: (2026)
Zero-Shot Video Translation via Token Warping
by: Zhu, Haiming, et al.
Published: (2024)
by: Zhu, Haiming, et al.
Published: (2024)
Zero-Shot Action Recognition in Surveillance Videos
by: Pereira, Joao, et al.
Published: (2024)
by: Pereira, Joao, et al.
Published: (2024)
ViD-GPT: Introducing GPT-style Autoregressive Generation in Video Diffusion Models
by: Gao, Kaifeng, et al.
Published: (2024)
by: Gao, Kaifeng, et al.
Published: (2024)
Spectral Prompt Tuning:Unveiling Unseen Classes for Zero-Shot Semantic Segmentation
by: Xu, Wenhao, et al.
Published: (2023)
by: Xu, Wenhao, et al.
Published: (2023)
Continual Learning Improves Zero-Shot Action Recognition
by: Gowda, Shreyank N, et al.
Published: (2024)
by: Gowda, Shreyank N, et al.
Published: (2024)
Novel Semantic Prompting for Zero-Shot Action Recognition
by: Iqbal, Salman, et al.
Published: (2026)
by: Iqbal, Salman, et al.
Published: (2026)
Representation-Level Counterfactual Calibration for Debiased Zero-Shot Recognition
by: Peng, Pei, et al.
Published: (2025)
by: Peng, Pei, et al.
Published: (2025)
Similar Items
-
$\text{H}^2$em: Learning Hierarchical Hyperbolic Embeddings for Compositional Zero-Shot Learning
by: Li, Lin, et al.
Published: (2025) -
SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
by: Li, Rong, et al.
Published: (2024) -
Compositional Zero-shot Learning via Progressive Language-based Observations
by: Li, Lin, et al.
Published: (2023) -
Solving Zero-Shot 3D Visual Grounding as Constraint Satisfaction Problems
by: Yuan, Qihao, et al.
Published: (2024) -
Zero-shot Compositional Action Recognition with Neural Logic Constraints
by: Ye, Gefan, et al.
Published: (2025)