Guardado en:
| Autores principales: | Dai, Fengyuan, Huang, Siteng, Zhang, Min, Gong, Biao, Wang, Donglin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2408.17083 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Troika: Multi-Path Cross-Modal Traction for Compositional Zero-Shot Learning
por: Huang, Siteng, et al.
Publicado: (2023)
por: Huang, Siteng, et al.
Publicado: (2023)
Learning Disentangled Identifiers for Action-Customized Text-to-Image Generation
por: Huang, Siteng, et al.
Publicado: (2023)
por: Huang, Siteng, et al.
Publicado: (2023)
VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL
por: Dai, Fengyuan, et al.
Publicado: (2025)
por: Dai, Fengyuan, et al.
Publicado: (2025)
VGDiffZero: Text-to-image Diffusion Models Can Be Zero-shot Visual Grounders
por: Liu, Xuyang, et al.
Publicado: (2023)
por: Liu, Xuyang, et al.
Publicado: (2023)
Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference
por: Zhao, Han, et al.
Publicado: (2024)
por: Zhao, Han, et al.
Publicado: (2024)
PiTe: Pixel-Temporal Alignment for Large Video-Language Model
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
Multi-Level Correlation Network For Few-Shot Image Classification
por: Dang, Yunkai, et al.
Publicado: (2024)
por: Dang, Yunkai, et al.
Publicado: (2024)
ProFD: Prompt-Guided Feature Disentangling for Occluded Person Re-Identification
por: Cui, Can, et al.
Publicado: (2024)
por: Cui, Can, et al.
Publicado: (2024)
Prompt-based Distribution Alignment for Unsupervised Domain Adaptation
por: Bai, Shuanghao, et al.
Publicado: (2023)
por: Bai, Shuanghao, et al.
Publicado: (2023)
Check, Locate, Rectify: A Training-Free Layout Calibration System for Text-to-Image Generation
por: Gong, Biao, et al.
Publicado: (2023)
por: Gong, Biao, et al.
Publicado: (2023)
CARP: Visuomotor Policy Learning via Coarse-to-Fine Autoregressive Prediction
por: Gong, Zhefei, et al.
Publicado: (2024)
por: Gong, Zhefei, et al.
Publicado: (2024)
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots
por: Ding, Pengxiang, et al.
Publicado: (2023)
por: Ding, Pengxiang, et al.
Publicado: (2023)
Learning Visual Proxy for Compositional Zero-Shot Learning
por: Zhang, Shiyu, et al.
Publicado: (2025)
por: Zhang, Shiyu, et al.
Publicado: (2025)
TsCA: On the Semantic Consistency Alignment via Conditional Transport for Compositional Zero-Shot Learning
por: Li, Miaoge, et al.
Publicado: (2024)
por: Li, Miaoge, et al.
Publicado: (2024)
Learning by Imagining: Debiased Feature Augmentation for Compositional Zero-Shot Learning
por: Zhang, Haozhe, et al.
Publicado: (2025)
por: Zhang, Haozhe, et al.
Publicado: (2025)
Prompting Language-Informed Distribution for Compositional Zero-Shot Learning
por: Bao, Wentao, et al.
Publicado: (2023)
por: Bao, Wentao, et al.
Publicado: (2023)
M2IST: Multi-Modal Interactive Side-Tuning for Efficient Referring Expression Comprehension
por: Liu, Xuyang, et al.
Publicado: (2024)
por: Liu, Xuyang, et al.
Publicado: (2024)
Compositional Zero-Shot Learning: A Survey
por: Munir, Ans, et al.
Publicado: (2025)
por: Munir, Ans, et al.
Publicado: (2025)
Exploring the Evolution of Physics Cognition in Video Generation: A Survey
por: Lin, Minghui, et al.
Publicado: (2025)
por: Lin, Minghui, et al.
Publicado: (2025)
SSR: Enhancing Depth Perception in Vision-Language Models via Rationale-Guided Spatial Reasoning
por: Liu, Yang, et al.
Publicado: (2025)
por: Liu, Yang, et al.
Publicado: (2025)
Semantically Guided Dynamic Visual Prototype Refinement for Compositional Zero-Shot Learning
por: Peng, Zhong, et al.
Publicado: (2025)
por: Peng, Zhong, et al.
Publicado: (2025)
Filter, Correlate, Compress: Training-Free Token Reduction for MLLM Acceleration
por: Han, Yuhang, et al.
Publicado: (2024)
por: Han, Yuhang, et al.
Publicado: (2024)
MSCI: Addressing CLIP's Inherent Limitations for Compositional Zero-Shot Learning
por: Wang, Yue, et al.
Publicado: (2025)
por: Wang, Yue, et al.
Publicado: (2025)
CSCNET: Class-Specified Cascaded Network for Compositional Zero-Shot Learning
por: Zhang, Yanyi, et al.
Publicado: (2024)
por: Zhang, Yanyi, et al.
Publicado: (2024)
Multi-Scale Memory Comparison for Zero-/Few-Shot Anomaly Detection
por: Huang, Chaoqin, et al.
Publicado: (2023)
por: Huang, Chaoqin, et al.
Publicado: (2023)
Bayesian Prompt Flow Learning for Zero-Shot Anomaly Detection
por: Qu, Zhen, et al.
Publicado: (2025)
por: Qu, Zhen, et al.
Publicado: (2025)
MAC: A Benchmark for Multiple Attributes Compositional Zero-Shot Learning
por: Xu, Shuo, et al.
Publicado: (2024)
por: Xu, Shuo, et al.
Publicado: (2024)
Multi-Granularity Mutual Refinement Network for Zero-Shot Learning
por: Wang, Ning, et al.
Publicado: (2025)
por: Wang, Ning, et al.
Publicado: (2025)
ANYPORTAL: Zero-Shot Consistent Video Background Replacement
por: Gao, Wenshuo, et al.
Publicado: (2025)
por: Gao, Wenshuo, et al.
Publicado: (2025)
Exploring Transferable Homogeneous Groups for Compositional Zero-Shot Learning
por: Rao, Zhijie, et al.
Publicado: (2025)
por: Rao, Zhijie, et al.
Publicado: (2025)
FlowComposer: Composable Flows for Compositional Zero-Shot Learning
por: He, Zhenqi, et al.
Publicado: (2026)
por: He, Zhenqi, et al.
Publicado: (2026)
3D Scene Change Modeling With Consistent Multi-View Aggregation
por: Zhou, Zirui, et al.
Publicado: (2025)
por: Zhou, Zirui, et al.
Publicado: (2025)
DARA: Domain- and Relation-aware Adapters Make Parameter-efficient Tuning for Visual Grounding
por: Liu, Ting, et al.
Publicado: (2024)
por: Liu, Ting, et al.
Publicado: (2024)
EVA: Mixture-of-Experts Semantic Variant Alignment for Compositional Zero-Shot Learning
por: Zhang, Xiao, et al.
Publicado: (2025)
por: Zhang, Xiao, et al.
Publicado: (2025)
Sparse-Tuning: Adapting Vision Transformers with Efficient Fine-tuning and Inference
por: Liu, Ting, et al.
Publicado: (2024)
por: Liu, Ting, et al.
Publicado: (2024)
Distributed Zero-Shot Learning for Visual Recognition
por: Chen, Zhi, et al.
Publicado: (2025)
por: Chen, Zhi, et al.
Publicado: (2025)
Learning Primitive Relations for Compositional Zero-Shot Learning
por: Lee, Insu, et al.
Publicado: (2025)
por: Lee, Insu, et al.
Publicado: (2025)
ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models
por: Gong, Bingchen, et al.
Publicado: (2024)
por: Gong, Bingchen, et al.
Publicado: (2024)
C2C: Component-to-Composition Learning for Zero-Shot Compositional Action Recognition
por: Li, Rongchang, et al.
Publicado: (2024)
por: Li, Rongchang, et al.
Publicado: (2024)
Context-based and Diversity-driven Specificity in Compositional Zero-Shot Learning
por: Li, Yun, et al.
Publicado: (2024)
por: Li, Yun, et al.
Publicado: (2024)
Ejemplares similares
-
Troika: Multi-Path Cross-Modal Traction for Compositional Zero-Shot Learning
por: Huang, Siteng, et al.
Publicado: (2023) -
Learning Disentangled Identifiers for Action-Customized Text-to-Image Generation
por: Huang, Siteng, et al.
Publicado: (2023) -
VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL
por: Dai, Fengyuan, et al.
Publicado: (2025) -
VGDiffZero: Text-to-image Diffusion Models Can Be Zero-shot Visual Grounders
por: Liu, Xuyang, et al.
Publicado: (2023) -
Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference
por: Zhao, Han, et al.
Publicado: (2024)