Guardado en:
| Autores principales: | Deng, Huilin, Luo, Hongchen, Zhai, Wei, Cao, Yang, Kang, Yu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2409.20146 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning
por: Deng, Huilin, et al.
Publicado: (2025)
por: Deng, Huilin, et al.
Publicado: (2025)
Bidirectional Progressive Transformer for Interaction Intention Anticipation
por: Zhang, Zichen, et al.
Publicado: (2024)
por: Zhang, Zichen, et al.
Publicado: (2024)
PEAR: Phrase-Based Hand-Object Interaction Anticipation
por: Zhang, Zichen, et al.
Publicado: (2024)
por: Zhang, Zichen, et al.
Publicado: (2024)
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
por: Xu, Jiacong, et al.
Publicado: (2025)
por: Xu, Jiacong, et al.
Publicado: (2025)
Visual-Geometric Collaborative Guidance for Affordance Learning
por: Luo, Hongchen, et al.
Publicado: (2024)
por: Luo, Hongchen, et al.
Publicado: (2024)
Intention-driven Ego-to-Exo Video Generation
por: Luo, Hongchen, et al.
Publicado: (2024)
por: Luo, Hongchen, et al.
Publicado: (2024)
Global-Regularized Neighborhood Regression for Efficient Zero-Shot Texture Anomaly Detection
por: Yao, Haiming, et al.
Publicado: (2024)
por: Yao, Haiming, et al.
Publicado: (2024)
CoPS: Conditional Prompt Synthesis for Zero-Shot Anomaly Detection
por: Chen, Qiyu, et al.
Publicado: (2025)
por: Chen, Qiyu, et al.
Publicado: (2025)
LEMON: Learning 3D Human-Object Interaction Relation from 2D Images
por: Yang, Yuhang, et al.
Publicado: (2023)
por: Yang, Yuhang, et al.
Publicado: (2023)
IAD-GPT: Advancing Visual Knowledge in Multimodal Large Language Model for Industrial Anomaly Detection
por: Li, Zewen, et al.
Publicado: (2025)
por: Li, Zewen, et al.
Publicado: (2025)
Leverage Task Context for Object Affordance Ranking
por: Huang, Haojie, et al.
Publicado: (2024)
por: Huang, Haojie, et al.
Publicado: (2024)
AG-VAS: Anchor-Guided Zero-Shot Visual Anomaly Segmentation with Large Multimodal Models
por: Qu, Zhen, et al.
Publicado: (2026)
por: Qu, Zhen, et al.
Publicado: (2026)
VisualAD: Language-Free Zero-Shot Anomaly Detection via Vision Transformer
por: Hou, Yanning, et al.
Publicado: (2026)
por: Hou, Yanning, et al.
Publicado: (2026)
Dual-Image Enhanced CLIP for Zero-Shot Anomaly Detection
por: Zhang, Zhaoxiang, et al.
Publicado: (2024)
por: Zhang, Zhaoxiang, et al.
Publicado: (2024)
Towards Zero-Shot Differential Morphing Attack Detection with Multimodal Large Language Models
por: Shekhawat, Ria, et al.
Publicado: (2025)
por: Shekhawat, Ria, et al.
Publicado: (2025)
SSVP: Synergistic Semantic-Visual Prompting for Industrial Zero-Shot Anomaly Detection
por: Fu, Chenhao, et al.
Publicado: (2026)
por: Fu, Chenhao, et al.
Publicado: (2026)
Learning Multi-view Multi-class Anomaly Detection
por: Yu, Qianzi, et al.
Publicado: (2025)
por: Yu, Qianzi, et al.
Publicado: (2025)
ZSG-IAD: A Multimodal Framework for Zero-Shot Grounded Industrial Anomaly Detection
por: Chen, Qiuhui, et al.
Publicado: (2026)
por: Chen, Qiuhui, et al.
Publicado: (2026)
Visual Context Window Extension: A New Perspective for Long Video Understanding
por: Wei, Hongchen, et al.
Publicado: (2024)
por: Wei, Hongchen, et al.
Publicado: (2024)
AnomalyAgent: Training-Free Agentic Models for Zero-/Few-Shot Anomaly Detection
por: Zhang, Yi, et al.
Publicado: (2026)
por: Zhang, Yi, et al.
Publicado: (2026)
AnyAnomaly: Zero-Shot Customizable Video Anomaly Detection with LVLM
por: Ahn, Sunghyun, et al.
Publicado: (2025)
por: Ahn, Sunghyun, et al.
Publicado: (2025)
Large Multilingual Models Pivot Zero-Shot Multimodal Learning across Languages
por: Hu, Jinyi, et al.
Publicado: (2023)
por: Hu, Jinyi, et al.
Publicado: (2023)
Bootstrap Fine-Grained Vision-Language Alignment for Unified Zero-Shot Anomaly Localization
por: Deng, Hanqiu, et al.
Publicado: (2023)
por: Deng, Hanqiu, et al.
Publicado: (2023)
VarAD: Lightweight High-Resolution Image Anomaly Detection via Visual Autoregressive Modeling
por: Cao, Yunkang, et al.
Publicado: (2024)
por: Cao, Yunkang, et al.
Publicado: (2024)
AnoRefiner: Anomaly-Aware Group-Wise Refinement for Zero-Shot Industrial Anomaly Detection
por: Huang, Dayou, et al.
Publicado: (2025)
por: Huang, Dayou, et al.
Publicado: (2025)
On the Problem of Consistent Anomalies in Zero-Shot Industrial Anomaly Detection
por: Le-Gia, Tai, et al.
Publicado: (2025)
por: Le-Gia, Tai, et al.
Publicado: (2025)
On the Problem of Consistent Anomalies in Zero-Shot Anomaly Detection
por: Le-Gia, Tai
Publicado: (2025)
por: Le-Gia, Tai
Publicado: (2025)
Seeing Is Believing? A Benchmark for Multimodal Large Language Models on Visual Illusions and Anomalies
por: Hou, Wenjin, et al.
Publicado: (2026)
por: Hou, Wenjin, et al.
Publicado: (2026)
GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding
por: Shao, Yawen, et al.
Publicado: (2024)
por: Shao, Yawen, et al.
Publicado: (2024)
Back to Point: Exploring Point-Language Models for Zero-Shot 3D Anomaly Detection
por: Li, Kaiqiang, et al.
Publicado: (2026)
por: Li, Kaiqiang, et al.
Publicado: (2026)
Zero-Shot Image Anomaly Detection Using Generative Foundation Models
por: Abdi, Lemar, et al.
Publicado: (2025)
por: Abdi, Lemar, et al.
Publicado: (2025)
Expanding Zero-Shot Object Counting with Rich Prompts
por: Zhu, Huilin, et al.
Publicado: (2025)
por: Zhu, Huilin, et al.
Publicado: (2025)
AnomalyPainter: Vision-Language-Diffusion Synergy for Zero-Shot Realistic and Diverse Industrial Anomaly Synthesis
por: Lai, Zhangyu, et al.
Publicado: (2025)
por: Lai, Zhangyu, et al.
Publicado: (2025)
AdaCLIP: Adapting CLIP with Hybrid Learnable Prompts for Zero-Shot Anomaly Detection
por: Cao, Yunkang, et al.
Publicado: (2024)
por: Cao, Yunkang, et al.
Publicado: (2024)
LongCaptioning: Unlocking the Power of Long Video Caption Generation in Large Multimodal Models
por: Wei, Hongchen, et al.
Publicado: (2025)
por: Wei, Hongchen, et al.
Publicado: (2025)
VETime: Vision Enhanced Zero-Shot Time Series Anomaly Detection
por: Yang, Yingyuan, et al.
Publicado: (2026)
por: Yang, Yingyuan, et al.
Publicado: (2026)
Distributed Zero-Shot Learning for Visual Recognition
por: Chen, Zhi, et al.
Publicado: (2025)
por: Chen, Zhi, et al.
Publicado: (2025)
Anomaly-Aware Vision-Language Adapters for Zero-Shot Anomaly Detection
por: Aqeel, Muhammad, et al.
Publicado: (2026)
por: Aqeel, Muhammad, et al.
Publicado: (2026)
Zero-Shot Scene Understanding with Multimodal Large Language Models for Automated Vehicles
por: Elhenawy, Mohammed, et al.
Publicado: (2025)
por: Elhenawy, Mohammed, et al.
Publicado: (2025)
No Need For Real Anomaly: MLLM Empowered Zero-Shot Video Anomaly Detection
por: Dai, Zunkai, et al.
Publicado: (2026)
por: Dai, Zunkai, et al.
Publicado: (2026)
Ejemplares similares
-
Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning
por: Deng, Huilin, et al.
Publicado: (2025) -
Bidirectional Progressive Transformer for Interaction Intention Anticipation
por: Zhang, Zichen, et al.
Publicado: (2024) -
PEAR: Phrase-Based Hand-Object Interaction Anticipation
por: Zhang, Zichen, et al.
Publicado: (2024) -
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
por: Xu, Jiacong, et al.
Publicado: (2025) -
Visual-Geometric Collaborative Guidance for Affordance Learning
por: Luo, Hongchen, et al.
Publicado: (2024)