Enregistré dans:
| Auteurs principaux: | Deng, Huilin, Luo, Hongchen, Zhai, Wei, Cao, Yang, Kang, Yu |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2409.20146 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning
par: Deng, Huilin, et autres
Publié: (2025)
par: Deng, Huilin, et autres
Publié: (2025)
Bidirectional Progressive Transformer for Interaction Intention Anticipation
par: Zhang, Zichen, et autres
Publié: (2024)
par: Zhang, Zichen, et autres
Publié: (2024)
PEAR: Phrase-Based Hand-Object Interaction Anticipation
par: Zhang, Zichen, et autres
Publié: (2024)
par: Zhang, Zichen, et autres
Publié: (2024)
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
par: Xu, Jiacong, et autres
Publié: (2025)
par: Xu, Jiacong, et autres
Publié: (2025)
Visual-Geometric Collaborative Guidance for Affordance Learning
par: Luo, Hongchen, et autres
Publié: (2024)
par: Luo, Hongchen, et autres
Publié: (2024)
Intention-driven Ego-to-Exo Video Generation
par: Luo, Hongchen, et autres
Publié: (2024)
par: Luo, Hongchen, et autres
Publié: (2024)
Global-Regularized Neighborhood Regression for Efficient Zero-Shot Texture Anomaly Detection
par: Yao, Haiming, et autres
Publié: (2024)
par: Yao, Haiming, et autres
Publié: (2024)
CoPS: Conditional Prompt Synthesis for Zero-Shot Anomaly Detection
par: Chen, Qiyu, et autres
Publié: (2025)
par: Chen, Qiyu, et autres
Publié: (2025)
LEMON: Learning 3D Human-Object Interaction Relation from 2D Images
par: Yang, Yuhang, et autres
Publié: (2023)
par: Yang, Yuhang, et autres
Publié: (2023)
IAD-GPT: Advancing Visual Knowledge in Multimodal Large Language Model for Industrial Anomaly Detection
par: Li, Zewen, et autres
Publié: (2025)
par: Li, Zewen, et autres
Publié: (2025)
Leverage Task Context for Object Affordance Ranking
par: Huang, Haojie, et autres
Publié: (2024)
par: Huang, Haojie, et autres
Publié: (2024)
AG-VAS: Anchor-Guided Zero-Shot Visual Anomaly Segmentation with Large Multimodal Models
par: Qu, Zhen, et autres
Publié: (2026)
par: Qu, Zhen, et autres
Publié: (2026)
VisualAD: Language-Free Zero-Shot Anomaly Detection via Vision Transformer
par: Hou, Yanning, et autres
Publié: (2026)
par: Hou, Yanning, et autres
Publié: (2026)
Dual-Image Enhanced CLIP for Zero-Shot Anomaly Detection
par: Zhang, Zhaoxiang, et autres
Publié: (2024)
par: Zhang, Zhaoxiang, et autres
Publié: (2024)
Towards Zero-Shot Differential Morphing Attack Detection with Multimodal Large Language Models
par: Shekhawat, Ria, et autres
Publié: (2025)
par: Shekhawat, Ria, et autres
Publié: (2025)
SSVP: Synergistic Semantic-Visual Prompting for Industrial Zero-Shot Anomaly Detection
par: Fu, Chenhao, et autres
Publié: (2026)
par: Fu, Chenhao, et autres
Publié: (2026)
Learning Multi-view Multi-class Anomaly Detection
par: Yu, Qianzi, et autres
Publié: (2025)
par: Yu, Qianzi, et autres
Publié: (2025)
ZSG-IAD: A Multimodal Framework for Zero-Shot Grounded Industrial Anomaly Detection
par: Chen, Qiuhui, et autres
Publié: (2026)
par: Chen, Qiuhui, et autres
Publié: (2026)
Visual Context Window Extension: A New Perspective for Long Video Understanding
par: Wei, Hongchen, et autres
Publié: (2024)
par: Wei, Hongchen, et autres
Publié: (2024)
AnomalyAgent: Training-Free Agentic Models for Zero-/Few-Shot Anomaly Detection
par: Zhang, Yi, et autres
Publié: (2026)
par: Zhang, Yi, et autres
Publié: (2026)
AnyAnomaly: Zero-Shot Customizable Video Anomaly Detection with LVLM
par: Ahn, Sunghyun, et autres
Publié: (2025)
par: Ahn, Sunghyun, et autres
Publié: (2025)
Large Multilingual Models Pivot Zero-Shot Multimodal Learning across Languages
par: Hu, Jinyi, et autres
Publié: (2023)
par: Hu, Jinyi, et autres
Publié: (2023)
Bootstrap Fine-Grained Vision-Language Alignment for Unified Zero-Shot Anomaly Localization
par: Deng, Hanqiu, et autres
Publié: (2023)
par: Deng, Hanqiu, et autres
Publié: (2023)
VarAD: Lightweight High-Resolution Image Anomaly Detection via Visual Autoregressive Modeling
par: Cao, Yunkang, et autres
Publié: (2024)
par: Cao, Yunkang, et autres
Publié: (2024)
AnoRefiner: Anomaly-Aware Group-Wise Refinement for Zero-Shot Industrial Anomaly Detection
par: Huang, Dayou, et autres
Publié: (2025)
par: Huang, Dayou, et autres
Publié: (2025)
On the Problem of Consistent Anomalies in Zero-Shot Industrial Anomaly Detection
par: Le-Gia, Tai, et autres
Publié: (2025)
par: Le-Gia, Tai, et autres
Publié: (2025)
On the Problem of Consistent Anomalies in Zero-Shot Anomaly Detection
par: Le-Gia, Tai
Publié: (2025)
par: Le-Gia, Tai
Publié: (2025)
Seeing Is Believing? A Benchmark for Multimodal Large Language Models on Visual Illusions and Anomalies
par: Hou, Wenjin, et autres
Publié: (2026)
par: Hou, Wenjin, et autres
Publié: (2026)
GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding
par: Shao, Yawen, et autres
Publié: (2024)
par: Shao, Yawen, et autres
Publié: (2024)
Back to Point: Exploring Point-Language Models for Zero-Shot 3D Anomaly Detection
par: Li, Kaiqiang, et autres
Publié: (2026)
par: Li, Kaiqiang, et autres
Publié: (2026)
Zero-Shot Image Anomaly Detection Using Generative Foundation Models
par: Abdi, Lemar, et autres
Publié: (2025)
par: Abdi, Lemar, et autres
Publié: (2025)
Expanding Zero-Shot Object Counting with Rich Prompts
par: Zhu, Huilin, et autres
Publié: (2025)
par: Zhu, Huilin, et autres
Publié: (2025)
AnomalyPainter: Vision-Language-Diffusion Synergy for Zero-Shot Realistic and Diverse Industrial Anomaly Synthesis
par: Lai, Zhangyu, et autres
Publié: (2025)
par: Lai, Zhangyu, et autres
Publié: (2025)
AdaCLIP: Adapting CLIP with Hybrid Learnable Prompts for Zero-Shot Anomaly Detection
par: Cao, Yunkang, et autres
Publié: (2024)
par: Cao, Yunkang, et autres
Publié: (2024)
LongCaptioning: Unlocking the Power of Long Video Caption Generation in Large Multimodal Models
par: Wei, Hongchen, et autres
Publié: (2025)
par: Wei, Hongchen, et autres
Publié: (2025)
VETime: Vision Enhanced Zero-Shot Time Series Anomaly Detection
par: Yang, Yingyuan, et autres
Publié: (2026)
par: Yang, Yingyuan, et autres
Publié: (2026)
Distributed Zero-Shot Learning for Visual Recognition
par: Chen, Zhi, et autres
Publié: (2025)
par: Chen, Zhi, et autres
Publié: (2025)
Anomaly-Aware Vision-Language Adapters for Zero-Shot Anomaly Detection
par: Aqeel, Muhammad, et autres
Publié: (2026)
par: Aqeel, Muhammad, et autres
Publié: (2026)
Zero-Shot Scene Understanding with Multimodal Large Language Models for Automated Vehicles
par: Elhenawy, Mohammed, et autres
Publié: (2025)
par: Elhenawy, Mohammed, et autres
Publié: (2025)
No Need For Real Anomaly: MLLM Empowered Zero-Shot Video Anomaly Detection
par: Dai, Zunkai, et autres
Publié: (2026)
par: Dai, Zunkai, et autres
Publié: (2026)
Documents similaires
-
Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning
par: Deng, Huilin, et autres
Publié: (2025) -
Bidirectional Progressive Transformer for Interaction Intention Anticipation
par: Zhang, Zichen, et autres
Publié: (2024) -
PEAR: Phrase-Based Hand-Object Interaction Anticipation
par: Zhang, Zichen, et autres
Publié: (2024) -
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
par: Xu, Jiacong, et autres
Publié: (2025) -
Visual-Geometric Collaborative Guidance for Affordance Learning
par: Luo, Hongchen, et autres
Publié: (2024)