Triad: Empowering LMM-based Anomaly Detection with Vision Expert-guided Visual Tokenizer and Manufacturing Process
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Yuanze, Yuan, Shihao, Wang, Haolin, Li, Qizhang, Liu, Ming, Xu, Chen, Shi, Guangming, Zuo, Wangmeng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Myriad: Large Multimodal Model by Applying Vision Experts for Industrial Anomaly Detection
di: Li, Yuanze, et al.
Pubblicazione: (2023)
di: Li, Yuanze, et al.
Pubblicazione: (2023)
Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation
di: Li, Qizhang, et al.
Pubblicazione: (2024)
di: Li, Qizhang, et al.
Pubblicazione: (2024)
Improved Generation of Adversarial Examples Against Safety-aligned LLMs
di: Li, Qizhang, et al.
Pubblicazione: (2024)
di: Li, Qizhang, et al.
Pubblicazione: (2024)
Empower Vision Applications with LoRA LMM
di: Mi, Liang, et al.
Pubblicazione: (2024)
di: Mi, Liang, et al.
Pubblicazione: (2024)
Improving Transferability of Adversarial Examples via Bayesian Attacks
di: Li, Qizhang, et al.
Pubblicazione: (2023)
di: Li, Qizhang, et al.
Pubblicazione: (2023)
Grounding-MD: Grounded Video-language Pre-training for Open-World Moment Detection
di: Zhuang, Weijun, et al.
Pubblicazione: (2025)
di: Zhuang, Weijun, et al.
Pubblicazione: (2025)
GLAD: Towards Better Reconstruction with Global and Local Adaptive Diffusion Models for Unsupervised Anomaly Detection
di: Yao, Hang, et al.
Pubblicazione: (2024)
di: Yao, Hang, et al.
Pubblicazione: (2024)
GLaVE-Cap: Global-Local Aligned Video Captioning with Vision Expert Integration
di: Xu, Wan, et al.
Pubblicazione: (2025)
di: Xu, Wan, et al.
Pubblicazione: (2025)
Unprejudiced Training Auxiliary Tasks Makes Primary Better: A Multi-Task Learning Perspective
di: Li, Yuanze, et al.
Pubblicazione: (2024)
di: Li, Yuanze, et al.
Pubblicazione: (2024)
LPT++: Efficient Training on Mixture of Long-tailed Experts
di: Dong, Bowen, et al.
Pubblicazione: (2024)
di: Dong, Bowen, et al.
Pubblicazione: (2024)
Retrieval Augmented Image Harmonization
di: Wang, Haolin, et al.
Pubblicazione: (2024)
di: Wang, Haolin, et al.
Pubblicazione: (2024)
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
di: Dong, Bowen, et al.
Pubblicazione: (2024)
di: Dong, Bowen, et al.
Pubblicazione: (2024)
Cross-Sensor RGB Spectrograms: A Visual Method for Anomaly Detection in Classical and Quantum Magnetometer Triads
di: Pandey, Manas
Pubblicazione: (2026)
di: Pandey, Manas
Pubblicazione: (2026)
Mixture of Nested Experts: Adaptive Processing of Visual Tokens
di: Jain, Gagan, et al.
Pubblicazione: (2024)
di: Jain, Gagan, et al.
Pubblicazione: (2024)
Pseudo Replay-based Class Continual Learning for Online New Category Anomaly Detection in Advanced Manufacturing
di: Li, Yuxuan, et al.
Pubblicazione: (2023)
di: Li, Yuxuan, et al.
Pubblicazione: (2023)
Responsible Visual Editing
di: Ni, Minheng, et al.
Pubblicazione: (2024)
di: Ni, Minheng, et al.
Pubblicazione: (2024)
Multi‐Task Learning Empowered Anomaly Detection for Internet of Power Systems
di: Xin Li, et al.
Pubblicazione: (2025)
di: Xin Li, et al.
Pubblicazione: (2025)
Empowering Visual Artists with Tokenized Digital Assets with NFTs
di: Li, Ruiqiang, et al.
Pubblicazione: (2024)
di: Li, Ruiqiang, et al.
Pubblicazione: (2024)
Rethinking Transformer-Based Blind-Spot Network for Self-Supervised Image Denoising
di: Li, Junyi, et al.
Pubblicazione: (2024)
di: Li, Junyi, et al.
Pubblicazione: (2024)
Lie Flow: Video Dynamic Fields Modeling and Predicting with Lie Algebra as Geometric Physics Principle
di: Qiao, Weidong, et al.
Pubblicazione: (2026)
di: Qiao, Weidong, et al.
Pubblicazione: (2026)
LMM-PCQA: Assisting Point Cloud Quality Assessment with LMM
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
Q-Adapt: Adapting LMM for Visual Quality Assessment with Progressive Instruction Tuning
di: Lu, Yiting, et al.
Pubblicazione: (2025)
di: Lu, Yiting, et al.
Pubblicazione: (2025)
Towards Token-Level Text Anomaly Detection
di: Cao, Yang, et al.
Pubblicazione: (2026)
di: Cao, Yang, et al.
Pubblicazione: (2026)
On-Device Continual Learning for Unsupervised Visual Anomaly Detection in Dynamic Manufacturing
di: Ren, Haoyu, et al.
Pubblicazione: (2025)
di: Ren, Haoyu, et al.
Pubblicazione: (2025)
An Unsupervised Time Series Anomaly Detection Approach for Efficient Online Process Monitoring of Additive Manufacturing
di: Cantu, Frida, et al.
Pubblicazione: (2025)
di: Cantu, Frida, et al.
Pubblicazione: (2025)
Image-Based Visual Servoing for Enhanced Cooperation of Dual-Arm Manipulation
di: Zhang, Zizhe, et al.
Pubblicazione: (2024)
di: Zhang, Zizhe, et al.
Pubblicazione: (2024)
No Need For Real Anomaly: MLLM Empowered Zero-Shot Video Anomaly Detection
di: Dai, Zunkai, et al.
Pubblicazione: (2026)
di: Dai, Zunkai, et al.
Pubblicazione: (2026)
AnomalyLMM: Bridging Generative Knowledge and Discriminative Retrieval for Text-Based Person Anomaly Search
di: Ju, Hao, et al.
Pubblicazione: (2025)
di: Ju, Hao, et al.
Pubblicazione: (2025)
Expert-Guided Extinction of Toxic Tokens for Debiased Generation
di: Sun, Xueyao, et al.
Pubblicazione: (2024)
di: Sun, Xueyao, et al.
Pubblicazione: (2024)
VisualAD: Language-Free Zero-Shot Anomaly Detection via Vision Transformer
di: Hou, Yanning, et al.
Pubblicazione: (2026)
di: Hou, Yanning, et al.
Pubblicazione: (2026)
Enhanced Generative Structure Prior for Chinese Text Image Super-resolution
di: Li, Xiaoming, et al.
Pubblicazione: (2025)
di: Li, Xiaoming, et al.
Pubblicazione: (2025)
LMM-Det: Make Large Multimodal Models Excel in Object Detection
di: Li, Jincheng, et al.
Pubblicazione: (2025)
di: Li, Jincheng, et al.
Pubblicazione: (2025)
Visual-O1: Understanding Ambiguous Instructions via Multi-modal Multi-turn Chain-of-thoughts Reasoning
di: Ni, Minheng, et al.
Pubblicazione: (2024)
di: Ni, Minheng, et al.
Pubblicazione: (2024)
An LMM for Efficient Video Understanding via Reinforced Compression of Video Cubes
di: Qi, Ji, et al.
Pubblicazione: (2025)
di: Qi, Ji, et al.
Pubblicazione: (2025)
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning
di: Li, Wenrui, et al.
Pubblicazione: (2025)
di: Li, Wenrui, et al.
Pubblicazione: (2025)
FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design
di: Lan, Kai, et al.
Pubblicazione: (2025)
di: Lan, Kai, et al.
Pubblicazione: (2025)
Unveiling Context-Related Anomalies: Knowledge Graph Empowered Decoupling of Scene and Action for Human-Related Video Anomaly Detection
di: Chen, Chenglizhao, et al.
Pubblicazione: (2024)
di: Chen, Chenglizhao, et al.
Pubblicazione: (2024)
Foundation Models for Anomaly Detection: Vision and Challenges
di: Ren, Jing, et al.
Pubblicazione: (2025)
di: Ren, Jing, et al.
Pubblicazione: (2025)
Foundation Models for Anomaly Detection: Vision and Challenges
di: Jing Ren, et al.
Pubblicazione: (2025)
di: Jing Ren, et al.
Pubblicazione: (2025)
KUKAloha: A General, Low-Cost, and Shared-Control based Teleoperation Framework for Construction Robot Arm
di: Xu, Yifan, et al.
Pubblicazione: (2026)
di: Xu, Yifan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Myriad: Large Multimodal Model by Applying Vision Experts for Industrial Anomaly Detection
di: Li, Yuanze, et al.
Pubblicazione: (2023) -
Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation
di: Li, Qizhang, et al.
Pubblicazione: (2024) -
Improved Generation of Adversarial Examples Against Safety-aligned LLMs
di: Li, Qizhang, et al.
Pubblicazione: (2024) -
Empower Vision Applications with LoRA LMM
di: Mi, Liang, et al.
Pubblicazione: (2024) -
Improving Transferability of Adversarial Examples via Bayesian Attacks
di: Li, Qizhang, et al.
Pubblicazione: (2023)