Triad: Empowering LMM-based Anomaly Detection with Vision Expert-guided Visual Tokenizer and Manufacturing Process
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yuanze, Yuan, Shihao, Wang, Haolin, Li, Qizhang, Liu, Ming, Xu, Chen, Shi, Guangming, Zuo, Wangmeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Myriad: Large Multimodal Model by Applying Vision Experts for Industrial Anomaly Detection
by: Li, Yuanze, et al.
Published: (2023)
by: Li, Yuanze, et al.
Published: (2023)
Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation
by: Li, Qizhang, et al.
Published: (2024)
by: Li, Qizhang, et al.
Published: (2024)
Improved Generation of Adversarial Examples Against Safety-aligned LLMs
by: Li, Qizhang, et al.
Published: (2024)
by: Li, Qizhang, et al.
Published: (2024)
Empower Vision Applications with LoRA LMM
by: Mi, Liang, et al.
Published: (2024)
by: Mi, Liang, et al.
Published: (2024)
Improving Transferability of Adversarial Examples via Bayesian Attacks
by: Li, Qizhang, et al.
Published: (2023)
by: Li, Qizhang, et al.
Published: (2023)
Grounding-MD: Grounded Video-language Pre-training for Open-World Moment Detection
by: Zhuang, Weijun, et al.
Published: (2025)
by: Zhuang, Weijun, et al.
Published: (2025)
GLAD: Towards Better Reconstruction with Global and Local Adaptive Diffusion Models for Unsupervised Anomaly Detection
by: Yao, Hang, et al.
Published: (2024)
by: Yao, Hang, et al.
Published: (2024)
GLaVE-Cap: Global-Local Aligned Video Captioning with Vision Expert Integration
by: Xu, Wan, et al.
Published: (2025)
by: Xu, Wan, et al.
Published: (2025)
Unprejudiced Training Auxiliary Tasks Makes Primary Better: A Multi-Task Learning Perspective
by: Li, Yuanze, et al.
Published: (2024)
by: Li, Yuanze, et al.
Published: (2024)
LPT++: Efficient Training on Mixture of Long-tailed Experts
by: Dong, Bowen, et al.
Published: (2024)
by: Dong, Bowen, et al.
Published: (2024)
Retrieval Augmented Image Harmonization
by: Wang, Haolin, et al.
Published: (2024)
by: Wang, Haolin, et al.
Published: (2024)
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
by: Dong, Bowen, et al.
Published: (2024)
by: Dong, Bowen, et al.
Published: (2024)
Cross-Sensor RGB Spectrograms: A Visual Method for Anomaly Detection in Classical and Quantum Magnetometer Triads
by: Pandey, Manas
Published: (2026)
by: Pandey, Manas
Published: (2026)
Mixture of Nested Experts: Adaptive Processing of Visual Tokens
by: Jain, Gagan, et al.
Published: (2024)
by: Jain, Gagan, et al.
Published: (2024)
Pseudo Replay-based Class Continual Learning for Online New Category Anomaly Detection in Advanced Manufacturing
by: Li, Yuxuan, et al.
Published: (2023)
by: Li, Yuxuan, et al.
Published: (2023)
Responsible Visual Editing
by: Ni, Minheng, et al.
Published: (2024)
by: Ni, Minheng, et al.
Published: (2024)
Multi‐Task Learning Empowered Anomaly Detection for Internet of Power Systems
by: Xin Li, et al.
Published: (2025)
by: Xin Li, et al.
Published: (2025)
Empowering Visual Artists with Tokenized Digital Assets with NFTs
by: Li, Ruiqiang, et al.
Published: (2024)
by: Li, Ruiqiang, et al.
Published: (2024)
Rethinking Transformer-Based Blind-Spot Network for Self-Supervised Image Denoising
by: Li, Junyi, et al.
Published: (2024)
by: Li, Junyi, et al.
Published: (2024)
Lie Flow: Video Dynamic Fields Modeling and Predicting with Lie Algebra as Geometric Physics Principle
by: Qiao, Weidong, et al.
Published: (2026)
by: Qiao, Weidong, et al.
Published: (2026)
LMM-PCQA: Assisting Point Cloud Quality Assessment with LMM
by: Zhang, Zicheng, et al.
Published: (2024)
by: Zhang, Zicheng, et al.
Published: (2024)
Q-Adapt: Adapting LMM for Visual Quality Assessment with Progressive Instruction Tuning
by: Lu, Yiting, et al.
Published: (2025)
by: Lu, Yiting, et al.
Published: (2025)
Towards Token-Level Text Anomaly Detection
by: Cao, Yang, et al.
Published: (2026)
by: Cao, Yang, et al.
Published: (2026)
On-Device Continual Learning for Unsupervised Visual Anomaly Detection in Dynamic Manufacturing
by: Ren, Haoyu, et al.
Published: (2025)
by: Ren, Haoyu, et al.
Published: (2025)
An Unsupervised Time Series Anomaly Detection Approach for Efficient Online Process Monitoring of Additive Manufacturing
by: Cantu, Frida, et al.
Published: (2025)
by: Cantu, Frida, et al.
Published: (2025)
Image-Based Visual Servoing for Enhanced Cooperation of Dual-Arm Manipulation
by: Zhang, Zizhe, et al.
Published: (2024)
by: Zhang, Zizhe, et al.
Published: (2024)
No Need For Real Anomaly: MLLM Empowered Zero-Shot Video Anomaly Detection
by: Dai, Zunkai, et al.
Published: (2026)
by: Dai, Zunkai, et al.
Published: (2026)
AnomalyLMM: Bridging Generative Knowledge and Discriminative Retrieval for Text-Based Person Anomaly Search
by: Ju, Hao, et al.
Published: (2025)
by: Ju, Hao, et al.
Published: (2025)
Expert-Guided Extinction of Toxic Tokens for Debiased Generation
by: Sun, Xueyao, et al.
Published: (2024)
by: Sun, Xueyao, et al.
Published: (2024)
VisualAD: Language-Free Zero-Shot Anomaly Detection via Vision Transformer
by: Hou, Yanning, et al.
Published: (2026)
by: Hou, Yanning, et al.
Published: (2026)
Enhanced Generative Structure Prior for Chinese Text Image Super-resolution
by: Li, Xiaoming, et al.
Published: (2025)
by: Li, Xiaoming, et al.
Published: (2025)
LMM-Det: Make Large Multimodal Models Excel in Object Detection
by: Li, Jincheng, et al.
Published: (2025)
by: Li, Jincheng, et al.
Published: (2025)
Visual-O1: Understanding Ambiguous Instructions via Multi-modal Multi-turn Chain-of-thoughts Reasoning
by: Ni, Minheng, et al.
Published: (2024)
by: Ni, Minheng, et al.
Published: (2024)
An LMM for Efficient Video Understanding via Reinforced Compression of Video Cubes
by: Qi, Ji, et al.
Published: (2025)
by: Qi, Ji, et al.
Published: (2025)
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning
by: Li, Wenrui, et al.
Published: (2025)
by: Li, Wenrui, et al.
Published: (2025)
FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design
by: Lan, Kai, et al.
Published: (2025)
by: Lan, Kai, et al.
Published: (2025)
Unveiling Context-Related Anomalies: Knowledge Graph Empowered Decoupling of Scene and Action for Human-Related Video Anomaly Detection
by: Chen, Chenglizhao, et al.
Published: (2024)
by: Chen, Chenglizhao, et al.
Published: (2024)
Foundation Models for Anomaly Detection: Vision and Challenges
by: Ren, Jing, et al.
Published: (2025)
by: Ren, Jing, et al.
Published: (2025)
Foundation Models for Anomaly Detection: Vision and Challenges
by: Jing Ren, et al.
Published: (2025)
by: Jing Ren, et al.
Published: (2025)
KUKAloha: A General, Low-Cost, and Shared-Control based Teleoperation Framework for Construction Robot Arm
by: Xu, Yifan, et al.
Published: (2026)
by: Xu, Yifan, et al.
Published: (2026)
Similar Items
-
Myriad: Large Multimodal Model by Applying Vision Experts for Industrial Anomaly Detection
by: Li, Yuanze, et al.
Published: (2023) -
Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation
by: Li, Qizhang, et al.
Published: (2024) -
Improved Generation of Adversarial Examples Against Safety-aligned LLMs
by: Li, Qizhang, et al.
Published: (2024) -
Empower Vision Applications with LoRA LMM
by: Mi, Liang, et al.
Published: (2024) -
Improving Transferability of Adversarial Examples via Bayesian Attacks
by: Li, Qizhang, et al.
Published: (2023)