TTL: Test-time Textual Learning for OOD Detection with Pretrained Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ye, Jinlun, Liao, Jiang, Lai, Runhe, Lu, Xinhua, Zhuang, Jiaxin, Gan, Zhiyong, Wang, Ruixuan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hierarchical Vision-Language Learning for Medical Out-of-Distribution Detection
by: Lai, Runhe, et al.
Published: (2025)
by: Lai, Runhe, et al.
Published: (2025)
FA: Forced Prompt Learning of Vision-Language Models for Out-of-Distribution Detection
by: Lu, Xinhua, et al.
Published: (2025)
by: Lu, Xinhua, et al.
Published: (2025)
Instruction Lens Score: Your Instruction Contributes a Powerful Object Hallucination Detector for Multimodal Large Language Models
by: Lai, Runhe, et al.
Published: (2026)
by: Lai, Runhe, et al.
Published: (2026)
VI-OOD: A Unified Representation Learning Framework for Textual Out-of-distribution Detection
by: Zhan, Li-Ming, et al.
Published: (2024)
by: Zhan, Li-Ming, et al.
Published: (2024)
Unsupervised Layer-wise Score Aggregation for Textual OOD Detection
by: Darrin, Maxime, et al.
Published: (2023)
by: Darrin, Maxime, et al.
Published: (2023)
DynProto: Dynamic Prototype Evolution for Out-of-Distribution Detection
by: Wu, Yanqi, et al.
Published: (2026)
by: Wu, Yanqi, et al.
Published: (2026)
Vision-and-Language Pretraining
by: Nguyen, Thong, et al.
Published: (2022)
by: Nguyen, Thong, et al.
Published: (2022)
Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings
by: Agrawal, Aakriti, et al.
Published: (2025)
by: Agrawal, Aakriti, et al.
Published: (2025)
Improving Low-Resource Sequence Labeling with Knowledge Fusion and Contextual Label Explanations
by: Lai, Peichao, et al.
Published: (2025)
by: Lai, Peichao, et al.
Published: (2025)
Textually Pretrained Speech Language Models
by: Hassid, Michael, et al.
Published: (2023)
by: Hassid, Michael, et al.
Published: (2023)
GITA: Graph to Visual and Textual Integration for Vision-Language Graph Reasoning
by: Wei, Yanbin, et al.
Published: (2024)
by: Wei, Yanbin, et al.
Published: (2024)
Negative Label Guided OOD Detection with Pretrained Vision-Language Models
by: Jiang, Xue, et al.
Published: (2024)
by: Jiang, Xue, et al.
Published: (2024)
2.5 Years in Class: A Multimodal Textbook for Vision-Language Pretraining
by: Zhang, Wenqi, et al.
Published: (2025)
by: Zhang, Wenqi, et al.
Published: (2025)
DCAC: Dynamic Class-Aware Cache Creates Stronger Out-of-Distribution Detectors
by: Wu, Yanqi, et al.
Published: (2026)
by: Wu, Yanqi, et al.
Published: (2026)
Strong Teacher Not Needed? On Distillation in LLM Pretraining
by: Lu, Taiming, et al.
Published: (2026)
by: Lu, Taiming, et al.
Published: (2026)
AgriGPT-VL: Agricultural Vision-Language Understanding Suite
by: Yang, Bo, et al.
Published: (2025)
by: Yang, Bo, et al.
Published: (2025)
Conceptual Contrastive Edits in Textual and Vision-Language Retrieval
by: Lymperaiou, Maria, et al.
Published: (2025)
by: Lymperaiou, Maria, et al.
Published: (2025)
SimLabel: Consistency-Guided OOD Detection with Pretrained Vision-Language Models
by: Zou, Shu, et al.
Published: (2025)
by: Zou, Shu, et al.
Published: (2025)
Node Level Graph Autoencoder: Unified Pretraining for Textual Graph Learning
by: Hu, Wenbin, et al.
Published: (2024)
by: Hu, Wenbin, et al.
Published: (2024)
Test-Time Alignment for Large Language Models via Textual Model Predictive Control
by: Wang, Kuang-Da, et al.
Published: (2025)
by: Wang, Kuang-Da, et al.
Published: (2025)
APP: Adaptive Prototypical Pseudo-Labeling for Few-shot OOD Detection
by: Wang, Pei, et al.
Published: (2023)
by: Wang, Pei, et al.
Published: (2023)
Think Visually, Reason Textually: Vision-Language Synergy in ARC
by: Zhang, Beichen, et al.
Published: (2025)
by: Zhang, Beichen, et al.
Published: (2025)
LangSAMP: Language-Script Aware Multilingual Pretraining
by: Liu, Yihong, et al.
Published: (2024)
by: Liu, Yihong, et al.
Published: (2024)
TransliCo: A Contrastive Learning Framework to Address the Script Barrier in Multilingual Pretrained Language Models
by: Liu, Yihong, et al.
Published: (2024)
by: Liu, Yihong, et al.
Published: (2024)
Vision-and-Language Navigation Generative Pretrained Transformer
by: Hanlin, Wen
Published: (2024)
by: Hanlin, Wen
Published: (2024)
Polysemantic Dropout: Conformal OOD Detection for Specialized LLMs
by: Gupta, Ayush, et al.
Published: (2025)
by: Gupta, Ayush, et al.
Published: (2025)
Language of Persuasion and Misrepresentation in Business Communication: A Textual Detection Approach
by: Hossen, Sayem, et al.
Published: (2025)
by: Hossen, Sayem, et al.
Published: (2025)
Local Background Features Matter in Out-of-Distribution Detection
by: Ye, Jinlun, et al.
Published: (2025)
by: Ye, Jinlun, et al.
Published: (2025)
Adam's Law: Textual Frequency Law on Large Language Models
by: Lu, Hongyuan Adam, et al.
Published: (2026)
by: Lu, Hongyuan Adam, et al.
Published: (2026)
Progressive Residual Warmup for Language Model Pretraining
by: Chen, Tianhao, et al.
Published: (2026)
by: Chen, Tianhao, et al.
Published: (2026)
Renaissance: Investigating the Pretraining of Vision-Language Encoders
by: Fields, Clayton, et al.
Published: (2024)
by: Fields, Clayton, et al.
Published: (2024)
Automatic Textual Normalization for Hate Speech Detection
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2023)
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2023)
Probabilistic Textual Time Series Depression Detection
by: Schmidt, Fabian, et al.
Published: (2025)
by: Schmidt, Fabian, et al.
Published: (2025)
Integrating Multi-view Analysis: Multi-view Mixture-of-Expert for Textual Personality Detection
by: Zhu, Haohao, et al.
Published: (2024)
by: Zhu, Haohao, et al.
Published: (2024)
FrEVL: Leveraging Frozen Pretrained Embeddings for Efficient Vision-Language Understanding
by: Bourigault, Emmanuelle, et al.
Published: (2025)
by: Bourigault, Emmanuelle, et al.
Published: (2025)
PhenoLIP: Integrating Phenotype Ontology Knowledge into Medical Vision-Language Pretraining
by: Liang, Cheng, et al.
Published: (2026)
by: Liang, Cheng, et al.
Published: (2026)
The Role of Mixed-Language Documents for Multilingual Large Language Model Pretraining
by: Shao, Jiandong, et al.
Published: (2026)
by: Shao, Jiandong, et al.
Published: (2026)
Language Model Distillation: A Temporal Difference Imitation Learning Perspective
by: Yu, Zishun, et al.
Published: (2025)
by: Yu, Zishun, et al.
Published: (2025)
Learning to Read Where to Look: Disease-Aware Vision-Language Pretraining for 3D CT
by: Ging, Simon, et al.
Published: (2026)
by: Ging, Simon, et al.
Published: (2026)
Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models
by: Xiong, Guangzhi, et al.
Published: (2026)
by: Xiong, Guangzhi, et al.
Published: (2026)
Similar Items
-
Hierarchical Vision-Language Learning for Medical Out-of-Distribution Detection
by: Lai, Runhe, et al.
Published: (2025) -
FA: Forced Prompt Learning of Vision-Language Models for Out-of-Distribution Detection
by: Lu, Xinhua, et al.
Published: (2025) -
Instruction Lens Score: Your Instruction Contributes a Powerful Object Hallucination Detector for Multimodal Large Language Models
by: Lai, Runhe, et al.
Published: (2026) -
VI-OOD: A Unified Representation Learning Framework for Textual Out-of-distribution Detection
by: Zhan, Li-Ming, et al.
Published: (2024) -
Unsupervised Layer-wise Score Aggregation for Textual OOD Detection
by: Darrin, Maxime, et al.
Published: (2023)