Out-of-Distribution Detection by Leveraging Between-Layer Transformation Smoothness
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jelenić, Fran, Jukić, Josip, Tutek, Martin, Puljiz, Mate, Šnajder, Jan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Disentangling Latent Shifts of In-Context Learning with Weak Supervision
von: Jukić, Josip, et al.
Veröffentlicht: (2024)
von: Jukić, Josip, et al.
Veröffentlicht: (2024)
Context Parametrization with Compositional Adapters
von: Jukić, Josip, et al.
Veröffentlicht: (2025)
von: Jukić, Josip, et al.
Veröffentlicht: (2025)
From Robustness to Improved Generalization and Calibration in Pre-trained Language Models
von: Jukić, Josip, et al.
Veröffentlicht: (2024)
von: Jukić, Josip, et al.
Veröffentlicht: (2024)
Leveraging Open Information Extraction for More Robust Domain Transfer of Event Trigger Detection
von: Dukić, David, et al.
Veröffentlicht: (2023)
von: Dukić, David, et al.
Veröffentlicht: (2023)
Characterizing Linguistic Shifts in Croatian News via Diachronic Word Embeddings
von: Dukić, David, et al.
Veröffentlicht: (2025)
von: Dukić, David, et al.
Veröffentlicht: (2025)
Improving Data and Parameter Efficiency of Neural Language Models Using Representation Analysis
von: Jukić, Josip
Veröffentlicht: (2025)
von: Jukić, Josip
Veröffentlicht: (2025)
Sequence Repetition Enhances Token Embeddings and Improves Sequence Labeling with Decoder-only Language Models
von: Kukić, Matija Luka, et al.
Veröffentlicht: (2026)
von: Kukić, Matija Luka, et al.
Veröffentlicht: (2026)
Smoothing Out Hallucinations: Mitigating LLM Hallucination with Smoothed Knowledge Distillation
von: Nguyen, Hieu, et al.
Veröffentlicht: (2025)
von: Nguyen, Hieu, et al.
Veröffentlicht: (2025)
Few-Shot Graph Out-of-Distribution Detection with LLMs
von: Xu, Haoyan, et al.
Veröffentlicht: (2025)
von: Xu, Haoyan, et al.
Veröffentlicht: (2025)
GEM: Gaussian Embedding Modeling for Out-of-Distribution Detection in GUI Agents
von: Wu, Zheng, et al.
Veröffentlicht: (2025)
von: Wu, Zheng, et al.
Veröffentlicht: (2025)
Bangla Grammatical Error Detection Leveraging Transformer-based Token Classification
von: Islam, Shayekh Bin, et al.
Veröffentlicht: (2024)
von: Islam, Shayekh Bin, et al.
Veröffentlicht: (2024)
Embedding Trajectory for Out-of-Distribution Detection in Mathematical Reasoning
von: Wang, Yiming, et al.
Veröffentlicht: (2024)
von: Wang, Yiming, et al.
Veröffentlicht: (2024)
Out-of-Distribution Detection using Synthetic Data Generation
von: Abbas, Momin, et al.
Veröffentlicht: (2025)
von: Abbas, Momin, et al.
Veröffentlicht: (2025)
Claim Check-Worthiness Detection: How Well do LLMs Grasp Annotation Guidelines?
von: Majer, Laura, et al.
Veröffentlicht: (2024)
von: Majer, Laura, et al.
Veröffentlicht: (2024)
CATfOOD: Counterfactual Augmented Training for Improving Out-of-Domain Performance and Calibration
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2023)
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2023)
'No' Matters: Out-of-Distribution Detection in Multimodality Long Dialogue
von: Gao, Rena, et al.
Veröffentlicht: (2024)
von: Gao, Rena, et al.
Veröffentlicht: (2024)
Generalization or Hallucination? Understanding Out-of-Context Reasoning in Transformers
von: Huang, Yixiao, et al.
Veröffentlicht: (2025)
von: Huang, Yixiao, et al.
Veröffentlicht: (2025)
Generalizing Reward Modeling for Out-of-Distribution Preference Learning
von: Jia, Chen
Veröffentlicht: (2024)
von: Jia, Chen
Veröffentlicht: (2024)
AMPLIFY:Attention-based Mixup for Performance Improvement and Label Smoothing in Transformer
von: Yang, Leixin, et al.
Veröffentlicht: (2023)
von: Yang, Leixin, et al.
Veröffentlicht: (2023)
From the Inside Out: Progressive Distribution Refinement for Confidence Calibration
von: Yang, Xizhong, et al.
Veröffentlicht: (2026)
von: Yang, Xizhong, et al.
Veröffentlicht: (2026)
Learning to Skip the Middle Layers of Transformers
von: Lawson, Tim, et al.
Veröffentlicht: (2025)
von: Lawson, Tim, et al.
Veröffentlicht: (2025)
Leveraging Language Models to Detect Greenwashing
von: Vinella, Avalon, et al.
Veröffentlicht: (2023)
von: Vinella, Avalon, et al.
Veröffentlicht: (2023)
Detecting and Understanding Vulnerabilities in Language Models via Mechanistic Interpretability
von: García-Carrasco, Jorge, et al.
Veröffentlicht: (2024)
von: García-Carrasco, Jorge, et al.
Veröffentlicht: (2024)
Human Texts Are Outliers: Detecting LLM-generated Texts via Out-of-distribution Detection
von: Zeng, Cong, et al.
Veröffentlicht: (2025)
von: Zeng, Cong, et al.
Veröffentlicht: (2025)
LayerNorm Induces Recency Bias in Transformer Decoders
von: Kim, Junu, et al.
Veröffentlicht: (2025)
von: Kim, Junu, et al.
Veröffentlicht: (2025)
Provable Knowledge Acquisition and Extraction in One-Layer Transformers
von: Xu, Ruichen, et al.
Veröffentlicht: (2025)
von: Xu, Ruichen, et al.
Veröffentlicht: (2025)
Word Sense Detection Leveraging Maximum Mean Discrepancy
von: Mitsuzawa, Kensuke
Veröffentlicht: (2025)
von: Mitsuzawa, Kensuke
Veröffentlicht: (2025)
Text Meets Topology: Rethinking Out-of-distribution Detection in Text-Rich Networks
von: Wang, Danny, et al.
Veröffentlicht: (2025)
von: Wang, Danny, et al.
Veröffentlicht: (2025)
Intent Recognition and Out-of-Scope Detection using LLMs in Multi-party Conversations
von: Castillo-López, Galo, et al.
Veröffentlicht: (2025)
von: Castillo-López, Galo, et al.
Veröffentlicht: (2025)
Mechanism and Emergence of Stacked Attention Heads in Multi-Layer Transformers
von: Musat, Tiberiu
Veröffentlicht: (2024)
von: Musat, Tiberiu
Veröffentlicht: (2024)
Leveraging Graph Structures to Detect Hallucinations in Large Language Models
von: Nonkes, Noa, et al.
Veröffentlicht: (2024)
von: Nonkes, Noa, et al.
Veröffentlicht: (2024)
MMD-Flagger: Leveraging Maximum Mean Discrepancy to Detect Hallucinations
von: Mitsuzawa, Kensuke, et al.
Veröffentlicht: (2025)
von: Mitsuzawa, Kensuke, et al.
Veröffentlicht: (2025)
Leveraging Big Data Frameworks for Spam Detection in Amazon Reviews
von: Khatun, Mst Eshita, et al.
Veröffentlicht: (2025)
von: Khatun, Mst Eshita, et al.
Veröffentlicht: (2025)
LayerBoost: Layer-Aware Attention Reduction for Efficient LLMs
von: Souibgui, Mohamed Ali, et al.
Veröffentlicht: (2026)
von: Souibgui, Mohamed Ali, et al.
Veröffentlicht: (2026)
Leveraging the true depth of LLMs
von: González, Ramón Calvo, et al.
Veröffentlicht: (2025)
von: González, Ramón Calvo, et al.
Veröffentlicht: (2025)
Looking Right is Sometimes Right: Investigating the Capabilities of Decoder-only LLMs for Sequence Labeling
von: Dukić, David, et al.
Veröffentlicht: (2024)
von: Dukić, David, et al.
Veröffentlicht: (2024)
Out-of-Distribution Detection with Attention Head Masking for Multimodal Document Classification
von: Constantinou, Christos, et al.
Veröffentlicht: (2024)
von: Constantinou, Christos, et al.
Veröffentlicht: (2024)
Interpretable Detection of Out-of-Context Misinformation with Neural-Symbolic-Enhanced Large Multimodal Model
von: Zhang, Yizhou, et al.
Veröffentlicht: (2023)
von: Zhang, Yizhou, et al.
Veröffentlicht: (2023)
PolarQuant: Leveraging Polar Transformation for Efficient Key Cache Quantization and Decoding Acceleration
von: Wu, Songhao, et al.
Veröffentlicht: (2025)
von: Wu, Songhao, et al.
Veröffentlicht: (2025)
Reducing Transformer Key-Value Cache Size with Cross-Layer Attention
von: Brandon, William, et al.
Veröffentlicht: (2024)
von: Brandon, William, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Disentangling Latent Shifts of In-Context Learning with Weak Supervision
von: Jukić, Josip, et al.
Veröffentlicht: (2024) -
Context Parametrization with Compositional Adapters
von: Jukić, Josip, et al.
Veröffentlicht: (2025) -
From Robustness to Improved Generalization and Calibration in Pre-trained Language Models
von: Jukić, Josip, et al.
Veröffentlicht: (2024) -
Leveraging Open Information Extraction for More Robust Domain Transfer of Event Trigger Detection
von: Dukić, David, et al.
Veröffentlicht: (2023) -
Characterizing Linguistic Shifts in Croatian News via Diachronic Word Embeddings
von: Dukić, David, et al.
Veröffentlicht: (2025)