Beyond Token Probes: Hallucination Detection via Activation Tensors with ACT-ViT
Fuente:
arXiv
Saved in:
| Main Authors: | Bar-Shalom, Guy, Frasca, Fabrizio, Galron, Yaniv, Ziser, Yftah, Maron, Haggai |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Neural Message-Passing on Attention Graphs for Hallucination Detection
by: Frasca, Fabrizio, et al.
Published: (2025)
by: Frasca, Fabrizio, et al.
Published: (2025)
Beyond Next Token Probabilities: Learnable, Fast Detection of Hallucinations and Data Contamination on LLM Output Distributions
by: Bar-Shalom, Guy, et al.
Published: (2025)
by: Bar-Shalom, Guy, et al.
Published: (2025)
A Flexible, Equivariant Framework for Subgraph GNNs via Graph Products and Graph Coarsening
by: Bar-Shalom, Guy, et al.
Published: (2024)
by: Bar-Shalom, Guy, et al.
Published: (2024)
Understanding and Improving Laplacian Positional Encodings For Temporal GNNs
by: Galron, Yaniv, et al.
Published: (2025)
by: Galron, Yaniv, et al.
Published: (2025)
FS-KAN: Permutation Equivariant Kolmogorov-Arnold Networks via Function Sharing
by: Elbaz, Ran, et al.
Published: (2025)
by: Elbaz, Ran, et al.
Published: (2025)
Learning from Historical Activations in Graph Neural Networks
by: Galron, Yaniv, et al.
Published: (2026)
by: Galron, Yaniv, et al.
Published: (2026)
Subgraphormer: Unifying Subgraph GNNs and Graph Transformers via Graph Products
by: Bar-Shalom, Guy, et al.
Published: (2024)
by: Bar-Shalom, Guy, et al.
Published: (2024)
On The Expressive Power of GNN Derivatives
by: Eitan, Yam, et al.
Published: (2025)
by: Eitan, Yam, et al.
Published: (2025)
Topological Blindspots: Understanding and Extending Topological Deep Learning Through the Lens of Expressivity
by: Eitan, Yam, et al.
Published: (2024)
by: Eitan, Yam, et al.
Published: (2024)
Balancing Efficiency and Expressiveness: Subgraph GNNs with Walk-Based Centrality
by: Southern, Joshua, et al.
Published: (2025)
by: Southern, Joshua, et al.
Published: (2025)
A Graph Meta-Network for Learning on Kolmogorov-Arnold Networks
by: Bar-Shalom, Guy, et al.
Published: (2026)
by: Bar-Shalom, Guy, et al.
Published: (2026)
Layer by Layer: Uncovering Where Multi-Task Learning Happens in Instruction-Tuned Large Language Models
by: Zhao, Zheng, et al.
Published: (2024)
by: Zhao, Zheng, et al.
Published: (2024)
SafeSteer: Interpretable Safety Steering with Refusal-Evasion in LLMs
by: Ghosh, Shaona, et al.
Published: (2025)
by: Ghosh, Shaona, et al.
Published: (2025)
More Test-Time Compute Can Hurt: Overestimation Bias in LLM Beam Search
by: Dalal, Gal, et al.
Published: (2026)
by: Dalal, Gal, et al.
Published: (2026)
ViT Registers and Fractal ViT
by: Chou, Jason Chuan-Chih, et al.
Published: (2026)
by: Chou, Jason Chuan-Chih, et al.
Published: (2026)
Towards Foundation Models on Graphs: An Analysis on Cross-Dataset Transfer of Pretrained GNNs
by: Frasca, Fabrizio, et al.
Published: (2024)
by: Frasca, Fabrizio, et al.
Published: (2024)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
by: Chattopadhyay, Nandish, et al.
Published: (2026)
by: Chattopadhyay, Nandish, et al.
Published: (2026)
It Takes a Graph to Know a Graph: Rewiring for Homophily with a Reference Graph
by: Mendelman, Harel, et al.
Published: (2025)
by: Mendelman, Harel, et al.
Published: (2025)
On the Expressive Power of Permutation-Equivariant Weight-Space Networks
by: Dayan, Adir, et al.
Published: (2026)
by: Dayan, Adir, et al.
Published: (2026)
Spectral Editing of Activations for Large Language Model Alignment
by: Qiu, Yifu, et al.
Published: (2024)
by: Qiu, Yifu, et al.
Published: (2024)
From Actions to Words: Towards Abstractive-Textual Policy Summarization in RL
by: Admoni, Sahar, et al.
Published: (2025)
by: Admoni, Sahar, et al.
Published: (2025)
On the Reconstruction of Training Data from Group Invariant Networks
by: Elbaz, Ran, et al.
Published: (2024)
by: Elbaz, Ran, et al.
Published: (2024)
Foldable SuperNets: Scalable Merging of Transformers with Different Initializations and Tasks
by: Kinderman, Edan, et al.
Published: (2024)
by: Kinderman, Edan, et al.
Published: (2024)
How to train your ViT for OOD Detection
by: Mueller, Maximilian, et al.
Published: (2024)
by: Mueller, Maximilian, et al.
Published: (2024)
CubistMerge: Spatial-Preserving Token Merging For Diverse ViT Backbones
by: Gong, Wenyi, et al.
Published: (2025)
by: Gong, Wenyi, et al.
Published: (2025)
Token Cropr: Faster ViTs for Quite a Few Tasks
by: Bergner, Benjamin, et al.
Published: (2024)
by: Bergner, Benjamin, et al.
Published: (2024)
ODE-ViT: Plug & Play Attention Layer from the Generalization of the ViT as an Ordinary Differential Equation
by: Riera, Carlos Boned, et al.
Published: (2025)
by: Riera, Carlos Boned, et al.
Published: (2025)
Efficient GNN Training Through Structure-Aware Randomized Mini-Batching
by: Balaji, Vignesh, et al.
Published: (2025)
by: Balaji, Vignesh, et al.
Published: (2025)
On the Expressive Power of Spectral Invariant Graph Neural Networks
by: Zhang, Bohang, et al.
Published: (2024)
by: Zhang, Bohang, et al.
Published: (2024)
Protected Test-Time Adaptation via Online Entropy Matching: A Betting Approach
by: Bar, Yarin, et al.
Published: (2024)
by: Bar, Yarin, et al.
Published: (2024)
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP
by: Balasubramanian, Sriram, et al.
Published: (2024)
by: Balasubramanian, Sriram, et al.
Published: (2024)
TPC-ViT: Token Propagation Controller for Efficient Vision Transformer
by: Zhu, Wentao
Published: (2024)
by: Zhu, Wentao
Published: (2024)
Provable Sparse Inversion and Token Relabel Enhanced One-shot Federated Learning with ViTs
by: Shen, Li, et al.
Published: (2026)
by: Shen, Li, et al.
Published: (2026)
ViTCAE: ViT-based Class-conditioned Autoencoder
by: Jebraeeli, Vahid, et al.
Published: (2025)
by: Jebraeeli, Vahid, et al.
Published: (2025)
Homomorphism Expressivity of Spectral Invariant Graph Neural Networks
by: Gai, Jingchu, et al.
Published: (2025)
by: Gai, Jingchu, et al.
Published: (2025)
Learning on LoRAs: GL-Equivariant Processing of Low-Rank Weight Spaces for Large Finetuned Models
by: Putterman, Theo, et al.
Published: (2024)
by: Putterman, Theo, et al.
Published: (2024)
Training Transformers for KV Cache Compressibility
by: Gelberg, Yoav, et al.
Published: (2026)
by: Gelberg, Yoav, et al.
Published: (2026)
Efficient Subgraph GNNs by Learning Effective Selection Policies
by: Bevilacqua, Beatrice, et al.
Published: (2023)
by: Bevilacqua, Beatrice, et al.
Published: (2023)
GRANOLA: Adaptive Normalization for Graph Neural Networks
by: Eliasof, Moshe, et al.
Published: (2024)
by: Eliasof, Moshe, et al.
Published: (2024)
HydraViT: Stacking Heads for a Scalable ViT
by: Haberer, Janek, et al.
Published: (2024)
by: Haberer, Janek, et al.
Published: (2024)
Similar Items
-
Neural Message-Passing on Attention Graphs for Hallucination Detection
by: Frasca, Fabrizio, et al.
Published: (2025) -
Beyond Next Token Probabilities: Learnable, Fast Detection of Hallucinations and Data Contamination on LLM Output Distributions
by: Bar-Shalom, Guy, et al.
Published: (2025) -
A Flexible, Equivariant Framework for Subgraph GNNs via Graph Products and Graph Coarsening
by: Bar-Shalom, Guy, et al.
Published: (2024) -
Understanding and Improving Laplacian Positional Encodings For Temporal GNNs
by: Galron, Yaniv, et al.
Published: (2025) -
FS-KAN: Permutation Equivariant Kolmogorov-Arnold Networks via Function Sharing
by: Elbaz, Ran, et al.
Published: (2025)