A BERTology View of LLM Orchestrations: Token- and Layer-Selective Probes for Efficient Single-Pass Classification
Fuente:
arXiv
Saved in:
| Main Authors: | Meyoyan, Gonzalo Ariel, Del Corro, Luciano |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Entropy Sentinel: Continuous LLM Accuracy Monitoring from Decoding Entropy Traces in STEM
by: Buffa, Pedro Memoli, et al.
Published: (2026)
by: Buffa, Pedro Memoli, et al.
Published: (2026)
Nested Named Entity Recognition as Single-Pass Sequence Labeling
by: Muñoz-Ortiz, Alberto, et al.
Published: (2025)
by: Muñoz-Ortiz, Alberto, et al.
Published: (2025)
BERTology of Molecular Property Prediction
by: Mostafanejad, Mohammad, et al.
Published: (2026)
by: Mostafanejad, Mohammad, et al.
Published: (2026)
Are Optimal Algorithms Still Optimal? Rethinking Sorting in LLM-Based Pairwise Ranking with Batching and Caching
by: Wisznia, Juan, et al.
Published: (2025)
by: Wisznia, Juan, et al.
Published: (2025)
Just Pass Twice: Efficient Token Classification with LLMs for Zero-Shot NER
by: Ewais, Ahmed, et al.
Published: (2026)
by: Ewais, Ahmed, et al.
Published: (2026)
Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference
by: Taniguchi, Rei, et al.
Published: (2026)
by: Taniguchi, Rei, et al.
Published: (2026)
A fast and sound tagging method for discontinuous named-entity recognition
by: Corro, Caio
Published: (2024)
by: Corro, Caio
Published: (2024)
BERTologyNavigator: Advanced Question Answering with BERT-based Semantics
by: Rajpal, Shreya, et al.
Published: (2024)
by: Rajpal, Shreya, et al.
Published: (2024)
Few-Shot Domain Adaptation for Named-Entity Recognition via Joint Constrained k-Means and Subspace Selection
by: Hammal, Ayoub, et al.
Published: (2024)
by: Hammal, Ayoub, et al.
Published: (2024)
Probe and Skip: Self-Predictive Token Skipping for Efficient Long-Context LLM Inference
by: Wu, Zimeng, et al.
Published: (2026)
by: Wu, Zimeng, et al.
Published: (2026)
Software Mention Recognition with a Three-Stage Framework Based on BERTology Models at SOMD 2024
by: Thi, Thuy Nguyen, et al.
Published: (2024)
by: Thi, Thuy Nguyen, et al.
Published: (2024)
A Study on Building Efficient Zero-Shot Relation Extraction Models
by: Thomas, Hugo, et al.
Published: (2026)
by: Thomas, Hugo, et al.
Published: (2026)
OrthoRank: Token Selection via Sink Token Orthogonality for Efficient LLM inference
by: Shin, Seungjun, et al.
Published: (2025)
by: Shin, Seungjun, et al.
Published: (2025)
OrchestraLLM: Efficient Orchestration of Language Models for Dialogue State Tracking
by: Lee, Chia-Hsuan, et al.
Published: (2023)
by: Lee, Chia-Hsuan, et al.
Published: (2023)
PromptDistill: Query-based Selective Token Retention in Intermediate Layers for Efficient Large Language Model Inference
by: Jin, Weisheng, et al.
Published: (2025)
by: Jin, Weisheng, et al.
Published: (2025)
The Greatest Good Benchmark: Measuring LLMs' Alignment with Utilitarian Moral Dilemmas
by: Marraffini, Giovanni Franco Gabriel, et al.
Published: (2025)
by: Marraffini, Giovanni Franco Gabriel, et al.
Published: (2025)
Single-Pass, Depth-Selective Reading for Multi-Aspect Sentiment Analysis
by: Xia, Yan, et al.
Published: (2026)
by: Xia, Yan, et al.
Published: (2026)
Token Prediction as Implicit Classification to Identify LLM-Generated Text
by: Chen, Yutian, et al.
Published: (2023)
by: Chen, Yutian, et al.
Published: (2023)
Training LayoutLM from Scratch for Efficient Named-Entity Recognition in the Insurance Domain
by: Uthayasooriyar, Benno, et al.
Published: (2024)
by: Uthayasooriyar, Benno, et al.
Published: (2024)
Active Learners as Efficient PRP Rerankers
by: Paschmann, Jeremías Figueiredo, et al.
Published: (2026)
by: Paschmann, Jeremías Figueiredo, et al.
Published: (2026)
DyLLM: Efficient Diffusion LLM Inference via Saliency-based Token Selection and Partial Attention
by: Lee, Younjoo, et al.
Published: (2026)
by: Lee, Younjoo, et al.
Published: (2026)
Efficient Training-Free Multi-Token Prediction via Embedding-Space Probing
by: Goel, Raghavv, et al.
Published: (2026)
by: Goel, Raghavv, et al.
Published: (2026)
Single-Pass Document Scanning for Question Answering
by: Cao, Weili, et al.
Published: (2025)
by: Cao, Weili, et al.
Published: (2025)
Vaporetto: Efficient Japanese Tokenization Based on Improved Pointwise Linear Classification
by: Akabe, Koichi, et al.
Published: (2024)
by: Akabe, Koichi, et al.
Published: (2024)
Kad: A Framework for Proxy-based Test-time Alignment with Knapsack Approximation Deferral
by: Hammal, Ayoub, et al.
Published: (2025)
by: Hammal, Ayoub, et al.
Published: (2025)
LCSB: Layer-Cyclic Selective Backpropagation for Memory-Efficient On-Device LLM Fine-Tuning
by: Park, Juneyoung, et al.
Published: (2026)
by: Park, Juneyoung, et al.
Published: (2026)
Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token Selection
by: Jo, Dongwon, et al.
Published: (2026)
by: Jo, Dongwon, et al.
Published: (2026)
Beyond the Next Token: Towards Prompt-Robust Zero-Shot Classification via Efficient Multi-Token Prediction
by: Qian, Junlang, et al.
Published: (2025)
by: Qian, Junlang, et al.
Published: (2025)
sPhinX: Sample Efficient Multilingual Instruction Fine-Tuning Through N-shot Guided Prompting
by: Ahuja, Sanchit, et al.
Published: (2024)
by: Ahuja, Sanchit, et al.
Published: (2024)
The Evolution of Tool Use in LLM Agents: From Single-Tool Call to Multi-Tool Orchestration
by: Xu, Haoyuan, et al.
Published: (2026)
by: Xu, Haoyuan, et al.
Published: (2026)
SENTRA: Selected-Next-Token Transformer for LLM Text Detection
by: Plyler, Mitchell, et al.
Published: (2025)
by: Plyler, Mitchell, et al.
Published: (2025)
DiffAdapt: Difficulty-Adaptive Reasoning for Token-Efficient LLM Inference
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
Beyond Tokens: Semantic-Aware Speculative Decoding for Efficient Inference by Probing Internal States
by: Dong, Ximing, et al.
Published: (2026)
by: Dong, Ximing, et al.
Published: (2026)
MaxPoolBERT: Enhancing BERT Classification via Layer- and Token-Wise Aggregation
by: Behrendt, Maike, et al.
Published: (2025)
by: Behrendt, Maike, et al.
Published: (2025)
On the Rejection Criterion for Proxy-based Test-time Alignment
by: Hammal, Ayoub, et al.
Published: (2026)
by: Hammal, Ayoub, et al.
Published: (2026)
Suffix-Constrained Greedy Search Algorithms for Causal Language Models
by: Hammal, Ayoub, et al.
Published: (2026)
by: Hammal, Ayoub, et al.
Published: (2026)
Sparse Logistic Regression with High-order Features for Automatic Grammar Rule Extraction from Treebanks
by: Herrera, Santiago, et al.
Published: (2024)
by: Herrera, Santiago, et al.
Published: (2024)
ONTO: A Token-Efficient Columnar Notation for LLM Input Optimization
by: Deekeswar, Harshavardhanan
Published: (2026)
by: Deekeswar, Harshavardhanan
Published: (2026)
Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving
by: Liu, Xunzhuo, et al.
Published: (2026)
by: Liu, Xunzhuo, et al.
Published: (2026)
Memory-Efficient Fine-Tuning of Transformers via Token Selection
by: Simoulin, Antoine, et al.
Published: (2025)
by: Simoulin, Antoine, et al.
Published: (2025)
Similar Items
-
Entropy Sentinel: Continuous LLM Accuracy Monitoring from Decoding Entropy Traces in STEM
by: Buffa, Pedro Memoli, et al.
Published: (2026) -
Nested Named Entity Recognition as Single-Pass Sequence Labeling
by: Muñoz-Ortiz, Alberto, et al.
Published: (2025) -
BERTology of Molecular Property Prediction
by: Mostafanejad, Mohammad, et al.
Published: (2026) -
Are Optimal Algorithms Still Optimal? Rethinking Sorting in LLM-Based Pairwise Ranking with Batching and Caching
by: Wisznia, Juan, et al.
Published: (2025) -
Just Pass Twice: Efficient Token Classification with LLMs for Zero-Shot NER
by: Ewais, Ahmed, et al.
Published: (2026)