Targeted Visualization of the Backbone of Encoder LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Roberts, Isaac, Schulz, Alexander, Hermes, Luca, Hammer, Barbara |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Intelligent Learning Rate Distribution to reduce Catastrophic Forgetting in Transformers
by: Kenneweg, Philip, et al.
Published: (2024)
by: Kenneweg, Philip, et al.
Published: (2024)
One-Class Intrusion Detection with Dynamic Graphs
by: Liuliakov, Aleksei, et al.
Published: (2025)
by: Liuliakov, Aleksei, et al.
Published: (2025)
DGPO: RL-Steered Graph Diffusion for Neural Architecture Generation
by: Liuliakov, Aleksei, et al.
Published: (2026)
by: Liuliakov, Aleksei, et al.
Published: (2026)
Conceptualizing Uncertainty: A Concept-based Approach to Explaining Uncertainty
by: Roberts, Isaac, et al.
Published: (2025)
by: Roberts, Isaac, et al.
Published: (2025)
Physics-Informed Graph Neural Networks for Water Distribution Systems
by: Ashraf, Inaam, et al.
Published: (2024)
by: Ashraf, Inaam, et al.
Published: (2024)
Attribution-Guided Pruning for Insight and Control: Circuit Discovery and Targeted Correction in Small-scale LLMs
by: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Published: (2025)
by: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Published: (2025)
Using LLMs to Model the Beliefs and Preferences of Targeted Populations
by: Namikoshi, Keiichi, et al.
Published: (2024)
by: Namikoshi, Keiichi, et al.
Published: (2024)
Language Models as Hierarchy Encoders
by: He, Yuan, et al.
Published: (2024)
by: He, Yuan, et al.
Published: (2024)
Can Visual Encoder Learn to See Arrows?
by: Terashita, Naoyuki, et al.
Published: (2025)
by: Terashita, Naoyuki, et al.
Published: (2025)
Large Language Models Are Overparameterized Text Encoders
by: K, Thennal D, et al.
Published: (2024)
by: K, Thennal D, et al.
Published: (2024)
ClusterUCB: Efficient Gradient-Based Data Selection for Targeted Fine-Tuning of LLMs
by: Wang, Zige, et al.
Published: (2025)
by: Wang, Zige, et al.
Published: (2025)
ContextGPT: Infusing LLMs Knowledge into Neuro-Symbolic Activity Recognition Models
by: Arrotta, Luca, et al.
Published: (2024)
by: Arrotta, Luca, et al.
Published: (2024)
RouteLLM: Learning to Route LLMs with Preference Data
by: Ong, Isaac, et al.
Published: (2024)
by: Ong, Isaac, et al.
Published: (2024)
Frozen Transformers in Language Models Are Effective Visual Encoder Layers
by: Pang, Ziqi, et al.
Published: (2023)
by: Pang, Ziqi, et al.
Published: (2023)
Languages are Modalities: Cross-Lingual Alignment via Encoder Injection
by: Agarwal, Rajan, et al.
Published: (2025)
by: Agarwal, Rajan, et al.
Published: (2025)
Large Language Models are Powerful Electronic Health Record Encoders
by: Hegselmann, Stefan, et al.
Published: (2025)
by: Hegselmann, Stefan, et al.
Published: (2025)
A Fuzzy Evaluation of Sentence Encoders on Grooming Risk Classification
by: Bihani, Geetanjali, et al.
Published: (2025)
by: Bihani, Geetanjali, et al.
Published: (2025)
Emergence and Effectiveness of Task Vectors in In-Context Learning: An Encoder Decoder Perspective
by: Han, Seungwook, et al.
Published: (2024)
by: Han, Seungwook, et al.
Published: (2024)
RAVEN: In-Context Learning with Retrieval-Augmented Encoder-Decoder Language Models
by: Huang, Jie, et al.
Published: (2023)
by: Huang, Jie, et al.
Published: (2023)
Dual Encoder: Exploiting the Potential of Syntactic and Semantic for Aspect Sentiment Triplet Extraction
by: Zhao, Xiaowei, et al.
Published: (2024)
by: Zhao, Xiaowei, et al.
Published: (2024)
TURNA: A Turkish Encoder-Decoder Language Model for Enhanced Understanding and Generation
by: Uludoğan, Gökçe, et al.
Published: (2024)
by: Uludoğan, Gökçe, et al.
Published: (2024)
MrBERT: Modern Multilingual Encoders via Vocabulary, Domain, and Dimensional Adaptation
by: Tamayo, Daniel, et al.
Published: (2026)
by: Tamayo, Daniel, et al.
Published: (2026)
1-800-SHARED-TASKS @ NLU of Devanagari Script Languages: Detection of Language, Hate Speech, and Targets using LLMs
by: Purbey, Jebish, et al.
Published: (2024)
by: Purbey, Jebish, et al.
Published: (2024)
SelfReflect: Can LLMs Communicate Their Internal Answer Distribution?
by: Kirchhof, Michael, et al.
Published: (2025)
by: Kirchhof, Michael, et al.
Published: (2025)
Evolutionary Strategies lead to Catastrophic Forgetting in LLMs
by: Abdi, Immanuel, et al.
Published: (2026)
by: Abdi, Immanuel, et al.
Published: (2026)
BTZSC: A Benchmark for Zero-Shot Text Classification Across Cross-Encoders, Embedding Models, Rerankers and LLMs
by: Aarab, Ilias
Published: (2026)
by: Aarab, Ilias
Published: (2026)
NuNER: Entity Recognition Encoder Pre-training via LLM-Annotated Data
by: Bogdanov, Sergei, et al.
Published: (2024)
by: Bogdanov, Sergei, et al.
Published: (2024)
Persona-Coded Poly-Encoder: Persona-Guided Multi-Stream Conversational Sentence Scoring
by: Liu, Junfeng, et al.
Published: (2023)
by: Liu, Junfeng, et al.
Published: (2023)
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
by: Nikitin, Alexander, et al.
Published: (2024)
by: Nikitin, Alexander, et al.
Published: (2024)
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
by: Wang, Yinong Oliver, et al.
Published: (2025)
by: Wang, Yinong Oliver, et al.
Published: (2025)
Evaluating the Effectiveness of XAI Techniques for Encoder-Based Language Models
by: Mersha, Melkamu Abay, et al.
Published: (2025)
by: Mersha, Melkamu Abay, et al.
Published: (2025)
New Encoders for German Trained from Scratch: Comparing ModernGBERT with Converted LLM2Vec Models
by: Wunderle, Julia, et al.
Published: (2025)
by: Wunderle, Julia, et al.
Published: (2025)
Steering Llama 2 via Contrastive Activation Addition
by: Panickssery, Nina, et al.
Published: (2023)
by: Panickssery, Nina, et al.
Published: (2023)
Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs
by: Laine, Rudolf, et al.
Published: (2024)
by: Laine, Rudolf, et al.
Published: (2024)
Explaining Similarity in Vision-Language Encoders with Weighted Banzhaf Interactions
by: Baniecki, Hubert, et al.
Published: (2025)
by: Baniecki, Hubert, et al.
Published: (2025)
GATech at AbjadMed: Bidirectional Encoders vs. Causal Decoders: Insights from 82-Class Arabic Medical Classification
by: Khamis, Ahmed Khaled
Published: (2026)
by: Khamis, Ahmed Khaled
Published: (2026)
Assessing Episodic Memory in LLMs with Sequence Order Recall Tasks
by: Pink, Mathis, et al.
Published: (2024)
by: Pink, Mathis, et al.
Published: (2024)
Matryoshka Pilot: Learning to Drive Black-Box LLMs with LLMs
by: Li, Changhao, et al.
Published: (2024)
by: Li, Changhao, et al.
Published: (2024)
Can LLMs Help Uncover Insights about LLMs? A Large-Scale, Evolving Literature Analysis of Frontier LLMs
by: Park, Jungsoo, et al.
Published: (2025)
by: Park, Jungsoo, et al.
Published: (2025)
NewsInterview: a Dataset and a Playground to Evaluate LLMs' Ground Gap via Informational Interviews
by: Spangher, Alexander, et al.
Published: (2024)
by: Spangher, Alexander, et al.
Published: (2024)
Similar Items
-
Intelligent Learning Rate Distribution to reduce Catastrophic Forgetting in Transformers
by: Kenneweg, Philip, et al.
Published: (2024) -
One-Class Intrusion Detection with Dynamic Graphs
by: Liuliakov, Aleksei, et al.
Published: (2025) -
DGPO: RL-Steered Graph Diffusion for Neural Architecture Generation
by: Liuliakov, Aleksei, et al.
Published: (2026) -
Conceptualizing Uncertainty: A Concept-based Approach to Explaining Uncertainty
by: Roberts, Isaac, et al.
Published: (2025) -
Physics-Informed Graph Neural Networks for Water Distribution Systems
by: Ashraf, Inaam, et al.
Published: (2024)