Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers
Fuente:
arXiv
Guardado en:
| Autores principales: | Vijayakumar, Soniya, van Genabith, Josef, Ostermann, Simon |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On Multilingual Encoder Language Model Compression for Low-Resource Languages
por: Gurgurov, Daniil, et al.
Publicado: (2025)
por: Gurgurov, Daniil, et al.
Publicado: (2025)
Sparse Subnetwork Enhancement for Underrepresented Languages in Large Language Models
por: Gurgurov, Daniil, et al.
Publicado: (2025)
por: Gurgurov, Daniil, et al.
Publicado: (2025)
Small Models, Big Impact: Efficient Corpus and Graph-Based Adaptation of Small Multilingual Language Models for Low-Resource Languages
por: Gurgurov, Daniil, et al.
Publicado: (2025)
por: Gurgurov, Daniil, et al.
Publicado: (2025)
Disentangling Mathematical Reasoning in LLMs: A Methodological Investigation of Internal Mechanisms
por: Baeumel, Tanja, et al.
Publicado: (2026)
por: Baeumel, Tanja, et al.
Publicado: (2026)
The Lookahead Limitation: Why Multi-Operand Addition is Hard for LLMs
por: Baeumel, Tanja, et al.
Publicado: (2025)
por: Baeumel, Tanja, et al.
Publicado: (2025)
The Latin Substrate: How Language Models Represent and Mediate Script Choice
por: Gurgurov, Daniil, et al.
Publicado: (2026)
por: Gurgurov, Daniil, et al.
Publicado: (2026)
Multilingual Political Views of Large Language Models: Identification and Steering
por: Gurgurov, Daniil, et al.
Publicado: (2025)
por: Gurgurov, Daniil, et al.
Publicado: (2025)
Modular Arithmetic: Language Models Solve Math Digit by Digit
por: Baeumel, Tanja, et al.
Publicado: (2025)
por: Baeumel, Tanja, et al.
Publicado: (2025)
Multilingual Steering by Design: Multilingual Sparse Autoencoders and Principled Layer Selection
por: Ghussin, Yusser Al, et al.
Publicado: (2026)
por: Ghussin, Yusser Al, et al.
Publicado: (2026)
Why Does Reinforcement Learning Generalize? A Feature-Level Mechanistic Study of Post-Training in Large Language Models
por: Shi, Dan, et al.
Publicado: (2026)
por: Shi, Dan, et al.
Publicado: (2026)
Language Arithmetics: Towards Systematic Language Neuron Identification and Manipulation
por: Gurgurov, Daniil, et al.
Publicado: (2025)
por: Gurgurov, Daniil, et al.
Publicado: (2025)
ReasonXL: Shifting LLM Reasoning Language Without Sacrificing Performance
por: Gurgurov, Daniil, et al.
Publicado: (2026)
por: Gurgurov, Daniil, et al.
Publicado: (2026)
AutoPsyC: Automatic Recognition of Psychodynamic Conflicts from Semi-structured Interviews with Large Language Models
por: Hossain, Sayed Muddashir, et al.
Publicado: (2025)
por: Hossain, Sayed Muddashir, et al.
Publicado: (2025)
Reverse Probing: Evaluating Knowledge Transfer via Finetuned Task Embeddings for Coreference Resolution
por: Anikina, Tatiana, et al.
Publicado: (2025)
por: Anikina, Tatiana, et al.
Publicado: (2025)
DFKI-MLT at SemEval-2026 TASK 7: Steering Multilingual Models Towards Cultural Knowledge
por: Ghussin, Yusser Al, et al.
Publicado: (2026)
por: Ghussin, Yusser Al, et al.
Publicado: (2026)
ViConBERT: Context-Gloss Aligned Vietnamese Word Embedding for Polysemous and Sense-Aware Representations
por: Huynh, Khang T., et al.
Publicado: (2025)
por: Huynh, Khang T., et al.
Publicado: (2025)
Sign Language Translation with Sentence Embedding Supervision
por: Hamidullah, Yasser, et al.
Publicado: (2025)
por: Hamidullah, Yasser, et al.
Publicado: (2025)
When Scale Meets Diversity: Evaluating Language Models on Fine-Grained Multilingual Claim Verification
por: Shcharbakova, Hanna, et al.
Publicado: (2025)
por: Shcharbakova, Hanna, et al.
Publicado: (2025)
Rethinking Evaluation of Sparse Autoencoders through the Representation of Polysemous Words
por: Minegishi, Gouki, et al.
Publicado: (2025)
por: Minegishi, Gouki, et al.
Publicado: (2025)
Seeing, Signing, and Saying: A Vision-Language Model-Assisted Pipeline for Sign Language Data Acquisition and Curation from Social Media
por: Yazdani, Shakib, et al.
Publicado: (2025)
por: Yazdani, Shakib, et al.
Publicado: (2025)
Spatio-temporal Sign Language Representation and Translation
por: Hamidullah, Yasser, et al.
Publicado: (2025)
por: Hamidullah, Yasser, et al.
Publicado: (2025)
CLaS-Bench: A Cross-Lingual Alignment and Steering Benchmark
por: Gurgurov, Daniil, et al.
Publicado: (2026)
por: Gurgurov, Daniil, et al.
Publicado: (2026)
From Ghazals to Sonnets: Decoding the Polysemous Expressions of Love Across Languages
por: Ali, Syed Mohammad Sualeh
Publicado: (2025)
por: Ali, Syed Mohammad Sualeh
Publicado: (2025)
SONAR-SLT: Multilingual Sign Language Translation via Language-Agnostic Sentence Embedding Supervision
por: Hamidullah, Yasser, et al.
Publicado: (2025)
por: Hamidullah, Yasser, et al.
Publicado: (2025)
A Critical Study of Automatic Evaluation in Sign Language Translation
por: Yazdani, Shakib, et al.
Publicado: (2025)
por: Yazdani, Shakib, et al.
Publicado: (2025)
When your Cousin has the Right Connections: Unsupervised Bilingual Lexicon Induction for Related Data-Imbalanced Languages
por: Bafna, Niyati, et al.
Publicado: (2023)
por: Bafna, Niyati, et al.
Publicado: (2023)
Probing Language Models for Pre-training Data Detection
por: Liu, Zhenhua, et al.
Publicado: (2024)
por: Liu, Zhenhua, et al.
Publicado: (2024)
When Flores Bloomz Wrong: Cross-Direction Contamination in Machine Translation Evaluation
por: Tan, David, et al.
Publicado: (2026)
por: Tan, David, et al.
Publicado: (2026)
Measuring Spurious Correlation in Classification: 'Clever Hans' in Translationese
por: Borah, Angana, et al.
Publicado: (2023)
por: Borah, Angana, et al.
Publicado: (2023)
Unpacking Ambiguity: The Interaction of Polysemous Discourse Markers and Non-DM Signals
por: Wu, Jingni, et al.
Publicado: (2025)
por: Wu, Jingni, et al.
Publicado: (2025)
LLMCheckup: Conversational Examination of Large Language Models via Interpretability Tools and Self-Explanations
por: Wang, Qianli, et al.
Publicado: (2024)
por: Wang, Qianli, et al.
Publicado: (2024)
Rewiring the Transformer with Depth-Wise LSTMs
por: Xu, Hongfei, et al.
Publicado: (2020)
por: Xu, Hongfei, et al.
Publicado: (2020)
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
por: Kadlčík, Marek, et al.
Publicado: (2025)
por: Kadlčík, Marek, et al.
Publicado: (2025)
Soft Language Prompts for Language Transfer
por: Vykopal, Ivan, et al.
Publicado: (2024)
por: Vykopal, Ivan, et al.
Publicado: (2024)
Efficient Knowledge Probing of Large Language Models by Adapting Pre-trained Embeddings
por: Sharma, Kartik, et al.
Publicado: (2025)
por: Sharma, Kartik, et al.
Publicado: (2025)
Unknown Word Detection for English as a Second Language (ESL) Learners Using Gaze and Pre-trained Language Models
por: Ding, Jiexin, et al.
Publicado: (2025)
por: Ding, Jiexin, et al.
Publicado: (2025)
PETra: A Multilingual Corpus of Pragmatic Explicitation in Translation
por: Osmelak, Doreen, et al.
Publicado: (2025)
por: Osmelak, Doreen, et al.
Publicado: (2025)
Teaching Old Tokenizers New Words: Efficient Tokenizer Adaptation for Pre-trained Models
por: Purason, Taido, et al.
Publicado: (2025)
por: Purason, Taido, et al.
Publicado: (2025)
Generative Large Language Models in Automated Fact-Checking: A Survey
por: Vykopal, Ivan, et al.
Publicado: (2024)
por: Vykopal, Ivan, et al.
Publicado: (2024)
CLOCR-C: Context Leveraging OCR Correction with Pre-trained Language Models
por: Bourne, Jonathan
Publicado: (2024)
por: Bourne, Jonathan
Publicado: (2024)
Ejemplares similares
-
On Multilingual Encoder Language Model Compression for Low-Resource Languages
por: Gurgurov, Daniil, et al.
Publicado: (2025) -
Sparse Subnetwork Enhancement for Underrepresented Languages in Large Language Models
por: Gurgurov, Daniil, et al.
Publicado: (2025) -
Small Models, Big Impact: Efficient Corpus and Graph-Based Adaptation of Small Multilingual Language Models for Low-Resource Languages
por: Gurgurov, Daniil, et al.
Publicado: (2025) -
Disentangling Mathematical Reasoning in LLMs: A Methodological Investigation of Internal Mechanisms
por: Baeumel, Tanja, et al.
Publicado: (2026) -
The Lookahead Limitation: Why Multi-Operand Addition is Hard for LLMs
por: Baeumel, Tanja, et al.
Publicado: (2025)