Through the LLM Looking Glass: A Socratic Probing of Donkeys, Elephants, and Markets
Fuente:
arXiv
Guardado en:
| Autores principales: | Kennedy, Molly, Imani, Ayyoob, Spinde, Timo, Aizawa, Akiko, Schütze, Hinrich |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
RET-LLM: Towards a General Read-Write Memory for Large Language Models
por: Modarressi, Ali, et al.
Publicado: (2023)
por: Modarressi, Ali, et al.
Publicado: (2023)
MemLLM: Finetuning LLMs to Use An Explicit Read-Write Memory
por: Modarressi, Ali, et al.
Publicado: (2024)
por: Modarressi, Ali, et al.
Publicado: (2024)
GlotLID: Language Identification for Low-Resource Languages
por: Kargaran, Amir Hossein, et al.
Publicado: (2023)
por: Kargaran, Amir Hossein, et al.
Publicado: (2023)
Left, Right, or Center? Evaluating LLM Framing in News Classification and Generation
por: Kennedy, Molly, et al.
Publicado: (2026)
por: Kennedy, Molly, et al.
Publicado: (2026)
Taxi1500: A Multilingual Dataset for Text Classification in 1500 Languages
por: Ma, Chunlan, et al.
Publicado: (2023)
por: Ma, Chunlan, et al.
Publicado: (2023)
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions
por: Köksal, Abdullatif, et al.
Publicado: (2024)
por: Köksal, Abdullatif, et al.
Publicado: (2024)
The Promises and Pitfalls of LLM Annotations in Dataset Labeling: a Case Study on Media Bias Detection
por: Horych, Tomas, et al.
Publicado: (2024)
por: Horych, Tomas, et al.
Publicado: (2024)
How Transliterations Improve Crosslingual Alignment
por: Liu, Yihong, et al.
Publicado: (2024)
por: Liu, Yihong, et al.
Publicado: (2024)
Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena
por: Weissweiler, Leonie, et al.
Publicado: (2024)
por: Weissweiler, Leonie, et al.
Publicado: (2024)
MAGPIE: Multi-Task Media-Bias Analysis Generalization for Pre-Trained Identification of Expressions
por: Horych, Tomáš, et al.
Publicado: (2024)
por: Horych, Tomáš, et al.
Publicado: (2024)
GKnow: Measuring the Entanglement of Gender Bias and Factual Gender
por: Veloso, Leonor, et al.
Publicado: (2026)
por: Veloso, Leonor, et al.
Publicado: (2026)
GLUScope: A Tool for Analyzing GLU Neurons in Transformer Language Models
por: Gerstner, Sebastian, et al.
Publicado: (2026)
por: Gerstner, Sebastian, et al.
Publicado: (2026)
Are Emotions Arranged in a Circle? Geometric Analysis of Emotion Representations via Hyperspherical Contrastive Learning
por: Yamauchi, Yusuke, et al.
Publicado: (2026)
por: Yamauchi, Yusuke, et al.
Publicado: (2026)
Unsupervised Domain Adaptation for Keyphrase Generation using Citation Contexts
por: Boudin, Florian, et al.
Publicado: (2024)
por: Boudin, Florian, et al.
Publicado: (2024)
Leveraging Large Language Models for Automated Definition Extraction with TaxoMatic A Case Study on Media Bias
por: Spinde, Timo, et al.
Publicado: (2025)
por: Spinde, Timo, et al.
Publicado: (2025)
LongForm: Effective Instruction Tuning with Reverse Instructions
por: Köksal, Abdullatif, et al.
Publicado: (2023)
por: Köksal, Abdullatif, et al.
Publicado: (2023)
Understanding Gated Neurons in Transformers from Their Input-Output Functionality
por: Gerstner, Sebastian, et al.
Publicado: (2025)
por: Gerstner, Sebastian, et al.
Publicado: (2025)
Refining and Reusing Annotation Guidelines for LLM Annotation
por: Kim, Kon Woo, et al.
Publicado: (2026)
por: Kim, Kon Woo, et al.
Publicado: (2026)
Harnessing PDF Data for Improving Japanese Large Multimodal Models
por: Baek, Jeonghun, et al.
Publicado: (2025)
por: Baek, Jeonghun, et al.
Publicado: (2025)
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
por: Zhao, Raoyuan, et al.
Publicado: (2025)
por: Zhao, Raoyuan, et al.
Publicado: (2025)
HYPEROFA: Expanding LLM Vocabulary to New Languages via Hypernetwork-Based Embedding Initialization
por: Özeren, Enes, et al.
Publicado: (2025)
por: Özeren, Enes, et al.
Publicado: (2025)
Reassessing Extractive QA Datasets at Scale: LLM-as-a-Judge and In-Depth Analyses
por: Ho, Xanh, et al.
Publicado: (2025)
por: Ho, Xanh, et al.
Publicado: (2025)
An Analysis of Datasets, Metrics and Models in Keyphrase Generation
por: Boudin, Florian, et al.
Publicado: (2025)
por: Boudin, Florian, et al.
Publicado: (2025)
Preface to the Special Issue of the TAL Journal on Scholarly Document Processing
por: Boudin, Florian, et al.
Publicado: (2025)
por: Boudin, Florian, et al.
Publicado: (2025)
Mechanistic Understanding and Mitigation of Language Confusion in English-Centric Large Language Models
por: Nie, Ercong, et al.
Publicado: (2025)
por: Nie, Ercong, et al.
Publicado: (2025)
Consistent Document-Level Relation Extraction via Counterfactuals
por: Modarressi, Ali, et al.
Publicado: (2024)
por: Modarressi, Ali, et al.
Publicado: (2024)
Evaluating Contextually Mediated Factual Recall in Multilingual Large Language Models
por: Liu, Yihong, et al.
Publicado: (2026)
por: Liu, Yihong, et al.
Publicado: (2026)
Why Better Cross-Lingual Alignment Fails for Better Cross-Lingual Transfer: Case of Encoders
por: Veitsman, Yana, et al.
Publicado: (2026)
por: Veitsman, Yana, et al.
Publicado: (2026)
The Anatomy of an Edit: Mechanism-Guided Activation Steering for Knowledge Editing
por: Cao, Yuan, et al.
Publicado: (2026)
por: Cao, Yuan, et al.
Publicado: (2026)
Breaking the Script Barrier in Multilingual Pre-Trained Language Models with Transliteration-Based Post-Training Alignment
por: Xhelili, Orgest, et al.
Publicado: (2024)
por: Xhelili, Orgest, et al.
Publicado: (2024)
JMedBench: A Benchmark for Evaluating Japanese Biomedical Large Language Models
por: Jiang, Junfeng, et al.
Publicado: (2024)
por: Jiang, Junfeng, et al.
Publicado: (2024)
A Recipe of Parallel Corpora Exploitation for Multilingual Large Language Models
por: Lin, Peiqin, et al.
Publicado: (2024)
por: Lin, Peiqin, et al.
Publicado: (2024)
GlotScript: A Resource and Tool for Low Resource Writing System Identification
por: Kargaran, Amir Hossein, et al.
Publicado: (2023)
por: Kargaran, Amir Hossein, et al.
Publicado: (2023)
CRAFT Your Dataset: Task-Specific Synthetic Dataset Generation Through Corpus Retrieval and Augmentation
por: Ziegler, Ingo, et al.
Publicado: (2024)
por: Ziegler, Ingo, et al.
Publicado: (2024)
Automatically Suggesting Diverse Example Sentences for L2 Japanese Learners Using Pre-Trained Language Models
por: Benedetti, Enrico, et al.
Publicado: (2025)
por: Benedetti, Enrico, et al.
Publicado: (2025)
Bring Your Own Knowledge: A Survey of Methods for LLM Knowledge Expansion
por: Wang, Mingyang, et al.
Publicado: (2025)
por: Wang, Mingyang, et al.
Publicado: (2025)
Through the Looking Glass, and what Horn Clause Programs Found There
por: Tarau, Paul
Publicado: (2024)
por: Tarau, Paul
Publicado: (2024)
MaskLID: Code-Switching Language Identification through Iterative Masking
por: Kargaran, Amir Hossein, et al.
Publicado: (2024)
por: Kargaran, Amir Hossein, et al.
Publicado: (2024)
XAMPLER: Learning to Retrieve Cross-Lingual In-Context Examples
por: Lin, Peiqin, et al.
Publicado: (2024)
por: Lin, Peiqin, et al.
Publicado: (2024)
Relational Linearity is a Predictor of Hallucinations
por: Lu, Yuetian, et al.
Publicado: (2026)
por: Lu, Yuetian, et al.
Publicado: (2026)
Ejemplares similares
-
RET-LLM: Towards a General Read-Write Memory for Large Language Models
por: Modarressi, Ali, et al.
Publicado: (2023) -
MemLLM: Finetuning LLMs to Use An Explicit Read-Write Memory
por: Modarressi, Ali, et al.
Publicado: (2024) -
GlotLID: Language Identification for Low-Resource Languages
por: Kargaran, Amir Hossein, et al.
Publicado: (2023) -
Left, Right, or Center? Evaluating LLM Framing in News Classification and Generation
por: Kennedy, Molly, et al.
Publicado: (2026) -
Taxi1500: A Multilingual Dataset for Text Classification in 1500 Languages
por: Ma, Chunlan, et al.
Publicado: (2023)