Measuring Spurious Correlation in Classification: 'Clever Hans' in Translationese
Fuente:
arXiv
Guardado en:
| Autores principales: | Borah, Angana, Pylypenko, Daria, Espana-Bonet, Cristina, van Genabith, Josef |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Sign Language Translation with Sentence Embedding Supervision
por: Hamidullah, Yasser, et al.
Publicado: (2025)
por: Hamidullah, Yasser, et al.
Publicado: (2025)
Spatio-temporal Sign Language Representation and Translation
por: Hamidullah, Yasser, et al.
Publicado: (2025)
por: Hamidullah, Yasser, et al.
Publicado: (2025)
Seeing, Signing, and Saying: A Vision-Language Model-Assisted Pipeline for Sign Language Data Acquisition and Curation from Social Media
por: Yazdani, Shakib, et al.
Publicado: (2025)
por: Yazdani, Shakib, et al.
Publicado: (2025)
When your Cousin has the Right Connections: Unsupervised Bilingual Lexicon Induction for Related Data-Imbalanced Languages
por: Bafna, Niyati, et al.
Publicado: (2023)
por: Bafna, Niyati, et al.
Publicado: (2023)
A Critical Study of Automatic Evaluation in Sign Language Translation
por: Yazdani, Shakib, et al.
Publicado: (2025)
por: Yazdani, Shakib, et al.
Publicado: (2025)
SONAR-SLT: Multilingual Sign Language Translation via Language-Agnostic Sentence Embedding Supervision
por: Hamidullah, Yasser, et al.
Publicado: (2025)
por: Hamidullah, Yasser, et al.
Publicado: (2025)
PETra: A Multilingual Corpus of Pragmatic Explicitation in Translation
por: Osmelak, Doreen, et al.
Publicado: (2025)
por: Osmelak, Doreen, et al.
Publicado: (2025)
DFKI-MLT at SemEval-2026 TASK 7: Steering Multilingual Models Towards Cultural Knowledge
por: Ghussin, Yusser Al, et al.
Publicado: (2026)
por: Ghussin, Yusser Al, et al.
Publicado: (2026)
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
por: Borah, Angana, et al.
Publicado: (2024)
por: Borah, Angana, et al.
Publicado: (2024)
Translationese-index: Using Likelihood Ratios for Graded and Generalizable Measurement of Translationese
por: Liu, Yikang, et al.
Publicado: (2025)
por: Liu, Yikang, et al.
Publicado: (2025)
The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models
por: Arif, Samee, et al.
Publicado: (2026)
por: Arif, Samee, et al.
Publicado: (2026)
Towards Region-aware Bias Evaluation Metrics
por: Borah, Angana, et al.
Publicado: (2024)
por: Borah, Angana, et al.
Publicado: (2024)
The Curious Case of Curiosity across Human Cultures and LLMs
por: Borah, Angana, et al.
Publicado: (2025)
por: Borah, Angana, et al.
Publicado: (2025)
Disentangling Mathematical Reasoning in LLMs: A Methodological Investigation of Internal Mechanisms
por: Baeumel, Tanja, et al.
Publicado: (2026)
por: Baeumel, Tanja, et al.
Publicado: (2026)
Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers
por: Vijayakumar, Soniya, et al.
Publicado: (2024)
por: Vijayakumar, Soniya, et al.
Publicado: (2024)
The Lookahead Limitation: Why Multi-Operand Addition is Hard for LLMs
por: Baeumel, Tanja, et al.
Publicado: (2025)
por: Baeumel, Tanja, et al.
Publicado: (2025)
HANS, are you clever? Clever Hans Effect Analysis of Neural Systems
por: Ranaldi, Leonardo, et al.
Publicado: (2023)
por: Ranaldi, Leonardo, et al.
Publicado: (2023)
Mind the (Belief) Gap: Group Identity in the World of LLMs
por: Borah, Angana, et al.
Publicado: (2025)
por: Borah, Angana, et al.
Publicado: (2025)
Application Specific Compression of Deep Learning Models
por: Rai, Rohit Raj, et al.
Publicado: (2024)
por: Rai, Rohit Raj, et al.
Publicado: (2024)
Persuasion at Play: Understanding Misinformation Dynamics in Demographic-Aware Human-LLM Interactions
por: Borah, Angana, et al.
Publicado: (2025)
por: Borah, Angana, et al.
Publicado: (2025)
Are Word Embedding Methods Stable and Should We Care About It?
por: Borah, Angana, et al.
Publicado: (2021)
por: Borah, Angana, et al.
Publicado: (2021)
The Clever Hans Mirage: A Comprehensive Survey on Spurious Correlations in Machine Learning
por: Ye, Wenqian, et al.
Publicado: (2024)
por: Ye, Wenqian, et al.
Publicado: (2024)
On Multilingual Encoder Language Model Compression for Low-Resource Languages
por: Gurgurov, Daniil, et al.
Publicado: (2025)
por: Gurgurov, Daniil, et al.
Publicado: (2025)
When Scale Meets Diversity: Evaluating Language Models on Fine-Grained Multilingual Claim Verification
por: Shcharbakova, Hanna, et al.
Publicado: (2025)
por: Shcharbakova, Hanna, et al.
Publicado: (2025)
Small Models, Big Impact: Efficient Corpus and Graph-Based Adaptation of Small Multilingual Language Models for Low-Resource Languages
por: Gurgurov, Daniil, et al.
Publicado: (2025)
por: Gurgurov, Daniil, et al.
Publicado: (2025)
Sparse Subnetwork Enhancement for Underrepresented Languages in Large Language Models
por: Gurgurov, Daniil, et al.
Publicado: (2025)
por: Gurgurov, Daniil, et al.
Publicado: (2025)
Understanding and Mitigating Spurious Correlations in Text Classification with Neighborhood Analysis
por: Chew, Oscar, et al.
Publicado: (2023)
por: Chew, Oscar, et al.
Publicado: (2023)
Pretraining Language Models Using Translationese
por: Doshi, Meet, et al.
Publicado: (2024)
por: Doshi, Meet, et al.
Publicado: (2024)
Leaving the barn door open for Clever Hans: Simple features predict LLM benchmark answers
por: Pacchiardi, Lorenzo, et al.
Publicado: (2024)
por: Pacchiardi, Lorenzo, et al.
Publicado: (2024)
When Flores Bloomz Wrong: Cross-Direction Contamination in Machine Translation Evaluation
por: Tan, David, et al.
Publicado: (2026)
por: Tan, David, et al.
Publicado: (2026)
The Power of Many: Multi-Agent Multimodal Models for Cultural Image Captioning
por: Bai, Longju, et al.
Publicado: (2024)
por: Bai, Longju, et al.
Publicado: (2024)
The Latin Substrate: How Language Models Represent and Mediate Script Choice
por: Gurgurov, Daniil, et al.
Publicado: (2026)
por: Gurgurov, Daniil, et al.
Publicado: (2026)
Rewiring the Transformer with Depth-Wise LSTMs
por: Xu, Hongfei, et al.
Publicado: (2020)
por: Xu, Hongfei, et al.
Publicado: (2020)
Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility
por: Borah, Angana, et al.
Publicado: (2026)
por: Borah, Angana, et al.
Publicado: (2026)
Explore Spurious Correlations at the Concept Level in Language Models for Text Classification
por: Zhou, Yuhang, et al.
Publicado: (2023)
por: Zhou, Yuhang, et al.
Publicado: (2023)
Reproducibility study on how to find Spurious Correlations, Shortcut Learning, Clever Hans or Group-Distributional non-robustness and how to fix them
por: Delzer, Ole, et al.
Publicado: (2026)
por: Delzer, Ole, et al.
Publicado: (2026)
Lost in Literalism: How Supervised Training Shapes Translationese in LLMs
por: Li, Yafu, et al.
Publicado: (2025)
por: Li, Yafu, et al.
Publicado: (2025)
A Dataset for Probing Translationese Preferences in English-to-Swedish Translation
por: Kunz, Jenny, et al.
Publicado: (2026)
por: Kunz, Jenny, et al.
Publicado: (2026)
Fighting Spurious Correlations in Text Classification via a Causal Learning Perspective
por: Zhou, Yuqing, et al.
Publicado: (2024)
por: Zhou, Yuqing, et al.
Publicado: (2024)
Multilingual Political Views of Large Language Models: Identification and Steering
por: Gurgurov, Daniil, et al.
Publicado: (2025)
por: Gurgurov, Daniil, et al.
Publicado: (2025)
Ejemplares similares
-
Sign Language Translation with Sentence Embedding Supervision
por: Hamidullah, Yasser, et al.
Publicado: (2025) -
Spatio-temporal Sign Language Representation and Translation
por: Hamidullah, Yasser, et al.
Publicado: (2025) -
Seeing, Signing, and Saying: A Vision-Language Model-Assisted Pipeline for Sign Language Data Acquisition and Curation from Social Media
por: Yazdani, Shakib, et al.
Publicado: (2025) -
When your Cousin has the Right Connections: Unsupervised Bilingual Lexicon Induction for Related Data-Imbalanced Languages
por: Bafna, Niyati, et al.
Publicado: (2023) -
A Critical Study of Automatic Evaluation in Sign Language Translation
por: Yazdani, Shakib, et al.
Publicado: (2025)