Measuring Spurious Correlation in Classification: 'Clever Hans' in Translationese
Fuente:
arXiv
Saved in:
| Main Authors: | Borah, Angana, Pylypenko, Daria, Espana-Bonet, Cristina, van Genabith, Josef |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sign Language Translation with Sentence Embedding Supervision
by: Hamidullah, Yasser, et al.
Published: (2025)
by: Hamidullah, Yasser, et al.
Published: (2025)
Spatio-temporal Sign Language Representation and Translation
by: Hamidullah, Yasser, et al.
Published: (2025)
by: Hamidullah, Yasser, et al.
Published: (2025)
Seeing, Signing, and Saying: A Vision-Language Model-Assisted Pipeline for Sign Language Data Acquisition and Curation from Social Media
by: Yazdani, Shakib, et al.
Published: (2025)
by: Yazdani, Shakib, et al.
Published: (2025)
When your Cousin has the Right Connections: Unsupervised Bilingual Lexicon Induction for Related Data-Imbalanced Languages
by: Bafna, Niyati, et al.
Published: (2023)
by: Bafna, Niyati, et al.
Published: (2023)
A Critical Study of Automatic Evaluation in Sign Language Translation
by: Yazdani, Shakib, et al.
Published: (2025)
by: Yazdani, Shakib, et al.
Published: (2025)
SONAR-SLT: Multilingual Sign Language Translation via Language-Agnostic Sentence Embedding Supervision
by: Hamidullah, Yasser, et al.
Published: (2025)
by: Hamidullah, Yasser, et al.
Published: (2025)
PETra: A Multilingual Corpus of Pragmatic Explicitation in Translation
by: Osmelak, Doreen, et al.
Published: (2025)
by: Osmelak, Doreen, et al.
Published: (2025)
DFKI-MLT at SemEval-2026 TASK 7: Steering Multilingual Models Towards Cultural Knowledge
by: Ghussin, Yusser Al, et al.
Published: (2026)
by: Ghussin, Yusser Al, et al.
Published: (2026)
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
by: Borah, Angana, et al.
Published: (2024)
by: Borah, Angana, et al.
Published: (2024)
Translationese-index: Using Likelihood Ratios for Graded and Generalizable Measurement of Translationese
by: Liu, Yikang, et al.
Published: (2025)
by: Liu, Yikang, et al.
Published: (2025)
The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models
by: Arif, Samee, et al.
Published: (2026)
by: Arif, Samee, et al.
Published: (2026)
Towards Region-aware Bias Evaluation Metrics
by: Borah, Angana, et al.
Published: (2024)
by: Borah, Angana, et al.
Published: (2024)
The Curious Case of Curiosity across Human Cultures and LLMs
by: Borah, Angana, et al.
Published: (2025)
by: Borah, Angana, et al.
Published: (2025)
Disentangling Mathematical Reasoning in LLMs: A Methodological Investigation of Internal Mechanisms
by: Baeumel, Tanja, et al.
Published: (2026)
by: Baeumel, Tanja, et al.
Published: (2026)
Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers
by: Vijayakumar, Soniya, et al.
Published: (2024)
by: Vijayakumar, Soniya, et al.
Published: (2024)
The Lookahead Limitation: Why Multi-Operand Addition is Hard for LLMs
by: Baeumel, Tanja, et al.
Published: (2025)
by: Baeumel, Tanja, et al.
Published: (2025)
HANS, are you clever? Clever Hans Effect Analysis of Neural Systems
by: Ranaldi, Leonardo, et al.
Published: (2023)
by: Ranaldi, Leonardo, et al.
Published: (2023)
Mind the (Belief) Gap: Group Identity in the World of LLMs
by: Borah, Angana, et al.
Published: (2025)
by: Borah, Angana, et al.
Published: (2025)
Application Specific Compression of Deep Learning Models
by: Rai, Rohit Raj, et al.
Published: (2024)
by: Rai, Rohit Raj, et al.
Published: (2024)
Persuasion at Play: Understanding Misinformation Dynamics in Demographic-Aware Human-LLM Interactions
by: Borah, Angana, et al.
Published: (2025)
by: Borah, Angana, et al.
Published: (2025)
Are Word Embedding Methods Stable and Should We Care About It?
by: Borah, Angana, et al.
Published: (2021)
by: Borah, Angana, et al.
Published: (2021)
The Clever Hans Mirage: A Comprehensive Survey on Spurious Correlations in Machine Learning
by: Ye, Wenqian, et al.
Published: (2024)
by: Ye, Wenqian, et al.
Published: (2024)
On Multilingual Encoder Language Model Compression for Low-Resource Languages
by: Gurgurov, Daniil, et al.
Published: (2025)
by: Gurgurov, Daniil, et al.
Published: (2025)
When Scale Meets Diversity: Evaluating Language Models on Fine-Grained Multilingual Claim Verification
by: Shcharbakova, Hanna, et al.
Published: (2025)
by: Shcharbakova, Hanna, et al.
Published: (2025)
Small Models, Big Impact: Efficient Corpus and Graph-Based Adaptation of Small Multilingual Language Models for Low-Resource Languages
by: Gurgurov, Daniil, et al.
Published: (2025)
by: Gurgurov, Daniil, et al.
Published: (2025)
Sparse Subnetwork Enhancement for Underrepresented Languages in Large Language Models
by: Gurgurov, Daniil, et al.
Published: (2025)
by: Gurgurov, Daniil, et al.
Published: (2025)
Understanding and Mitigating Spurious Correlations in Text Classification with Neighborhood Analysis
by: Chew, Oscar, et al.
Published: (2023)
by: Chew, Oscar, et al.
Published: (2023)
Pretraining Language Models Using Translationese
by: Doshi, Meet, et al.
Published: (2024)
by: Doshi, Meet, et al.
Published: (2024)
Leaving the barn door open for Clever Hans: Simple features predict LLM benchmark answers
by: Pacchiardi, Lorenzo, et al.
Published: (2024)
by: Pacchiardi, Lorenzo, et al.
Published: (2024)
When Flores Bloomz Wrong: Cross-Direction Contamination in Machine Translation Evaluation
by: Tan, David, et al.
Published: (2026)
by: Tan, David, et al.
Published: (2026)
The Power of Many: Multi-Agent Multimodal Models for Cultural Image Captioning
by: Bai, Longju, et al.
Published: (2024)
by: Bai, Longju, et al.
Published: (2024)
The Latin Substrate: How Language Models Represent and Mediate Script Choice
by: Gurgurov, Daniil, et al.
Published: (2026)
by: Gurgurov, Daniil, et al.
Published: (2026)
Rewiring the Transformer with Depth-Wise LSTMs
by: Xu, Hongfei, et al.
Published: (2020)
by: Xu, Hongfei, et al.
Published: (2020)
Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility
by: Borah, Angana, et al.
Published: (2026)
by: Borah, Angana, et al.
Published: (2026)
Explore Spurious Correlations at the Concept Level in Language Models for Text Classification
by: Zhou, Yuhang, et al.
Published: (2023)
by: Zhou, Yuhang, et al.
Published: (2023)
Reproducibility study on how to find Spurious Correlations, Shortcut Learning, Clever Hans or Group-Distributional non-robustness and how to fix them
by: Delzer, Ole, et al.
Published: (2026)
by: Delzer, Ole, et al.
Published: (2026)
Lost in Literalism: How Supervised Training Shapes Translationese in LLMs
by: Li, Yafu, et al.
Published: (2025)
by: Li, Yafu, et al.
Published: (2025)
A Dataset for Probing Translationese Preferences in English-to-Swedish Translation
by: Kunz, Jenny, et al.
Published: (2026)
by: Kunz, Jenny, et al.
Published: (2026)
Fighting Spurious Correlations in Text Classification via a Causal Learning Perspective
by: Zhou, Yuqing, et al.
Published: (2024)
by: Zhou, Yuqing, et al.
Published: (2024)
Multilingual Political Views of Large Language Models: Identification and Steering
by: Gurgurov, Daniil, et al.
Published: (2025)
by: Gurgurov, Daniil, et al.
Published: (2025)
Similar Items
-
Sign Language Translation with Sentence Embedding Supervision
by: Hamidullah, Yasser, et al.
Published: (2025) -
Spatio-temporal Sign Language Representation and Translation
by: Hamidullah, Yasser, et al.
Published: (2025) -
Seeing, Signing, and Saying: A Vision-Language Model-Assisted Pipeline for Sign Language Data Acquisition and Curation from Social Media
by: Yazdani, Shakib, et al.
Published: (2025) -
When your Cousin has the Right Connections: Unsupervised Bilingual Lexicon Induction for Related Data-Imbalanced Languages
by: Bafna, Niyati, et al.
Published: (2023) -
A Critical Study of Automatic Evaluation in Sign Language Translation
by: Yazdani, Shakib, et al.
Published: (2025)