Hatevolution: What Static Benchmarks Don't Tell Us
Fuente:
arXiv
Guardado en:
| Autores principales: | Di Bonaventura, Chiara, McGillivray, Barbara, He, Yulan, Meroño-Peñuela, Albert |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DWUG: A large Resource of Diachronic Word Usage Graphs in Four Languages
por: Schlechtweg, Dominik, et al.
Publicado: (2021)
por: Schlechtweg, Dominik, et al.
Publicado: (2021)
Semantic Journeys: Quantifying Change in Emoji Meaning from 2012-2018
por: Robertson, Alexander, et al.
Publicado: (2021)
por: Robertson, Alexander, et al.
Publicado: (2021)
Ascensão e queda do pacto populista em Cuba, 1934-1959
por: Gillian McGillivray
Publicado: (2012)
por: Gillian McGillivray
Publicado: (2012)
Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell
por: Lu, Taiming, et al.
Publicado: (2024)
por: Lu, Taiming, et al.
Publicado: (2024)
Show, Don't Tell: Uncovering Implicit Character Portrayal using LLMs
por: Jaipersaud, Brandon, et al.
Publicado: (2024)
por: Jaipersaud, Brandon, et al.
Publicado: (2024)
Tell, Don't Show: Leveraging Language Models' Abstractive Retellings to Model Literary Themes
por: Lucy, Li, et al.
Publicado: (2025)
por: Lucy, Li, et al.
Publicado: (2025)
What Can String Probability Tell Us About Grammaticality?
por: Hu, Jennifer, et al.
Publicado: (2025)
por: Hu, Jennifer, et al.
Publicado: (2025)
Can AI Assistants Know What They Don't Know?
por: Cheng, Qinyuan, et al.
Publicado: (2024)
por: Cheng, Qinyuan, et al.
Publicado: (2024)
NeIn: Telling What You Don't Want
por: Bui, Nhat-Tan, et al.
Publicado: (2024)
por: Bui, Nhat-Tan, et al.
Publicado: (2024)
What Artificial Neural Networks Can Tell Us About Human Language Acquisition
por: Warstadt, Alex, et al.
Publicado: (2022)
por: Warstadt, Alex, et al.
Publicado: (2022)
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
por: Kalluri, Tarun, et al.
Publicado: (2024)
por: Kalluri, Tarun, et al.
Publicado: (2024)
Show, Don't Tell: Evaluating Large Language Models Beyond Textual Understanding with ChildPlay
por: de Carvalho, Gonçalo Hora, et al.
Publicado: (2024)
por: de Carvalho, Gonçalo Hora, et al.
Publicado: (2024)
Reasoning Models Don't Always Say What They Think
por: Chen, Yanda, et al.
Publicado: (2025)
por: Chen, Yanda, et al.
Publicado: (2025)
Don't Learn, Ground: A Case for Natural Language Inference with Visual Grounding
por: Ignatev, Daniil, et al.
Publicado: (2025)
por: Ignatev, Daniil, et al.
Publicado: (2025)
When Words Don't Mean What They Say: Figurative Understanding in Bengali Idioms
por: Sakhawat, Adib, et al.
Publicado: (2026)
por: Sakhawat, Adib, et al.
Publicado: (2026)
Don't Touch My Diacritics
por: Gorman, Kyle, et al.
Publicado: (2024)
por: Gorman, Kyle, et al.
Publicado: (2024)
Don't Pay Attention
por: Hammoud, Mohammad, et al.
Publicado: (2025)
por: Hammoud, Mohammad, et al.
Publicado: (2025)
Convomem Benchmark: Why Your First 150 Conversations Don't Need RAG
por: Pakhomov, Egor, et al.
Publicado: (2025)
por: Pakhomov, Egor, et al.
Publicado: (2025)
What Prompts Don't Say: Understanding and Managing Underspecification in LLM Prompts
por: Yang, Chenyang, et al.
Publicado: (2025)
por: Yang, Chenyang, et al.
Publicado: (2025)
Variance-Hawkes Process and its Application to Energy Markets
por: McGillivray, Joshua, et al.
Publicado: (2024)
por: McGillivray, Joshua, et al.
Publicado: (2024)
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
por: Zhao, Raoyuan, et al.
Publicado: (2025)
por: Zhao, Raoyuan, et al.
Publicado: (2025)
EpiCaR: Knowing What You Don't Know Matters for Better Reasoning in LLMs
por: Yeom, Jewon, et al.
Publicado: (2026)
por: Yeom, Jewon, et al.
Publicado: (2026)
Who Are All The Stochastic Parrots Imitating? They Should Tell Us!
por: Shaier, Sagi, et al.
Publicado: (2023)
por: Shaier, Sagi, et al.
Publicado: (2023)
If You Don't Understand It, Don't Use It: Eliminating Trojans with Filters Between Layers
por: Hernandez, Adriano
Publicado: (2024)
por: Hernandez, Adriano
Publicado: (2024)
Don't Sweat the Small Stuff: Segment-Level Meta-Evaluation Based on Pairwise Difference Correlation
por: DiIanni, Colten, et al.
Publicado: (2025)
por: DiIanni, Colten, et al.
Publicado: (2025)
Large Language Models Must Be Taught to Know What They Don't Know
por: Kapoor, Sanyam, et al.
Publicado: (2024)
por: Kapoor, Sanyam, et al.
Publicado: (2024)
Don't Throw Away Your Pretrained Model
por: Feng, Shangbin, et al.
Publicado: (2025)
por: Feng, Shangbin, et al.
Publicado: (2025)
Don't Say No: Jailbreaking LLM by Suppressing Refusal
por: Zhou, Yukai, et al.
Publicado: (2024)
por: Zhou, Yukai, et al.
Publicado: (2024)
Don't Trust Generative Agents to Mimic Communication on Social Networks Unless You Benchmarked their Empirical Realism
por: Münker, Simon, et al.
Publicado: (2025)
por: Münker, Simon, et al.
Publicado: (2025)
What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization
por: Rezaeimanesh, Sara, et al.
Publicado: (2026)
por: Rezaeimanesh, Sara, et al.
Publicado: (2026)
Think, But Don't Overthink: Reproducing Recursive Language Models
por: Wang, Daren
Publicado: (2026)
por: Wang, Daren
Publicado: (2026)
Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users
por: Balepur, Nishant, et al.
Publicado: (2026)
por: Balepur, Nishant, et al.
Publicado: (2026)
How Much Do Circuits Tell Us? Measuring the Consistency and Specificity of Language Model Circuits
por: Li, Michael, et al.
Publicado: (2026)
por: Li, Michael, et al.
Publicado: (2026)
User Experience In Dataset Search
por: Zhao, Yihang, et al.
Publicado: (2024)
por: Zhao, Yihang, et al.
Publicado: (2024)
Reasoning Models Reason Well, Until They Don't
por: Rameshkumar, Revanth, et al.
Publicado: (2025)
por: Rameshkumar, Revanth, et al.
Publicado: (2025)
What Don't You Understand? Using Large Language Models to Identify and Characterize Student Misconceptions About Challenging Topics
por: Parker, Michael J., et al.
Publicado: (2026)
por: Parker, Michael J., et al.
Publicado: (2026)
Don't Throw Away Data: Better Sequence Knowledge Distillation
por: Wang, Jun, et al.
Publicado: (2024)
por: Wang, Jun, et al.
Publicado: (2024)
Don't Command, Cultivate: An Exploratory Study of System-2 Alignment
por: Wang, Yuhang, et al.
Publicado: (2024)
por: Wang, Yuhang, et al.
Publicado: (2024)
Don't Go To Extremes: Revealing the Excessive Sensitivity and Calibration Limitations of LLMs in Implicit Hate Speech Detection
por: Zhang, Min, et al.
Publicado: (2024)
por: Zhang, Min, et al.
Publicado: (2024)
Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness
por: Cheang, Chi Seng, et al.
Publicado: (2025)
por: Cheang, Chi Seng, et al.
Publicado: (2025)
Ejemplares similares
-
DWUG: A large Resource of Diachronic Word Usage Graphs in Four Languages
por: Schlechtweg, Dominik, et al.
Publicado: (2021) -
Semantic Journeys: Quantifying Change in Emoji Meaning from 2012-2018
por: Robertson, Alexander, et al.
Publicado: (2021) -
Ascensão e queda do pacto populista em Cuba, 1934-1959
por: Gillian McGillivray
Publicado: (2012) -
Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell
por: Lu, Taiming, et al.
Publicado: (2024) -
Show, Don't Tell: Uncovering Implicit Character Portrayal using LLMs
por: Jaipersaud, Brandon, et al.
Publicado: (2024)