Enregistré dans:
| Auteurs principaux: | Sandoval, Sandra C., Acquaye, Christabel, Cobbina, Kwesi, Teli, Mohammad Nayeem, Daumé III, Hal |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2502.04564 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
When Stereotypes GTG: The Impact of Predictive Text Suggestions on Gender Bias in Human-AI Co-Writing
par: Baumler, Connor, et autres
Publié: (2024)
par: Baumler, Connor, et autres
Publié: (2024)
Susu Box or Piggy Bank: Assessing Cultural Commonsense Knowledge between Ghana and the U.S
par: Acquaye, Christabel, et autres
Publié: (2024)
par: Acquaye, Christabel, et autres
Publié: (2024)
Where to show Demos in Your Prompt: A Positional Bias of In-Context Learning
par: Cobbina, Kwesi, et autres
Publié: (2025)
par: Cobbina, Kwesi, et autres
Publié: (2025)
A Necessary Step toward Faithfulness: Measuring and Improving Consistency in Free-Text Explanations
par: Zhao, Lingjun, et autres
Publié: (2025)
par: Zhao, Lingjun, et autres
Publié: (2025)
Steering Safely or Off a Cliff? Rethinking Specificity and Robustness in Inference-Time Interventions
par: Goyal, Navita, et autres
Publié: (2026)
par: Goyal, Navita, et autres
Publié: (2026)
HateCOT: An Explanation-Enhanced Dataset for Generalizable Offensive Speech Detection via Large Language Models
par: Nghiem, Huy, et autres
Publié: (2024)
par: Nghiem, Huy, et autres
Publié: (2024)
Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations
par: Acquaye, Christabel, et autres
Publié: (2026)
par: Acquaye, Christabel, et autres
Publié: (2026)
SMARTER: A Data-efficient Framework to Improve Toxicity Detection with Explanation via Self-augmenting Large Language Models
par: Nghiem, Huy, et autres
Publié: (2025)
par: Nghiem, Huy, et autres
Publié: (2025)
Language Models Predict Empathy Gaps Between Social In-groups and Out-groups
par: Hou, Yu, et autres
Publié: (2025)
par: Hou, Yu, et autres
Publié: (2025)
Successfully Guiding Humans with Imperfect Instructions by Highlighting Potential Errors and Suggesting Corrections
par: Zhao, Lingjun, et autres
Publié: (2024)
par: Zhao, Lingjun, et autres
Publié: (2024)
Do Large Language Models Discriminate in Hiring Decisions on the Basis of Race, Ethnicity, and Gender?
par: An, Haozhe, et autres
Publié: (2024)
par: An, Haozhe, et autres
Publié: (2024)
"You Gotta be a Doctor, Lin": An Investigation of Name-Based Bias of Large Language Models in Employment Recommendations
par: Nghiem, Huy, et autres
Publié: (2024)
par: Nghiem, Huy, et autres
Publié: (2024)
Pragmatics Meets Culture: Culturally-adapted Artwork Description Generation and Evaluation
par: Zhao, Lingjun, et autres
Publié: (2026)
par: Zhao, Lingjun, et autres
Publié: (2026)
Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring
par: Nghiem, Huy, et autres
Publié: (2026)
par: Nghiem, Huy, et autres
Publié: (2026)
Causal Effect of Group Diversity on Redundancy and Coverage in Peer-Reviewing
par: Goyal, Navita, et autres
Publié: (2024)
par: Goyal, Navita, et autres
Publié: (2024)
Can You Make It Sound Like You? Post-Editing LLM-Generated Text for Personal Style
par: Baumler, Connor, et autres
Publié: (2026)
par: Baumler, Connor, et autres
Publié: (2026)
Do great minds think alike? Investigating Human-AI Complementarity in Question Answering with CAIMIRA
par: Gor, Maharshi, et autres
Publié: (2024)
par: Gor, Maharshi, et autres
Publié: (2024)
Large Language Models Help Humans Verify Truthfulness -- Except When They Are Convincingly Wrong
par: Si, Chenglei, et autres
Publié: (2023)
par: Si, Chenglei, et autres
Publié: (2023)
'Rich Dad, Poor Lad': How do Large Language Models Contextualize Socioeconomic Factors in College Admission ?
par: Nghiem, Huy, et autres
Publié: (2025)
par: Nghiem, Huy, et autres
Publié: (2025)
Can Hallucination Correction Improve Video-Language Alignment?
par: Zhao, Lingjun, et autres
Publié: (2025)
par: Zhao, Lingjun, et autres
Publié: (2025)
Natural Language Inference Improves Compositionality in Vision-Language Models
par: Cascante-Bonilla, Paola, et autres
Publié: (2024)
par: Cascante-Bonilla, Paola, et autres
Publié: (2024)
Reheat Nachos for Dinner? Evaluating AI Support for Cross-Cultural Communication of Neologisms
par: Ki, Dayeon, et autres
Publié: (2026)
par: Ki, Dayeon, et autres
Publié: (2026)
Balancing Safety and Helpfulness in Healthcare AI Assistants through Iterative Preference Alignment
par: Nghiem, Huy, et autres
Publié: (2025)
par: Nghiem, Huy, et autres
Publié: (2025)
Multilingual large language models leak human stereotypes across language boundaries
par: Cao, Yang Trista, et autres
Publié: (2023)
par: Cao, Yang Trista, et autres
Publié: (2023)
Bringing together invertible UNets with invertible attention modules for memory-efficient diffusion models
par: Jain, Karan, et autres
Publié: (2025)
par: Jain, Karan, et autres
Publié: (2025)
Improving Deep Generative Models on Many-To-One Image-to-Image Translation
par: Saxena, Sagar, et autres
Publié: (2024)
par: Saxena, Sagar, et autres
Publié: (2024)
Toxicity Detection is NOT all you Need: Measuring the Gaps to Supporting Volunteer Content Moderators
par: Cao, Yang Trista, et autres
Publié: (2023)
par: Cao, Yang Trista, et autres
Publié: (2023)
Say It My Way: Exploring Control in Conversational Visual Question Answering with Blind Users
par: Zeraati, Farnaz Zamiri, et autres
Publié: (2026)
par: Zeraati, Farnaz Zamiri, et autres
Publié: (2026)
ASL STEM Wiki: Dataset and Benchmark for Interpreting STEM Articles
par: Yin, Kayo, et autres
Publié: (2024)
par: Yin, Kayo, et autres
Publié: (2024)
The Impact of Explanations on Fairness in Human-AI Decision-Making: Protected vs Proxy Features
par: Goyal, Navita, et autres
Publié: (2023)
par: Goyal, Navita, et autres
Publié: (2023)
All-in-One Conditioning for Text-to-Image Synthesis
par: Jayasekara, Hirunima, et autres
Publié: (2026)
par: Jayasekara, Hirunima, et autres
Publié: (2026)
ColorBench: Can VLMs See and Understand the Colorful World? A Comprehensive Benchmark for Color Perception, Reasoning, and Robustness
par: Liang, Yijun, et autres
Publié: (2025)
par: Liang, Yijun, et autres
Publié: (2025)
A Heterogeneous Ensemble for Multi-Center COVID-19 Classification from Chest CT Scans
par: Nilay, Aadit, et autres
Publié: (2026)
par: Nilay, Aadit, et autres
Publié: (2026)
PRISE: LLM-Style Sequence Compression for Learning Temporal Action Abstractions in Control
par: Zheng, Ruijie, et autres
Publié: (2024)
par: Zheng, Ruijie, et autres
Publié: (2024)
Evaluating the Semantic Profiling Abilities of LLMs for Natural Language Utterances in Data Visualization
par: Bako, Hannah K., et autres
Publié: (2024)
par: Bako, Hannah K., et autres
Publié: (2024)
Seamful XAI: Operationalizing Seamful Design in Explainable AI
par: Ehsan, Upol, et autres
Publié: (2022)
par: Ehsan, Upol, et autres
Publié: (2022)
Steer Like the LLM: Activation Steering that Mimics Prompting
par: Heyman, Geert, et autres
Publié: (2026)
par: Heyman, Geert, et autres
Publié: (2026)
V-REX: Benchmarking Exploratory Visual Reasoning via Chain-of-Questions
par: Fan, Chenrui, et autres
Publié: (2025)
par: Fan, Chenrui, et autres
Publié: (2025)
An Examination of the Compositionality of Large Generative Vision-Language Models
par: Ma, Teli, et autres
Publié: (2023)
par: Ma, Teli, et autres
Publié: (2023)
KidLM: Advancing Language Models for Children -- Early Insights and Future Directions
par: Nayeem, Mir Tafseer, et autres
Publié: (2024)
par: Nayeem, Mir Tafseer, et autres
Publié: (2024)
Documents similaires
-
When Stereotypes GTG: The Impact of Predictive Text Suggestions on Gender Bias in Human-AI Co-Writing
par: Baumler, Connor, et autres
Publié: (2024) -
Susu Box or Piggy Bank: Assessing Cultural Commonsense Knowledge between Ghana and the U.S
par: Acquaye, Christabel, et autres
Publié: (2024) -
Where to show Demos in Your Prompt: A Positional Bias of In-Context Learning
par: Cobbina, Kwesi, et autres
Publié: (2025) -
A Necessary Step toward Faithfulness: Measuring and Improving Consistency in Free-Text Explanations
par: Zhao, Lingjun, et autres
Publié: (2025) -
Steering Safely or Off a Cliff? Rethinking Specificity and Robustness in Inference-Time Interventions
par: Goyal, Navita, et autres
Publié: (2026)