Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Shen, Hanwen, Ying, Ting, Lu, Jiajie, Wang, Shanshan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DeCAP: Context-Adaptive Prompt Generation for Debiasing Zero-shot Question Answering in Large Language Models
por: Bae, Suyoung, et al.
Publicado: (2025)
por: Bae, Suyoung, et al.
Publicado: (2025)
On the Effectiveness and Generalization of Race Representations for Debiasing High-Stakes Decisions
por: Nguyen, Dang, et al.
Publicado: (2025)
por: Nguyen, Dang, et al.
Publicado: (2025)
Beyond Pairwise Comparisons: A Distributional Test of Distinctiveness for Machine-Generated Works in Intellectual Property Law
por: Mukherjee, Anirban, et al.
Publicado: (2026)
por: Mukherjee, Anirban, et al.
Publicado: (2026)
Measuring Information Distortion in Hierarchical Ultra long Novel Reconstruction:The Optimal Expansion Ratio
por: Shen, Hanwen, et al.
Publicado: (2025)
por: Shen, Hanwen, et al.
Publicado: (2025)
InsideOut: Measuring and Mitigating Insider-Outsider Bias in Interview Script Generation
por: Wan, Yixin, et al.
Publicado: (2025)
por: Wan, Yixin, et al.
Publicado: (2025)
A Multi-LLM Debiasing Framework
por: Owens, Deonna M., et al.
Publicado: (2024)
por: Owens, Deonna M., et al.
Publicado: (2024)
MirrorStories: Reflecting Diversity through Personalized Narrative Generation with Large Language Models
por: Yunusov, Sarfaroz, et al.
Publicado: (2024)
por: Yunusov, Sarfaroz, et al.
Publicado: (2024)
Science Out of Its Ivory Tower: Improving Accessibility with Reinforcement Learning
por: Wang, Haining, et al.
Publicado: (2024)
por: Wang, Haining, et al.
Publicado: (2024)
Raising the Bar: Investigating the Values of Large Language Models via Generative Evolving Testing
por: Jiang, Han, et al.
Publicado: (2024)
por: Jiang, Han, et al.
Publicado: (2024)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
por: Xu, Shanshan, et al.
Publicado: (2025)
por: Xu, Shanshan, et al.
Publicado: (2025)
Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations
por: Park, Eunkyu, et al.
Publicado: (2025)
por: Park, Eunkyu, et al.
Publicado: (2025)
LLM-Assisted Content Conditional Debiasing for Fair Text Embedding
por: Deng, Wenlong, et al.
Publicado: (2024)
por: Deng, Wenlong, et al.
Publicado: (2024)
SHIELD: Evaluation and Defense Strategies for Copyright Compliance in LLM Text Generation
por: Liu, Xiaoze, et al.
Publicado: (2024)
por: Liu, Xiaoze, et al.
Publicado: (2024)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
por: Xu, Xin, et al.
Publicado: (2025)
por: Xu, Xin, et al.
Publicado: (2025)
Rethinking Test-Time Scaling for Medical AI: Model and Task-Aware Strategies for LLMs and VLMs
por: Oh, Gyutaek, et al.
Publicado: (2025)
por: Oh, Gyutaek, et al.
Publicado: (2025)
Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes
por: Gallegos, Isabel O., et al.
Publicado: (2024)
por: Gallegos, Isabel O., et al.
Publicado: (2024)
AXOLOTL: Fairness through Assisted Self-Debiasing of Large Language Model Outputs
por: Ebrahimi, Sana, et al.
Publicado: (2024)
por: Ebrahimi, Sana, et al.
Publicado: (2024)
Wait! There's a Way Out: A Decision Mechanism for Forecasting Conversational Derailment
por: Kim, Laerdon, et al.
Publicado: (2026)
por: Kim, Laerdon, et al.
Publicado: (2026)
Computational Phenomenology of Temporal Experience in Autism: Quantifying the Emotional and Narrative Characteristics of Lived Unpredictability
por: Dudzic, Kacper, et al.
Publicado: (2026)
por: Dudzic, Kacper, et al.
Publicado: (2026)
A Close Reading Approach to Gender Narrative Biases in AI-Generated Stories
por: Raffini, Daniel, et al.
Publicado: (2025)
por: Raffini, Daniel, et al.
Publicado: (2025)
Metamorpheus: Interactive, Affective, and Creative Dream Narration Through Metaphorical Visual Storytelling
por: Wan, Qian, et al.
Publicado: (2024)
por: Wan, Qian, et al.
Publicado: (2024)
Developmental trajectories of decision making and affective dynamics in large language models
por: Wang, Zhihao, et al.
Publicado: (2025)
por: Wang, Zhihao, et al.
Publicado: (2025)
Narrative over Numbers: The Identifiable Victim Effect and its Amplification Under Alignment and Reasoning in Large Language Models
por: Raiyan, Syed Rifat
Publicado: (2026)
por: Raiyan, Syed Rifat
Publicado: (2026)
Planning Beyond Text: Graph-based Reasoning for Complex Narrative Generation
por: Gu, Hanwen, et al.
Publicado: (2026)
por: Gu, Hanwen, et al.
Publicado: (2026)
Responsible AI for Test Equity and Quality: The Duolingo English Test as a Case Study
por: Burstein, Jill, et al.
Publicado: (2024)
por: Burstein, Jill, et al.
Publicado: (2024)
Deep Value Benchmark: Measuring Whether Models Generalize Deep Values or Shallow Preferences
por: Ashkinaze, Joshua, et al.
Publicado: (2025)
por: Ashkinaze, Joshua, et al.
Publicado: (2025)
REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading
por: Zhao, Chengshuai, et al.
Publicado: (2026)
por: Zhao, Chengshuai, et al.
Publicado: (2026)
From General Reasoning to Domain Expertise: Uncovering the Limits of Generalization in Large Language Models
por: Alsagheer, Dana, et al.
Publicado: (2025)
por: Alsagheer, Dana, et al.
Publicado: (2025)
Causal Stories from Sensor Traces: Auditing Epistemic Overreach in LLM-Generated Personal Sensing Explanations
por: Zhu, Shanshan, et al.
Publicado: (2026)
por: Zhu, Shanshan, et al.
Publicado: (2026)
Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement
por: Kersting, Nicholas S., et al.
Publicado: (2026)
por: Kersting, Nicholas S., et al.
Publicado: (2026)
In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores
por: Tang, Zeyu, et al.
Publicado: (2026)
por: Tang, Zeyu, et al.
Publicado: (2026)
MCTSr-Zero: Self-Reflective Psychological Counseling Dialogues Generation via Principles and Adaptive Exploration
por: Lu, Hao, et al.
Publicado: (2025)
por: Lu, Hao, et al.
Publicado: (2025)
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
por: Morabito, Robert, et al.
Publicado: (2024)
por: Morabito, Robert, et al.
Publicado: (2024)
Attributions toward Artificial Agents in a modified Moral Turing Test
por: Aharoni, Eyal, et al.
Publicado: (2024)
por: Aharoni, Eyal, et al.
Publicado: (2024)
Multi-Agent Comedy Club: Investigating Community Discussion Effects on LLM Humor Generation
por: Hong, Shiwei, et al.
Publicado: (2026)
por: Hong, Shiwei, et al.
Publicado: (2026)
ML-EAT: A Multilevel Embedding Association Test for Interpretable and Transparent Social Science
por: Wolfe, Robert, et al.
Publicado: (2024)
por: Wolfe, Robert, et al.
Publicado: (2024)
Out-of-distribution Evidence-aware Fake News Detection via Dual Adversarial Debiasing
por: Liu, Qiang, et al.
Publicado: (2023)
por: Liu, Qiang, et al.
Publicado: (2023)
Societal Impacts Research Requires Benchmarks for Creative Composition Tasks
por: Shen, Judy Hanwen, et al.
Publicado: (2025)
por: Shen, Judy Hanwen, et al.
Publicado: (2025)
Towards Next-Generation Medical Agent: How o1 is Reshaping Decision-Making in Medical Scenarios
por: Xu, Shaochen, et al.
Publicado: (2024)
por: Xu, Shaochen, et al.
Publicado: (2024)
Interactive Narrative Analytics: Bridging Computational Narrative Extraction and Human Sensemaking
por: Keith, Brian
Publicado: (2026)
por: Keith, Brian
Publicado: (2026)
Ejemplares similares
-
DeCAP: Context-Adaptive Prompt Generation for Debiasing Zero-shot Question Answering in Large Language Models
por: Bae, Suyoung, et al.
Publicado: (2025) -
On the Effectiveness and Generalization of Race Representations for Debiasing High-Stakes Decisions
por: Nguyen, Dang, et al.
Publicado: (2025) -
Beyond Pairwise Comparisons: A Distributional Test of Distinctiveness for Machine-Generated Works in Intellectual Property Law
por: Mukherjee, Anirban, et al.
Publicado: (2026) -
Measuring Information Distortion in Hierarchical Ultra long Novel Reconstruction:The Optimal Expansion Ratio
por: Shen, Hanwen, et al.
Publicado: (2025) -
InsideOut: Measuring and Mitigating Insider-Outsider Bias in Interview Script Generation
por: Wan, Yixin, et al.
Publicado: (2025)