Is the Top Still Spinning? Evaluating Subjectivity in Narrative Understanding
Fuente:
arXiv
Salvato in:
| Autori principali: | Subbiah, Melanie, Mishra, Akankshya, Kim, Grace, Tang, Liyan, Durrett, Greg, McKeown, Kathleen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
STORYSUMM: Evaluating Faithfulness in Story Summarization
di: Subbiah, Melanie, et al.
Pubblicazione: (2024)
di: Subbiah, Melanie, et al.
Pubblicazione: (2024)
Counterfactual Simulatability of LLM Explanations for Generation Tasks
di: Limpijankit, Marvin, et al.
Pubblicazione: (2025)
di: Limpijankit, Marvin, et al.
Pubblicazione: (2025)
MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents
di: Tang, Liyan, et al.
Pubblicazione: (2024)
di: Tang, Liyan, et al.
Pubblicazione: (2024)
Enhancing Multimodal Affective Analysis with Learned Live Comment Features
di: Deng, Zhaoyuan, et al.
Pubblicazione: (2024)
di: Deng, Zhaoyuan, et al.
Pubblicazione: (2024)
Reading Subtext: Evaluating Large Language Models on Short Story Summarization with Writers
di: Subbiah, Melanie, et al.
Pubblicazione: (2024)
di: Subbiah, Melanie, et al.
Pubblicazione: (2024)
See It from My Perspective: How Language Affects Cultural Bias in Image Understanding
di: Ananthram, Amith, et al.
Pubblicazione: (2024)
di: Ananthram, Amith, et al.
Pubblicazione: (2024)
Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives
di: Subbiah, Melanie, et al.
Pubblicazione: (2026)
di: Subbiah, Melanie, et al.
Pubblicazione: (2026)
Molecular Facts: Desiderata for Decontextualization in LLM Fact Verification
di: Gunjal, Anisha, et al.
Pubblicazione: (2024)
di: Gunjal, Anisha, et al.
Pubblicazione: (2024)
ParaGuide: Guided Diffusion Paraphrasers for Plug-and-Play Textual Style Transfer
di: Horvitz, Zachary, et al.
Pubblicazione: (2023)
di: Horvitz, Zachary, et al.
Pubblicazione: (2023)
Computational Representations of Character Significance in Novels
di: Mian, Haaris, et al.
Pubblicazione: (2026)
di: Mian, Haaris, et al.
Pubblicazione: (2026)
Parallel Structures in Pre-training Data Yield In-Context Learning
di: Chen, Yanda, et al.
Pubblicazione: (2024)
di: Chen, Yanda, et al.
Pubblicazione: (2024)
On the Relation between Sensitivity and Accuracy in In-context Learning
di: Chen, Yanda, et al.
Pubblicazione: (2022)
di: Chen, Yanda, et al.
Pubblicazione: (2022)
SynthesizRR: Generating Diverse Datasets with Retrieval Augmentation
di: Divekar, Abhishek, et al.
Pubblicazione: (2024)
di: Divekar, Abhishek, et al.
Pubblicazione: (2024)
AdvSumm: Adversarial Training for Bias Mitigation in Text Summarization
di: Gupta, Mukur, et al.
Pubblicazione: (2025)
di: Gupta, Mukur, et al.
Pubblicazione: (2025)
Artificial Impressions: Evaluating Large Language Model Behavior Through the Lens of Trait Impressions
di: Deas, Nicholas, et al.
Pubblicazione: (2025)
di: Deas, Nicholas, et al.
Pubblicazione: (2025)
Calibrate-Then-Act: Cost-Aware Exploration in LLM Agents
di: Ding, Wenxuan, et al.
Pubblicazione: (2026)
di: Ding, Wenxuan, et al.
Pubblicazione: (2026)
TofuEval: Evaluating Hallucinations of LLMs on Topic-Focused Dialogue Summarization
di: Tang, Liyan, et al.
Pubblicazione: (2024)
di: Tang, Liyan, et al.
Pubblicazione: (2024)
Summarization of Opinionated Political Documents with Varied Perspectives
di: Deas, Nicholas, et al.
Pubblicazione: (2024)
di: Deas, Nicholas, et al.
Pubblicazione: (2024)
Getting Serious about Humor: Crafting Humor Datasets with Unfunny Large Language Models
di: Horvitz, Zachary, et al.
Pubblicazione: (2024)
di: Horvitz, Zachary, et al.
Pubblicazione: (2024)
Mining Contextualized Visual Associations from Images for Creativity Understanding
di: Sahu, Ananya, et al.
Pubblicazione: (2025)
di: Sahu, Ananya, et al.
Pubblicazione: (2025)
PropMEND: Hypernetworks for Knowledge Propagation in LLMs
di: Liu, Zeyu Leo, et al.
Pubblicazione: (2025)
di: Liu, Zeyu Leo, et al.
Pubblicazione: (2025)
Adaptive Margin RLHF via Preference over Preferences
di: Chittepu, Yaswanth, et al.
Pubblicazione: (2025)
di: Chittepu, Yaswanth, et al.
Pubblicazione: (2025)
Contrastive Learning to Improve Retrieval for Real-world Fact Checking
di: Sriram, Aniruddh, et al.
Pubblicazione: (2024)
di: Sriram, Aniruddh, et al.
Pubblicazione: (2024)
Learning Composable Chains-of-Thought
di: Yin, Fangcong, et al.
Pubblicazione: (2025)
di: Yin, Fangcong, et al.
Pubblicazione: (2025)
ModelingAgent: Bridging LLMs and Mathematical Modeling for Real-World Challenges
di: Qian, Cheng, et al.
Pubblicazione: (2025)
di: Qian, Cheng, et al.
Pubblicazione: (2025)
SkillFactory: Self-Distillation For Learning Cognitive Behaviors
di: Sprague, Zayne, et al.
Pubblicazione: (2025)
di: Sprague, Zayne, et al.
Pubblicazione: (2025)
LiveNewsBench: Evaluating LLM Web Search Capabilities with Freshly Curated News
di: Zhang, Yunfan, et al.
Pubblicazione: (2026)
di: Zhang, Yunfan, et al.
Pubblicazione: (2026)
MemGuard: Preventing Memory Contamination in Long-Term Memory-Augmented Large Language Models
di: Ha, Hyeonjeong, et al.
Pubblicazione: (2026)
di: Ha, Hyeonjeong, et al.
Pubblicazione: (2026)
Forecasting Conversation Derailments Through Generation
di: Zhang, Yunfan, et al.
Pubblicazione: (2025)
di: Zhang, Yunfan, et al.
Pubblicazione: (2025)
Reranking-based Generation for Unbiased Perspective Summarization
di: Ri, Narutatsu, et al.
Pubblicazione: (2025)
di: Ri, Narutatsu, et al.
Pubblicazione: (2025)
iBERT: Interpretable Embeddings via Sense Decomposition
di: Anand, Vishal, et al.
Pubblicazione: (2025)
di: Anand, Vishal, et al.
Pubblicazione: (2025)
Northeastern Uni at Multilingual Counterspeech Generation: Enhancing Counter Speech Generation with LLM Alignment through Direct Preference Optimization
di: Wadhwa, Sahil, et al.
Pubblicazione: (2024)
di: Wadhwa, Sahil, et al.
Pubblicazione: (2024)
ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving
di: Thakur, Amitayush, et al.
Pubblicazione: (2025)
di: Thakur, Amitayush, et al.
Pubblicazione: (2025)
Exploring Chain-of-Thought Reasoning for Steerable Pluralistic Alignment
di: Zhang, Yunfan, et al.
Pubblicazione: (2025)
di: Zhang, Yunfan, et al.
Pubblicazione: (2025)
PoSh: Using Scene Graphs To Guide LLMs-as-a-Judge For Detailed Image Descriptions
di: Ananthram, Amith, et al.
Pubblicazione: (2025)
di: Ananthram, Amith, et al.
Pubblicazione: (2025)
Still Between Us? Evaluating and Improving Voice Assistant Robustness to Third-Party Interruptions
di: Lee, Dongwook, et al.
Pubblicazione: (2026)
di: Lee, Dongwook, et al.
Pubblicazione: (2026)
LLMs as Science Journalists: Supporting Early-stage Researchers in Communicating Their Science to the Public
di: Alshomary, Milad, et al.
Pubblicazione: (2026)
di: Alshomary, Milad, et al.
Pubblicazione: (2026)
Still Not Quite There! Evaluating Large Language Models for Comorbid Mental Health Diagnosis
di: Hengle, Amey, et al.
Pubblicazione: (2024)
di: Hengle, Amey, et al.
Pubblicazione: (2024)
Still Fresh? Evaluating Temporal Drift in Retrieval Benchmarks
di: Kuissi, Nathan, et al.
Pubblicazione: (2026)
di: Kuissi, Nathan, et al.
Pubblicazione: (2026)
Understand the Implication: Learning to Think for Pragmatic Understanding
di: Sravanthi, Settaluri Lakshmi, et al.
Pubblicazione: (2025)
di: Sravanthi, Settaluri Lakshmi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
STORYSUMM: Evaluating Faithfulness in Story Summarization
di: Subbiah, Melanie, et al.
Pubblicazione: (2024) -
Counterfactual Simulatability of LLM Explanations for Generation Tasks
di: Limpijankit, Marvin, et al.
Pubblicazione: (2025) -
MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents
di: Tang, Liyan, et al.
Pubblicazione: (2024) -
Enhancing Multimodal Affective Analysis with Learned Live Comment Features
di: Deng, Zhaoyuan, et al.
Pubblicazione: (2024) -
Reading Subtext: Evaluating Large Language Models on Short Story Summarization with Writers
di: Subbiah, Melanie, et al.
Pubblicazione: (2024)