KRISTEVA: Close Reading as a Novel Task for Benchmarking Interpretive Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Sui, Peiqi, Rodriguez, Juan Diego, Laban, Philippe, Murphy, Dean, Dexter, Joseph P., So, Richard Jean, Baker, Samuel, Chaudhuri, Pramit |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AcrosticSleuth: Probabilistic Identification and Ranking of Acrostics in Multilingual Corpora
by: Fedchin, Aleksandr, et al.
Published: (2024)
by: Fedchin, Aleksandr, et al.
Published: (2024)
Confabulation: The Surprising Value of Large Language Model Hallucinations
by: Sui, Peiqi, et al.
Published: (2024)
by: Sui, Peiqi, et al.
Published: (2024)
Critical Confabulation: Can LLMs Hallucinate for Social Good?
by: Sui, Peiqi, et al.
Published: (2025)
by: Sui, Peiqi, et al.
Published: (2025)
What Does AI Do for Cultural Interpretation? A Randomized Experiment on Close Reading Poems with Exposure to AI Interpretation
by: Zhi, Jiayin, et al.
Published: (2026)
by: Zhi, Jiayin, et al.
Published: (2026)
LLMs Exhibit Significantly Lower Uncertainty in Creative Writing Than Professional Writers
by: Sui, Peiqi
Published: (2026)
by: Sui, Peiqi
Published: (2026)
Spoiler Alert: Narrative Forecasting as a Metric for Tension in LLM Storytelling
by: Sui, Peiqi, et al.
Published: (2026)
by: Sui, Peiqi, et al.
Published: (2026)
SummExecEdit: A Factual Consistency Benchmark in Summarization with Executable Edits
by: Thorat, Onkar, et al.
Published: (2024)
by: Thorat, Onkar, et al.
Published: (2024)
Generative AI & Fictionality: How Novels Power Large Language Models
by: Roland, Edwin, et al.
Published: (2026)
by: Roland, Edwin, et al.
Published: (2026)
Hybrid star within $f(\mathcal{G})$ gravity
by: Rej, Pramit
Published: (2024)
by: Rej, Pramit
Published: (2024)
Measuring Fractal Dimension using Discrete Global Grid Systems
by: Ghosh, Pramit
Published: (2025)
by: Ghosh, Pramit
Published: (2025)
Charged analog of anisotropic dark energy star in Rastall gravity
by: Rej, Pramit
Published: (2024)
by: Rej, Pramit
Published: (2024)
MMTU: A Massive Multi-Task Table Understanding and Reasoning Benchmark
by: Xing, Junjie, et al.
Published: (2025)
by: Xing, Junjie, et al.
Published: (2025)
LLMs Corrupt Your Documents When You Delegate
by: Laban, Philippe, et al.
Published: (2026)
by: Laban, Philippe, et al.
Published: (2026)
MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents
by: Tang, Liyan, et al.
Published: (2024)
by: Tang, Liyan, et al.
Published: (2024)
INDIC DIALECT: A Multi Task Benchmark to Evaluate and Translate in Indian Language Dialects
by: Sharma, Tarun, et al.
Published: (2026)
by: Sharma, Tarun, et al.
Published: (2026)
La tiranía de la mala teoría económica
by: Dean Baker
Published: (2009)
by: Dean Baker
Published: (2009)
La "Nueva Economía" de EEUU. Alguien a quien culpar cuando la próxima burbuja estalle
by: Dean Baker
Published: (2009)
by: Dean Baker
Published: (2009)
Beyond Vision: How Large Language Models Interpret Facial Expressions from Valence-Arousal Values
by: Mehra, Vaibhav, et al.
Published: (2025)
by: Mehra, Vaibhav, et al.
Published: (2025)
LEXI: Large Language Models Experimentation Interface
by: Laban, Guy, et al.
Published: (2024)
by: Laban, Guy, et al.
Published: (2024)
NLIP_Lab-IITH Multilingual MT System for WAT24 MT Shared Task
by: Brahma, Maharaj, et al.
Published: (2024)
by: Brahma, Maharaj, et al.
Published: (2024)
La inversión con reformas increíbles / Raúl Laban, Holger C. Wolf
by: Laban, Raúl
by: Laban, Raúl
NLIP_Lab-IITH Low-Resource MT System for WMT24 Indic MT Shared Task
by: Sahoo, Pramit, et al.
Published: (2024)
by: Sahoo, Pramit, et al.
Published: (2024)
AI-Slop to AI-Polish? Aligning Language Models through Edit-Based Writing Rewards and Test-time Computation
by: Chakrabarty, Tuhin, et al.
Published: (2025)
by: Chakrabarty, Tuhin, et al.
Published: (2025)
Can AI writing be salvaged? Mitigating Idiosyncrasies and Improving Human-AI Alignment in the Writing Process through Edits
by: Chakrabarty, Tuhin, et al.
Published: (2024)
by: Chakrabarty, Tuhin, et al.
Published: (2024)
Jane Austen’s Novels: A Study from Feminist Perspective
by: Anuradha Chaudhuri
Published: (2021)
by: Anuradha Chaudhuri
Published: (2021)
Weighted Support Points from Random Measures: An Interpretable Alternative for Generative Modeling
by: Zhao, Peiqi, et al.
Published: (2025)
by: Zhao, Peiqi, et al.
Published: (2025)
Flipping the Dialogue: Training and Evaluating User Language Models
by: Naous, Tarek, et al.
Published: (2025)
by: Naous, Tarek, et al.
Published: (2025)
Attribution Gradients: Incrementally Unfolding Citations for Critical Examination of Attributed AI Answers
by: Kambhamettu, Hita, et al.
Published: (2025)
by: Kambhamettu, Hita, et al.
Published: (2025)
LLMs Get Lost In Multi-Turn Conversation
by: Laban, Philippe, et al.
Published: (2025)
by: Laban, Philippe, et al.
Published: (2025)
Supplementary Data 1
by: Wan, Peiqi
Published: (2025)
by: Wan, Peiqi
Published: (2025)
Uncertainty Quantification Via the Posterior Predictive Variance
by: Chaudhuri, Sanjay, et al.
Published: (2026)
by: Chaudhuri, Sanjay, et al.
Published: (2026)
Man or Monster?
by: Hinton, Alexander Laban
Published: (2017)
by: Hinton, Alexander Laban
Published: (2017)
One Model for Teaching Research Skills.
by: Laban, Lawrence F.
Published: (1977)
by: Laban, Lawrence F.
Published: (1977)
AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions
by: Kirichenko, Polina, et al.
Published: (2025)
by: Kirichenko, Polina, et al.
Published: (2025)
Lexical and Statistical Analysis of Bangla Newspaper and Literature: A Corpus-Driven Study on Diversity, Readability, and NLP Adaptation
by: Bhattacharyya, Pramit, et al.
Published: (2025)
by: Bhattacharyya, Pramit, et al.
Published: (2025)
Isotropic Buchdahl's relativistic fluid sphere within $f(R,\,T)$ gravity
by: Bhar, Piyali, et al.
Published: (2022)
by: Bhar, Piyali, et al.
Published: (2022)
Video-based Exercise Classification and Activated Muscle Group Prediction with Hybrid X3D-SlowFast Network
by: Pasula, Manvik, et al.
Published: (2024)
by: Pasula, Manvik, et al.
Published: (2024)
Model of hybrid star with baryonic and strange quark matter in Tolman-Kuchowicz spacetime
by: Rej, Pramit, et al.
Published: (2022)
by: Rej, Pramit, et al.
Published: (2022)
Well behaved class of Heintzmann's solution within $f(R,\,T)$ framework
by: Rej, Pramit, et al.
Published: (2024)
by: Rej, Pramit, et al.
Published: (2024)
Investigating Deep Learning Models for Ejection Fraction Estimation from Echocardiography Videos
by: Saranyan, Shravan, et al.
Published: (2025)
by: Saranyan, Shravan, et al.
Published: (2025)
Similar Items
-
AcrosticSleuth: Probabilistic Identification and Ranking of Acrostics in Multilingual Corpora
by: Fedchin, Aleksandr, et al.
Published: (2024) -
Confabulation: The Surprising Value of Large Language Model Hallucinations
by: Sui, Peiqi, et al.
Published: (2024) -
Critical Confabulation: Can LLMs Hallucinate for Social Good?
by: Sui, Peiqi, et al.
Published: (2025) -
What Does AI Do for Cultural Interpretation? A Randomized Experiment on Close Reading Poems with Exposure to AI Interpretation
by: Zhi, Jiayin, et al.
Published: (2026) -
LLMs Exhibit Significantly Lower Uncertainty in Creative Writing Than Professional Writers
by: Sui, Peiqi
Published: (2026)