Guardado en:
| Autores principales: | Zhu, Xiaomeng, Zhou, Zhenghao, Charlow, Simon, Frank, Robert |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2502.14119 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LIEDER: Linguistically-Informed Evaluation for Discourse Entity Recognition
por: Zhu, Xiaomeng, et al.
Publicado: (2024)
por: Zhu, Xiaomeng, et al.
Publicado: (2024)
What Exactly do Children Receive in Language Acquisition? A Case Study on CHILDES with Automated Detection of Filler-Gap Dependencies
por: Zhou, Zhenghao Herbert, et al.
Publicado: (2026)
por: Zhou, Zhenghao Herbert, et al.
Publicado: (2026)
Effect-driven interpretation: Functors for natural language composition
por: Bumford, Dylan, et al.
Publicado: (2025)
por: Bumford, Dylan, et al.
Publicado: (2025)
The Structural Sources of Verb Meaning Revisited: Large Language Models Display Syntactic Bootstrapping
por: Zhu, Xiaomeng, et al.
Publicado: (2025)
por: Zhu, Xiaomeng, et al.
Publicado: (2025)
Towards Generating Automatic Anaphora Annotations
por: Taji, Dima, et al.
Publicado: (2025)
por: Taji, Dima, et al.
Publicado: (2025)
Evaluating Prompt-Based and Fine-Tuned Approaches to Czech Anaphora Resolution
por: Stano, Patrik, et al.
Publicado: (2025)
por: Stano, Patrik, et al.
Publicado: (2025)
Subjectivity in the Annotation of Bridging Anaphora
por: Levine, Lauren, et al.
Publicado: (2025)
por: Levine, Lauren, et al.
Publicado: (2025)
Truth Knows No Language: Evaluating Truthfulness Beyond English
por: Figueras, Blanca Calvo, et al.
Publicado: (2025)
por: Figueras, Blanca Calvo, et al.
Publicado: (2025)
GUMBridge: a Corpus for Varieties of Bridging Anaphora
por: Levine, Lauren, et al.
Publicado: (2025)
por: Levine, Lauren, et al.
Publicado: (2025)
The Detection and Understanding of Fictional Discourse
por: Piper, Andrew, et al.
Publicado: (2024)
por: Piper, Andrew, et al.
Publicado: (2024)
Evaluating LLM Understanding via Structured Tabular Decision Simulations
por: Li, Sichao, et al.
Publicado: (2025)
por: Li, Sichao, et al.
Publicado: (2025)
Is In-Context Learning a Type of Error-Driven Learning? Evidence from the Inverse Frequency Effect in Structural Priming
por: Zhou, Zhenghao, et al.
Publicado: (2024)
por: Zhou, Zhenghao, et al.
Publicado: (2024)
Anaphora Resolution and Text Retrieval
por: Schmolz, Helene
Publicado: (2020)
por: Schmolz, Helene
Publicado: (2020)
Causal Interventions on Continuous Variables: A Case Study on Verb Bias in Steering Vectors for In-Context Learning
por: Zhou, Zhenghao Herbert, et al.
Publicado: (2026)
por: Zhou, Zhenghao Herbert, et al.
Publicado: (2026)
A Multi-Level Benchmark for Causal Language Understanding in Social Media Discourse
por: Ding, Xiaohan, et al.
Publicado: (2025)
por: Ding, Xiaohan, et al.
Publicado: (2025)
Beyond Single Ground Truth: Reference Monism as Epistemic Injustice in ASR Evaluation
por: Choi, Anna Seo Gyeong, et al.
Publicado: (2026)
por: Choi, Anna Seo Gyeong, et al.
Publicado: (2026)
Enhancing Dialogue Systems with Discourse-Level Understanding Using Deep Canonical Correlation Analysis
por: Mehndiratta, Akanksha, et al.
Publicado: (2025)
por: Mehndiratta, Akanksha, et al.
Publicado: (2025)
Understanding the Effects of Iterative Prompting on Truthfulness
por: Krishna, Satyapriya, et al.
Publicado: (2024)
por: Krishna, Satyapriya, et al.
Publicado: (2024)
Unifying the Scope of Bridging Anaphora Types in English: Bridging Annotations in ARRAU and GUM
por: Levine, Lauren, et al.
Publicado: (2024)
por: Levine, Lauren, et al.
Publicado: (2024)
Understanding Risk and Dependency in AI Chatbot Use from User Discourse
por: Zhu, Jianfeng, et al.
Publicado: (2026)
por: Zhu, Jianfeng, et al.
Publicado: (2026)
From Ground Trust to Truth: Disparities in Offensive Language Judgments on Contemporary Korean Political Discourse
por: Yu, Seunguk, et al.
Publicado: (2025)
por: Yu, Seunguk, et al.
Publicado: (2025)
ImpScore: A Learnable Metric For Quantifying The Implicitness Level of Sentence
por: Wang, Yuxin, et al.
Publicado: (2024)
por: Wang, Yuxin, et al.
Publicado: (2024)
Truth Neurons
por: Li, Haohang, et al.
Publicado: (2025)
por: Li, Haohang, et al.
Publicado: (2025)
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
por: Khatun, Aisha, et al.
Publicado: (2024)
por: Khatun, Aisha, et al.
Publicado: (2024)
ChatGPT Evaluation on Sentence Level Relations: A Focus on Temporal, Causal, and Discourse Relations
por: Chan, Chunkit, et al.
Publicado: (2023)
por: Chan, Chunkit, et al.
Publicado: (2023)
Findings of the WMT 2024 Shared Task on Discourse-Level Literary Translation
por: Wang, Longyue, et al.
Publicado: (2024)
por: Wang, Longyue, et al.
Publicado: (2024)
Measuring Hong Kong Massive Multi-Task Language Understanding
por: Cao, Chuxue, et al.
Publicado: (2025)
por: Cao, Chuxue, et al.
Publicado: (2025)
Competition-Level Problems are Effective LLM Evaluators
por: Huang, Yiming, et al.
Publicado: (2023)
por: Huang, Yiming, et al.
Publicado: (2023)
Beyond Screenshots: Evaluating VLMs' Understanding of UI Animations
por: Liang, Chen, et al.
Publicado: (2026)
por: Liang, Chen, et al.
Publicado: (2026)
Improving Dialogue Discourse Parsing through Discourse-aware Utterance Clarification
por: Fan, Yaxin, et al.
Publicado: (2025)
por: Fan, Yaxin, et al.
Publicado: (2025)
Discourse Features Enhance Detection of Document-Level Machine-Generated Content
por: Li, Yupei, et al.
Publicado: (2024)
por: Li, Yupei, et al.
Publicado: (2024)
TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
por: Wei, Zhepei, et al.
Publicado: (2025)
por: Wei, Zhepei, et al.
Publicado: (2025)
Show, Don't Tell: Evaluating Large Language Models Beyond Textual Understanding with ChildPlay
por: de Carvalho, Gonçalo Hora, et al.
Publicado: (2024)
por: de Carvalho, Gonçalo Hora, et al.
Publicado: (2024)
KatotohananQA: Evaluating Truthfulness of Large Language Models in Filipino
por: Nery, Lorenzo Alfred, et al.
Publicado: (2025)
por: Nery, Lorenzo Alfred, et al.
Publicado: (2025)
Evaluating Discourse Cohesion in Pre-trained Language Models
por: He, Jie, et al.
Publicado: (2025)
por: He, Jie, et al.
Publicado: (2025)
Fine-Grained Evaluation for Implicit Discourse Relation Recognition
por: Cai, Xinyi
Publicado: (2025)
por: Cai, Xinyi
Publicado: (2025)
Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark
por: Turetzky, Arnon, et al.
Publicado: (2026)
por: Turetzky, Arnon, et al.
Publicado: (2026)
TruthFlow: Truthful LLM Generation via Representation Flow Correction
por: Wang, Hanyu, et al.
Publicado: (2025)
por: Wang, Hanyu, et al.
Publicado: (2025)
Understanding Emotion in Discourse: Recognition Insights and Linguistic Patterns for Generation
por: Jeong, Cheonkam, et al.
Publicado: (2026)
por: Jeong, Cheonkam, et al.
Publicado: (2026)
BeDiscovER: The Benchmark of Discourse Understanding in the Era of Reasoning Language Models
por: Li, Chuyuan, et al.
Publicado: (2025)
por: Li, Chuyuan, et al.
Publicado: (2025)
Ejemplares similares
-
LIEDER: Linguistically-Informed Evaluation for Discourse Entity Recognition
por: Zhu, Xiaomeng, et al.
Publicado: (2024) -
What Exactly do Children Receive in Language Acquisition? A Case Study on CHILDES with Automated Detection of Filler-Gap Dependencies
por: Zhou, Zhenghao Herbert, et al.
Publicado: (2026) -
Effect-driven interpretation: Functors for natural language composition
por: Bumford, Dylan, et al.
Publicado: (2025) -
The Structural Sources of Verb Meaning Revisited: Large Language Models Display Syntactic Bootstrapping
por: Zhu, Xiaomeng, et al.
Publicado: (2025) -
Towards Generating Automatic Anaphora Annotations
por: Taji, Dima, et al.
Publicado: (2025)