Saved in:
| Main Authors: | Faille, Juliette, Gatt, Albert, Gardent, Claire |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2409.16707 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Question Generation in Knowledge-Driven Dialog: Explainability and Evaluation
by: Faille, Juliette, et al.
Published: (2024)
by: Faille, Juliette, et al.
Published: (2024)
Evaluating Document Simplification: On the Importance of Separately Assessing Simplicity and Meaning Preservation
by: Cripwell, Liam, et al.
Published: (2024)
by: Cripwell, Liam, et al.
Published: (2024)
CV-Probes: Studying the interplay of lexical and world knowledge in visually grounded verb understanding
by: Beňová, Ivana, et al.
Published: (2024)
by: Beňová, Ivana, et al.
Published: (2024)
ModelWriter: Text & Model-Synchronized Document Engineering Platform
by: Erata, Ferhat, et al.
Published: (2024)
by: Erata, Ferhat, et al.
Published: (2024)
Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning
by: Schaffelder, Max, et al.
Published: (2025)
by: Schaffelder, Max, et al.
Published: (2025)
A Systematic Analysis of Large Language Models as Soft Reasoners: The Case of Syllogistic Inferences
by: Bertolazzi, Leonardo, et al.
Published: (2024)
by: Bertolazzi, Leonardo, et al.
Published: (2024)
Morphological Analysis for the Maltese Language: The Challenges of a Hybrid System
by: Borg, Claudia, et al.
Published: (2017)
by: Borg, Claudia, et al.
Published: (2017)
When Models Decide and When They Bind: A Two-Stage Computation for Multiple-Choice Question-Answering
by: Wong, Hugh Mee, et al.
Published: (2026)
by: Wong, Hugh Mee, et al.
Published: (2026)
Context-aware Visual Storytelling with Visual Prefix Tuning and Contrastive Learning
by: Song, Yingjin, et al.
Published: (2024)
by: Song, Yingjin, et al.
Published: (2024)
From Image Captioning to Visual Storytelling
by: Passadakis, Admitos, et al.
Published: (2025)
by: Passadakis, Admitos, et al.
Published: (2025)
How and where does CLIP process negation?
by: Quantmeyer, Vincent, et al.
Published: (2024)
by: Quantmeyer, Vincent, et al.
Published: (2024)
Grounded Misunderstandings in Asymmetric Dialogue: A Perspectivist Annotation Scheme for MapTask
by: Li, Nan, et al.
Published: (2025)
by: Li, Nan, et al.
Published: (2025)
Evaluating LLM-Generated Versus Human-Authored Responses in Role-Play Dialogues
by: Lu, Dongxu, et al.
Published: (2025)
by: Lu, Dongxu, et al.
Published: (2025)
Contrast Is All You Need
by: Kilic, Burak, et al.
Published: (2023)
by: Kilic, Burak, et al.
Published: (2023)
FTFT: Efficient and Robust Fine-Tuning by Transferring Training Dynamics
by: Du, Yupei, et al.
Published: (2023)
by: Du, Yupei, et al.
Published: (2023)
VAQUUM: Are Vague Quantifiers Grounded in Visual Data?
by: Wong, Hugh Mee, et al.
Published: (2025)
by: Wong, Hugh Mee, et al.
Published: (2025)
Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences?
by: Song, Yingjin, et al.
Published: (2025)
by: Song, Yingjin, et al.
Published: (2025)
Don't Learn, Ground: A Case for Natural Language Inference with Visual Grounding
by: Ignatev, Daniil, et al.
Published: (2025)
by: Ignatev, Daniil, et al.
Published: (2025)
Ta-G-T: Subjectivity Capture in Table to Text Generation via RDF Graphs
by: Upasham, Ronak, et al.
Published: (2025)
by: Upasham, Ronak, et al.
Published: (2025)
Common Objects Out of Context (COOCo): Investigating Multimodal Context and Semantic Scene Violations in Referential Communication
by: Merlo, Filippo, et al.
Published: (2025)
by: Merlo, Filippo, et al.
Published: (2025)
Do LLMs exhibit the same commonsense capabilities across languages?
by: Martínez-Murillo, Ivan, et al.
Published: (2025)
by: Martínez-Murillo, Ivan, et al.
Published: (2025)
Summarizing long regulatory documents with a multi-step pipeline
by: Sie, Mika, et al.
Published: (2024)
by: Sie, Mika, et al.
Published: (2024)
Extrinsically-Focused Evaluation of Omissions in Medical Summarization
by: Schumacher, Elliot, et al.
Published: (2023)
by: Schumacher, Elliot, et al.
Published: (2023)
That's Optional: A Contemporary Exploration of "that" Omission in English Subordinate Clauses
by: Rabinovich, Ella
Published: (2024)
by: Rabinovich, Ella
Published: (2024)
Reasoning About the Unsaid: Misinformation Detection with Omission-Aware Graph Inference
by: Wang, Zhengjia, et al.
Published: (2025)
by: Wang, Zhengjia, et al.
Published: (2025)
UNIQORN: Unified Question Answering over RDF Knowledge Graphs and Natural Language Text
by: Pramanik, Soumajit, et al.
Published: (2021)
by: Pramanik, Soumajit, et al.
Published: (2021)
References Matter: Investigating the Impact of Reference Set Variation on Summarization Evaluation
by: Casola, Silvia, et al.
Published: (2025)
by: Casola, Silvia, et al.
Published: (2025)
LLM Agents Implement an NLG System from Scratch: Building Interpretable Rule-Based RDF-to-Text Generators
by: Lango, Mateusz, et al.
Published: (2025)
by: Lango, Mateusz, et al.
Published: (2025)
What's in the News? Towards Identification of Bias by Commission, Omission, and Source Selection (COSS)
by: Zhukova, Anastasia, et al.
Published: (2025)
by: Zhukova, Anastasia, et al.
Published: (2025)
WEBDial, a Multi-domain, Multitask Statistical Dialogue Framework with RDF
by: Veyret, Morgan, et al.
Published: (2024)
by: Veyret, Morgan, et al.
Published: (2024)
Disentangling the Roles of Representation and Selection in Data Pruning
by: Du, Yupei, et al.
Published: (2025)
by: Du, Yupei, et al.
Published: (2025)
OTTAWA: Optimal TransporT Adaptive Word Aligner for Hallucination and Omission Translation Errors Detection
by: Huang, Chenyang, et al.
Published: (2024)
by: Huang, Chenyang, et al.
Published: (2024)
LIME-LLM: Probing Models with Fluent Counterfactuals, Not Broken Text
by: Mihaila, George, et al.
Published: (2026)
by: Mihaila, George, et al.
Published: (2026)
VALSE: A Task-Independent Benchmark for Vision and Language Models Centered on Linguistic Phenomena
by: Parcalabescu, Letitia, et al.
Published: (2021)
by: Parcalabescu, Letitia, et al.
Published: (2021)
Disentangling Prompt Element Level Risk Factors for Hallucinations and Omissions in Mental Health LLM Responses
by: Ni, Congning, et al.
Published: (2026)
by: Ni, Congning, et al.
Published: (2026)
Blending Human and LLM Expertise to Detect Hallucinations and Omissions in Mental Health Chatbot Responses
by: Hussain, Khizar, et al.
Published: (2026)
by: Hussain, Khizar, et al.
Published: (2026)
Enhancing Knowledge Graph Construction: Evaluating with Emphasis on Hallucination, Omission, and Graph Similarity Metrics
by: Ghanem, Hussam, et al.
Published: (2025)
by: Ghanem, Hussam, et al.
Published: (2025)
RDF-Based Structured Quality Assessment Representation of Multilingual LLM Evaluations
by: Gwozdz, Jonas, et al.
Published: (2025)
by: Gwozdz, Jonas, et al.
Published: (2025)
Interpretable Recognition of Cognitive Distortions in Natural Language Texts
by: Kolonin, Anton, et al.
Published: (2025)
by: Kolonin, Anton, et al.
Published: (2025)
Automatic Metrics in Natural Language Generation: A Survey of Current Evaluation Practices
by: Schmidtová, Patrícia, et al.
Published: (2024)
by: Schmidtová, Patrícia, et al.
Published: (2024)
Similar Items
-
Question Generation in Knowledge-Driven Dialog: Explainability and Evaluation
by: Faille, Juliette, et al.
Published: (2024) -
Evaluating Document Simplification: On the Importance of Separately Assessing Simplicity and Meaning Preservation
by: Cripwell, Liam, et al.
Published: (2024) -
CV-Probes: Studying the interplay of lexical and world knowledge in visually grounded verb understanding
by: Beňová, Ivana, et al.
Published: (2024) -
ModelWriter: Text & Model-Synchronized Document Engineering Platform
by: Erata, Ferhat, et al.
Published: (2024) -
Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning
by: Schaffelder, Max, et al.
Published: (2025)