Free-text Rationale Generation under Readability Level Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hsu, Yi-Sheng, Feldhus, Nils, Hakimov, Sherzod |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TurkicNLP: An NLP Toolkit for Turkic Languages
von: Hakimov, Sherzod
Veröffentlicht: (2026)
von: Hakimov, Sherzod
Veröffentlicht: (2026)
Retrieval-Augmented Code Generation for Situated Action Generation: A Case Study on Minecraft
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2024)
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2024)
Unveiling Global Narratives: A Multilingual Twitter Dataset of News Media on the Russo-Ukrainian Conflict
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2023)
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2023)
From Templates to Natural Language: Generalization Challenges in Instruction-Tuned LLMs for Spatial Reasoning
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2025)
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2025)
M2SA: Multimodal and Multilingual Model for Sentiment Analysis of Tweets
von: Thakkar, Gaurish, et al.
Veröffentlicht: (2024)
von: Thakkar, Gaurish, et al.
Veröffentlicht: (2024)
Learning Communication Policies for Different Follower Behaviors in a Collaborative Reference Game
von: Sadler, Philipp, et al.
Veröffentlicht: (2024)
von: Sadler, Philipp, et al.
Veröffentlicht: (2024)
Ad-hoc Concept Forming in the Game Codenames as a Means for Evaluating Large Language Models
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2025)
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2025)
clem:todd: A Framework for the Systematic Benchmarking of LLM-Based Task-Oriented Dialogue System Realisations
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2025)
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2025)
Plant in Cupboard, Orange on Rably, Inat Aphone. Benchmarking Incremental Learning of Situation and Language Model using a Text-Simulated Situated Environment
von: Jordan, Jonathan, et al.
Veröffentlicht: (2025)
von: Jordan, Jonathan, et al.
Veröffentlicht: (2025)
Towards No-Code Programming of Cobots: Experiments with Code Synthesis by Large Code Models for Conversational Programming
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2024)
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2024)
Multi-Turn Multi-Agent Dialogue for Collaborative Reconstruction Improves VLM Performance on Spatial Reasoning, But Only Barely
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2026)
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2026)
Sharing the Cost of Success: A Game for Evaluating and Learning Collaborative Multi-Agent Instruction Giving and Following Policies
von: Sadler, Philipp, et al.
Veröffentlicht: (2024)
von: Sadler, Philipp, et al.
Veröffentlicht: (2024)
How Many Parameters Does it Take to Change a Light Bulb? Evaluating Performance in Self-Play of Conversational Games as a Function of Model Characteristics
von: Bhavsar, Nidhir, et al.
Veröffentlicht: (2024)
von: Bhavsar, Nidhir, et al.
Veröffentlicht: (2024)
What Are We Measuring in NLG? A Meta-Analysis of Evaluation Trends 2020-2025
von: Yang, Jing, et al.
Veröffentlicht: (2026)
von: Yang, Jing, et al.
Veröffentlicht: (2026)
Interpreting Language Models Through Concept Descriptions: A Survey
von: Feldhus, Nils, et al.
Veröffentlicht: (2025)
von: Feldhus, Nils, et al.
Veröffentlicht: (2025)
A Third Paradigm for LLM Evaluation: Dialogue Game-Based Evaluation using clembench
von: Schlangen, David, et al.
Veröffentlicht: (2025)
von: Schlangen, David, et al.
Veröffentlicht: (2025)
The Image Reconstruction Game: Drawing Common Ground Through Iterative Multimodal Dialogue
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2026)
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2026)
Simplifying Outcomes of Language Model Component Analyses with ELIA
von: Eidt, Aaron Louis, et al.
Veröffentlicht: (2026)
von: Eidt, Aaron Louis, et al.
Veröffentlicht: (2026)
iFlip: Iterative Feedback-driven Counterfactual Example Refinement
von: Wang, Yilong, et al.
Veröffentlicht: (2026)
von: Wang, Yilong, et al.
Veröffentlicht: (2026)
Proceedings of the ISCA/ITG Workshop on Diversity in Large Speech and Language Models
von: Möller, Sebastian, et al.
Veröffentlicht: (2025)
von: Möller, Sebastian, et al.
Veröffentlicht: (2025)
clembench-2024: A Challenging, Dynamic, Complementary, Multilingual Benchmark and Underlying Flexible Framework for LLMs as Multi-Action Agents
von: Beyer, Anne, et al.
Veröffentlicht: (2024)
von: Beyer, Anne, et al.
Veröffentlicht: (2024)
Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulness under Optimization
von: Sun, Jingyi, et al.
Veröffentlicht: (2026)
von: Sun, Jingyi, et al.
Veröffentlicht: (2026)
Infherno: End-to-end Agent-based FHIR Resource Synthesis from Free-form Clinical Notes
von: Frei, Johann, et al.
Veröffentlicht: (2025)
von: Frei, Johann, et al.
Veröffentlicht: (2025)
CoXQL: A Dataset for Parsing Explanation Requests in Conversational XAI Systems
von: Wang, Qianli, et al.
Veröffentlicht: (2024)
von: Wang, Qianli, et al.
Veröffentlicht: (2024)
Generating Educational Materials with Different Levels of Readability using LLMs
von: Huang, Chieh-Yang, et al.
Veröffentlicht: (2024)
von: Huang, Chieh-Yang, et al.
Veröffentlicht: (2024)
Using Game Play to Investigate Multimodal and Conversational Grounding in Large Multimodal Models
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2024)
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2024)
Cross-Refine: Improving Natural Language Explanation Generation by Learning in Tandem
von: Wang, Qianli, et al.
Veröffentlicht: (2024)
von: Wang, Qianli, et al.
Veröffentlicht: (2024)
Gender Bias in Explainability: Investigating Performance Disparity in Post-hoc Methods
von: Dhaini, Mahdi, et al.
Veröffentlicht: (2025)
von: Dhaini, Mahdi, et al.
Veröffentlicht: (2025)
RORA: Robust Free-Text Rationale Evaluation
von: Jiang, Zhengping, et al.
Veröffentlicht: (2024)
von: Jiang, Zhengping, et al.
Veröffentlicht: (2024)
The Price of Thought: A Multilingual Analysis of Reasoning, Performance, and Cost of Negotiation in Large Language Models
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2025)
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2025)
FitCF: A Framework for Automatic Feature Importance-guided Counterfactual Example Generation
von: Wang, Qianli, et al.
Veröffentlicht: (2025)
von: Wang, Qianli, et al.
Veröffentlicht: (2025)
Rationale Behind Essay Scores: Enhancing S-LLM's Multi-Trait Essay Scoring with Rationale Generated by LLMs
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024)
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024)
Persona Prompting as a Lens on LLM Social Reasoning
von: Yang, Jing, et al.
Veröffentlicht: (2026)
von: Yang, Jing, et al.
Veröffentlicht: (2026)
Persuasiveness of Generated Free-Text Rationales in Subjective Decisions: A Case Study on Pairwise Argument Ranking
von: Elaraby, Mohamed, et al.
Veröffentlicht: (2024)
von: Elaraby, Mohamed, et al.
Veröffentlicht: (2024)
Analysing Zero-Shot Readability-Controlled Sentence Simplification
von: Barayan, Abdullah, et al.
Veröffentlicht: (2024)
von: Barayan, Abdullah, et al.
Veröffentlicht: (2024)
Zero-shot Large Language Models for Automatic Readability Assessment
von: Grossman, Riley, et al.
Veröffentlicht: (2026)
von: Grossman, Riley, et al.
Veröffentlicht: (2026)
Analyzing Feedback Mechanisms in AI-Generated MCQs: Insights into Readability, Lexical Properties, and Levels of Challenge
von: Yaacoub, Antoun, et al.
Veröffentlicht: (2025)
von: Yaacoub, Antoun, et al.
Veröffentlicht: (2025)
Rationales Are Not Silver Bullets: Measuring the Impact of Rationales on Model Performance and Reliability
von: Zhu, Chiwei, et al.
Veröffentlicht: (2025)
von: Zhu, Chiwei, et al.
Veröffentlicht: (2025)
Explaining Russian-German code-mixing
von: Hakimov, Nikolay
Veröffentlicht: (2022)
von: Hakimov, Nikolay
Veröffentlicht: (2022)
Readability Reconsidered: A Cross-Dataset Analysis of Reference-Free Metrics
von: Belem, Catarina G, et al.
Veröffentlicht: (2025)
von: Belem, Catarina G, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TurkicNLP: An NLP Toolkit for Turkic Languages
von: Hakimov, Sherzod
Veröffentlicht: (2026) -
Retrieval-Augmented Code Generation for Situated Action Generation: A Case Study on Minecraft
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2024) -
Unveiling Global Narratives: A Multilingual Twitter Dataset of News Media on the Russo-Ukrainian Conflict
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2023) -
From Templates to Natural Language: Generalization Challenges in Instruction-Tuned LLMs for Spatial Reasoning
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2025) -
M2SA: Multimodal and Multilingual Model for Sentiment Analysis of Tweets
von: Thakkar, Gaurish, et al.
Veröffentlicht: (2024)