Training data generation for context-dependent rubric-based short answer grading
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Šindelář, Pavel, Slivka, Dávid, Bouma, Christopher, Prášil, Filip, Bojar, Ondřej |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Overview of the Sensemaking Task at the ELOQUENT 2025 Lab: LLMs as Teachers, Students and Evaluators
von: Šindelář, Pavel, et al.
Veröffentlicht: (2025)
von: Šindelář, Pavel, et al.
Veröffentlicht: (2025)
Finetuning LLMs for EvaCun 2025 token prediction shared task
von: Jon, Josef, et al.
Veröffentlicht: (2025)
von: Jon, Josef, et al.
Veröffentlicht: (2025)
Intrinsic vs. Extrinsic Evaluation of Czech Sentence Embeddings: Semantic Relevance Doesn't Help with MT Evaluation
von: Barančíková, Petra, et al.
Veröffentlicht: (2025)
von: Barančíková, Petra, et al.
Veröffentlicht: (2025)
Understanding the role of FFNs in driving multilingual behaviour in LLMs
von: Bhattacharya, Sunit, et al.
Veröffentlicht: (2024)
von: Bhattacharya, Sunit, et al.
Veröffentlicht: (2024)
Quality and Quantity of Machine Translation References for Automatic Metrics
von: Zouhar, Vilém, et al.
Veröffentlicht: (2024)
von: Zouhar, Vilém, et al.
Veröffentlicht: (2024)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
von: Luu, Nam, et al.
Veröffentlicht: (2025)
von: Luu, Nam, et al.
Veröffentlicht: (2025)
Continuous Rating as Reliable Human Evaluation of Simultaneous Speech Translation
von: Javorský, Dávid, et al.
Veröffentlicht: (2022)
von: Javorský, Dávid, et al.
Veröffentlicht: (2022)
Prompting LLMs: Length Control for Isometric Machine Translation
von: Javorský, Dávid, et al.
Veröffentlicht: (2025)
von: Javorský, Dávid, et al.
Veröffentlicht: (2025)
MockConf: A Student Interpretation Dataset: Analysis, Word- and Span-level Alignment and Baselines
von: Javorský, Dávid, et al.
Veröffentlicht: (2025)
von: Javorský, Dávid, et al.
Veröffentlicht: (2025)
ChatGPT for automated grading of short answer questions in mechanical ventilation
von: Jade, Tejas, et al.
Veröffentlicht: (2025)
von: Jade, Tejas, et al.
Veröffentlicht: (2025)
ParCzech4Speech: A New Speech Corpus Derived from Czech Parliamentary Data
von: Stankov, Vladislav, et al.
Veröffentlicht: (2025)
von: Stankov, Vladislav, et al.
Veröffentlicht: (2025)
Multimodal Shannon Game with Images
von: Zouhar, Vilém, et al.
Veröffentlicht: (2023)
von: Zouhar, Vilém, et al.
Veröffentlicht: (2023)
Long-Form End-to-End Speech Translation via Latent Alignment Segmentation
von: Polák, Peter, et al.
Veröffentlicht: (2023)
von: Polák, Peter, et al.
Veröffentlicht: (2023)
Evaluating Optimal Reference Translations
von: Zouhar, Vilém, et al.
Veröffentlicht: (2023)
von: Zouhar, Vilém, et al.
Veröffentlicht: (2023)
Better Late Than Never: Meta-Evaluation of Latency Metrics for Simultaneous Speech-to-Text Translation
von: Polák, Peter, et al.
Veröffentlicht: (2025)
von: Polák, Peter, et al.
Veröffentlicht: (2025)
Corpus of Cross-lingual Dialogues with Minutes and Detection of Misunderstandings
von: Čechovič, Marko, et al.
Veröffentlicht: (2025)
von: Čechovič, Marko, et al.
Veröffentlicht: (2025)
Findings of the Third Automatic Minuting (AutoMin) Challenge
von: Shinde, Kartik, et al.
Veröffentlicht: (2025)
von: Shinde, Kartik, et al.
Veröffentlicht: (2025)
Ratas framework: A comprehensive genai-based approach to rubric-based marking of real-world textual exams
von: Safilian, Masoud, et al.
Veröffentlicht: (2025)
von: Safilian, Masoud, et al.
Veröffentlicht: (2025)
How "Real" is Your Real-Time Simultaneous Speech-to-Text Translation System?
von: Papi, Sara, et al.
Veröffentlicht: (2024)
von: Papi, Sara, et al.
Veröffentlicht: (2024)
FusionMind -- Improving question and answering with external context fusion
von: Verma, Shreyas, et al.
Veröffentlicht: (2023)
von: Verma, Shreyas, et al.
Veröffentlicht: (2023)
ConSens: Assessing context grounding in open-book question answering
von: Vankov, Ivan, et al.
Veröffentlicht: (2025)
von: Vankov, Ivan, et al.
Veröffentlicht: (2025)
Czech Dataset for Complex Aspect-Based Sentiment Analysis Tasks
von: Šmíd, Jakub, et al.
Veröffentlicht: (2025)
von: Šmíd, Jakub, et al.
Veröffentlicht: (2025)
CLIPPER: Compression enables long-context synthetic data generation
von: Pham, Chau Minh, et al.
Veröffentlicht: (2025)
von: Pham, Chau Minh, et al.
Veröffentlicht: (2025)
Extract, Match, and Score: An Evaluation Paradigm for Long Question-context-answer Triplets in Financial Analysis
von: Hu, Bo, et al.
Veröffentlicht: (2025)
von: Hu, Bo, et al.
Veröffentlicht: (2025)
Retrieval augmented text-to-SQL generation for epidemiological question answering using electronic health records
von: Ziletti, Angelo, et al.
Veröffentlicht: (2024)
von: Ziletti, Angelo, et al.
Veröffentlicht: (2024)
How effective are VLMs in assisting humans in inferring the quality of mental models from Multimodal short answers?
von: Sil, Pritam, et al.
Veröffentlicht: (2026)
von: Sil, Pritam, et al.
Veröffentlicht: (2026)
PIAST: Rapid Prompting with In-context Augmentation for Scarce Training data
von: Batorski, Pawel, et al.
Veröffentlicht: (2025)
von: Batorski, Pawel, et al.
Veröffentlicht: (2025)
LLM Compression: How Far Can We Go in Balancing Size and Performance?
von: Sk, Sahil, et al.
Veröffentlicht: (2025)
von: Sk, Sahil, et al.
Veröffentlicht: (2025)
SRS-Stories: Vocabulary-constrained multilingual story generation for language learning
von: Kamzela, Wiktor, et al.
Veröffentlicht: (2025)
von: Kamzela, Wiktor, et al.
Veröffentlicht: (2025)
Evaluating the IWSLT2023 Speech Translation Tasks: Human Annotations, Automatic Metrics, and Segmentation
von: Sperber, Matthias, et al.
Veröffentlicht: (2024)
von: Sperber, Matthias, et al.
Veröffentlicht: (2024)
From text to multimodal: a survey of adversarial example generation in question answering systems
von: Yigit, Gulsum, et al.
Veröffentlicht: (2023)
von: Yigit, Gulsum, et al.
Veröffentlicht: (2023)
Enhancing textual textbook question answering with large language models and retrieval augmented generation
von: Alawwad, Hessa Abdulrahman, et al.
Veröffentlicht: (2024)
von: Alawwad, Hessa Abdulrahman, et al.
Veröffentlicht: (2024)
CMRAG: Co-modality-based visual document retrieval and question answering
von: Chen, Wang, et al.
Veröffentlicht: (2025)
von: Chen, Wang, et al.
Veröffentlicht: (2025)
Exploring Multiple Strategies to Improve Multilingual Coreference Resolution in CorefUD
von: Pražák, Ondřej, et al.
Veröffentlicht: (2024)
von: Pražák, Ondřej, et al.
Veröffentlicht: (2024)
factgenie: A Framework for Span-based Evaluation of Generated Texts
von: Kasner, Zdeněk, et al.
Veröffentlicht: (2024)
von: Kasner, Zdeněk, et al.
Veröffentlicht: (2024)
MEEDAV: A Synchronous Web Viewer for EEG, Eye-Tracking and Speech Data
von: Pijálek, Jan, et al.
Veröffentlicht: (2026)
von: Pijálek, Jan, et al.
Veröffentlicht: (2026)
What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering
von: Maar, Jim, et al.
Veröffentlicht: (2026)
von: Maar, Jim, et al.
Veröffentlicht: (2026)
LADM: Long-context Training Data Selection with Attention-based Dependency Measurement for LLMs
von: Chen, Jianghao, et al.
Veröffentlicht: (2025)
von: Chen, Jianghao, et al.
Veröffentlicht: (2025)
TANQ: An open domain dataset of table answered questions
von: Akhtar, Mubashara, et al.
Veröffentlicht: (2024)
von: Akhtar, Mubashara, et al.
Veröffentlicht: (2024)
A dependently-typed calculus of event telicity and culminativity
von: Kovalev, Pavel, et al.
Veröffentlicht: (2025)
von: Kovalev, Pavel, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Overview of the Sensemaking Task at the ELOQUENT 2025 Lab: LLMs as Teachers, Students and Evaluators
von: Šindelář, Pavel, et al.
Veröffentlicht: (2025) -
Finetuning LLMs for EvaCun 2025 token prediction shared task
von: Jon, Josef, et al.
Veröffentlicht: (2025) -
Intrinsic vs. Extrinsic Evaluation of Czech Sentence Embeddings: Semantic Relevance Doesn't Help with MT Evaluation
von: Barančíková, Petra, et al.
Veröffentlicht: (2025) -
Understanding the role of FFNs in driving multilingual behaviour in LLMs
von: Bhattacharya, Sunit, et al.
Veröffentlicht: (2024) -
Quality and Quantity of Machine Translation References for Automatic Metrics
von: Zouhar, Vilém, et al.
Veröffentlicht: (2024)