Towards Better Open-Ended Text Generation: A Multicriteria Evaluation Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Arias, Esteban Garces, Blocher, Hannah, Rodemann, Julian, Li, Meimingwei, Heumann, Christian, Aßenmacher, Matthias |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Contrastive Search: Uncertainty-Guided Decoding for Open-Ended Text Generation
by: Arias, Esteban Garces, et al.
Published: (2024)
by: Arias, Esteban Garces, et al.
Published: (2024)
Decoding Decoded: Understanding Hyperparameter Effects in Open-Ended Text Generation
by: Arias, Esteban Garces, et al.
Published: (2024)
by: Arias, Esteban Garces, et al.
Published: (2024)
Statistical Multicriteria Evaluation of LLM-Generated Text
by: Arias, Esteban Garces, et al.
Published: (2025)
by: Arias, Esteban Garces, et al.
Published: (2025)
GUARD: Glocal Uncertainty-Aware Robust Decoding for Effective and Efficient Open-Ended Text Generation
by: Ding, Yuanhao, et al.
Published: (2025)
by: Ding, Yuanhao, et al.
Published: (2025)
Min-$k$ Sampling: Decoupling Truncation from Temperature Scaling via Relative Logit Dynamics
by: Ding, Yuanhao, et al.
Published: (2026)
by: Ding, Yuanhao, et al.
Published: (2026)
Unveiling Factors for Enhanced POS Tagging: A Study of Low-Resource Medieval Romance Languages
by: Schöffel, Matthias, et al.
Published: (2025)
by: Schöffel, Matthias, et al.
Published: (2025)
The Truncation Blind Spot: How Decoding Strategies Systematically Exclude Human-Like Token Choices
by: Arias, Esteban Garces, et al.
Published: (2026)
by: Arias, Esteban Garces, et al.
Published: (2026)
Beyond Temperature: Hyperfitting as a Late-Stage Geometric Expansion
by: Li, Meimingwei, et al.
Published: (2026)
by: Li, Meimingwei, et al.
Published: (2026)
Can OpenSource beat ChatGPT? -- A Comparative Study of Large Language Models for Text-to-Code Generation
by: Mayer, Luis, et al.
Published: (2024)
by: Mayer, Luis, et al.
Published: (2024)
Partial Rankings of Optimizers
by: Rodemann, Julian, et al.
Published: (2024)
by: Rodemann, Julian, et al.
Published: (2024)
The Geometry of Creative Variability: How Credal Sets Expose Calibration Gaps in Language Models
by: Arias, Esteban Garces, et al.
Published: (2025)
by: Arias, Esteban Garces, et al.
Published: (2025)
Statistical Multicriteria Benchmarking via the GSD-Front
by: Jansen, Christoph, et al.
Published: (2024)
by: Jansen, Christoph, et al.
Published: (2024)
Modern Models, Medieval Texts: A POS Tagging Study of Old Occitan
by: Schöffel, Matthias, et al.
Published: (2025)
by: Schöffel, Matthias, et al.
Published: (2025)
A Statistical Case Against Empirical Human-AI Alignment
by: Rodemann, Julian, et al.
Published: (2025)
by: Rodemann, Julian, et al.
Published: (2025)
How Prevalent is Gender Bias in ChatGPT? -- Exploring German and English ChatGPT Responses
by: Urchs, Stefanie, et al.
Published: (2023)
by: Urchs, Stefanie, et al.
Published: (2023)
A Bayesian approach to modeling topic-metadata relationships
by: Schulze, P., et al.
Published: (2021)
by: Schulze, P., et al.
Published: (2021)
Robust Statistical Comparison of Random Variables with Locally Varying Scale of Measurement
by: Jansen, Christoph, et al.
Published: (2023)
by: Jansen, Christoph, et al.
Published: (2023)
A Semantic-Sampling Framework for Evaluating Calibration in Open-Ended Question Answering
by: Wang, Zhanliang, et al.
Published: (2026)
by: Wang, Zhanliang, et al.
Published: (2026)
CausalEvolve: Towards Open-Ended Discovery with Causal Scratchpad
by: Chen, Yongqiang, et al.
Published: (2026)
by: Chen, Yongqiang, et al.
Published: (2026)
Towards Bayesian Data Selection
by: Rodemann, Julian
Published: (2024)
by: Rodemann, Julian
Published: (2024)
Self-Reinforcing Controllable Synthesis of Rare Relational Data via Bayesian Calibration
by: Zhang, Chongsheng, et al.
Published: (2026)
by: Zhang, Chongsheng, et al.
Published: (2026)
Reinforcement Learning for Latent-Space Thinking in LLMs
by: Özeren, Enes, et al.
Published: (2025)
by: Özeren, Enes, et al.
Published: (2025)
ShinkaEvolve: Towards Open-Ended And Sample-Efficient Program Evolution
by: Lange, Robert Tjarko, et al.
Published: (2025)
by: Lange, Robert Tjarko, et al.
Published: (2025)
MCU: An Evaluation Framework for Open-Ended Game Agents
by: Zheng, Xinyue, et al.
Published: (2023)
by: Zheng, Xinyue, et al.
Published: (2023)
XL-Suite: Cross-Lingual Synthetic Training and Evaluation Data for Open-Ended Generation
by: Iyer, Vivek, et al.
Published: (2025)
by: Iyer, Vivek, et al.
Published: (2025)
Fair Play in the Newsroom: Actor-Based Filtering Gender Discrimination in Text Corpora
by: Urchs, Stefanie, et al.
Published: (2025)
by: Urchs, Stefanie, et al.
Published: (2025)
Semantically-Aware Rewards for Open-Ended R1 Training in Free-Form Generation
by: Li, Zongxia, et al.
Published: (2025)
by: Li, Zongxia, et al.
Published: (2025)
Scaling Open-Ended Reasoning to Predict the Future
by: Chandak, Nikhil, et al.
Published: (2025)
by: Chandak, Nikhil, et al.
Published: (2025)
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
by: Lu, Chris, et al.
Published: (2024)
by: Lu, Chris, et al.
Published: (2024)
MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning Research
by: Chen, Hui, et al.
Published: (2025)
by: Chen, Hui, et al.
Published: (2025)
Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts
by: Samvelyan, Mikayel, et al.
Published: (2024)
by: Samvelyan, Mikayel, et al.
Published: (2024)
InfoQuest: Evaluating Multi-Turn Dialogue Agents for Open-Ended Conversations with Hidden Context
by: de Oliveira, Bryan L. M., et al.
Published: (2025)
by: de Oliveira, Bryan L. M., et al.
Published: (2025)
REAL Sampling: Boosting Factuality and Diversity of Open-Ended Generation via Asymptotic Entropy
by: Chang, Haw-Shiuan, et al.
Published: (2024)
by: Chang, Haw-Shiuan, et al.
Published: (2024)
GRLO: Towards Generalizable Reinforcement Learning in Open-Ended Environments from Zero
by: Yin, Shangjian, et al.
Published: (2026)
by: Yin, Shangjian, et al.
Published: (2026)
G-Zero: Self-Play for Open-Ended Generation from Zero Data
by: Huang, Chengsong, et al.
Published: (2026)
by: Huang, Chengsong, et al.
Published: (2026)
Dreaming in Code for Curriculum Learning in Open-Ended Worlds
by: Mitsides, Konstantinos, et al.
Published: (2026)
by: Mitsides, Konstantinos, et al.
Published: (2026)
Lost in Translation? Exploring the Shift in Grammatical Gender from Latin to Occitan
by: Chatterjee, Ahan, et al.
Published: (2026)
by: Chatterjee, Ahan, et al.
Published: (2026)
Generalization Bounds and Stopping Rules for Learning with Self-Selected Data
by: Rodemann, Julian, et al.
Published: (2025)
by: Rodemann, Julian, et al.
Published: (2025)
Self-Rewarding Rubric-Based Reinforcement Learning for Open-Ended Reasoning
by: Ye, Zhiling, et al.
Published: (2025)
by: Ye, Zhiling, et al.
Published: (2025)
DIVERGE: Diversity-Enhanced RAG for Open-Ended Information Seeking
by: Hu, Tianyi, et al.
Published: (2026)
by: Hu, Tianyi, et al.
Published: (2026)
Similar Items
-
Adaptive Contrastive Search: Uncertainty-Guided Decoding for Open-Ended Text Generation
by: Arias, Esteban Garces, et al.
Published: (2024) -
Decoding Decoded: Understanding Hyperparameter Effects in Open-Ended Text Generation
by: Arias, Esteban Garces, et al.
Published: (2024) -
Statistical Multicriteria Evaluation of LLM-Generated Text
by: Arias, Esteban Garces, et al.
Published: (2025) -
GUARD: Glocal Uncertainty-Aware Robust Decoding for Effective and Efficient Open-Ended Text Generation
by: Ding, Yuanhao, et al.
Published: (2025) -
Min-$k$ Sampling: Decoupling Truncation from Temperature Scaling via Relative Logit Dynamics
by: Ding, Yuanhao, et al.
Published: (2026)