Measuring LLM Novelty As The Frontier Of Original And High-Quality Output
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Padmakumar, Vishakh, Yueh-Han, Chen, Pan, Jane, Chen, Valerie, He, He |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Does Writing with Language Models Reduce Content Diversity?
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2023)
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2023)
Beyond the Binary: Capturing Diverse Preferences With Reward Regularization
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2024)
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2024)
Evaluating the Diversity and Quality of LLM Generated Content
von: Shypula, Alexander, et al.
Veröffentlicht: (2025)
von: Shypula, Alexander, et al.
Veröffentlicht: (2025)
Principled Content Selection to Generate Diverse and Personalized Multi-Document Summaries
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2025)
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2025)
Creativity Support in the Age of Large Language Models: An Empirical Study Involving Emerging Writers
von: Chakrabarty, Tuhin, et al.
Veröffentlicht: (2023)
von: Chakrabarty, Tuhin, et al.
Veröffentlicht: (2023)
No Single Best Model for Diversity: Learning a Router for Sample Diversity
von: Liu, Yuhan, et al.
Veröffentlicht: (2026)
von: Liu, Yuhan, et al.
Veröffentlicht: (2026)
Offloading Score: Measuring AI Reliance Through Counterfactual Workflows
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2026)
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2026)
Two Failures of Self-Consistency in the Multi-Step Reasoning of LLMs
von: Chen, Angelica, et al.
Veröffentlicht: (2023)
von: Chen, Angelica, et al.
Veröffentlicht: (2023)
Intent-Aware Schema Generation And Refinement For Literature Review Tables
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2025)
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2025)
Modifying Large Language Model Post-Training for Diverse Creative Writing
von: Chung, John Joon Young, et al.
Veröffentlicht: (2025)
von: Chung, John Joon Young, et al.
Veröffentlicht: (2025)
Whose Boat Does it Float? Improving Personalization in Preference Tuning via Inferred User Personas
von: Balepur, Nishant, et al.
Veröffentlicht: (2025)
von: Balepur, Nishant, et al.
Veröffentlicht: (2025)
SparkMe: Adaptive Semi-Structured Interviewing for Qualitative Insight Discovery
von: Anugraha, David, et al.
Veröffentlicht: (2026)
von: Anugraha, David, et al.
Veröffentlicht: (2026)
Transformers Struggle to Learn to Search
von: Saparov, Abulhair, et al.
Veröffentlicht: (2024)
von: Saparov, Abulhair, et al.
Veröffentlicht: (2024)
Control the Temperature: Selective Sampling for Diverse and High-Quality LLM Outputs
von: Troshin, Sergey, et al.
Veröffentlicht: (2025)
von: Troshin, Sergey, et al.
Veröffentlicht: (2025)
Minimum Tuning to Unlock Long Output from LLMs with High Quality Data as the Key
von: Chen, Yingda, et al.
Veröffentlicht: (2024)
von: Chen, Yingda, et al.
Veröffentlicht: (2024)
AgentFrontier: Expanding the Capability Frontier of LLM Agents with ZPD-Guided Data Synthesis
von: Chen, Xuanzhong, et al.
Veröffentlicht: (2025)
von: Chen, Xuanzhong, et al.
Veröffentlicht: (2025)
NoveltyRank: A Retrieval-Augmented Framework for Conceptual Novelty Estimation in AI Research
von: Yan, Zhengxu, et al.
Veröffentlicht: (2025)
von: Yan, Zhengxu, et al.
Veröffentlicht: (2025)
OpenNovelty: An LLM-powered Agentic System for Verifiable Scholarly Novelty Assessment
von: Zhang, Ming, et al.
Veröffentlicht: (2026)
von: Zhang, Ming, et al.
Veröffentlicht: (2026)
Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units
von: Chen, Jianhui, et al.
Veröffentlicht: (2026)
von: Chen, Jianhui, et al.
Veröffentlicht: (2026)
Perspective Dial: Measuring Perspective of Text and Guiding LLM Outputs
von: Kim, Taejin, et al.
Veröffentlicht: (2025)
von: Kim, Taejin, et al.
Veröffentlicht: (2025)
LLM as a Scorer: The Impact of Output Order on Dialogue Evaluation
von: Chen, Yi-Pei, et al.
Veröffentlicht: (2024)
von: Chen, Yi-Pei, et al.
Veröffentlicht: (2024)
LiteraryTaste: A Preference Dataset for Creative Writing Personalization
von: Chung, John Joon Young, et al.
Veröffentlicht: (2025)
von: Chung, John Joon Young, et al.
Veröffentlicht: (2025)
Cross-Context Review: Improving LLM Output Quality by Separating Production and Review Sessions
von: Song, Tae-Eun
Veröffentlicht: (2026)
von: Song, Tae-Eun
Veröffentlicht: (2026)
Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification
von: Zhang, Anqi, et al.
Veröffentlicht: (2025)
von: Zhang, Anqi, et al.
Veröffentlicht: (2025)
HardTests: Synthesizing High-Quality Test Cases for LLM Coding
von: He, Zhongmou, et al.
Veröffentlicht: (2025)
von: He, Zhongmou, et al.
Veröffentlicht: (2025)
Nova: An Iterative Planning and Search Approach to Enhance Novelty and Diversity of LLM Generated Ideas
von: Hu, Xiang, et al.
Veröffentlicht: (2024)
von: Hu, Xiang, et al.
Veröffentlicht: (2024)
GraphMind: Interactive Novelty Assessment System for Accelerating Scientific Discovery
von: da Silva, Italo Luis, et al.
Veröffentlicht: (2025)
von: da Silva, Italo Luis, et al.
Veröffentlicht: (2025)
Ensuring Safe and High-Quality Outputs: A Guideline Library Approach for Language Models
von: Luo, Yi, et al.
Veröffentlicht: (2024)
von: Luo, Yi, et al.
Veröffentlicht: (2024)
Novelty-based Tree-of-Thought Search for LLM Reasoning and Planning
von: Hamm, Leon, et al.
Veröffentlicht: (2026)
von: Hamm, Leon, et al.
Veröffentlicht: (2026)
Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement
von: Kersting, Nicholas S., et al.
Veröffentlicht: (2026)
von: Kersting, Nicholas S., et al.
Veröffentlicht: (2026)
Spontaneous Reward Hacking in Iterative Self-Refinement
von: Pan, Jane, et al.
Veröffentlicht: (2024)
von: Pan, Jane, et al.
Veröffentlicht: (2024)
A Survey on LLM-based Multi-Agent System: Recent Advances and New Frontiers in Application
von: Chen, Shuaihang, et al.
Veröffentlicht: (2024)
von: Chen, Shuaihang, et al.
Veröffentlicht: (2024)
Benchmarking LLM-as-a-Judge for Long-Form Output Evaluation
von: Chen, Junjie, et al.
Veröffentlicht: (2026)
von: Chen, Junjie, et al.
Veröffentlicht: (2026)
ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests
von: Xu, Shiyi, et al.
Veröffentlicht: (2025)
von: Xu, Shiyi, et al.
Veröffentlicht: (2025)
A Better LLM Evaluator for Text Generation: The Impact of Prompt Output Sequencing and Optimization
von: Chu, KuanChao, et al.
Veröffentlicht: (2024)
von: Chu, KuanChao, et al.
Veröffentlicht: (2024)
MiCEval: Unveiling Multimodal Chain of Thought's Quality via Image Description and Reasoning Steps
von: Zhou, Xiongtao, et al.
Veröffentlicht: (2024)
von: Zhou, Xiongtao, et al.
Veröffentlicht: (2024)
FANNO: Augmenting High-Quality Instruction Data with Open-Sourced LLMs Only
von: Zhu, He, et al.
Veröffentlicht: (2024)
von: Zhu, He, et al.
Veröffentlicht: (2024)
UCSC at SemEval-2025 Task 3: Context, Models and Prompt Optimization for Automated Hallucination Detection in LLM Output
von: Huang, Sicong, et al.
Veröffentlicht: (2025)
von: Huang, Sicong, et al.
Veröffentlicht: (2025)
A Content-Based Novelty Measure for Scholarly Publications: A Proof of Concept
von: Wang, Haining
Veröffentlicht: (2024)
von: Wang, Haining
Veröffentlicht: (2024)
STED and Consistency Scoring: A Framework for Evaluating LLM Structured Output Reliability
von: Wang, Guanghui, et al.
Veröffentlicht: (2025)
von: Wang, Guanghui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Does Writing with Language Models Reduce Content Diversity?
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2023) -
Beyond the Binary: Capturing Diverse Preferences With Reward Regularization
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2024) -
Evaluating the Diversity and Quality of LLM Generated Content
von: Shypula, Alexander, et al.
Veröffentlicht: (2025) -
Principled Content Selection to Generate Diverse and Personalized Multi-Document Summaries
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2025) -
Creativity Support in the Age of Large Language Models: An Empirical Study Involving Emerging Writers
von: Chakrabarty, Tuhin, et al.
Veröffentlicht: (2023)