What Shapes a Creative Machine Mind? Comprehensively Benchmarking Creativity in Foundation Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | He, Zicong, Zhang, Boxuan, Liu, Weihao, Tang, Ruixiang, Cheng, Lu |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Shakespearean Sparks: The Dance of Hallucination and Creativity in LLMs' Decoding Layers
par: He, Zicong, et autres
Publié: (2025)
par: He, Zicong, et autres
Publié: (2025)
CreativeBench: Benchmarking and Enhancing Machine Creativity via Self-Evolving Challenges
par: Wang, Zi-Han, et autres
Publié: (2026)
par: Wang, Zi-Han, et autres
Publié: (2026)
Weaver: Foundation Models for Creative Writing
par: Wang, Tiannan, et autres
Publié: (2024)
par: Wang, Tiannan, et autres
Publié: (2024)
CreativityPrism: A Holistic Evaluation Framework for Large Language Model Creativity
par: Hou, Zhaoyi Joey, et autres
Publié: (2025)
par: Hou, Zhaoyi Joey, et autres
Publié: (2025)
AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems
par: Zhang, Boxuan, et autres
Publié: (2026)
par: Zhang, Boxuan, et autres
Publié: (2026)
On the Creativity of Large Language Models
par: Franceschelli, Giorgio, et autres
Publié: (2023)
par: Franceschelli, Giorgio, et autres
Publié: (2023)
Assessing and Understanding Creativity in Large Language Models
par: Zhao, Yunpu, et autres
Publié: (2024)
par: Zhao, Yunpu, et autres
Publié: (2024)
Creative Preference Optimization
par: Ismayilzada, Mete, et autres
Publié: (2025)
par: Ismayilzada, Mete, et autres
Publié: (2025)
Can AI Be as Creative as Humans?
par: Wang, Haonan, et autres
Publié: (2024)
par: Wang, Haonan, et autres
Publié: (2024)
CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing
par: Qian, Cheng, et autres
Publié: (2026)
par: Qian, Cheng, et autres
Publié: (2026)
LitBench: A Benchmark and Dataset for Reliable Evaluation of Creative Writing
par: Fein, Daniel, et autres
Publié: (2025)
par: Fein, Daniel, et autres
Publié: (2025)
Is Temperature the Creativity Parameter of Large Language Models?
par: Peeperkorn, Max, et autres
Publié: (2024)
par: Peeperkorn, Max, et autres
Publié: (2024)
Divergent Creativity in Humans and Large Language Models
par: Bellemare-Pepin, Antoine, et autres
Publié: (2024)
par: Bellemare-Pepin, Antoine, et autres
Publié: (2024)
Beyond Reproduction: A Paired-Task Framework for Assessing LLM Comprehension and Creativity in Literary Translation
par: Zhang, Ran, et autres
Publié: (2026)
par: Zhang, Ran, et autres
Publié: (2026)
Creativity in AI: Progresses and Challenges
par: Ismayilzada, Mete, et autres
Publié: (2024)
par: Ismayilzada, Mete, et autres
Publié: (2024)
Evaluating Progress in Graph Foundation Models: A Comprehensive Benchmark and New Insights
par: Yu, Xingtong, et autres
Publié: (2026)
par: Yu, Xingtong, et autres
Publié: (2026)
CresOWLve: Benchmarking Creative Problem-Solving Over Real-World Knowledge
par: Ismayilzada, Mete, et autres
Publié: (2026)
par: Ismayilzada, Mete, et autres
Publié: (2026)
DPWriter: Reinforcement Learning with Diverse Planning Branching for Creative Writing
par: Cao, Qian, et autres
Publié: (2026)
par: Cao, Qian, et autres
Publié: (2026)
EscapeBench: Towards Advancing Creative Intelligence of Language Model Agents
par: Qian, Cheng, et autres
Publié: (2024)
par: Qian, Cheng, et autres
Publié: (2024)
Steering Large Language Models to Evaluate and Amplify Creativity
par: Olson, Matthew Lyle, et autres
Publié: (2024)
par: Olson, Matthew Lyle, et autres
Publié: (2024)
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning
par: Liu, Wei, et autres
Publié: (2023)
par: Liu, Wei, et autres
Publié: (2023)
OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models
par: Xu, Hainiu, et autres
Publié: (2024)
par: Xu, Hainiu, et autres
Publié: (2024)
Collective Critics for Creative Story Generation
par: Bae, Minwook, et autres
Publié: (2024)
par: Bae, Minwook, et autres
Publié: (2024)
Cooking Up Creativity: Enhancing LLM Creativity through Structured Recombination
par: Mizrahi, Moran, et autres
Publié: (2025)
par: Mizrahi, Moran, et autres
Publié: (2025)
Igniting Creative Writing in Small Language Models: LLM-as-a-Judge versus Multi-Agent Refined Rewards
par: Wei, Xiaolong, et autres
Publié: (2025)
par: Wei, Xiaolong, et autres
Publié: (2025)
RLMR: Reinforcement Learning with Mixed Rewards for Creative Writing
par: Liao, Jianxing, et autres
Publié: (2025)
par: Liao, Jianxing, et autres
Publié: (2025)
NegotiationToM: A Benchmark for Stress-testing Machine Theory of Mind on Negotiation Surrounding
par: Chan, Chunkit, et autres
Publié: (2024)
par: Chan, Chunkit, et autres
Publié: (2024)
Semantic-Driven Topic Modeling for Analyzing Creativity in Virtual Brainstorming
par: Mersha, Melkamu Abay, et autres
Publié: (2025)
par: Mersha, Melkamu Abay, et autres
Publié: (2025)
MacGyver: Are Large Language Models Creative Problem Solvers?
par: Tian, Yufei, et autres
Publié: (2023)
par: Tian, Yufei, et autres
Publié: (2023)
Creativity Has Left the Chat: The Price of Debiasing Language Models
par: Mohammadi, Behnam
Publié: (2024)
par: Mohammadi, Behnam
Publié: (2024)
Creativity Benchmark: A benchmark for marketing creativity for large language models
par: Bhat, Ninad, et autres
Publié: (2025)
par: Bhat, Ninad, et autres
Publié: (2025)
BILLY: Steering Large Language Models via Merging Persona Vectors for Creative Generation
par: Pai, Tsung-Min, et autres
Publié: (2025)
par: Pai, Tsung-Min, et autres
Publié: (2025)
Creativity or Brute Force? Using Brainteasers as a Window into the Problem-Solving Abilities of Large Language Models
par: Han, Simeng, et autres
Publié: (2025)
par: Han, Simeng, et autres
Publié: (2025)
Enhancing Creativity in Large Language Models through Associative Thinking Strategies
par: Mehrotra, Pronita, et autres
Publié: (2024)
par: Mehrotra, Pronita, et autres
Publié: (2024)
Evaluating Creative Short Story Generation in Humans and Large Language Models
par: Ismayilzada, Mete, et autres
Publié: (2024)
par: Ismayilzada, Mete, et autres
Publié: (2024)
LLM Discussion: Enhancing the Creativity of Large Language Models via Discussion Framework and Role-Play
par: Lu, Li-Chun, et autres
Publié: (2024)
par: Lu, Li-Chun, et autres
Publié: (2024)
Algebraic Quantum Intelligence: A New Framework for Reproducible Machine Creativity
par: Yano, Kazuo, et autres
Publié: (2026)
par: Yano, Kazuo, et autres
Publié: (2026)
Do LLMs Agree on the Creativity Evaluation of Alternative Uses?
par: Rabeyah, Abdullah Al, et autres
Publié: (2024)
par: Rabeyah, Abdullah Al, et autres
Publié: (2024)
Probing and Inducing Combinational Creativity in Vision-Language Models
par: Peng, Yongqian, et autres
Publié: (2025)
par: Peng, Yongqian, et autres
Publié: (2025)
Automated Creativity Evaluation for Large Language Models: A Reference-Based Approach
par: Li, Ruizhe, et autres
Publié: (2025)
par: Li, Ruizhe, et autres
Publié: (2025)
Documents similaires
-
Shakespearean Sparks: The Dance of Hallucination and Creativity in LLMs' Decoding Layers
par: He, Zicong, et autres
Publié: (2025) -
CreativeBench: Benchmarking and Enhancing Machine Creativity via Self-Evolving Challenges
par: Wang, Zi-Han, et autres
Publié: (2026) -
Weaver: Foundation Models for Creative Writing
par: Wang, Tiannan, et autres
Publié: (2024) -
CreativityPrism: A Holistic Evaluation Framework for Large Language Model Creativity
par: Hou, Zhaoyi Joey, et autres
Publié: (2025) -
AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems
par: Zhang, Boxuan, et autres
Publié: (2026)