Sample theme detail

Sign in to create your own trend reports and explore all themes

Sign In

LLM Creativity Evaluation and Generation

Theme #2
35 papers 17% of analysis Jul 2025 - Nov 2025

Scoping Review: LLM Creativity Evaluation and Generation

Overview

This research theme examines how to measure and enhance creative capabilities in large language models (LLMs)—AI systems trained on vast amounts of text data. The field addresses a fundamental challenge: while LLMs can generate fluent and coherent text, researchers are working to understand whether this output is truly creative or simply recombination of familiar patterns, and how to evaluate and improve creative performance across diverse domains like writing, science, and advertising.

Research Landscape

Knowledge Gaps

  • Theoretical Grounding of Creativity Metrics: While multiple evaluation frameworks exist, there is limited consensus on which metrics best capture genuine creativity versus surface-level novelty, and how different metrics relate to human creative judgment across cultures and contexts.

  • Cross-Domain Creativity Consistency: Most research focuses on isolated domains (writing, science, marketing), leaving unclear how creative capabilities transfer across different types of tasks and whether improvements in one domain benefit others.

  • Process-Level Understanding: The field lacks deep investigation into how LLMs generate creative outputs—the internal reasoning and decision-making processes—compared to extensive focus on evaluating final outputs.

  • Personalization and Subjective Creativity: Limited research addresses how to evaluate creativity relative to individual preferences and cultural contexts, as most benchmarks assume universal definitions of what constitutes creative work.

Methodological Approaches

Future Directions

  • Unified Creativity Framework: Developing a standardized, theoretically-grounded creativity evaluation system that works across domains and languages could enable more meaningful comparisons between models and clearer progress tracking in the field.

  • Real-World Creative Collaboration: Research should explore how LLMs can effectively collaborate with human creators in practical settings (design, research, marketing), moving beyond benchmarks to understand actual creative value and user satisfaction.

  • Interpretability of Creative Processes: Investigating what makes certain model architectures, training approaches, or prompting strategies more conducive to creativity could reveal actionable insights for improving creative capabilities at scale.

Papers in this Theme (35)

LLMscape

Gottfried Haider, Jie Zhang Nov 2025 2511.07161v1
View →

Large Language Models for Scientific Idea Generation: A Creativity-Centered Survey

Fatemeh Shahhosseini, Arash Marioriyad, Ali Momen, Mahdieh Soleymani Baghshah, Mohammad Hossein Rohban, Shaghayegh Haghjooy Javanmard Nov 2025 2511.07448v1
View →

Generating Creative Chess Puzzles

Xidong Feng, Vivek Veeriah, Marcus Chiam, Michael Dennis, Ryan Pachauri, Thomas Tumiel, Federico Barbero, Johan Obando-Ceron, Jiaxin Shi, Satinder Singh, Shaobo Hou, Nenad Tomašev, Tom Zahavy Oct 2025 2510.23881v1
View →

Magellan: Guided MCTS for Latent Space Exploration and Novelty Generation

Lufan Chang Oct 2025 2510.21341v1
View →

A computational model and tool for generating more novel opportunities in professional innovation processes

Neil Maiden, Konstantinos Zachos, James Lockerbie, Kostas Petrianakis, Amanda Brown Oct 2025 2510.20402v1
View →

CreativityPrism: A Holistic Benchmark for Large Language Model Creativity

Zhaoyi Joey Hou, Bowei Alvin Zhang, Yining Lu, Bhiman Kumar Baghel, Anneliese Brei, Ximing Lu, Meng Jiang, Faeze Brahman, Snigdha Chaturvedi, Haw-Shiuan Chang, Daniel Khashabi, Xiang Lorraine Li Oct 2025 2510.20091v1
View →

Cultural Alien Sampler: Open-ended art generation balancing originality and coherence

Alejandro H. Artiles, Hiromu Yakura, Levin Brinkmann, Mar Canet Sola, Hassan Abu Alhaija, Ignacio Serna, Nasim Rahaman, Bernhard Schölkopf, Iyad Rahwan Oct 2025 2510.20849v1
View →

CLAWS:Creativity detection for LLM-generated solutions using Attention Window of Sections

Keuntae Kim, Eunhye Jeong, Sehyeon Lee, Seohee Yoon, Yong Suk Choi Oct 2025 2510.17921v1
View →

Automated Composition of Agents: A Knapsack Approach for Agentic Component Selection

Michelle Yuan, Khushbu Pahwa, Shuaichen Chang, Mustafa Kaba, Jiarong Jiang, Xiaofei Ma, Yi Zhang, Monica Sunkara Oct 2025 2510.16499v1
View →

HypoSpace: Evaluating LLM Creativity as Set-Valued Hypothesis Generators under Underdetermination

Tingting Chen, Beibei Lin, Zifeng Yuan, Qiran Zou, Hongyu He, Yew-Soon Ong, Anirudh Goyal, Dianbo Liu Oct 2025 2510.15614v1
View →

COIG-Writer: A High-Quality Dataset for Chinese Creative Writing with Thought Processes

Yunwen Li, Shuangshuang Ying, Xingwei Qu, Xin Li, Sheng Jin, Minghao Liu, Zhoufutu Wen, Tianyu Zheng, Xeron Du, Qiguang Chen, Jiajun Shi, Wangchunshu Zhou, Jiazhan Feng, Wanjun Zhong, Libo Qin, Stephen Huang, Wanxiang Che, Chenghua Lin, Eli Zhang Oct 2025 2510.14763v1
View →

Deep Associations, High Creativity: A Simple yet Effective Metric for Evaluating Large Language Models

Ziliang Qiu, Renfen Hu Oct 2025 2510.12110v1
View →

Confidence, Not Perplexity: A Better Metric for the Creative Era of LLMs

V. S. Raghu Parupudi Oct 2025 2510.08596v1
View →

What Shapes a Creative Machine Mind? Comprehensively Benchmarking Creativity in Foundation Models

Zicong He, Boxuan Zhang, Weihao Liu, Ruixiang Tang, Lu Cheng Oct 2025 2510.04009v1
View →

Algorithm Generation via Creative Ideation

Ruiying Ma, Chieh-Jan Mike Liang, Yanjie Gao, Francis Y. Yan Oct 2025 2510.03851v1
View →

Style Over Story: A Process-Oriented Study of Authorial Creativity in Large Language Models

Donghoon Jung, Jiwoo Choi, Songeun Chae, Seohyon Jung Oct 2025 2510.02025v1
View →

Complex System Exploration with Interactive Human Guidance

Bastien Morel, Clément Moulin-Frier, Pascal Barla Oct 2025 2510.00794v1
View →

Curiosity-Driven LLM-as-a-judge for Personalized Creative Judgment

Vanya Bannihatti Kumar, Divyanshu Goyal, Akhil Eppa, Neel Bhandari Oct 2025 2510.05135v1
View →

CreAgentive: An Agent Workflow Driven Multi-Category Creative Generation Engine

Yuyang Cheng, Linyue Cai, Changwei Peng, Yumiao Xu, Rongfang Bie, Yong Zhao Sep 2025 2509.26461v1
View →

Galton's Law of Mediocrity: Why Large Language Models Regress to the Mean and Fail at Creativity in Advertising

Matt Keon, Aabid Karim, Bhoomika Lohana, Abdul Karim, Thai Nguyen, Tara Hamilton, Ali Abbas Sep 2025 2509.25767v1
View →

The Geometry of Creative Variability: How Credal Sets Expose Calibration Gaps in Language Models

Esteban Garces Arias, Julian Rodemann, Christian Heumann Sep 2025 2509.23088v1
View →

Death of the Novel(ty): Beyond n-Gram Novelty as a Metric for Textual Creativity

Arkadiy Saakyan, Najoung Kim, Smaranda Muresan, Tuhin Chakrabarty Sep 2025 2509.22641v1
View →

Finding your MUSE: Mining Unexpected Solutions Engine

Nir Sweed, Hanit Hakim, Ben Wolfson, Hila Lifshitz, Dafna Shahaf Sep 2025 2509.05072v1
View →

Creativity Benchmark: A benchmark for marketing creativity for large language models

Ninad Bhat, Kieran Browne, Pip Bingemann Sep 2025 2509.09702v2
View →

AMCR: A Framework for Assessing and Mitigating Copyright Risks in Generative Models

Zhipeng Yin, Zichong Wang, Avash Palikhe, Zhen Liu, Jun Liu, Wenbin Zhang Aug 2025 2509.00641v1
View →

Adaptive Originality Filtering: Rejection Based Prompting and RiddleScore for Culturally Grounded Multilingual Riddle Generation

Duy Le, Kent Ziti, Evan Girard-Sun, Bakr Bouhaya, Sean O'Brien, Vasu Sharma, Kevin Zhu Aug 2025 2508.18709v3
View →

Spacer: Towards Engineered Scientific Inspiration

Minhyeong Lee, Suyoung Hwang, Seunghyun Moon, Geonho Nah, Donghyun Koh, Youngjun Cho, Johyun Park, Hojin Yoo, Jiho Park, Haneul Choi, Sungbin Moon, Taehoon Hwang, Seungwon Kim, Jaeyeong Kim, Seongjun Kim, Juneau Jung Aug 2025 2508.17661v1
View →

Generative Modeling with Multi-Instance Reward Learning for E-commerce Creative Optimization

Qiaolei Gu, Yu Li, DingYi Zeng, Lu Wang, Ming Pang, Changping Peng, Zhangang Lin, Ching Law, Jingping Shao Aug 2025 2508.09730v1
View →

Rethinking Creativity Evaluation: A Critical Analysis of Existing Creativity Evaluations

Li-Chun Lu, Miri Liu, Pin-Chun Lu, Yufei Tian, Shao-Hua Sun, Nanyun Peng Aug 2025 2508.05470v2
View →

SMART-Editor: A Multi-Agent Framework for Human-Like Design Editing with Structural Integrity

Ishani Mondal, Meera Bharadwaj, Ayush Roy, Aparna Garimella, Jordan Lee Boyd-Graber Jul 2025 2507.23095v2
View →

Mining Contextualized Visual Associations from Images for Creativity Understanding

Ananya Sahu, Amith Ananthram, Kathleen McKeown Jul 2025 2507.18915v1
View →

From Seed to Harvest: Augmenting Human Creativity with AI for Red-teaming Text-to-Image Models

Jessica Quaye, Charvi Rastogi, Alicia Parrish, Oana Inel, Minsuk Kahng, Lora Aroyo, Vijay Janapa Reddi Jul 2025 2507.17922v1
View →

Multi-Agent Synergy-Driven Iterative Visual Narrative Synthesis

Wang Xi, Quan Shi, Tian Yu, Yujie Peng, Jiayi Sun, Mengxing Ren, Zenghui Ding, Ningguang Yao Jul 2025 2507.13285v1
View →

A Comparative Approach to Assessing Linguistic Creativity of Large Language Models and Humans

Anca Dinu, Andra-Maria Florescu, Alina Resceanu Jul 2025 2507.12039v2
View →

Participatory Evolution of Artificial Life Systems via Semantic Feedback

Shuowen Li, Kexin Wang, Minglu Fang, Danqi Huang, Ali Asadipour, Haipeng Mi, Yitong Sun Jul 2025 2507.03839v1
View →

Create your own trend reports

Free beta access — explore any AI/ML research topic with automated trend analysis.

Get Started Now