https://doi.org/10.1140/epjds/s13688-026-00672-z
Research
Can large language models generate novel scientific ideas? A comprehensive study on data-driven astronomy
1
State Key Laboratory of Public Big Data, College of Computer Science and Technology, Guizhou University, 550025, Guiyang, China
2
School of Data Science and Engineering, East China Normal University, 200062, Shanghai, China
3
School of Astronomy and Space Sciences, University of Chinese Academy of Sciences, 100049, Beijing, China
4
National Astronomical Observatories, Chinese Academy of Sciences, 100012, Beijing, China
a
This email address is being protected from spambots. You need JavaScript enabled to view it.
b
This email address is being protected from spambots. You need JavaScript enabled to view it.
Received:
1
January
2026
Accepted:
25
May
2026
Published online:
18
June
2026
Abstract
Scientific discovery is a cornerstone of societal advancement, and its rapid development demands innovative tools that can facilitate the generation of research ideas. Recent breakthroughs in Generative Artificial Intelligence (GenAI), particularly in Large Language Models (LLMs), offer transformative potential for scientific idea generation. However, existing LLM-based idea generation methods are limited to computer science and closely related domains, and their application in other scientific fields, e.g., astronomy, remains largely underexplored. In this paper, we investigate the applicability of LLMs for scientific idea generation in data-driven astronomy, an interdisciplinary field that applies advanced data science methods to analyze massive datasets collected by modern telescopes and satellites to drive astronomical discoveries. Specifically, we implement a novel framework, AstroInsight, that integrates conception, iterative refinement, expert validation, and knowledge integration for idea generation. Through extensive experiments with human expert assessments and model self-evaluation, we show that AstroInsight effectively generates new research concepts and substantially accelerates discovery cycles, with the generated drafts achieving a novelty score of 3+/6 in both human- and model-based evaluations. Additionally, its generated ideas match or exceed human-generated ones in terms of originality and feasibility across multiple topics. In summary, we provide researchers with tools to boost productivity while maintaining rigor in a human-AI collaborative framework, thereby illuminating pathways to building LLM-assisted systems for autonomous scientific ideation.
Key words: Research idea generation / Data-driven astronomy / Large language models / Automated research
Handling Editor: Santo Fortunato
© The Author(s) 2026
Open Access This article is licensed under a Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License, which permits any non-commercial use, sharing, distribution and reproduction in any medium or format, as long as you give appropriate credit to the original author(s) and the source, provide a link to the Creative Commons licence, and indicate if you modified the licensed material. You do not have permission under this licence to share adapted material derived from this article or parts of it. The images or other third party material in this article are included in the article’s Creative Commons licence, unless indicated otherwise in a credit line to the material. If material is not included in the article’s Creative Commons licence and your intended use is not permitted by statutory regulation or exceeds the permitted use, you will need to obtain permission directly from the copyright holder. To view a copy of this licence, visit http://creativecommons.org/licenses/by-nc-nd/4.0/.

