Tag
This paper evaluates how well LLMs (ChatGPT, Claude, DeepSeek) can generate one-page project plans in physics, astrophysics, and cosmology, and how human and AI reviewers assess them. Results show that human reviewers rate AI and human proposals similarly, while AI reviewers prefer AI-written proposals and can perfectly distinguish them from human-written ones.