Tag
This paper develops an LLM pipeline for automating systematic literature reviews in disease spread modeling, comparing the performance of GPT-4.1 and GPT-5.0 against human-conducted reviews.