Tag
OpenAI has paused model training to prioritize hardening its research systems, likely for enhanced safety and security measures.
This paper introduces a formal vocabulary for describing and comparing multi-agent automated research systems, covering design choices such as agent identity, operations, communication, and evaluation. It distinguishes between generative and evaluative taste and instantiates the vocabulary on recent systems.