Tag
SpecFirst proposes making behavioral specification elicitation a first-class step in agent-based program synthesis, separating it from exploration and coding. Evaluated on 200 program instances, it significantly improves test pass rates and binary exploration coverage across multiple models.
This paper presents a mixed-method survey and expert interview study examining how LLM-based validation tools can help organizations translate EU AI Act obligations into testable, auditable requirements and evidence artifacts.