Tag
A tweet criticizes AI models for writing ineffective unit tests and shares a prompt to enforce better testing practices, emphasizing E2E tests and avoiding unit tests written after code.
Noam Brown announces that GPT-5.6 Sol Ultra proved a 50-year-old math conjecture, demonstrating impressive prompt engineering and agent prompting.