@FudanUniversity: Final exam at Fudan: students don't answer questions. They write them — to stump AI. 51 students, 10 questions each, 3 …
Summary
Fudan University held a novel final exam where 51 students each wrote 10 questions designed to stump three AI models (Claude, DeepSeek, MiniMax), with grades based on how difficult the questions were for the AI.
View Cached Full Text
Cached at: 07/01/26, 01:58 AM
Final exam at Fudan: students don’t answer questions. They write them — to stump AI. 51 students, 10 questions each, 3 AI models (Claude, DeepSeek, MiniMax) on the hot seat. The harder you make AI fail, the higher your grade. https://t.co/ysQ00ww7dD
Similar Articles
@Phoenixyin13: This kind of assessment method where students create questions that stump AI is indeed very innovative and highly forward-looking. Students need to explore the strengths and weaknesses of the three models: Claude, DeepSeek, and MiniMax. In this process, students no longer blindly trust AI outputs but learn to review AI responses with a critical and discerning eye, which...
This educational assessment method encourages students to explore the strengths and weaknesses of Claude, DeepSeek, and MiniMax, creating questions that defeat AI, thereby cultivating critical thinking and competitiveness needed in the AI era.
Suspecting AI cheating, Ivy League prof ordered an in-person final; scores fell 50%
Brown University economics professor Roberto Serrano, suspecting AI cheating after record-high scores on a take-home midterm, administered an in-person final exam; scores plummeted by 50%, revealing widespread AI-assisted cheating among Ivy League students.
Professor denounces mass AI fraud on an exam at Brown
Professor Roberto Serrano at Brown University detected at least 50 students cheating using AI on a midterm exam, sparking a debate about academic integrity in higher education.
@FinanceYF5: 2. 回答那些能难倒大多数 AI 的问题
A user tests Claude Fable on classic AI-stumping questions like counting 'r's in strawberry and comparing 5.11 and 5.1, jokingly claiming AGI is achieved.
humanity's last exam current benchmarks thoughts?
Discussion of recent AI model scores on the 'humanity's last exam' benchmark, noting improvement from GPT-4o's 2.7% in May 2024 to around 45% by June 2026, questioning the exam's difficulty.