@FudanUniversity: Final exam at Fudan: students don't answer questions. They write them — to stump AI. 51 students, 10 questions each, 3 …

X AI KOLs Following News

Summary

Fudan University held a novel final exam where 51 students each wrote 10 questions designed to stump three AI models (Claude, DeepSeek, MiniMax), with grades based on how difficult the questions were for the AI.

Final exam at Fudan: students don't answer questions. They write them — to stump AI. 51 students, 10 questions each, 3 AI models (Claude, DeepSeek, MiniMax) on the hot seat. The harder you make AI fail, the higher your grade. https://t.co/ysQ00ww7dD
Original Article
View Cached Full Text

Cached at: 07/01/26, 01:58 AM

Final exam at Fudan: students don’t answer questions. They write them — to stump AI. 51 students, 10 questions each, 3 AI models (Claude, DeepSeek, MiniMax) on the hot seat. The harder you make AI fail, the higher your grade. https://t.co/ysQ00ww7dD

Similar Articles

@Phoenixyin13: This kind of assessment method where students create questions that stump AI is indeed very innovative and highly forward-looking. Students need to explore the strengths and weaknesses of the three models: Claude, DeepSeek, and MiniMax. In this process, students no longer blindly trust AI outputs but learn to review AI responses with a critical and discerning eye, which...

X AI KOLs Timeline

This educational assessment method encourages students to explore the strengths and weaknesses of Claude, DeepSeek, and MiniMax, creating questions that defeat AI, thereby cultivating critical thinking and competitiveness needed in the AI era.

humanity's last exam current benchmarks thoughts?

Reddit r/singularity

Discussion of recent AI model scores on the 'humanity's last exam' benchmark, noting improvement from GPT-4o's 2.7% in May 2024 to around 45% by June 2026, questioning the exam's difficulty.