Tag
In his talk, Carson Gross discussed the impact of AI on university computer science education, arguing that in the AI era, students still need to be taught to write and read code, while also noting that AI brings an assessment crisis and opportunities for pedagogical change.
This paper investigates whether LLMs can accurately predict item difficulty levels in large-scale reading and writing tests, finding that GPT-4.1 achieves moderate accuracy but is outperformed by ConvBERT, and that LLMs tend to underestimate difficulty for hard items.