item-difficulty

Tag

Cards List
#item-difficulty

Can LLMs Really Understand Item Difficulty Levels? Implications for Automated Item Generation Using LLMs

arXiv cs.CL · 6d ago Cached

This paper investigates whether LLMs can accurately predict item difficulty levels in large-scale reading and writing tests, finding that GPT-4.1 achieves moderate accuracy but is outperformed by ConvBERT, and that LLMs tend to underestimate difficulty for hard items.

0 favorites 0 likes
← Back to home

Submit Feedback