arXiv AI By Amanda La Hadi, Muhammad Johan Alibasa, Guanliang Chen, A. Taufiq Asyhari

The Easy Trap: Why LLMs Underestimate Misconception-Driven Difficulty

Read the original on arXiv AI →

arXiv:2607. 26067v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for estimating item difficulty in educational assessment.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.