arXiv Machine Learning By Joshua Mitton, Prarthana Bhattacharyya, Ralph Abboud, Simon Woodhead

Knowing When to Defer: Selective Prediction for Responsible Knowledge Tracing

Read the original on arXiv Machine Learning →

arXiv:2509. 21514v4 Announce Type: replace Abstract: Research on Knowledge Tracing (KT) models traditionally focuses on improving predictive accuracy.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
1d ago

One Mastery Threshold Does Not Fit All Knowledge Tracing Models

The study investigates how a single mastery threshold can produce divergent outcomes across different knowledge tracing (KT) models. By evaluating six KT models on four datasets with thresholds ranging from 0.50 to 0.99, the authors find that Bayesian Knowledge Tracing (BKT) is relatively insensitive to threshold changes, whereas neural models become increasingly selective as thresholds rise. The optimal threshold varies widely across models and instructional settings, and stricter thresholds can disproportionately limit advancement for weaker students.

By Xianghui Meng, Yujing Zhang, Jionghao Lin
arXiv AI
Sep 17

Confidence Comes from Experience: Experiential Confidence Estimation from Reasoning to Agents

The paper introduces XConf, an experiential confidence estimator that augments a language model’s current inference with a record of its past graded episodes. By recalling similar past tasks and reflecting on past outcomes, XConf generates confidence scores without accessing logits or updating weights, achieving superior discrimination and calibration across diverse benchmarks. The method demonstrates significant gains in selective prediction, improving success rates on agent tasks by up to 8.7 points.

By Caiqi Zhang, Xiaochen Zhu, Chengzu Li, Yulong Chen, Dharshan Kumaran, Nigel Collier