arXiv Computation and Language

Beyond ID Embeddings: Process-Grounded Language Modeling for Cognitive Diagnosis

The paper introduces Process-aware Language Cognitive Diagnosis (PLCD), a framework that replaces traditional ID-based embeddings in Cognitive Diagnosis Models with language-derived structures and response records. PLCD employs large language models to build concept schemas and cognitive process graphs, and uses a Language-to-Cognition Mapper with DA-MoE experts and contrastive learning to map textual evidence into a unified cognitive space. Experiments demonstrate that PLCD outperforms conventional baselines in student performance prediction and shows strong cognitive transfer, improving cold-start robustness and cognitive grounding.

arXiv AI
Aug 26

Incorporating Cognitive Load and Knowledge Transfer for Multi-Domain Knowledge Tracing

The paper introduces LT‑MKT, a new method for multi‑domain knowledge tracing that incorporates cognitive load and knowledge transfer. It constructs a multi‑domain hierarchical graph using textual information from questions and concepts, then explicitly models cross‑domain temporal and knowledge features to capture cognitive load effects. A knowledge transfer module further captures propagation of knowledge states within and across domains, leading to more accurate predictions of students’ future performance.

By Haotian Zhang, Shucun Wang, Jinze Wu, Liang Ding, Shuochen Liu, Zhenya Huang, Jing Sha, Shijin Wang, Qi Liu
arXiv Machine Learning
Jun 30

MOSAIC: Orchestrating Collaborative Knowledge Tracing with Hierarchical Semantic Alignment

arXiv:2606. 29049v1 Announce Type: new Abstract: Knowledge Tracing (KT) is important for personalized education but traditionally suffers from two key limitations: a reliance on shallow ID-based representations that neglect semantic depth and a restriction to single-granularity mastery estimation that overlooks hierarchical knowledge dependencies.

By Xinjin Li, Mengyue Wang, Yuzhen Lin, Pengbin Feng, Ziqi Sha, Yeyang Zhou, Yu Ma
arXiv AI
Jul 3

Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

arXiv:2607. 02234v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) has emerged as a promising paradigm for improving LLM reasoning, where a privileged teacher with access to reference solutions provides token-level supervision on the student's own generated trajectories.

By Zhanming Shen, Jintao Tong, Shaotian Yan, Chen Shen, Hao Chen, Wentao Ye, Xiaomeng Hu, Rui Miao, Haobo Wang, Junbo Zhao, Gang Chen, Jieping Ye
arXiv AI
Aug 14

EgoMonth: A Month-Level Egocentric Video Benchmark for Long-Term Spatiotemporal Memory

arXiv:2608. 13113v1 Announce Type: cross Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have led to substantial progress in video understanding, accompanied by a growing number of long video benchmarks.

By Weitao Chen, Hu Jiaxin, Xie Tianyidan, Yang Li, Yuyi Qian, Banghao Xu, Ziheng Tang, Shenyi Wang, Mingyue Yu, Duo Li, Jiacheng Shi, Gao Wang, Zhan Xu, Zhicheng Qiu, Xuanfu Li, Jian Yang, Lanjun Wang, Zili Yi