arXiv AI

Context-Aware Prediction of Student Quiz Performance with Multimodal Textbook Features

arXiv:2606. 24770v1 Announce Type: cross Abstract: Educational platforms often predict student performance from prior interactions, but the assessment content itself also varies in linguistic and visual complexity.

Hugging Face Trending Papers
Sep 3

KnowVis: Knowledge-Centric Visual Summarization for Video Lectures

KnowVis is a framework that converts linear video lectures into knowledge‑centric visual summaries. It first builds a detailed concept map from multimodal video content to identify key and challenging concepts, then organizes these into structured knowledge units before synthesizing engaging visual narratives. The authors also provide a dataset of 125 educational videos with 1,079 visual summaries and show through automated metrics and a human study that KnowVis outperforms existing methods in accuracy, clarity, and learning outcomes.

arXiv Computation and Language
Sep 4

KnowVis: Knowledge-Centric Visual Summarization for Video Lectures

KnowVis is a framework that converts linear video lectures into knowledge‑centric visual narratives. It first extracts a detailed concept map from multimodal video content to identify key and challenging concepts, then builds structured knowledge units and synthesizes engaging visual summaries. The authors also provide a curated dataset of 125 educational videos across 10 disciplines, paired with 1,079 visual summaries, and show through automated evaluations and a human study that KnowVis produces more accurate, clear visuals that reduce cognitive load and improve learning effectiveness and knowledge retention.

By Yi Xu, Yifan Hou, Xiaoyu Zhang
arXiv AI
Sep 17

Enhancing knowledge tracing robustness for new question cold start in Intelligent Tutoring Systems

The paper introduces Practical Integrated Cross-consistent Knowledge Tracing (PICKT), a model that incorporates multiple feature types to improve Knowledge Tracing robustness when new questions lack interaction history. It evaluates the impact of difficulty, textual, and knowledge‑map relational features, finding that difficulty is especially informative for hard questions, while fused text and map features help estimate unseen questions by leveraging similar ones seen during training. The study concludes that prioritizing feature annotation aligned with educational service characteristics is essential for maintaining robust diagnostics in Intelligent Tutoring Systems.

By Wonbeen Lee, Channyoung Lee, Junho Sohn, Hansam Cho
arXiv Machine Learning
Jun 9

Structure-Aware Modeling of Multiple-Choice Questions Improves Automatic Difficulty Estimation

arXiv:2606. 08988v1 Announce Type: cross Abstract: Automatic Question Difficulty Estimation (AQDE) holds growing promise for educational assessment because it has the potential to yield difficulty estimates that are competitive with expert judgment, while helping reduce the time and financial burden associated with pilot administrations and scaling to digital testing contexts.

By Gabriel Ortega, Abelino Jim\'enez, S\'everin Lions, Pablo Dartnell
arXiv AI
Jun 16

When RAG Hurts: Diagnosing and Mitigating Attention Distraction in Retrieval-Augmented LVLMs

arXiv:2602. 00344v2 Announce Type: replace-cross Abstract: While Retrieval-Augmented Generation (RAG) is one of the dominant paradigms for enhancing Large Vision-Language Models (LVLMs) on knowledge-based VQA tasks, recent work attributes RAG failures to insufficient attention towards the retrieved context, proposing to reduce the attention allocated to image tokens.

By Beidi Zhao, Wenlong Deng, Xinting Liao, Yushu Li, Nazim Shaikh, Yao Nie, Xiaoxiao Li