arXiv AI By Adrian Marius Dumitran, Theodor-Pierre Moroianu, Mihnea-Vicentiu Buca

MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks

Read the original on arXiv AI →

arXiv:2507. 03162v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has transformed various domains, particularly computer science (CS) education.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.