arXiv AI By Adrian Marius Dumitran, Theodor-Pierre Moroianu, Mihnea-Vicentiu Buca

MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks

Read the original on arXiv AI →

arXiv:2507. 03162v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has transformed various domains, particularly computer science (CS) education.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.