Hugging Face Trending Papers

GSM-Plus-BN: A Perturbation-Based Benchmark for Bangla Mathematical Reasoning in Large Language Models

Read the original on Hugging Face Trending Papers →

The evaluation of mathematical reasoning in large language models (LLMs) has predominantly focused on high-resource languages like English. This has created a significant barrier to the equitable development and deployment of AI in linguistically diverse regions such as Bangladesh, where over 230 million people speak Bengali.

Summary generated by The Flow from the publisher's feed. The full article lives at Hugging Face Trending Papers.