| dc.contributor.advisor | Dr. AKM Mahbubur Rahman | en_US |
| dc.contributor.author | Nondon, Nishorgo | |
| dc.contributor.author | Naurin, Sabah | |
| dc.contributor.author | Muzaffar, Shams Habib | |
| dc.date.accessioned | 2026-09-23T14:11:25Z | |
| dc.date.available | 2026-09-23T14:11:25Z | |
| dc.date.issued | 2026-08 | |
| dc.identifier.other | ID 2231450 | |
| dc.identifier.other | ID 2230677 | |
| dc.identifier.other | ID 2231353 | |
| dc.identifier.uri | https://ar.iub.edu.bd/handle/11348/1614 | |
| dc.description | This thesis is submitted in partial fulfillment of the requirements for the degree of Bachelor of Science (BSc) in Computer Science and Engineering (CSC), 2026. | |
| dc.description.abstract | This thesis introduces the Bangladesh Mathematical Olympiad (BdMO 2022–2024, 2026) benchmark, containing 1,336 mathematical problems, including 201 image-based problems, to evaluate the mathematical and visual reasoning abilities of Large Language Models (LLMs) and Vision-Language Models (VLMs) in Bangla and English. Four AI models were assessed across seven mathematical topics and four educational levels. Results showed that image-based problems reduced accuracy by 18–19%, while switching from English to Bangla caused only a 0.46% decrease, indicating that visual reasoning presents a greater challenge than language processing. The study also examined bilingual consistency, model performance, and common reasoning failures, providing insights into the limitations of current AI models in mathematical and multimodal reasoning. | en_US |
| dc.format.extent | 97 pages | |
| dc.language.iso | en | en_US |
| dc.publisher | Independent University, Bangladesh (IUB) | en_US |
| dc.rights | Theses submitted to Independent University, Bangladesh, are protected by copyright. They may be accessed for academic and research purposes; however, reproduction, distribution, or use of the material in any form requires prior written permission from the University. | |
| dc.subject | Mathematical Reasoning | en_US |
| dc.subject | Large Language Models (LLMs) | en_US |
| dc.subject | Vision-Language Models (VLMs) | en_US |
| dc.subject | Bangla-English Benchmark | en_US |
| dc.subject | Multimodal AI Evaluation | en_US |
| dc.title | Cross-lingual multimodal mathematical reasoning: benchmarking frontier models in Bengali | en_US |
| dc.type | Thesis | en_US |
| dc.contributor.department | Department of Computer Science and Engineering | |