Bengali text to international phonetic alphabet (IPA) transcription: a comprehensive BYT5 based approach
| bracu.degree.level | Undergraduate | |
| bracu.type.group | Student Works | |
| datacite.rights | Open Access | |
| dc.contributor.advisor | Ahmed, Md. Sabbir | |
| dc.contributor.author | Kamal, S M Sazzad Bin | |
| dc.contributor.department | Department of Computer Science and Engineering | |
| dc.date.accessioned | 2025-12-29T09:46:03Z | |
| dc.date.available | 2025-12-29T09:46:03Z | |
| dc.date.copyright | 2025 | |
| dc.date.issued | 2025-10 | |
| dc.description | Cataloged from PDF version of thesis. | |
| dc.description | Includes bibliographical references (pages 25-26). | |
| dc.description | This thesis is submitted in partial fulfillment of the requirements for the degree of Bachelor of Science in Computer Science, 2025. | en_US |
| dc.description.abstract | The International Phonetic Alphabet (IPA) is an alphabetic system of phonetic notation based primarily on the Latin script. The IPA is widely used by lexicographers, foreign language students and teachers, linguists, speech-language pathologists, singers, actors, language contractors, and translators. In the rapidly evolving digital era, the use of the International Phonetic Alphabet for the Bangla language is crucial to maintaining linguistic precision, aiding language learning, fostering research, preserving linguistic diversity, and facilitating communication and speech therapy. It plays a vital role in various aspects of the study of the Bangla language and related fields. In the work, it is planned to train T5-based Large Language Model (LLM) architectures on the collected dataset to convert Bangla texts into IPA notation and compare the LLM results to find the best language model. Our work extends beyond training, delving into the exploration of potential enhancement avenues. Another aspect is to investigate the impact of Bangla numerals, abbreviations, English texts, and out-of-vocabulary words. We trained and evaluated three T5-based architectures—ByT5, UMT5, and mT5—on a merged dataset of over 50,000 Bangla sentences representing both standard and dialectal speech. Through these explorations, we observe a spectrum of outcomes, where some modifications result in tangible performance improvements, while others offer unique insights for future endeavors in the fields of Natural Language Processing (NLP) and linguistic preservation. | en_US |
| dc.description.degree | Bachelor of Science in Computer Science | |
| dc.description.statementofresponsibility | S M Sazzad Bin Kamal | |
| dc.format.extent | 35 pages | |
| dc.identifier.other | ID 17101529 | |
| dc.identifier.uri | http://hdl.handle.net/10361/27382 | |
| dc.language.iso | en | en_US |
| dc.publisher | BRAC University | en_US |
| dc.rights | BRAC University theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. | |
| dc.subject | NLP | en_US |
| dc.subject | Natural language processing | en_US |
| dc.subject | International phonetic alphabet | en_US |
| dc.subject | IPA | en_US |
| dc.subject | Bengali language | en_US |
| dc.subject | Phonetic alphabet | en_US |
| dc.subject | T5 model | en_US |
| dc.subject | Large language models | en_US |
| dc.subject | Bengali text | en_US |
| dc.subject | Language processing | en_US |
| dc.subject.lcsh | Natural language processing (Computer science). | |
| dc.subject.lcsh | Phonetic alphabet. | |
| dc.subject.lcsh | International Phonetic Alphabet (IPA). | |
| dc.subject.lcsh | Bengali language--Alphabet. | |
| dc.subject.lcsh | Bengali language--Phonetics. | |
| dc.subject.lcsh | Speech synthesis. | |
| dc.title | Bengali text to international phonetic alphabet (IPA) transcription: a comprehensive BYT5 based approach | en_US |
| dc.type | Thesis | en_US |