Bengali text to international phonetic alphabet (IPA) transcription: a comprehensive BYT5 based approach

bracu.degree.levelUndergraduate
bracu.type.groupStudent Works
datacite.rightsOpen Access
dc.contributor.advisorAhmed, Md. Sabbir
dc.contributor.authorKamal, S M Sazzad Bin
dc.contributor.departmentDepartment of Computer Science and Engineering
dc.date.accessioned2025-12-29T09:46:03Z
dc.date.available2025-12-29T09:46:03Z
dc.date.copyright2025
dc.date.issued2025-10
dc.descriptionCataloged from PDF version of thesis.
dc.descriptionIncludes bibliographical references (pages 25-26).
dc.descriptionThis thesis is submitted in partial fulfillment of the requirements for the degree of Bachelor of Science in Computer Science, 2025.en_US
dc.description.abstractThe International Phonetic Alphabet (IPA) is an alphabetic system of phonetic notation based primarily on the Latin script. The IPA is widely used by lexicographers, foreign language students and teachers, linguists, speech-language pathologists, singers, actors, language contractors, and translators. In the rapidly evolving digital era, the use of the International Phonetic Alphabet for the Bangla language is crucial to maintaining linguistic precision, aiding language learning, fostering research, preserving linguistic diversity, and facilitating communication and speech therapy. It plays a vital role in various aspects of the study of the Bangla language and related fields. In the work, it is planned to train T5-based Large Language Model (LLM) architectures on the collected dataset to convert Bangla texts into IPA notation and compare the LLM results to find the best language model. Our work extends beyond training, delving into the exploration of potential enhancement avenues. Another aspect is to investigate the impact of Bangla numerals, abbreviations, English texts, and out-of-vocabulary words. We trained and evaluated three T5-based architectures—ByT5, UMT5, and mT5—on a merged dataset of over 50,000 Bangla sentences representing both standard and dialectal speech. Through these explorations, we observe a spectrum of outcomes, where some modifications result in tangible performance improvements, while others offer unique insights for future endeavors in the fields of Natural Language Processing (NLP) and linguistic preservation.en_US
dc.description.degreeBachelor of Science in Computer Science
dc.description.statementofresponsibilityS M Sazzad Bin Kamal
dc.format.extent35 pages
dc.identifier.otherID 17101529
dc.identifier.urihttp://hdl.handle.net/10361/27382
dc.language.isoenen_US
dc.publisherBRAC Universityen_US
dc.rightsBRAC University theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission.
dc.subjectNLPen_US
dc.subjectNatural language processingen_US
dc.subjectInternational phonetic alphabeten_US
dc.subjectIPAen_US
dc.subjectBengali languageen_US
dc.subjectPhonetic alphabeten_US
dc.subjectT5 modelen_US
dc.subjectLarge language modelsen_US
dc.subjectBengali texten_US
dc.subjectLanguage processingen_US
dc.subject.lcshNatural language processing (Computer science).
dc.subject.lcshPhonetic alphabet.
dc.subject.lcshInternational Phonetic Alphabet (IPA).
dc.subject.lcshBengali language--Alphabet.
dc.subject.lcshBengali language--Phonetics.
dc.subject.lcshSpeech synthesis.
dc.titleBengali text to international phonetic alphabet (IPA) transcription: a comprehensive BYT5 based approachen_US
dc.typeThesisen_US

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
17101529_CSE.pdf
Size:
269.88 KB
Format:
Adobe Portable Document Format
Description:

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.71 KB
Format:
Item-specific license agreed upon to submission
Description: