Welcome to the upgraded BRAC University Institutional Repository. We are currently organizing collections after a recent system upgrade. Homepage category counters may temporarily show lower numbers while syncing, but over 27,000 repository items remain safe and accessible. Please use the search bar to find theses, scholarly outputs, and institutional documents.

From impersonation to authentication: techniques for identifying deep fake voices

bracu.degree.levelUndergraduate
bracu.type.groupStudent Works
datacite.rightsOpen Access
dc.contributor.advisorReza, Md Tanzim
dc.contributor.advisorRahman, Rafeed
dc.contributor.authorRabbi, B.M Saqlain
dc.contributor.authorHaque, Erfanul
dc.contributor.authorHuri, Humayera Shayera
dc.contributor.authorTabassum, Nuzhut
dc.contributor.departmentDepartment of Computer Science and Engineering
dc.date.accessioned2025-01-21T09:27:15Z
dc.date.available2025-01-21T09:27:15Z
dc.date.copyright©2024
dc.date.issued2024-10
dc.descriptionCataloged from PDF version of thesis.
dc.descriptionIncludes bibliographical references (pages 55-56).
dc.descriptionThis thesis is submitted in partial fulfillment of the requirements for the degree of Bachelor of Science in Computer Science, 2024.en_US
dc.description.abstractThe proliferation of fake voices has become a concerning issue, making it increasingly challenging to distinguish between authentic and fabricated audio recordings. Notable examples include the replication of the voices of prominent U.S. President through the use of AI technology. These manipulated audio clips have been disseminated across various YouTube channels, serving both benign and malicious purposes. While fake audio can be entertaining in content creation, it also carries a darker potential, including threats to political leaders and diplomatic relations between nations, potentially leading to conflict. In Bangladesh, the surge in fake audio content has left its populace in a state of skepticism, casting doubt on the authenticity of online videos and news reports. The citizens of Bangladesh have become susceptible to accepting counterfeit content as real, amplifying the need for a solution to this issue. Dealing with this problem needs the use of machine learning, AI, and various algorithmic techniques to mitigate the spread of fake audio within Bangladesh. To tackle this problem, we propose utilizing diverse audio datasets and trained models, running them through specific algorithms to enhance accuracy in discerning genuine from manipulated audio. The main goal of our thesis is to implement three features that will help to detect the deep-fake audio spread across various social media affecting the lives of people and prevent scamming across various online banking transactions. In order to complete our desired project we had to work on various machine-learning models – namely Recurrent Neural Network (RNN), Long Short Term Memory (LSTM), bi LSTM and LSTM based RNN to detect fake audios across social media. Also using these models along with feature extraction like Mel-Frequency Cepstral Coefficients (MFCC) and Short-Term Fourier Transform (STFT), our study aims to create a strong system to differentiate between real and deepfake-audio. In spite of facing multiple challenges and achieving a better accuracy, we were successfully able to reach our desired goal.We accomplished our accuracy at 99% using Bi-LSTM for deepfake audio detection.en_US
dc.description.degreeBachelor of Science in Computer Science
dc.description.statementofresponsibilityB.M Saqlain Rabbi
dc.description.statementofresponsibilityErfanul Haque
dc.description.statementofresponsibilityHumayera Shayera Huri
dc.description.statementofresponsibilityNuzhut Tabassum
dc.format.extent63 pages
dc.identifier.otherID 19101587
dc.identifier.otherID 19201009
dc.identifier.otherID 21341026
dc.identifier.otherID 23241091
dc.identifier.urihttp://hdl.handle.net/10361/25249
dc.language.isoenen_US
dc.publisherBRAC Universityen_US
dc.rightsBRAC University theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission.
dc.subjectAudio detectionen_US
dc.subjectFake voiceen_US
dc.subjectManipulated audioen_US
dc.subjectMachine learningen_US
dc.subjectRNNen_US
dc.subjectLSTMen_US
dc.subjectBiLSTMen_US
dc.subjectMFCCen_US
dc.subject.lcshDeepfakes--Detection.
dc.subject.lcshDeep learning (Machine learning).
dc.titleFrom impersonation to authentication: techniques for identifying deep fake voicesen_US
dc.typeThesisen_US

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
19201009, 19101587, 21341026, 23241091_CSE.pdf
Size:
611.88 KB
Format:
Adobe Portable Document Format
Description:

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.71 KB
Format:
Item-specific license agreed upon to submission
Description: