Bengali text summarization using TextRank, Fuzzy C-Means and aggregate scoring methods

bracu.type.groupResearch Publications
datacite.rightsMetadata Only
dc.contributor.authorRahman, Alvee
dc.contributor.authorRafiq, Fahim Md
dc.contributor.authorSaha, Ramkrishna
dc.contributor.authorRafian, Ruhit
dc.contributor.authorArif, Hossain
dc.contributor.departmentDepartment of Computer Science and Engineering
dc.date.accessioned2026-08-30T07:38:31Z
dc.date.available2026-08-30T07:38:31Z
dc.date.issued2019-06-01
dc.description.abstractIn this world, it is very difficult and time consuming for humans to summarize large documents, reports, news and research articles. Multiple text summarization techniques play vital roles in picking the important points and sentences thus reducing the time and effort required to read a whole article. Numerous summarization techniques have been applied to the English language but works on Bangla text summarization is still limited. Furthermore, in our country, Bangladesh, all summarization is mainly done by humans. Keeping that in mind we aim to find a simple way of summarizing Bengali texts with the technology at hand. Text summarization can be of two types, abstractive and extractive. In this paper, we will use extractive text summarization to summarize Bengali passages, using Fuzzy C-Means, TextRank and Aggregate Sentence Scoring methodologies. We have also done a comparative study, among the three methodologies to find out that Fuzzy C-Means out performs the other two methods to generate a more concise and accurate summary.
dc.description.versionPublished
dc.format.extent331-336
dc.identifier.doi10.1109/TENSYMP46218.2019.8971039
dc.identifier.issn9781728102979
dc.identifier.other2-s2.0-85079268136
dc.identifier.urihttps://hdl.handle.net/10361/29600
dc.language.isoen_US
dc.publisherInstitute of Electrical and Electronics Engineers Inc.
dc.relation.hasversion10.1109/TENSYMP46218.2019.8971039
dc.relation.ispartofProceedings of 2019 IEEE Region 10 Symposium Tensymp 2019
dc.relation.ispartofseriesProceedings of 2019 IEEE Region 10 Symposium Tensymp 2019
dc.relation.urihttps://ieeexplore.ieee.org/document/8971039
dc.rightsfalse
dc.subjectBengali
dc.subjectExtractive text summarization
dc.subjectFCM
dc.subjectROUGE
dc.subjectText summarization
dc.subjectTextRank
dc.subject.lcshBengali language.
dc.subject.lcshInformation storage and retrieval.
dc.titleBengali text summarization using TextRank, Fuzzy C-Means and aggregate scoring methods
dc.typeConference Proceeding
person.affiliation.nameBRAC University
person.affiliation.nameBRAC University
person.affiliation.nameBRAC University
person.affiliation.nameBRAC University
person.affiliation.nameBRAC University
person.identifier.scopus-author-id57208007267
person.identifier.scopus-author-id57215124164
person.identifier.scopus-author-id57207916409
person.identifier.scopus-author-id57215134782
person.identifier.scopus-author-id55843238200

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
IMG_8345.jpg
Size:
27.35 KB
Format:
Joint Photographic Experts Group/JPEG File Interchange Format (JFIF)

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.71 KB
Format:
Item-specific license agreed upon to submission
Description: