BanglaSarc: a dataset for sarcasm detection
| bracu.type.group | Research Publications | |
| datacite.rights | Metadata Only | |
| dc.contributor.author | Apon, Tasnim Sakib | |
| dc.contributor.author | Anan, Ramisa | |
| dc.contributor.author | Modhu, Elizabeth Antora | |
| dc.contributor.author | Suter, Arjun | |
| dc.contributor.author | Sneha, Ifrit Jamal | |
| dc.contributor.author | Alam, Md. Golam Rabiul | |
| dc.contributor.department | Department of Computer Science and Engineering | |
| dc.date.accessioned | 2026-08-13T09:37:39Z | |
| dc.date.available | 2026-08-13T09:37:39Z | |
| dc.date.issued | 2022-01-01 | |
| dc.description.abstract | Being one of the most widely spoken language in the world, the use of Bangla has been increasing in the world of social media as well. Sarcasm is a positive statement or remark with an underlying negative motivation that is extensively employed in today's social media platforms. There has been a significant improvement in sarcasm detection in English over the previous many years, however the situation regarding Bangla sarcasm detection remains unchanged. As a result, it is still difficult to identify sarcasm in bangla, and a lack of high-quality data is a major contributing factor. This article proposes BanglaSarc, a dataset constructed specifically for bangla textual data sarcasm detection. This dataset contains of 5112 comments/status and contents collected from various online social platforms such as Facebook, YouTube, along with a few online blogs. Due to the limited amount of data collection of categorized comments in Bengali, this dataset will aid in the of study identifying sarcasm, recognizing people's emotion, detecting various types of Bengali expressions, and other domains. The dataset is publicly available at https://www.kaggle.com/datasets/sakibapon/banglasarc. | |
| dc.description.version | Published | |
| dc.format.extent | 5 Pages | |
| dc.identifier.citation | T. S. Apon, R. Anan, E. A. Modhu, A. Suter, I. J. Sneha and M. G. R. Alam, "BanglaSarc: A Dataset for Sarcasm Detection," 2022 IEEE Asia-Pacific Conference on Computer Science and Data Engineering (CSDE), Gold Coast, Australia, 2022, pp. 1-5, doi: 10.1109/CSDE56538.2022.10089322. | |
| dc.identifier.doi | 10.1109/CSDE56538.2022.10089322 | |
| dc.identifier.issn | 9781665453059 | |
| dc.identifier.other | 2-s2.0-85151081700 | |
| dc.identifier.uri | https://hdl.handle.net/10361/29051 | |
| dc.language.iso | en_US | |
| dc.publisher | Institute of Electrical and Electronics Engineers Inc. | |
| dc.relation.hasversion | 10.1109/CSDE56538.2022.10089322 | |
| dc.relation.ispartof | Proceedings of IEEE Asia Pacific Conference on Computer Science and Data Engineering Csde 2022 | |
| dc.relation.ispartofseries | Proceedings of IEEE Asia Pacific Conference on Computer Science and Data Engineering Csde 2022 | |
| dc.relation.uri | https://ieeexplore.ieee.org/document/10089322 | |
| dc.subject | Bangla Natural Langauge Processing (BNLP) | |
| dc.subject | Bangla sarcasm detection | |
| dc.subject | Emotion recognition | |
| dc.subject | Data engineering | |
| dc.subject.lcsh | Bengali language--Data processing. | |
| dc.subject.lcsh | Computational linguistics. | |
| dc.subject.lcsh | Natural language processing (Computer science). | |
| dc.title | BanglaSarc: a dataset for sarcasm detection | |
| dc.type | Conference Proceeding | |
| person.affiliation.name | BRAC University | |
| person.affiliation.name | BRAC University | |
| person.affiliation.name | BRAC University | |
| person.affiliation.name | BRAC University | |
| person.affiliation.name | BRAC University | |
| person.affiliation.name | BRAC University | |
| person.identifier.scopus-author-id | 57348873600 | |
| person.identifier.scopus-author-id | 57913845400 | |
| person.identifier.scopus-author-id | 57913421800 | |
| person.identifier.scopus-author-id | 57912994600 | |
| person.identifier.scopus-author-id | 57913845500 | |
| person.identifier.scopus-author-id | 26434126600 |