Welcome to the upgraded BRAC University Institutional Repository. We are currently organizing collections after a recent system upgrade. Homepage category counters may temporarily show lower numbers while syncing, but over 27,000 repository items remain safe and accessible. Please use the search bar to find theses, scholarly outputs, and institutional documents.

Important keywords extraction from documents using semantic analysis

bracu.degree.levelUndergraduate
bracu.type.groupStudent Works
datacite.rightsOpen Access
dc.contributor.advisorAli, Md. Haider
dc.contributor.advisorChaki, Dipankar
dc.contributor.authorHasan, H. M. Mahedi
dc.contributor.authorSanyal, Falguni
dc.contributor.departmentDepartment of Computer Science and Engineering
dc.date.accessioned2017-06-15T05:16:02Z
dc.date.available2017-06-15T05:16:02Z
dc.date.copyright2017
dc.date.issued2017-04
dc.descriptionCataloged from PDF version of thesis report.
dc.descriptionIncludes bibliographical references (page 50 - 51).
dc.descriptionThis thesis report is submitted in partial fulfillment of the requirements for the degree of Bachelor of Science in Computer Science and Engineering, 2017.en_US
dc.description.abstractKeyword extraction is an automatic selection of terms which describes the content of a document. Keywords define the terms that represent the core information from the documents. In order to go through massive amount of documents to find out the relevant information, keyword extraction will be the key approach. This approach will help us to understand the depth of a document even before we read it. In this research, we have found out different approaches and algorithms that have been used in keyword extraction technique. Conditional random fields (CRF), Support vector machine (SVM), NP-chunk, N-grams, Multiple linear regression, Logistic regression, and semantic analysis has been used to find out important keywords from a document. Immense research shows us that SVM and CRF gives better results where CRF accuracy is greater than SVM based on F1 score (The balance between precision and recall). According to precision, SVM shows better result than CRF. But, in case of recall, logit shows the greater result. Semantic relation between words is also another key feature in keyword extraction techniques. Semantic analysis is very effective field in natural language processing and using semantic relation, it is possible to find out the relation between words as well as between the lines. In this thesis paper, we have used semantic analysis and processing the documents to find out the important keywords from documents.en_US
dc.description.degreeBachelor of Science in Computer Science and Engineering
dc.format.extent51 pages
dc.identifier.otherID 13101270
dc.identifier.otherID 13301058
dc.identifier.urihttp://hdl.handle.net/10361/8245
dc.language.isoenen_US
dc.rightsBRAC University thesis are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission.
dc.subjectNatural Language Processing (NLP)en_US
dc.subjectSemantic analysisen_US
dc.subjectTextBloben_US
dc.subjectPOS-taggingen_US
dc.subjectN-gramsen_US
dc.subjectKeyworden_US
dc.titleImportant keywords extraction from documents using semantic analysisen_US
dc.typeThesisen_US

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
13101270, 13301058 _CSE.pdf
Size:
940.58 KB
Format:
Adobe Portable Document Format
Description:

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.71 KB
Format:
Item-specific license agreed upon to submission
Description: