Evaluating question generation models using QA systems and semantic textual similarity
| bracu.type.group | Research Publications | |
| datacite.rights | Metadata Only | |
| dc.contributor.author | Shaheer, Safwan | |
| dc.contributor.author | Hossain, Ishmam | |
| dc.contributor.author | Sarna, Sudipta Nandi | |
| dc.contributor.author | Kabir Mehedi, Md Humaion | |
| dc.contributor.author | Rasel, Annajiat Alim | |
| dc.contributor.department | Department of Computer Science and Engineering | |
| dc.date.accessioned | 2026-07-26T09:51:15Z | |
| dc.date.available | 2026-07-26T09:51:15Z | |
| dc.date.issued | 2023-01-01 | |
| dc.description.abstract | Question generation based on conversational context is a difficult problem to solve. A widely used technique for generating quality questions using fine-tuned models relies on a suitable answer and the context, usually the passage. But when it comes to conversational settings, the questions generated are not of the highest quality as they lack the contextual element in the question, especially due to the lack of co-reference resolution of the entity. Furthermore, in most of the evaluation techniques for generating questions, there seems to be a lack of utilizing powerful question-answering systems to judge the answerability of the questions generated. The most prevalent metric used for judging machine-generated text against the human gold standard, BLUE, unfortunately doesn't factor in whether a question answering system would be able to answer the question, but instead focuses mostly on the number of substrings that match against each other. Various question generation models following a generalized encoder-decoder architecture were evaluated using semantic textual similarity for both the generated questions and the generated answers. Although higher parameters in a model usually lend to better performance, our experiment displayed that such is not always the case, at least when there is a massive amount of context missing. | |
| dc.description.version | Publisher | |
| dc.format.extent | 431-435 | |
| dc.identifier.citation | S. Shaheer, I. Hossain, S. N. Sarna, M. H. Kabir Mehedi and A. A. Rasel, "Evaluating Question generation models using QA systems and Semantic Textual Similarity," 2023 IEEE 13th Annual Computing and Communication Workshop and Conference (CCWC), Las Vegas, NV, USA, 2023, pp. 0431-0435, doi: 10.1109/CCWC57344.2023.10099244. | |
| dc.identifier.doi | 10.1109/CCWC57344.2023.10099244 | |
| dc.identifier.issn | 9798350332865 | |
| dc.identifier.other | 2-s2.0-85156246478 | |
| dc.identifier.uri | https://hdl.handle.net/10361/28646 | |
| dc.language.iso | en_US | |
| dc.publisher | Institute of Electrical and Electronics Engineers Inc. | |
| dc.relation.hasversion | 10.1109/CCWC57344.2023.10099244 | |
| dc.relation.ispartof | 2023 IEEE 13th Annual Computing and Communication Workshop and Conference Ccwc 2023 | |
| dc.relation.ispartofseries | 2023 IEEE 13th Annual Computing and Communication Workshop and Conference Ccwc 2023 | |
| dc.relation.uri | https://ieeexplore.ieee.org/document/10099244 | |
| dc.subject | BLEU | |
| dc.subject | Question answering | |
| dc.subject | Question generation | |
| dc.subject | Semantic textual similarity | |
| dc.subject.lcsh | Question-answering systems. | |
| dc.subject.lcsh | Natural language processing (Computer science). | |
| dc.subject.lcsh | Text processing (Computer science). | |
| dc.title | Evaluating question generation models using QA systems and semantic textual similarity | |
| dc.type | Conference Proceeding | |
| person.affiliation.name | BRAC University | |
| person.affiliation.name | BRAC University | |
| person.affiliation.name | BRAC University | |
| person.affiliation.name | BRAC University | |
| person.affiliation.name | BRAC University | |
| person.identifier.scopus-author-id | 58223344100 | |
| person.identifier.scopus-author-id | 57203035707 | |
| person.identifier.scopus-author-id | 58222396200 | |
| person.identifier.scopus-author-id | 57971673000 | |
| person.identifier.scopus-author-id | 56495276900 |