<?xml version="1.0" encoding="UTF-8"?><!DOCTYPE article  PUBLIC "-//NLM//DTD Journal Publishing DTD v3.0 20080202//EN" "http://dtd.nlm.nih.gov/publishing/3.0/journalpublishing3.dtd"><article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" dtd-version="3.0" xml:lang="en" article-type="research article"><front><journal-meta><journal-id journal-id-type="publisher-id">JSSM</journal-id><journal-title-group><journal-title>Journal of Service Science and Management</journal-title></journal-title-group><issn pub-type="epub">1940-9893</issn><publisher><publisher-name>Scientific Research Publishing</publisher-name></publisher></journal-meta><article-meta><article-id pub-id-type="doi">10.4236/jssm.2021.143024</article-id><article-id pub-id-type="publisher-id">JSSM-110281</article-id><article-categories><subj-group subj-group-type="heading"><subject>Articles</subject></subj-group><subj-group subj-group-type="Discipline-v2"><subject>Business&amp;Economics</subject></subj-group></article-categories><title-group><article-title>
 
 
  Development of Answer Validation System Using Responders’ Attributes and Crowd Ranking
 
</article-title></title-group><contrib-group><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Mercy</surname><given-names>Adebisi</given-names></name><xref ref-type="aff" rid="aff1"><sup>1</sup></xref></contrib><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Bolanle</surname><given-names>Ojokoh</given-names></name><xref ref-type="aff" rid="aff2"><sup>2</sup></xref><xref ref-type="corresp" rid="cor1"><sup>*</sup></xref></contrib><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Tolulope</surname><given-names>Adebayo</given-names></name><xref ref-type="aff" rid="aff1"><sup>1</sup></xref></contrib><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Akintoba</surname><given-names>Akinwonmi</given-names></name><xref ref-type="aff" rid="aff1"><sup>1</sup></xref></contrib><contrib contrib-type="author" xlink:type="simple"><name name-style="western"><surname>Fatai</surname><given-names>Sunmola</given-names></name><xref ref-type="aff" rid="aff3"><sup>3</sup></xref></contrib></contrib-group><aff id="aff3"><addr-line>Department of Information Technology, Federal University of Technology, Akure, Nigeria</addr-line></aff><aff id="aff2"><addr-line>Department of Information Systems, Federal University of Technology, Akure, Nigeria</addr-line></aff><aff id="aff1"><addr-line>Department of Computer Science, Federal University of Technology, Akure, Nigeria</addr-line></aff><pub-date pub-type="epub"><day>11</day><month>05</month><year>2021</year></pub-date><volume>14</volume><issue>03</issue><fpage>382</fpage><lpage>398</lpage><history><date date-type="received"><day>22,</day>	<month>February</month>	<year>2021</year></date><date date-type="rev-recd"><day>27,</day>	<month>June</month>	<year>2021</year>	</date><date date-type="accepted"><day>30,</day>	<month>June</month>	<year>2021</year></date></history><permissions><copyright-statement>&#169; Copyright  2014 by authors and Scientific Research Publishing Inc. </copyright-statement><copyright-year>2014</copyright-year><license><license-p>This work is licensed under the Creative Commons Attribution International License (CC BY). http://creativecommons.org/licenses/by/4.0/</license-p></license></permissions><abstract><p>
 
 
  Crowdsourcing has found a wide range of application in Community Question Answering (CQA). However, one of its biggest challenges is the need to address the quality of crowd answers contributions. Therefore, this work proposed a system that seeks to validate answers to questions provided by respondents using responders’ attributes and crowd ranking technique. Weights were assigned to respondent answers based on their academic records, experience and understanding of the question to obtain valid answers. Thereafter, valid answers were ranked by the crowd using Borda Count algorithm. The proposed system was evaluated using Usability and User experience (UX) measurement. The result obtained demonstrated the effectiveness of the applied technique.
 
</p></abstract><kwd-group><kwd>Askers</kwd><kwd> Answerers</kwd><kwd> Community Question Answering (CQA)</kwd><kwd> Question Answering (QA)</kwd><kwd> Crowd Sourcing</kwd><kwd> Answer Validation</kwd></kwd-group></article-meta></front><body><sec id="s1"><title>1. Introduction</title><p>The new information era provides readily available access to information, especially with the advent of the internet. Different questions requiring correct answers are uploaded on the internet on daily basis which leads to the development of question answering (QA) systems, with the aim of providing accurate answers to explicit questions which are contrasting to document retrieval (Ojokoh &amp; Adebisi, 2019</p><p>Several studies have been carried out on how to make better the quality of the answers provided by QA system, focusing on textual entailment, question type analysis, answer ranking by the crowd workers and domain experts and personal and community features (past history) of the answerer to determine the quality of the answers (R&#237;os-Gaona et al., 2012; Su et al., 2007; Ishikawa et al., 2011; Ojokoh &amp; Ayokunle, 2012; Anderson et al., 2012; Schofield &amp; Thielscher, 2019). Since past history alone may not be fitting enough to determine the quality of an answer, level of confidence in the answer provided is introduced in order to obtain credible answers from respondents. The proposed system is aimed at using community presence interaction as one of the basis for quality answer selection; capturing crowd specialty as part of the personal features used to validate answers; modelling the criteria used in evaluation automatically and preventing bias crowd ranking of answers by enabling them to specify their preferential schedule using Na&#239;ve Bayes Spam filter and Borda count ranking Algorithm.</p><p>The remaining part of this paper is structured as follows: Section 2 presents the review of related works. Section 3 presents the proposed system architecture, and the description of the components that make up the architecture. Section 4 is dedicated to the experimental setup and results while Section 5, concludes the paper and presents some future works.</p></sec><sec id="s2"><title>2. Related Works</title><p>Question Answering (QA) according to Chandra et al. (2017) is a computer science discipline concerned with developing a system that automatically provide answers to questions requested by human in a natural language. QA study attempts to deal with a wide-ranging question types that consist of facts, lists, definitions, how, why, putative, semantically constrained, and cross lingual questions (Cimiano et al., 2014</p><p>Dobšovič et al. (2014) proposed and developed a CQA system “Askalot” which is focused on the area of education by implementing a functionality that encompasses the educational goal and specifics of universities, based on open source technologies. Answers to questions are verified by other students, comments are however provided by a teacher using a five-grade scale on which the assessment of the quality of question or answer can be done. Toba et al. (2014) proposed a hybrid hierarchy-of-classifiers framework to model QA pairs and integrate the question type analysis and answer quality information in an integrated framework. The quality classifier gives two probabilities each, showing the probability of good or bad-quality. They tested the framework on a dataset of about 50 thousand QA pairs from Yahoo! Answers and an effective identification of high quality answers was realized based on their evaluation of the system. Tran et al. (2015) presented a method to detect the right or possible right answers from the answer thread in Community Question Answering pools. They used multiple features for quality answer selection which exploits the surface word-based similarity between the question and answer to allot score using a regression model. Afterwards, translation probabilities were computed via IBM and Hidden Markov Models to obtain the likelihood of an answer being the translation of the question. Savenkov et al. (2016) presented a system that could be used to filter or re-rank the candidate answers by providing validation for the answers. They specifically focused on knowing the effect of time restrictions in the close real-time QA setting, thereby developing a way in which crowd will be able to create the answer candidates directly within a limited amount of time and also the way in which crowd will be able to rank sets of given answers to a question within a specified amount of time. Hung et al. (2017) developed a probabilistic model that helps to recognise the most valuable validation questions in improving results’ accuracy and detecting faulty workers in their quest to validate and control the quality of crowd answers to reduce cost incurred from utilizing experts.</p><p>Nie et al. (2017) presented a novel scheme to rank answer candidates via pairwise comparisons consisting of one offline learning and one online search component. In the online search component, a pool of candidate answers for the given question was extracted via finding its similar questions. The extracted answers were then sorted by leveraging the offline trained model to judge the preference orders.</p><p>Fan et al. (2019) proposed to enhance answer selection in CQA using multidimensional feature combination and similarity order. They made full use of the information in answers to questions to determine the similarity between questions and answers, and use the text-based description of the answer to determine its sensibility. Le et al. (2019) proposed a framework for automatically assessing answer quality by integrating different groups of features such as personal, community-based, textual, and contextual, to build a classification model and determine what constitutes answer quality. Experiments conducted on Brainly and stack overflow datasets show that the random forest model achieves high accuracy in identifying high-quality answers. Also indicating that personal and community-based features have more prediction power in assessing answer quality.</p><p>In this paper, we leverage on the fact that the performance of the crowd workers determines the quality of the result of a crowdsourcing task, and hence the need to develop an effective and reliable question answering system that is capable of validating and evaluating the answers provided by the crowd because of their varying reliability as established in past works (Hung et al., 2017; Savenkov et al., 2016). All these are important issues to be addressed in Artificial Intelligence.</p></sec><sec id="s3"><title>3. The Proposed System</title><p>The architectural overview of the proposed system is presented in <xref ref-type="fig" rid="fig1">Figure 1</xref>. The subsections that follow describes each of the segments.</p><sec id="s3_1"><title>3.1. User Interface</title><p>The user interface module consists of four (4) components listed as follows:</p><p>1) Ask Question: This component enables the asker (that is someone who wishes to ask any computer-related questions) to post his/her questions on the platform.</p><p>2) Answer Question: This component enables experts or anyone familiar with the question asked to provide answers.</p><p>3) Rank Answers: This allow users from the crowd to rank answers provided by other users based on their knowledge of the question.</p><p>4) View Recent Questions: This component provides a view of the list of the most recently posted questions.</p></sec><sec id="s3_2"><title>3.2. Database</title><p>The database is the component of the Answer Validation model that stores information about the system and its users. It stores both legitimate questions and answers from web users, and most importantly, answerers’ personal information for the purpose of validating their answers which is obtained the first time a respondent uses the system.</p></sec><sec id="s3_3"><title>3.3. Na&#239;ve Bayes Spam Filter</title><p>Na&#239;ve Bayes (NB) Spam Filter, a machine learning algorithm, which is one of the powerful tools for Artificial Intelligence was used in this work to filter inconsequential and redundant messages from the collection of messages or information provided by the crowd. Every incoming text (both question and answer) pass through the trained Na&#239;ve Bayes Spam filter to determine the probability of the message being a legitimate message or spam. The NB spam filter is trained with the commonly used online spam words and spam dataset downloaded from kaggle.com. A sample is shown in <xref ref-type="fig" rid="fig2">Figure 2</xref>.</p><p>From Bayes’ theorem, the probability that a message with vector X = ( X 1 , ⋯ , X m ) belongs in category c is:</p><p>P ( c | x ) = p ( c ) ⋅ p ( x | c ) p ( x ) (1)</p><p>Using Na&#239;ve Bayes Spam filter, a message is classified as spam whenever</p><p>P = p ( c s ) ⋅ p ( x | c s ) p ( c s ) ⋅ p ( x | c s ) + p ( c h ) ⋅ p ( x | c h ) (2)</p><p>P { &gt; T ,     message is spam ≤ T ,     message not spam (3)</p><p>where c s is a message in spam category; c h is a message in ham category;</p><p>p ( c s ) is the probability that the response x belongs to spam category, c s ;</p><p>p ( c h ) is the probability that the response x belongs to ham category, c h ;</p><p>p ( x | c s ) is the likelihood of response x given the spam category;</p><p>p ( x | c h ) is the likelihood of response x given the ham category and;</p><p>T is a threshold value.</p><p>If P is greater than T, the incoming message is being classified as spam message and will be discarded else if P is less than or equal to T, the message will be accepted by the system and presented as a question or accepted as an incoming answer.</p></sec><sec id="s3_4"><title>3.4. Separate Question from Answer</title><p>This is the component of the system where a legitimate message from the user is being identified as either a question or answer. If the incoming message is a</p><p>question, this component ensures that the question is presented at the User Interface for the answerers to provide answers, and if otherwise, the system will pass it to the next component where the criteria for quality answers will be implemented.</p></sec><sec id="s3_5"><title>3.5. Criteria for Quality Answers</title><p>The quality of the result of a question answering system rest on the source of the answers provided by the system. Since the aim of the question answering system is to provide a precise answer in natural language; it is therefore important to provide quality assurance on every answer obtained from the web users, as these users can vary in reliability. The criteria employed for validation and used to ensure quality answers in this work are User attributes, Area of Specialization, Understandability and Confidence (displayed in <xref ref-type="table" rid="table1">Table 1</xref>).</p></sec><sec id="s3_6"><title>3.6. Weighted Voting System</title><p>A game playing situation is applied for ranking answers using a collection of weighted players P i together with a quotaq, which is the total number of votes required to pass a motion. This is used to determine the level of reliability of the users that provide answers. A player is a user attribute that is used to allot point to answerers. In a weighted voting system, a player’s weight w i refers to the number of points allotted to that player and is always a positive integer value. A weighted voting system is described by specifying the voting weights, w 1 , w 2 , ⋯ , w n of the players P 1 , P 2 , ⋯ , P n , and the quota, q. A coalition is called winning if the sum of the players’ weights is greater or equal to the quota, and losing if otherwise. The coalitions, which are the criteria used in this work to ensure quality answers from the web users are User attributes, area of specialization, Understandability and Confidence. User attributes that are used comprises of user Course of study, Grade point, number of years of experience in computing and the general level of knowledge of computing. Point is added to the weight of the responder based on their selections from the range of value of the attributes. A user is also allowed to choose any area of specialization such as Networking, Cyber Security and hardware and repairs and so on. Users’ understandability of the given question is measured based on a five-level rating scale, as well as the Confidence which is a way in which the answerer can infer how much the system can trust the answer provided. This is also measured based on a five level rating scale. Combining these and the weighted voting system, this phase of the system is represented by:</p><p>q : P 1 , P 2 , P 3 , P 4</p><p>where,</p><p>P<sub>1 </sub>is User’s personal attribute, P<sub>2</sub> is Specialization;</p><p>P<sub>3</sub> is Understandability, P<sub>4</sub> is Confidence.</p><p>The totality of weights, T w per Answerer is computed as:</p><table-wrap id="table1" ><label><xref ref-type="table" rid="table1">Table 1</xref></label><caption><title> Weight distribution table</title></caption><table><tbody><thead><tr><th align="center" valign="middle" >S/N</th><th align="center" valign="middle" >Criteria for quality Assurance</th><th align="center" valign="middle" >Description</th><th align="center" valign="middle" >Metrics</th><th align="center" valign="middle" >Points</th></tr></thead><tr><td align="center" valign="middle"  rowspan="17"  >1</td><td align="center" valign="middle"  rowspan="17"  >User attributes (P<sub>1</sub>).</td><td align="center" valign="middle"  rowspan="5"  >Grade point.</td><td align="center" valign="middle" >First class</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" >Second class upper</td><td align="center" valign="middle" >4</td></tr><tr><td align="center" valign="middle" >Second class lower</td><td align="center" valign="middle" >3</td></tr><tr><td align="center" valign="middle" >Third class</td><td align="center" valign="middle" >2</td></tr><tr><td align="center" valign="middle" >Pass</td><td align="center" valign="middle" >1</td></tr><tr><td align="center" valign="middle"  rowspan="2"  >Course of study during Undergraduate.</td><td align="center" valign="middle" >Computer Science related course</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" >Other science related course</td><td align="center" valign="middle" >3</td></tr><tr><td align="center" valign="middle"  rowspan="5"  >Numbers of Years of experience.</td><td align="center" valign="middle" >21 yrs and above</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" >16 - 20 yrs</td><td align="center" valign="middle" >4</td></tr><tr><td align="center" valign="middle" >11 - 15 yrs</td><td align="center" valign="middle" >3</td></tr><tr><td align="center" valign="middle" >6 - 10 yrs</td><td align="center" valign="middle" >2</td></tr><tr><td align="center" valign="middle" >1 - 5 yrs</td><td align="center" valign="middle" >1</td></tr><tr><td align="center" valign="middle"  rowspan="5"  >General level of computing.</td><td align="center" valign="middle" >Very high</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" >High</td><td align="center" valign="middle" >4</td></tr><tr><td align="center" valign="middle" >Medium</td><td align="center" valign="middle" >3</td></tr><tr><td align="center" valign="middle" >Low</td><td align="center" valign="middle" >2</td></tr><tr><td align="center" valign="middle" >Very low</td><td align="center" valign="middle" >1</td></tr><tr><td align="center" valign="middle"  rowspan="2"  >2</td><td align="center" valign="middle"  rowspan="2"  >Area of specialization (P<sub>2</sub>)</td><td align="center" valign="middle"  rowspan="2"  >Area of specialization of answerers.</td><td align="center" valign="middle" >Specialize area</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" >Non-specialize area</td><td align="center" valign="middle" >3</td></tr><tr><td align="center" valign="middle"  rowspan="5"  >3</td><td align="center" valign="middle"  rowspan="5"  >Understandability (P<sub>3</sub>)</td><td align="center" valign="middle"  rowspan="5"  >Level of Understanding of the question by the answerers</td><td align="center" valign="middle" >Very high</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" >High</td><td align="center" valign="middle" >4</td></tr><tr><td align="center" valign="middle" >Medium</td><td align="center" valign="middle" >3</td></tr><tr><td align="center" valign="middle" >Low</td><td align="center" valign="middle" >2</td></tr><tr><td align="center" valign="middle" >Very low</td><td align="center" valign="middle" >1</td></tr><tr><td align="center" valign="middle"  rowspan="5"  >4</td><td align="center" valign="middle"  rowspan="5"  >Confidentiality (P<sub>4</sub>).</td><td align="center" valign="middle"  rowspan="5"  >Level of Confidence of the answerers in their answer</td><td align="center" valign="middle" >Very high</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" >High</td><td align="center" valign="middle" >4</td></tr><tr><td align="center" valign="middle" >Medium</td><td align="center" valign="middle" >3</td></tr><tr><td align="center" valign="middle" >Low</td><td align="center" valign="middle" >2</td></tr><tr><td align="center" valign="middle" >Very low</td><td align="center" valign="middle" >1</td></tr></tbody></table></table-wrap><p>T w = ∑ i = 1 4     w i (4)</p><p>where w i is the weight corresponding to each player, P i . The maximum weight, N obtainable by an answerer with q being the minimum weight required for an acceptable (valid) answer is expressed as:</p><p>N = w 1 + w 2 + w 3 + w 4 (5)</p><p>then, N 2 &lt; q ≤ N holds for equation (6)</p><p>In this work, q was obtained by calculating the 70% of N as follows:</p><p>q = 70 % N</p><p>From Equation (6), q can be said to be less than or equal to N but greater than N 2 . This means that 35 2 &lt; q ≤ 35 . Since this work is based on quality answer validation, 70% of N was used as the quota q.</p><p>Quota ( q ) = 35 100 &#215; 70 = 24.5 = 25 ( approx . ) .</p><p>Therefore the quota, q will be 25. <xref ref-type="table" rid="table2">Table 2</xref> depicts the different criteria considered in this work with the respective maximum weight obtainable.</p><p>Depending on the point obtained from each criterion by the Responders (Answerers), these points are aggregated based on their selection. The total weight of the answer is calculated to check whether the weight meets up to the quota. If the total weight of the answer is greater or equal to the quota, the answer is considered valid and is passed to the next phase which is the ranking phase,and if not the answer is discarded.</p></sec><sec id="s3_7"><title>3.7. Crowd Ranking</title><p>The last phase employs a crowdsourcing ranking algorithm called Borda count. The algorithm ranks all the valid answers from phase two using a preference schedule point. It awards points to candidates based on preference schedule, then the candidate with the highest points is declared the winner. For instance, given M, the number of candidate answers, each first-place, second-place and third-place votes is worth M , M − 1 , M − 2 points respectively. Consequently, each Mth-place (that is, last-place) vote is worth 1 point. Now, suppose there are n voters, every voter ranks the M candidates according to his preference, and a candidate answer has an average rank score, s n .</p><p>s n = ∑ i = 1 n     r i (7)</p><p>where r i is the point assigned by n crowd (ranker).</p><table-wrap id="table2" ><label><xref ref-type="table" rid="table2">Table 2</xref></label><caption><title> Maximum weight obtainable (N)</title></caption><table><tbody><thead><tr><th align="center" valign="middle" >S/N</th><th align="center" valign="middle" >Criteria</th><th align="center" valign="middle" >Maximum weight Obtainable (N)</th></tr></thead><tr><td align="center" valign="middle" >1.</td><td align="center" valign="middle" >User Attributes (P<sub>1</sub>)</td><td align="center" valign="middle" >20</td></tr><tr><td align="center" valign="middle" >2.</td><td align="center" valign="middle" >Area of specialization (P<sub>2</sub>)</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" >3.</td><td align="center" valign="middle" >Users understandability (P<sub>3</sub>)</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" >4.</td><td align="center" valign="middle" >Users confidentiality (P<sub>4</sub>)</td><td align="center" valign="middle" >5</td></tr><tr><td align="center" valign="middle" ></td><td align="center" valign="middle" >Total</td><td align="center" valign="middle" >35</td></tr></tbody></table></table-wrap><p>The candidate answers will be ranked according to their performance starting from the best on top of the list (answer with the highest point) to the worst (answer with the lowest point).</p></sec></sec><sec id="s4"><title>4. Experiments and Evaluation</title><sec id="s4_1"><title>4.1. Data and Tools</title><p>A dataset consisting of 185 Spam messages was downloaded from Kaggle.com and was used to train the Na&#239;ve Bayes Filter in order to distinguish between legitimate and inconsequential information provided by the crowd. The system was implemented using HTML, Python Script and Djangoweb framework.</p></sec><sec id="s4_2"><title>4.2. Experimental Setup</title><p>Experiments were conducted to verify the system performance and to determine how useful and precise the answers provided were. The users of the system are allowed to post questions which will be answered by responders who are vast in the field of the question being asked. However, before the responders would be allowed to provide answers, they will be required to sigin/sign up as the case may be, verifying their Course of study, Area of specialization, Grade point, number of years of experience in Computing, general level of Computing knowledge and the level of understanding of the question. Also, the confidence level of the responder will be confirmed before posting the answer. In cases where a minimum of five different answers are provided to a particular question, they are ranked by the crowd starting from the most correct to the least correct answer. A sample of asked questions and answers provided is shown in <xref ref-type="fig" rid="fig3">Figure 3</xref>.</p></sec><sec id="s4_3"><title>4.3. Evaluation</title><p>The method of evaluation used in this work is based on ISO/IEC 9126 standard metrics and the Usability and User experience (UX) measurement instruments adopted in (Tan et al., 2010). The model consists of 21 subcharacteristics distributed on six main characteristics of software measurement metrics. Using the common Goal Question Metric (GQM) approach, a nomenclature for usability and UX attributes were defined and were able to identify an extensive set of questions and measures for each attribute. The metrics used for this work are shown in <xref ref-type="table" rid="table3">Table 3</xref>.</p><p>From the above stated metrics, twenty (20) questions were formed in order to evaluate the Answer Validation system by Users. Eighty five users out of One hundred sample size evaluated the system, with each question (Q<sub>1</sub>, …, Q<sub>20</sub>) answered using four-level rating scale; Very High, High, Medium and Low respectively. Ratings obtained from the Users were analyzed using weight means techniques in which weights are added (such that Very High = 4, High = 3, Medium = 2 and Low = 1) to users feedback. A sample of the questionnaire is shown in <xref ref-type="table" rid="table4">Table 4</xref>.</p></sec><sec id="s4_4"><title>4.4. Results and Discussion</title><p>The ratings were analyzed and the frequency at which each point occurs was obtained. The metrics were measured and analyzed to form a continuous score in percentage (%). <xref ref-type="table" rid="table5">Table 5</xref> illustrates the number of users out of eighty-five (85) that rated the system either Very high, High, Medium or Low based on the given questionnaire. <xref ref-type="fig" rid="fig4">Figure 4</xref> and <xref ref-type="fig" rid="fig5">Figure 5</xref> shows the graphical representation of the obtained results. <xref ref-type="table" rid="table6">Table 6</xref> shows the Combination of Very High and High ratings in order to define the User ratings as High, Medium, Low. <xref ref-type="fig" rid="fig6">Figure 6</xref> and <xref ref-type="fig" rid="fig7">Figure 7</xref> show the Combination of Very High and High ratings for Usability and User Experience respectively.</p><table-wrap id="table3" ><label><xref ref-type="table" rid="table3">Table 3</xref></label><caption><title> Usability and user experience metrics</title></caption><table><tbody><thead><tr><th align="center" valign="middle" >S/N</th><th align="center" valign="middle" >Usability metrics</th><th align="center" valign="middle" >User Experience (UX) metrics</th></tr></thead><tr><td align="center" valign="middle" >1.</td><td align="center" valign="middle" >Understandability</td><td align="center" valign="middle" >Correctness</td></tr><tr><td align="center" valign="middle" >2.</td><td align="center" valign="middle" >Efficiency</td><td align="center" valign="middle" >Satisfaction</td></tr><tr><td align="center" valign="middle" >3.</td><td align="center" valign="middle" >Error tolerance</td><td align="center" valign="middle" >Simplicity</td></tr><tr><td align="center" valign="middle" >4.</td><td align="center" valign="middle" >Ease of use</td><td align="center" valign="middle" >Validation</td></tr><tr><td align="center" valign="middle" >5.</td><td align="center" valign="middle" >Attractiveness</td><td align="center" valign="middle" >Effectiveness</td></tr><tr><td align="center" valign="middle" >6.</td><td align="center" valign="middle" >Time response</td><td align="center" valign="middle" >Quality of outcome</td></tr><tr><td align="center" valign="middle" >7.</td><td align="center" valign="middle" >Visualization</td><td align="center" valign="middle" >Reliability</td></tr><tr><td align="center" valign="middle" >8.</td><td align="center" valign="middle" >Navigability</td><td align="center" valign="middle" >Consistency</td></tr><tr><td align="center" valign="middle" >9.</td><td align="center" valign="middle" >Reusability</td><td align="center" valign="middle" >Accessibility</td></tr><tr><td align="center" valign="middle" >10.</td><td align="center" valign="middle" >Feedback</td><td align="center" valign="middle" >Preferability</td></tr></tbody></table></table-wrap><table-wrap id="table4" ><label><xref ref-type="table" rid="table4">Table 4</xref></label><caption><title> Questionnaire for answer validation system evaluation</title></caption><table><tbody><thead><tr><th align="center" valign="middle" >S/N</th><th align="center" valign="middle" >Usability metrics</th><th align="center" valign="middle" >User Experience (UX) metrics</th></tr></thead><tr><td align="center" valign="middle" >1.</td><td align="center" valign="middle" >What is the rate at which you understand the system?</td><td align="center" valign="middle" >What is the rate at which the answers provided by the system are correct?</td></tr><tr><td align="center" valign="middle" >2.</td><td align="center" valign="middle" >What is the rate at which you think the system is efficient?</td><td align="center" valign="middle" >What is the rate at which you are satisfied with the answers provided by the system</td></tr><tr><td align="center" valign="middle" >3.</td><td align="center" valign="middle" >What is the rate at which the system tolerates error and corrects you when you made mistakes?</td><td align="center" valign="middle" >What is the rate at which the language used by the system is simple?</td></tr><tr><td align="center" valign="middle" >4.</td><td align="center" valign="middle" >What is the rate at which the system is easy to use?</td><td align="center" valign="middle" >What is the rate at which the answers provided by the system in corresponding to their questions are valid?</td></tr><tr><td align="center" valign="middle" >5.</td><td align="center" valign="middle" >What is the rate at which the system design is attractive?</td><td align="center" valign="middle" >What is the rate at which the system is effective enough in providing valid answer to questions?</td></tr><tr><td align="center" valign="middle" >6.</td><td align="center" valign="middle" >What is the rate at which you are satisfied with the time response of the system?</td><td align="center" valign="middle" >What is the rate at which the system can provide high quality answers?</td></tr><tr><td align="center" valign="middle" >7.</td><td align="center" valign="middle" >What is the rate at which you are satisfied with the visual content of the system?</td><td align="center" valign="middle" >What is the rate at which the system is reliable in providing answers to computing related questions?</td></tr><tr><td align="center" valign="middle" >8.</td><td align="center" valign="middle" >What is the rate at which you find it easy to Navigate through the system?</td><td align="center" valign="middle" >What is the rate at which the system is consistent in performing its functions?</td></tr><tr><td align="center" valign="middle" >9.</td><td align="center" valign="middle" >What is the rate at which you will like to use the system the next time?</td><td align="center" valign="middle" >What is the rate at which the system is accessible from your end?</td></tr><tr><td align="center" valign="middle" >10.</td><td align="center" valign="middle" >What is the rate at which you are satisfied with the system feedback?</td><td align="center" valign="middle" >What is the rate at which you prefer the system to others?</td></tr></tbody></table></table-wrap><table-wrap id="table5" ><label><xref ref-type="table" rid="table5">Table 5</xref></label><caption><title> User rating frequency table and their percentage</title></caption><table><tbody><thead><tr><th align="center" valign="middle" >Metrics/Ratings</th><th align="center" valign="middle"  colspan="2"  >Very High</th><th align="center" valign="middle"  colspan="2"  >High</th><th align="center" valign="middle"  colspan="2"  >Medium</th><th align="center" valign="middle"  colspan="2"  >Low</th></tr></thead><tr><td align="center" valign="middle" >Correctness</td><td align="center" valign="middle" >44</td><td align="center" valign="middle" >51.76%</td><td align="center" valign="middle" >38</td><td align="center" valign="middle" >44.71%</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3.53%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Satisfaction</td><td align="center" valign="middle" >46</td><td align="center" valign="middle" >54.12%</td><td align="center" valign="middle" >39</td><td align="center" valign="middle" >45.88%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Validation</td><td align="center" valign="middle" >43</td><td align="center" valign="middle" >50.59%</td><td align="center" valign="middle" >40</td><td align="center" valign="middle" >47.05%</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2.35%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Effectiveness</td><td align="center" valign="middle" >30</td><td align="center" valign="middle" >35.29%</td><td align="center" valign="middle" >50</td><td align="center" valign="middle" >58.82%</td><td align="center" valign="middle" >5</td><td align="center" valign="middle" >5.88%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Quality of Outcome</td><td align="center" valign="middle" >35</td><td align="center" valign="middle" >41.17%</td><td align="center" valign="middle" >48</td><td align="center" valign="middle" >56.47%</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2.35%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Reliability</td><td align="center" valign="middle" >24</td><td align="center" valign="middle" >28.24%</td><td align="center" valign="middle" >60</td><td align="center" valign="middle" >70.58%</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >1.18%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Consistency</td><td align="center" valign="middle" >58</td><td align="center" valign="middle" >68.24%</td><td align="center" valign="middle" >25</td><td align="center" valign="middle" >29.41%</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2.35%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Accessibility</td><td align="center" valign="middle" >42</td><td align="center" valign="middle" >49.41%</td><td align="center" valign="middle" >42</td><td align="center" valign="middle" >49.41%</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >1.18%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Fault Tolerance</td><td align="center" valign="middle" >20</td><td align="center" valign="middle" >23.53%</td><td align="center" valign="middle" >57</td><td align="center" valign="middle" >67.05%</td><td align="center" valign="middle" >8</td><td align="center" valign="middle" >9.41%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Preferability</td><td align="center" valign="middle" >18</td><td align="center" valign="middle" >21.17%</td><td align="center" valign="middle" >65</td><td align="center" valign="middle" >76.47%</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2.35%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Ease of use</td><td align="center" valign="middle" >34</td><td align="center" valign="middle" >40.0%</td><td align="center" valign="middle" >50</td><td align="center" valign="middle" >58.82%</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >1.18%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Navigability</td><td align="center" valign="middle" >48</td><td align="center" valign="middle" >56.47%</td><td align="center" valign="middle" >34</td><td align="center" valign="middle" >40.00%</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3.53%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Simplicity</td><td align="center" valign="middle" >30</td><td align="center" valign="middle" >35.29%</td><td align="center" valign="middle" >53</td><td align="center" valign="middle" >62.35%</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2.35%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Understandability</td><td align="center" valign="middle" >39</td><td align="center" valign="middle" >45.88%</td><td align="center" valign="middle" >45</td><td align="center" valign="middle" >52.94%</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3.53%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Attractiveness</td><td align="center" valign="middle" >38</td><td align="center" valign="middle" >44.71%</td><td align="center" valign="middle" >39</td><td align="center" valign="middle" >45.88%</td><td align="center" valign="middle" >8</td><td align="center" valign="middle" >9.41%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Time Response</td><td align="center" valign="middle" >32</td><td align="center" valign="middle" >37.65%</td><td align="center" valign="middle" >28</td><td align="center" valign="middle" >32.94%</td><td align="center" valign="middle" >25</td><td align="center" valign="middle" >29.41%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Visualization</td><td align="center" valign="middle" >20</td><td align="center" valign="middle" >23.53%</td><td align="center" valign="middle" >56</td><td align="center" valign="middle" >65.88%</td><td align="center" valign="middle" >9</td><td align="center" valign="middle" >10.58%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Reusability</td><td align="center" valign="middle" >20</td><td align="center" valign="middle" >23.53%</td><td align="center" valign="middle" >62</td><td align="center" valign="middle" >72.94%</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3.53%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Feedback</td><td align="center" valign="middle" >21</td><td align="center" valign="middle" >24.71%</td><td align="center" valign="middle" >54</td><td align="center" valign="middle" >63.53%</td><td align="center" valign="middle" >10</td><td align="center" valign="middle" >11.76%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Efficiency</td><td align="center" valign="middle" >38</td><td align="center" valign="middle" >44.71%</td><td align="center" valign="middle" >44</td><td align="center" valign="middle" >51.76%</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3.53%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Total</td><td align="center" valign="middle" >680</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >929</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >93</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0</td><td align="center" valign="middle" ></td></tr></tbody></table></table-wrap><table-wrap id="table6" ><label><xref ref-type="table" rid="table6">Table 6</xref></label><caption><title> Combined very high and high rating</title></caption><table><tbody><thead><tr><th align="center" valign="middle" >Metrics/Ratings</th><th align="center" valign="middle"  colspan="2"  >Combined High</th><th align="center" valign="middle"  colspan="2"  >Medium</th><th align="center" valign="middle"  colspan="2"  >Low</th></tr></thead><tr><td align="center" valign="middle" >Correctness</td><td align="center" valign="middle" >82</td><td align="center" valign="middle" >96.47%</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3.53%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Satisfaction</td><td align="center" valign="middle" >85</td><td align="center" valign="middle" >100%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Validation</td><td align="center" valign="middle" >83</td><td align="center" valign="middle" >97.65%</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2.35%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Effectiveness</td><td align="center" valign="middle" >80</td><td align="center" valign="middle" >94.11%</td><td align="center" valign="middle" >5</td><td align="center" valign="middle" >5.88%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Quality of Outcome</td><td align="center" valign="middle" >83</td><td align="center" valign="middle" >97.65%</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2.35%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Reliability</td><td align="center" valign="middle" >84</td><td align="center" valign="middle" >98.82%</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >1.18%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Consistency</td><td align="center" valign="middle" >83</td><td align="center" valign="middle" >97.65%</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2.35%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Accessibility</td><td align="center" valign="middle" >84</td><td align="center" valign="middle" >98.82%</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >1.18%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Fault Tolerance</td><td align="center" valign="middle" >77</td><td align="center" valign="middle" >90.59%</td><td align="center" valign="middle" >8</td><td align="center" valign="middle" >9.41%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Preferability</td><td align="center" valign="middle" >83</td><td align="center" valign="middle" >97.65%</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2.35%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Ease of use</td><td align="center" valign="middle" >84</td><td align="center" valign="middle" >98.82%</td><td align="center" valign="middle" >1</td><td align="center" valign="middle" >3.33%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Navigability</td><td align="center" valign="middle" >82</td><td align="center" valign="middle" >96.47%</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3.53%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Simplicity</td><td align="center" valign="middle" >83</td><td align="center" valign="middle" >97.65%</td><td align="center" valign="middle" >2</td><td align="center" valign="middle" >2.35%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Understandability</td><td align="center" valign="middle" >82</td><td align="center" valign="middle" >96.47%</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3.53%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Attractiveness</td><td align="center" valign="middle" >77</td><td align="center" valign="middle" >90.59%</td><td align="center" valign="middle" >8</td><td align="center" valign="middle" >9.41%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Time Response</td><td align="center" valign="middle" >60</td><td align="center" valign="middle" >70.59%</td><td align="center" valign="middle" >25</td><td align="center" valign="middle" >29.41%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Visualization</td><td align="center" valign="middle" >76</td><td align="center" valign="middle" >89.41%</td><td align="center" valign="middle" >9</td><td align="center" valign="middle" >10.58%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Reusability</td><td align="center" valign="middle" >82</td><td align="center" valign="middle" >96.47%</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3.53%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Feedback</td><td align="center" valign="middle" >75</td><td align="center" valign="middle" >88.23%</td><td align="center" valign="middle" >10</td><td align="center" valign="middle" >11.76%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Efficiency</td><td align="center" valign="middle" >82</td><td align="center" valign="middle" >96.47%</td><td align="center" valign="middle" >3</td><td align="center" valign="middle" >3.53%</td><td align="center" valign="middle" >0</td><td align="center" valign="middle" >0%</td></tr><tr><td align="center" valign="middle" >Total</td><td align="center" valign="middle" >1607</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >93</td><td align="center" valign="middle" ></td><td align="center" valign="middle" >0</td><td align="center" valign="middle" ></td></tr></tbody></table></table-wrap><p>The overall results show that the user experience evaluations of the system based on the metrics given are excellent. This is because in most case of the metrics used “Very High” and “High” (which are good scale to measure superior or improved opinion ) are rated up to 90% and above, Medium are rated less than 10% respectively.</p><p>The Relevance of the system is calculated thus:</p><p>Relevance = ∑ i = 1 N = 4     k i ∗ r i N ∗ ∑ i = 1 N = 4     k i ,</p><p>where N is the total number of rate point, r = 1 , ⋯ , N and k i is the sum of user that selected a given rate point for all the metrics.</p><p>= ( 680 &#215; 4 ) + ( 929 &#215; 3 ) + ( 93 &#215; 2 ) + ( 0 &#215; 1 ) 1702 &#215; 4 = 2720 + 2787 + 186 6808 = 5693 6808 = 0.8362 = 83.62 % .</p></sec></sec><sec id="s5"><title>5. Conclusion and Future Works</title><p>An answer validation system for answers using answerers attributes and crowd ranking has been developed. For the effectiveness of the system, illegitimate questions and answers were filtered out using a trained Na&#239;ve Bayes spam filter with a threshold of 0.5. Answerers’ personal attributes (such as Grade points, Area of specialization, Years of experience Level of Computing, Course of study, Question Understandability and the answer confidence level (trustworthiness)) were used to ensure high quality answers by employing a weighted system that assigns weights to individual attributes in order to know the weight of the answers for validation. Answers are ranked by the crowd to get the best four answers from the candidate answers obtained from the answerers using Borda count ranking algorithm and least best answer is discarded. The system correctness is 96.47%, Answer satisfaction is 100%, answer Validation is 97.65%, system Simplicity is 97.6%, system Feedback is 88.23% and the system efficiency is 96.47%. Future works could include more User attributes such as age, qualification and so on and ensure that there is an improvement in the system feedback so that users can receive instant live answers to their respective questions. There should be a way in which the answerers are motivated for the task performed in order to enhance their performance. In addition, the system should be more general to accommodate questions from other science related domain.</p></sec><sec id="s6"><title>Conflicts of Interest</title><p>The authors declare no conflicts of interest regarding the publication of this paper.</p></sec><sec id="s7"><title>Cite this paper</title><p>Adebisi, M., Ojokoh, B., Adebayo, T., Akinwonmi, A., &amp; Sunmola, F. (2021). Development of Answer Validation System Using Responders’ Attributes and Crowd Ranking. Journal of Service Science and Management, 14, 382-398. https://doi.org/10.4236/jssm.2021.143024</p></sec></body><back><ref-list><title>References</title><ref id="scirp.110281-ref1"><label>1</label><mixed-citation publication-type="other" xlink:type="simple">Anderson, A., Huttenlocher, D., Kleinberg, J., &amp; Leskovec, J. (2012). Discovering Value from Community Activity on Focused Question Answering Sites: A Case Study of Stack Overflow. Proceedings the 18th ACM International Conference on Knowledge Discovery and Data Mining, Beijing, 12-16 August 2012, 850-858. https://doi.org/10.1145/2339530.2339665</mixed-citation></ref><ref id="scirp.110281-ref2"><label>2</label><mixed-citation publication-type="other" xlink:type="simple">Aydin, B., Yilmaz, Y., Li, Y., Li, Q., Gao, J., &amp; Demirbas, M. (2014). Crowdsourcing for Multiple-Choice Question Answering. Proceedings the 26th Annual Conference on Innovative Applications of Artificial Intelligence, Québec, 29-31 July 2014, 1-12.</mixed-citation></ref><ref id="scirp.110281-ref3"><label>3</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Chandra</surname><given-names> A.</given-names></name>,<name name-style="western"><surname> Reddy</surname><given-names> O.</given-names></name>,<name name-style="western"><surname> &amp; Madhavi</surname><given-names> K. </given-names></name>,<etal>et al</etal>. (<year>2017</year>)<article-title>. A Survey on Types of Question Answering System</article-title><source> IOSR Journal of Computer Engineering</source><volume> 19</volume>,<fpage> 19</fpage>-<lpage>23</lpage>.<pub-id pub-id-type="doi"></pub-id></mixed-citation></ref><ref id="scirp.110281-ref4"><label>4</label><mixed-citation publication-type="other" xlink:type="simple">Cimiano, P., Unger, C., &amp; McCrae, J. (2014). Ontology-Based Interpretation of Natural Language. San Rafael, CA: Morgan &amp; Claypool Publishers. https://doi.org/10.2200/S00561ED1V01Y201401HLT024</mixed-citation></ref><ref id="scirp.110281-ref5"><label>5</label><mixed-citation publication-type="other" xlink:type="simple">Dob&amp;#353;ovi&amp;#269;, R., Grznar, M., Harinek, J., Molnar, S., Palenik, P., Poizl, D., &amp; Zbell, P. (2014). Askalot: An Educational Community Question Answering System. Unpublished PhD Thesis, Bratislava: Slovak University of Technology.</mixed-citation></ref><ref id="scirp.110281-ref6"><label>6</label><mixed-citation publication-type="other" xlink:type="simple">Fan, H., Ma, Z., Li, H., Wang, D., &amp; Liu, J. (2019). Enhanced Answer Selection in CQA Using Multi-Dimensional Features Combination. Tsinghua Science and Technology, 24, 346-359. https://doi.org/10.26599/TST.2018.9010050</mixed-citation></ref><ref id="scirp.110281-ref7"><label>7</label><mixed-citation publication-type="other" xlink:type="simple">Harper, F., Raban, D., Rafaeli, S., &amp; Konstan, J. (2008). Predictors of Answer Quality in Online Q&amp;A Sites. Proceedings of the Twenty-Sixth Annual SIGCHI Conference on Human Factors in Computing Systems, Florence, 5-10 April 2008, 865-874. https://doi.org/10.1145/1357054.1357191</mixed-citation></ref><ref id="scirp.110281-ref8"><label>8</label><mixed-citation publication-type="other" xlink:type="simple">Howe, J. (2006). Crowdsourcing: A Definition. Wired Blog Network: Crowdsourcing. http://crowdsourcing.typepad.com/cs/2006/06/crowdsourcing_a.html</mixed-citation></ref><ref id="scirp.110281-ref9"><label>9</label><mixed-citation publication-type="other" xlink:type="simple">Hung, N., Thang, D., Tam, N., Weidlich, M., Aberer, K., Yin, H., &amp; Zhou, X. (2017). Answer Validation for Generic Crowdsourcing Tasks with Minimal Efforts. The VLDB Journal, 26, 855-880. https://doi.org/10.1007/s00778-017-0484-3</mixed-citation></ref><ref id="scirp.110281-ref10"><label>10</label><mixed-citation publication-type="other" xlink:type="simple">Hung, N., Thang, D., Weidlich, M., &amp; Aberer, K. (2015). Minimizing Efforts in Validating Crowd Answers. Proceedings of the Association for Computer Machinery’s Special Interest Group on Management of Data, Melbourne, May 2015, 999-1014. https://doi.org/10.1145/2723372.2723731</mixed-citation></ref><ref id="scirp.110281-ref11"><label>11</label><mixed-citation publication-type="other" xlink:type="simple">Ishikawa, D., Kando, N., &amp; Sakai, T. (2011). What Makes a Good Answer in Community Question Answering? An Analysis of Assessors’ Criteria. The Fourth International Workshop on Evaluating Information Access, Tokyo, December 2011, 169-181.</mixed-citation></ref><ref id="scirp.110281-ref12"><label>12</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Le</surname><given-names> L.</given-names></name>,<name name-style="western"><surname> Shah</surname><given-names> C.</given-names></name>,<name name-style="western"><surname> &amp; Choi</surname><given-names> E. </given-names></name>,<etal>et al</etal>. (<year>2019</year>)<article-title>. Assessing the Quality of Answers Autonomously in Community Question-Answering</article-title><source> International Journal on Digital Libraries</source><volume> 20</volume>,<fpage> 1</fpage>-<lpage>17</lpage>.<pub-id pub-id-type="doi"></pub-id></mixed-citation></ref><ref id="scirp.110281-ref13"><label>13</label><mixed-citation publication-type="other" xlink:type="simple">Magnini, B., Negri, M., Prevete, R., &amp; Tanev, H. (2002). Comparing Statistical and Content-Based Techniques for Answer Validation on the Web. Proceedings of the 8th Convegno AI&amp;IA, Siena, September 2002, 413-427.</mixed-citation></ref><ref id="scirp.110281-ref14"><label>14</label><mixed-citation publication-type="other" xlink:type="simple">Magnini, B., Negri, M., Prevete, R., &amp; Tanev, H. (2005). Is It the Right Answer? Exploiting Web Redundancy for Answer Validation. ACL-02: Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics, Philadelphia, July, 425-432.</mixed-citation></ref><ref id="scirp.110281-ref15"><label>15</label><mixed-citation publication-type="other" xlink:type="simple">Nie, L., Wei, X., Zhang, D., Wang, X., Gao, Z., &amp; Yang, Y. (2017). Data-Driven Answer Selection in Community QA Systems. IEEE Transactions on Knowledge and Data Engineering, 29, 1186-1198. https://doi.org/10.1109/TKDE.2017.2669982</mixed-citation></ref><ref id="scirp.110281-ref16"><label>16</label><mixed-citation publication-type="other" xlink:type="simple">Ojokoh, B., &amp; Adebisi, E. (2019). A Review of Question Answering Systems. Journal of Web Engineering, 17, 717-758. https://doi.org/10.13052/jwe1540-9589.1785</mixed-citation></ref><ref id="scirp.110281-ref17"><label>17</label><mixed-citation publication-type="other" xlink:type="simple">Ojokoh, B., &amp; Ayokunle, P. (2012). Fuzzy-Based Answer Ranking in Question Answering Communities. International Journal of Digital Library Systems, 3, 47-63. https://doi.org/10.4018/jdls.2012070105</mixed-citation></ref><ref id="scirp.110281-ref18"><label>18</label><mixed-citation publication-type="other" xlink:type="simple">Ríos-Gaona, M., Gelbukh, A., &amp; Bandyopadhyay, S. (2012). Recognizing Textual Entailment Using a Machine Learning Approach. In Mexican International Conference on Artificial Intelligence (pp. 177-185). Berlin: Springer. https://doi.org/10.1007/978-3-642-16773-7_15</mixed-citation></ref><ref id="scirp.110281-ref19"><label>19</label><mixed-citation publication-type="other" xlink:type="simple">Savenkov, D., Weitzner, S., &amp; Agichtein, E. (2016). Crowdsourcing for (Almost) Real-Time Question Answering. NAACL 2016: Proceedings of the Workshop on Human-Computer Question Answering, San-Diego, 12-17 June 2016, 8-14. https://doi.org/10.18653/v1/W16-0102</mixed-citation></ref><ref id="scirp.110281-ref20"><label>20</label><mixed-citation publication-type="other" xlink:type="simple">Schofield, M., &amp; Thielscher, M. (2019). General Game Playing with Imperfect Information. Artificial Intelligence Journal, 66, 901-935. https://doi.org/10.1613/jair.1.11844</mixed-citation></ref><ref id="scirp.110281-ref21"><label>21</label><mixed-citation publication-type="other" xlink:type="simple">&amp;#352;imko, J., Simko, M., Bieliková, M., Sevcech, J., &amp; Burger, R. (2013). Classsourcing: Crowd-Based Validation of Question-Answer Learning Objects. In International Conference on Computational Collective Intelligence (pp. 62-71). Berlin: Springer. https://doi.org/10.1007/978-3-642-40495-5_7</mixed-citation></ref><ref id="scirp.110281-ref22"><label>22</label><mixed-citation publication-type="other" xlink:type="simple">Su, Q., Pavlov, D., Chow, J., &amp; Baker, W. (2007). Internet-Scale Collection of Human-Reviewed Data. Proceedings of the 16th International Conference on World Wide Web, Banff, 8-12 May 2007, 231-240. https://doi.org/10.1145/1242572.1242604</mixed-citation></ref><ref id="scirp.110281-ref23"><label>23</label><mixed-citation publication-type="other" xlink:type="simple">Tan, J., R&amp;#246;nkk&amp;#246;, K., &amp; Gencel, C. (2010). A Framework for Software Usability and User Experience Measurement in Mobile Industry. MSc Thesis, Karlshamn: Blekinge Institute of Technology, Sweden.</mixed-citation></ref><ref id="scirp.110281-ref24"><label>24</label><mixed-citation publication-type="other" xlink:type="simple">Toba, H., Ming, Z., Adriani, M., &amp; Chua, T. (2014). Discovering High Quality Answers in Community Question Answering Archives Using a Hierarchy of Classifiers. Information Sciences, 261, 101-115. https://doi.org/10.1016/j.ins.2013.10.030</mixed-citation></ref><ref id="scirp.110281-ref25"><label>25</label><mixed-citation publication-type="other" xlink:type="simple">Tran, Q., Tran, V., Vu, T., Nguyen, M., &amp; Pham, S. (2015). JAIST: Combining Multiple Features for Answer Selection in Community Question Answering. Proceedings of the 9th International Workshop on Semantic Evaluation, Denver, 4-5 June 2015, 215-219. https://doi.org/10.18653/v1/S15-2038</mixed-citation></ref></ref-list></back></article>