<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2022.865841</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Explainable Personality Prediction Using Answers to Open-Ended Interview Questions</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name><surname>Dai</surname> <given-names>Yimeng</given-names></name>
<uri xlink:href="http://loop.frontiersin.org/people/1658448/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Jayaratne</surname> <given-names>Madhura</given-names></name>
<uri xlink:href="http://loop.frontiersin.org/people/1647884/overview"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name><surname>Jayatilleke</surname> <given-names>Buddhi</given-names></name>
<xref ref-type="corresp" rid="c001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/1536982/overview"/>
</contrib>
</contrib-group>
<aff><institution>Sapia&#x00026;Co Pty Ltd.</institution>, <addr-line>Melbourne</addr-line>, <addr-line>VIC</addr-line>, <country>Australia</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Petar &#x0010C;olovi&#x00107;, University of Novi Sad, Serbia</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Alexander P. Christensen, University of Pennsylvania, United States; Andry Chowanda, Binus University, Indonesia</p></fn>
<corresp id="c001">&#x0002A;Correspondence: Buddhi Jayatilleke  <email>buddhi&#x00040;sapia.ai</email></corresp>
<fn fn-type="other" id="fn001"><p>This article was submitted to Personality and Social Psychology, a section of the journal Frontiers in Psychology</p></fn></author-notes>
<pub-date pub-type="epub">
<day>18</day>
<month>11</month>
<year>2022</year>
</pub-date>
<pub-date pub-type="collection">
<year>2022</year>
</pub-date>
<volume>13</volume>
<elocation-id>865841</elocation-id>
<history>
<date date-type="received">
<day>30</day>
<month>01</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>19</day>
<month>04</month>
<year>2022</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2022 Dai, Jayaratne and Jayatilleke.</copyright-statement>
<copyright-year>2022</copyright-year>
<copyright-holder>Dai, Jayaratne and Jayatilleke</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license> </permissions>
<abstract>
<p>In this work, we demonstrate how textual content from answers to interview questions related to past behavior and situational judgement can be used to infer personality traits. We analyzed responses from over 58,000 job applicants who completed an online text-based interview that also included a personality questionnaire based on the HEXACO personality model to self-rate their personality. The inference model training utilizes a fine-tuned version of InterviewBERT, a pre-trained Bidirectional Encoder Representations from Transformers (BERT) model extended with a large interview answer corpus of over 3 million answers (over 330 million words). InterviewBERT is able to better contextualize interview responses based on the interview specific knowledge learnt from the answer corpus in addition to the general language knowledge already encoded in the initial pre-trained BERT. Further, the &#x0201C;Attention-based&#x0201D; learning approaches in InterviewBERT enable the development of explainable personality inference models that can address concerns of model explainability, a frequently raised issue when using machine learning models. We obtained an average correlation of <italic>r</italic> = 0.37 (<italic>p</italic> &#x0003C; 0.001) across the six HEXACO dimensions between the self-rated and the language-inferred trait scores with the highest correlation of <italic>r</italic> = 0.45 for Openness and the lowest of <italic>r</italic> = 0.28 for Agreeableness. We also show that the mean differences in inferred trait scores between male and female groups are similar to that reported by others using standard self-rated item inventories. Our results show the potential of using InterviewBERT to infer personality in an explainable manner using only the textual content of interview responses, making personality assessments more accessible and removing the subjective biases involved in human interviewer judgement of candidate personality.</p></abstract>
<kwd-group>
<kwd>personality prediction</kwd>
<kwd>HEXACO personality model</kwd>
<kwd>linguistic analysis</kwd>
<kwd>NLP</kwd>
<kwd>BERT</kwd>
</kwd-group>
<counts>
<fig-count count="4"/>
<table-count count="5"/>
<equation-count count="5"/>
<ref-count count="88"/>
<page-count count="14"/>
<word-count count="11045"/>
</counts>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>1. Introduction</title>
<p>Understanding personality plays a critical role in making sense of one&#x00027;s own self and their relationships with others, especially within a work environment. To that end, personality is widely accepted as an indicator of job performance, job satisfaction, and tenure intention (Barrick and Mount, <xref ref-type="bibr" rid="B6">1991</xref>; Salgado, <xref ref-type="bibr" rid="B73">2002</xref>; Rothmann and Coetzer, <xref ref-type="bibr" rid="B72">2003</xref>; Lounsbury et al., <xref ref-type="bibr" rid="B45">2008</xref>, <xref ref-type="bibr" rid="B44">2012</xref>; Ariyabuddhiphongs and Marican, <xref ref-type="bibr" rid="B2">2015</xref>). The most common approach for assessing personality is to use a self-report personality questionnaire such as the NEO-PI-R (Costa and McCrae, <xref ref-type="bibr" rid="B14">2008</xref>) or the HEXACO-PI-R (Lee and Ashton, <xref ref-type="bibr" rid="B38">2018</xref>) that consists of a large number of personality related statements rated by the individual on a Likert scale. While decades of research have shown the validity and improved on the traditional approach of assessing personality (Morgeson et al., <xref ref-type="bibr" rid="B54">2007a</xref>,<xref ref-type="bibr" rid="B55">b</xref>; Ones et al., <xref ref-type="bibr" rid="B58">2007</xref>), adding a personality test to the recruitment process tends to increase the cost-to-hire and diminishes candidate experience since most personality tests are lengthy and tedious (Mcdaniel et al., <xref ref-type="bibr" rid="B51">1994</xref>; Macan, <xref ref-type="bibr" rid="B48">2009</xref>). Hence, personality assessments are not frequently included in hiring for most roles, especially in high-volume recruitment, despite its validity.</p>
<p>On the other hand, job interview remains the most common form of assessment in candidate selection and the ability to automatically infer personality from answers to job interview questions could replace lengthy personality assessments (Jayaratne and Jayatilleke, <xref ref-type="bibr" rid="B32">2020</xref>). Moreover, a data-driven approach can help counter flaws in human judgement due to personal factors such as mood, own personality, and biases that unavoidably affect interview outcomes (Uleman, <xref ref-type="bibr" rid="B81">1999</xref>; Ham and Vonk, <xref ref-type="bibr" rid="B28">2003</xref>; Ma et al., <xref ref-type="bibr" rid="B47">2011</xref>; Ferreira et al., <xref ref-type="bibr" rid="B20">2012</xref>). When conducting a large number of interviews, human interviewers can hardly infer personality accurately and efficiently with clear explanations for each candidate. The ability to automate the inference of personality can offer more candidates the opportunity to express themselves and be heard, especially in high-volume recruitment where only a small fraction typically progress to a face-to-face interview.</p>
<p>In this work, we demonstrate how textual content from answers to interview questions related to past behavior and situational judgement can be used to infer personality traits reliably. <xref ref-type="fig" rid="F1">Figure 1</xref> shows the overview of our methodology. We used data from over 58,000 job applicants who completed an online chat interview that also included a personality questionnaire based on the six-factor HEXACO personality model (Ashton and Lee, <xref ref-type="bibr" rid="B4">2007</xref>) to self-rate their personality. We proposed InterviewBERT, a variant of the state-of-the-art Bidirectional Encoder Representations from Transformers (BERT) (Devlin et al., <xref ref-type="bibr" rid="B16">2019</xref>), which is a transformer (Vaswani et al., <xref ref-type="bibr" rid="B84">2017</xref>) based machine learning technique for NLP pre-training. We extended BERT with a large interview answer corpus of over 3 million answers consisting of over 330 million words. InterviewBERT is able to better contextualize interview responses based on the interview specific knowledge learnt from the answer corpus in addition to the general language knowledge already encoded in the initial pre-trained BERT. We show the advantage of using context-specific answer representations to infer personality compared to context-free methods, and study different ways of using InterviewBERT to achieve context-specific answer representations. The use of InterviewBERT also differentiates our approach from previous work related to the inference of personality from interview responses such as Jayaratne and Jayatilleke (<xref ref-type="bibr" rid="B32">2020</xref>), that uses context-free NLP approaches.</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p>Illustration of InterviewBERT based HEXACO personality prediction.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-865841-g0001.tif"/>
</fig>
<p>Moreover, we show how the self-attention mechanism (Vaswani et al., <xref ref-type="bibr" rid="B84">2017</xref>) in InterviewBERT can be used to develop more explainable personality inference models. Attention in this context is motivated by human behaviors seen in activities such as vision and reading comprehension where people pay varying levels of attention to different regions in an image or words in a text, according to different situations and goals. We use self-attention to capture the relationships between words and use the attention to provide a basis for providing explanations for personality inference outcomes.</p>
<p>Our results show the potential of algorithms to objectively infer a candidate&#x00027;s personality in an explainable manner using only the textual content of interview responses, presenting significant opportunities to remove the subjective biases involved in human interviewer judgement of candidate personality.</p>
<p>The main contributions of this paper are listed as follows.</p>
<list list-type="order">
<list-item><p>We demonstrate that the textual content from answers to standard interview questions can be used to infer one&#x00027;s personality.</p></list-item>
<list-item><p>We propose the use of context-specific text representations for interview answers and propose InterviewBERT that extends the BERT model with a large interview response corpus.</p></list-item>
<list-item><p>We empirically investigate the performance of InterviewBERT based personality prediction using a real online interview dataset.</p></list-item>
<list-item><p>We show the language-level explainability of the InterviewBERT based prediction results.</p></list-item>
<list-item><p>We investigate the gender differences in personality traits inferred from interviews.</p></list-item>
</list>
<p>The rest of the paper is organized as follows. In Section 2, we provide a review of related work and introduce the details of our methodology preliminaries. In Section 3, we introduce the way we construct our data set (Section 3.1), present different methods of answer representations (Section 3.2), and outline the model to infer personality from answer representations (Section 3.3). In Section 4, we present the experimental results followed by a discussion in Section 5. In Section 6, we conclude with suggestions for future directions.</p>
</sec>
<sec id="s2">
<title>2. Background</title>
<p>In this section, we introduce the preliminaries of the HEXACO personality model used as the underlying personality model in our study (Section 2.1), and the related work around language and personality (Section 2.2). We also provide an overview of the methods we use to infer personality from textual content of interview responses. These include the different word and document representation approaches found in natural language processing (Section 2.3), the BERT model architecture and the self-attention mechanism (Section 2.4) that form the basis for the InterviewBERT model. We find that a lengthy discussion of the technical details of the above topics is out of the scope of this paper and refer the reader to the related work we reference under each topic.</p>
<sec>
<title>2.1. HEXACO Model</title>
<p>HEXACO (Ashton and Lee, <xref ref-type="bibr" rid="B4">2007</xref>) is a six-dimensional model of personality consisting of Honesty-humility (H), Emotionality (E), eXtraversion (X), Agreeableness (A), Conscientiousness (C), and Openness (O) as dimensions. Similar to the Big Five model (Goldberg, <xref ref-type="bibr" rid="B25">1993</xref>) of personality, HEXACO model has its origins in lexical studies and subsequent factor analysis used to identify a minimal set of independent dimensions or personality traits and their underlying facets. It&#x00027;s relevant to note here that the use of lexical studies are grounded on the <italic>lexical hypothesis</italic> that claims descriptors of personality characteristics are encoded in language (Saucier and Goldberg, <xref ref-type="bibr" rid="B74">1996</xref>), a fact we will re-visit in the next section. While there are similarities and subtle differences in the dimensions in HEXACO and the Big Five model, a key difference is the addition of the Honesty-Humility (H) dimension or the H-factor. The H-factor is especially important in the employment assessment context given it represents characteristics desired in a workplace environment such as modesty, fairness, and honesty. Previous studies have shown that the H-factor can help explain and predict workplace deviance (Pletzer et al., <xref ref-type="bibr" rid="B66">2019</xref>), delinquency (Lee et al., <xref ref-type="bibr" rid="B41">2005</xref>; de Vries and van Gelder, <xref ref-type="bibr" rid="B15">2015</xref>), integrity (Lee et al., <xref ref-type="bibr" rid="B40">2008</xref>), counterproductive work behavior and organizational citizenship (Anglim et al., <xref ref-type="bibr" rid="B1">2018</xref>), and job performance (Johnson et al., <xref ref-type="bibr" rid="B34">2011</xref>).</p>
</sec>
<sec>
<title>2.2. Language and Personality</title>
<p>Language analysis is a first-principles approach to understanding psychological constructs as studied in <italic>psycholinguistics</italic> and the application of <italic>lexical hypothesis</italic> in discovering personality dimensions. The field of psycholinguistics is dedicated to the study of the relationship between language and various psychological aspects related to language acquisition, understanding and human thought (Pinker, <xref ref-type="bibr" rid="B64">2007</xref>; Gleitman and Papafragou, <xref ref-type="bibr" rid="B23">2012</xref>). In Pinker (<xref ref-type="bibr" rid="B64">2007</xref>), the author details with extensive research on how we speak reveals what we think. More importantly personality models such as HEXACO and Big Five are grounded on the <italic>lexical hypothesis</italic>, which states that personality characteristics that are salient in people&#x00027;s daily transactions and relates to important social outcomes are encoded in language (John et al., <xref ref-type="bibr" rid="B33">1988</xref>; Saucier and Goldberg, <xref ref-type="bibr" rid="B74">1996</xref>). Advances in machine learning and natural language processing (NLP) have catalyzed the growing body of evidence showing the relationship between one&#x00027;s language use and personality (Boyd and Pennebaker, <xref ref-type="bibr" rid="B9">2017</xref>). This relationship has been demonstrated in both informal contexts such as social media (Gill et al., <xref ref-type="bibr" rid="B21">2009</xref>; Golbeck et al., <xref ref-type="bibr" rid="B24">2011</xref>; Iacobelli et al., <xref ref-type="bibr" rid="B31">2011</xref>; Park et al., <xref ref-type="bibr" rid="B60">2015</xref>; Christian et al., <xref ref-type="bibr" rid="B12">2021</xref>; Lucky and Suhartono, <xref ref-type="bibr" rid="B46">2021</xref>) as well as in formal contexts such as self-narratives (Fast and Funder, <xref ref-type="bibr" rid="B18">2008</xref>; Hirsh and Peterson, <xref ref-type="bibr" rid="B29">2009</xref>), and job interviews (Jayaratne and Jayatilleke, <xref ref-type="bibr" rid="B32">2020</xref>).</p>
<p>The language-personality relationship has been utilized to develop predictive machine learning models to accurately infer personality traits from blogs (Iacobelli et al., <xref ref-type="bibr" rid="B31">2011</xref>), essays (Neuman and Cohen, <xref ref-type="bibr" rid="B57">2014</xref>), microblogs (Twitter, Sina Weibo) (Golbeck et al., <xref ref-type="bibr" rid="B24">2011</xref>; Sumner et al., <xref ref-type="bibr" rid="B77">2012</xref>; Xue et al., <xref ref-type="bibr" rid="B88">2017</xref>; Lucky and Suhartono, <xref ref-type="bibr" rid="B46">2021</xref>), social media posts (Tadesse et al., <xref ref-type="bibr" rid="B78">2018</xref>; Wang et al., <xref ref-type="bibr" rid="B87">2019</xref>), etc. The success of such attempts has led researchers to propose computer generated personality predictions to &#x0201C;complement&#x02014;and in some instances replace&#x02014;traditional self-report measures, which suffer from well-known response biases and are difficult to scale&#x0201D; (Hall and Matz, <xref ref-type="bibr" rid="B27">2020</xref>).</p>
<p>Language modeling within psychological sciences typically involves two types of approaches: the <italic>closed-vocabulary</italic> approach and the <italic>open-vocabulary</italic> approach. In closed-vocabulary approaches, words are assigned to psycho-socio-educational relevant categories to create dictionaries that are considered to represent that category. For example, words such as happiness, joy, etc. can be part of a dictionary for positive emotions. Linguistic Inquiry and Word Count (LIWC) (Pennebaker et al., <xref ref-type="bibr" rid="B61">2015</xref>) is one such lexicon. Using the LIWC, researchers have found correlations among language patterns and personality (Fast and Funder, <xref ref-type="bibr" rid="B18">2008</xref>; Gill et al., <xref ref-type="bibr" rid="B21">2009</xref>; Hirsh and Peterson, <xref ref-type="bibr" rid="B29">2009</xref>; Golbeck et al., <xref ref-type="bibr" rid="B24">2011</xref>; Qiu et al., <xref ref-type="bibr" rid="B69">2012</xref>). On the other hand, open-vocabulary approaches are more data-driven. In an open-vocabulary NLP system, algorithms process a large set of linguistic data and identify semantically related words through numerical word representation methods (We detail these methods in Section 2.3), which can be used to predict outcomes using supervised machine learning algorithms or gain further insights through exploration using unsupervised algorithms such as clustering. Compared to the closed-vocabulary methods, the open-vocabulary methods build upon the idea that words can be represented with numerical values based on how they co-occur, yielding to powerful language models that allow us to model words according to the contexts in which they appear rather than relying on assumptions about word-category relations. It eliminates the need for a human to have created categories and related dictionaries that limits the vocabulary known to learning algorithms. Open-vocabulary approaches are the current de facto standard for modeling language data and usually require a large amount of training data to learn the relationship between personality and language representation. Such predictive models have been demonstrated on textual data from social media with success (Schwartz et al., <xref ref-type="bibr" rid="B75">2013</xref>; Park et al., <xref ref-type="bibr" rid="B60">2015</xref>; Liu et al., <xref ref-type="bibr" rid="B43">2016</xref>; Christian et al., <xref ref-type="bibr" rid="B12">2021</xref>; Lucky and Suhartono, <xref ref-type="bibr" rid="B46">2021</xref>).</p>
</sec>
<sec>
<title>2.3. Word and Document Representations</title>
<p>Natural language processing (NLP) requires representation of language and broadly two types of representations are used: context-free representations and context-specific representations. Traditional context-free representation methods include Bag of Words (BoW) and term frequency-inverse document frequency (TF-IDF) (Christopher et al., <xref ref-type="bibr" rid="B13">2008</xref>) where BoW represents a document using the raw count of a term or n-gram (sequence of terms) in the corpus, while TF-IDF evaluates the importance of a term within a single document based on its occurrences across the document corpus. An obvious limitation of BoW and TF-IDF is that the meaning and term similarity are not encoded leaving unseen words as &#x0201C;out of vocabulary&#x0201D; when a trained model is applied on a new document. Further, they introduce very long and sparse input vectors, especially when the vocabulary is large. These context-free representations (in some instances along with other features) have been used in personality prediction from textual content (Iacobelli et al., <xref ref-type="bibr" rid="B31">2011</xref>; Schwartz et al., <xref ref-type="bibr" rid="B75">2013</xref>; Plank and Hovy, <xref ref-type="bibr" rid="B65">2015</xref>; Verhoeven et al., <xref ref-type="bibr" rid="B85">2016</xref>; Gjurkovi&#x00107; and &#x00160;najder, <xref ref-type="bibr" rid="B22">2018</xref>; Jayaratne and Jayatilleke, <xref ref-type="bibr" rid="B32">2020</xref>). Neural word embedding methods such as Word2Vec (Mikolov et al., <xref ref-type="bibr" rid="B53">2013</xref>) and GloVe (Pennington et al., <xref ref-type="bibr" rid="B62">2014</xref>) attempt to address the capturing of contextual similarity of terms by providing a term level representation (called an embedding) by pre-training on a large corpus of documents (e.g., Wikipedia, open web crawl). For example, Word2Vec learns embeddings by predicting the current word based on its surrounding words or predicting the surrounding words given a current word (Skip-Gram). GloVe uses a count-based model, which learns embeddings by looking at how often a word appears in the context of another word within the corpus, focusing on the co-occurrence probabilities of words within a large training corpus of documents such as Wikipedia. Studies of personality inferences that use neural word embeddings include (Kamijo et al., <xref ref-type="bibr" rid="B35">2016</xref>; Arnoux et al., <xref ref-type="bibr" rid="B3">2017</xref>; Majumder et al., <xref ref-type="bibr" rid="B49">2017</xref>; Jayaratne and Jayatilleke, <xref ref-type="bibr" rid="B32">2020</xref>). Though pre-trained neural word embeddings are widely used, they assume that a word&#x00027;s meaning is relatively stable and does not change across different sentences. Hence these word embedding are not context-specific at the level of different uses of the same word. We used both TF-IDF and GloVe as context-free representations of language in our study to compare the outcomes against context-specific representations introduced below (see Sections 3.2.1, 3.2.2 for details).</p>
<p>Recent work such as ELMo (Peters et al., <xref ref-type="bibr" rid="B63">2018</xref>), BERT (Devlin et al., <xref ref-type="bibr" rid="B16">2019</xref>), and OpenAI GPT (Radford et al., <xref ref-type="bibr" rid="B70">2018</xref>) use fine-tuning methods to further improve the pre-trained word embeddings. Instead of directly using fixed pre-trained neural word embeddings (as in the case of GloVe), these models fine-tune the pre-trained models on downstream tasks and target data to achieve context-dependent word embeddings. For example, the pre-training stage of BERT is typically task-agnostic and the models cannot always capture the domain-specific language patterns well. To improve the pre-trained language models for specific domains, some studies extend BERT on specialty corpora to generate a domain-specific BERT, such as BioBERT (Lee et al., <xref ref-type="bibr" rid="B37">2020</xref>) for biomedical text, SciBERT (Beltagy et al., <xref ref-type="bibr" rid="B7">2019</xref>) for scientific text, ClinicalBERT (Huang et al., <xref ref-type="bibr" rid="B30">2019</xref>) for clinical text. Similar to these studies, we extended BERT with a large interview answer corpus of over 3 million answers (over 330 million words) collected from online candidate interviews. The resulting InterviewBERT contains the general language knowledge already encoded in the initial BERT with the addition of job interview specific knowledge learnt from interview answers. We introduce the details of using InterviewBERT for personality prediction in Section 3.2.3. Many NLP tasks achieve state-of-the-art performance with BERT based methods, including text-based personality predictions using social media text (Christian et al., <xref ref-type="bibr" rid="B12">2021</xref>; Lucky and Suhartono, <xref ref-type="bibr" rid="B46">2021</xref>). Our work remains novel in the context of predicting personality from interview responses using an extended version of the BERT model trained on a very large corpus of interview responses.</p>
</sec>
<sec>
<title>2.4. BERT and Self-Attention</title>
<p><xref ref-type="fig" rid="F2">Figure 2</xref> illustrates the overall architecture of BERT, which is a stack of six layers, and each layer has a multi-head self-attention layer and a fully connected feed-forward network. The first token of every input sentence is a special token identified as <italic>[CLS]</italic>. The final hidden state of this token within the BERT model is typically used as the aggregated context-aware representation for the input sentence.</p>
<fig id="F2" position="float">
<label>Figure 2</label>
<caption><p>Illustration of BERT and self-attention.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-865841-g0002.tif"/>
</fig>
<p>During pre-training, BERT is trained on unlabeled data over different pre-training tasks, including: (1) <italic>predicting the original vocabulary of a randomly masked word in input based only on its context</italic>, and (2) <italic>whether a given sentence is the next sentence of a input sentence</italic>. During this fine-tuning, the BERT model is first initialized based on the general corpus, and then fine-tuned using training data from the downstream tasks. In the case of personality inference, each personality trait has a separate fine-tuned model. We&#x00027;ll introduce the detail of pre-training and fine-tuning of InterviewBERT in Section 3.2.3.</p>
<p>The multi-head self-attention in BERT is illustrated in the right side of <xref ref-type="fig" rid="F2">Figure 2</xref>. Attention is a mechanism to find the words of importance for a given query word in a sentence and multi-head attentions combine the knowledge explored by multiple heads instead of using one. Mathematically, it repeats the self-attention computations multiple times in parallel, and each of them computes attentions based on different aspects of the meanings of each word. With this multi-head self-attention mechanism and the learned context-specific representations of the answer (i.e., <italic>[CLS]</italic>), we can interpret the relationship between the word usage in answers and the predicted personality scores. We show how this ability in BERT can be used to provide better explainability to the personality predictions made by InterviewBERT.</p>
</sec>
</sec>
<sec sec-type="materials and methods" id="s3">
<title>3. Materials and Methods</title>
<p>In this section, we discuss the dataset, algorithms and experimental methodology used in achieving the two key aims of this study, namely, training of InterviewBERT and training of individual trait inference models using InterviewBERT. While a lengthy technical discussion of InterviewBERT is out of the scope of this paper, we provide a brief overview in Section 3.2.3. We then demonstrate the use of both context-free (TF-IDF and GloVe) and context-specific (InterviewBERT) representations in building regression models to predict personality traits and compare their accuracy. Given that a candidate typically answers multiple questions in an interview, we explore different ways of aggregating the multiple answers in order to achieve the highest accuracy in the regression task.</p>
<sec>
<title>3.1. Dataset Construction</title>
<p>Our training data comes from the Sapia<xref ref-type="fn" rid="fn0001"><sup>1</sup></xref> FirstInterview&#x02122; product, which is an online chat-based interview platform where candidates answer 5&#x02013;7 open-ended interview questions related to past behavior and situational judgement. The larger data set used to build InterviewBERT included 3,030,018 individual interview question responses from 505,013 candidates. Following are some examples of the open ended questions answered by the candidates.</p>
<list list-type="bullet">
<list-item><p><italic>Tell us about a problem you solved in a unique or unusual way. What was the outcome?</italic></p></list-item>
<list-item><p><italic>Describe a time when you missed a deadline or personal commitment. How did that make you feel?</italic></p></list-item>
<list-item><p><italic>Give an example of a time you have gone over and above to achieve something. Why was it important for you to achieve this?</italic></p></list-item>
<list-item><p><italic>Tell us about a time when you have rolled up your sleeves to help out your team or someone else</italic>.</p></list-item>
</list>
<p>The 5&#x02013;7 questions in each interview were selected based on the requirements of the role (e.g., retail assistant, sales, call center agent, engineer etc.) and the values sought by the employer. It&#x00027;s important to note that questions are rotated regularly to address gaming risk and plagiarized answers are flagged. On average candidates wrote 110 words per question and were encouraged to write at least 50 words per answer.</p>
<p>Following are two example answers to the question <italic>Tell us about a time when you have rolled up your sleeves to help out your team or someone else?</italic></p>
<list list-type="bullet">
<list-item><p><italic>As captain of my football team I always had to aid my team week in and week out. I had to communicate with the team to ensure everyone was happy in their positions and to ensure our cohesion was at a perfect level to ensure top performance. I would advise each player on one thing they can improve on for the next game and one thing they did particularly well on during the game. Through this method our team was highly successful and was always developing. I thoroughly enjoy working in a team because I love communicating with new people and learning from others</italic>.</p></list-item>
<list-item><p><italic>Whilst working as a tutor whenever another member of staff was sick or unable to come to work that day I was always happy to share out their work load and take on more children than usual for that session. During exam periods we were often spread thin but I was happy to do some extra marking, work a little later and come earlier to help set up the tables and chairs</italic>.</p></list-item>
</list>
<p>A subset of the candidates (<italic>N</italic> = 58,000) also self-rated themselves on a HEXACO-based personality inventory that provided us with the ground truth to train individual HEXACO trait inference models. It is important to note that not all 58,000 candidates were presented with self-rating items for all six traits due to the strain on candidates to answer both open ended text questions and a further set of close to 50 self-rating questions. Instead, candidates were presented with inventory items to cover at least two traits and a maximum of six. <xref ref-type="table" rid="T1">Table 1</xref> shows the number of candidates who answered self-rating items for each of the HEXACO traits and other important statistics.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p>Statistics of the dataset used in experiments. Std. is in short for &#x0201C;<italic>standard deviation</italic>&#x0201D; and Ave. is in short for &#x0201C;<italic>average</italic>.&#x0201D; Gender information was not available for all participants.</p></caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th valign="top" align="center"><bold>H</bold></th>
<th valign="top" align="center"><bold>E</bold></th>
<th valign="top" align="center"><bold>X</bold></th>
<th valign="top" align="center"><bold>A</bold></th>
<th valign="top" align="center"><bold>C</bold></th>
<th valign="top" align="center"><bold>O</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Participants</td>
<td valign="top" align="center">8,317</td>
<td valign="top" align="center">13,831</td>
<td valign="top" align="center">23,293</td>
<td valign="top" align="center">15,683</td>
<td valign="top" align="center">12,524</td>
<td valign="top" align="center">15,995</td>
</tr>
<tr>
<td valign="top" align="left">Female %</td>
<td valign="top" align="center">36</td>
<td valign="top" align="center">49</td>
<td valign="top" align="center">49</td>
<td valign="top" align="center">45</td>
<td valign="top" align="center">34</td>
<td valign="top" align="center">53</td>
</tr>
<tr>
<td valign="top" align="left">Male %</td>
<td valign="top" align="center">41</td>
<td valign="top" align="center">38</td>
<td valign="top" align="center">51</td>
<td valign="top" align="center">55</td>
<td valign="top" align="center">42</td>
<td valign="top" align="center">47</td>
</tr>
<tr>
<td valign="top" align="left">Ave. trait score</td>
<td valign="top" align="center">4.23</td>
<td valign="top" align="center">2.87</td>
<td valign="top" align="center">3.79</td>
<td valign="top" align="center">3.89</td>
<td valign="top" align="center">4.31</td>
<td valign="top" align="center">3.30</td>
</tr>
<tr>
<td valign="top" align="left">Std. trait score</td>
<td valign="top" align="center">0.56</td>
<td valign="top" align="center">0.47</td>
<td valign="top" align="center">0.54</td>
<td valign="top" align="center">0.48</td>
<td valign="top" align="center">0.43</td>
<td valign="top" align="center">0.49</td>
</tr>
<tr>
<td valign="top" align="left">Female&#x02014;Ave. trait score</td>
<td valign="top" align="center">4.28</td>
<td valign="top" align="center">2.90</td>
<td valign="top" align="center">3.74</td>
<td valign="top" align="center">3.91</td>
<td valign="top" align="center">4.37</td>
<td valign="top" align="center">3.25</td>
</tr>
<tr>
<td valign="top" align="left">Male&#x02014;Ave. trait score</td>
<td valign="top" align="center">4.18</td>
<td valign="top" align="center">2.83</td>
<td valign="top" align="center">3.85</td>
<td valign="top" align="center">3.88</td>
<td valign="top" align="center">4.32</td>
<td valign="top" align="center">3.35</td>
</tr>
<tr>
<td valign="top" align="left">Ave. word length</td>
<td valign="top" align="center">87.25</td>
<td valign="top" align="center">100.68</td>
<td valign="top" align="center">84.95</td>
<td valign="top" align="center">80.98</td>
<td valign="top" align="center">83.77</td>
<td valign="top" align="center">95.75</td>
</tr>
<tr>
<td valign="top" align="left">Std. word length</td>
<td valign="top" align="center">39.30</td>
<td valign="top" align="center">56.74</td>
<td valign="top" align="center">33.74</td>
<td valign="top" align="center">28.17</td>
<td valign="top" align="center">35.01</td>
<td valign="top" align="center">52.63</td>
</tr>
<tr>
<td valign="top" align="left">Female&#x02014;Ave. word length</td>
<td valign="top" align="center">109.56</td>
<td valign="top" align="center">137.74</td>
<td valign="top" align="center">103.06</td>
<td valign="top" align="center">97.68</td>
<td valign="top" align="center">102.47</td>
<td valign="top" align="center">126.94</td>
</tr>
<tr>
<td valign="top" align="left">Male&#x02014;Ave. word length</td>
<td valign="top" align="center">107.74</td>
<td valign="top" align="center">126.35</td>
<td valign="top" align="center">101.15</td>
<td valign="top" align="center">95.71</td>
<td valign="top" align="center">99.80</td>
<td valign="top" align="center">113.33</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>In the model training process for each trait, 80% of the data was used for training, 10% as the development data set for selecting the hyperparameters, and the remaining 10% of the data to validate the accuracy of the trained models. Answers with length less than 50 words were excluded from the training, development and testing data sets as candidate were instructed to answer questions with more than 50 words to provide enough context. More than 70% of candidates provided their gender information and female candidates tended to write longer answers than males on average with respect to the word count.</p>
</sec>
<sec>
<title>3.2. Answer Representations</title>
<p>We evaluated two commonly used open-vocabulary approaches for text representation, namely, TF-IDF &#x0002B; LDA and GloVe word embedding to compare the outcomes with the proposed InterviewBERT approach. Here we briefly describe how each approach was implemented in the experiment and details of the IntervewBERT development.</p>
<sec>
<title>3.2.1. TF-IDF and LDA</title>
<p>In this approach, we first remove special characters, numbers and stop words from answers. Then each answer is converted to lowercase and lemmatized before being tokenized. These are typical pre-processing steps in NLP for a context-free representation of textual data. Subsequently, 2,000-dimensional vector representations <inline-formula><mml:math id="M1"><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>v</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mi>f</mml:mi><mml:mi>i</mml:mi><mml:mi>d</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> are formed based on the term frequency-inverse document frequency (TF-IDF) scheme using the 2,000 most common unigrams, bigrams, and trigrams. In the TF-IDF scheme, the value for an answer-term combination increases with the number of times the term is used in the response while offsetting for the overall usage of the term in the whole training dataset. We implemented TF-IDF using sklearn<xref ref-type="fn" rid="fn0002"><sup>2</sup></xref> package.</p>
<p>We also used the Latent Dirichlet Allocation (LDA) (Blei et al., <xref ref-type="bibr" rid="B8">2003</xref>) to derive 100 topics from the answer. LDA assumes the existence of latent topics in a given set of documents and tries to probabilistically uncover these topics. Once uncovered, an answer can be represented with a 100-dimensional vector <inline-formula><mml:math id="M2"><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>v</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>l</mml:mi><mml:mi>d</mml:mi><mml:mi>a</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula>. We use the Gensim software package<xref ref-type="fn" rid="fn0003"><sup>3</sup></xref> for topic modeling. The combined use of TF-IDF with LDA was shown by Jayaratne and Jayatilleke (<xref ref-type="bibr" rid="B32">2020</xref>) to produce the best accuracy in predicting personality from a similar dataset related to interview responses.</p>
<p>We used the same approach to obtain the final answer representation <italic><bold>V</bold></italic><sub><italic><bold>i</bold></italic></sub> by concatenating the representation based on terms <inline-formula><mml:math id="M3"><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>v</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mi>f</mml:mi><mml:mi>i</mml:mi><mml:mi>d</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> and the answer representation based on topics <inline-formula><mml:math id="M4"><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>v</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>l</mml:mi><mml:mi>d</mml:mi><mml:mi>a</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> for answer <italic>a</italic><sub><italic>i</italic></sub>:</p>
<disp-formula id="E1"><label>(1)</label><mml:math id="M5"><mml:mtable class="eqnarray" columnalign="right center left"><mml:mtr><mml:mtd><mml:mstyle mathvariant="bold"><mml:msub><mml:mrow><mml:mi>V</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub></mml:mstyle><mml:mo>=</mml:mo><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>v</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>l</mml:mi><mml:mi>d</mml:mi><mml:mi>a</mml:mi></mml:mrow></mml:msubsup><mml:mo>&#x02295;</mml:mo><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>v</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mi>f</mml:mi><mml:mi>i</mml:mi><mml:mi>d</mml:mi><mml:mi>f</mml:mi></mml:mrow></mml:msubsup></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where &#x02295; is the concatenating operation and <italic><bold>V</bold></italic><sub><italic><bold>i</bold></italic></sub> is a 2100-dimensional vector.</p>
</sec>
<sec>
<title>3.2.2. GloVe</title>
<p>GloVe model uses the co-occurrence probabilities of words within a text corpus in order to embed them in meaningful vectors. It first collects word co-occurrence statistics in the form of a word co-occurrence matrix <italic>X</italic>. Element <italic>X</italic><sub><italic>ij</italic></sub> represents how often the main word <italic>i</italic> appears in the context of word <italic>j</italic> by scanning the corpus with a fixed window size for the main word <italic>i</italic>. Then it learns vectors by doing dimensional reduction on the co-occurrence counts matrix. In this paper, we use the GloVe embeddings that are pre-trained on Common Crawl<xref ref-type="fn" rid="fn0004"><sup>4</sup></xref>. The pre-trained model contains 840B tokens with each token represented as a 300-dimensional vector.</p>
<p>We first tokenize answers based on whitespace, newline characters, and punctuation as delimiters. Then, we represent each token as a GloVe embedding <inline-formula><mml:math id="M6"><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>v</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mi>g</mml:mi><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mi>v</mml:mi><mml:mi>e</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> using torchtext&#x00027;s glove embedding tool<xref ref-type="fn" rid="fn0005"><sup>5</sup></xref>. All the out-of-vocabulary tokens are represented as the same vector of <italic>[UNK]</italic>. To get the final answer representation <italic><bold>V</bold></italic><sub><italic><bold>i</bold></italic></sub>, we averaged across all token representations:</p>
<disp-formula id="E2"><label>(2)</label><mml:math id="M7"><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>V</mml:mi></mml:mstyle><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>i</mml:mi></mml:mstyle></mml:msub><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mi>m</mml:mi></mml:mfrac><mml:mstyle displaystyle='true'><mml:munderover><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>j</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mi>m</mml:mi></mml:munderover><mml:mrow><mml:msubsup><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>v</mml:mi></mml:mstyle><mml:mrow><mml:msub><mml:mi>i</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mi>g</mml:mi><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mi>v</mml:mi><mml:mi>e</mml:mi></mml:mrow></mml:msubsup></mml:mrow></mml:mstyle></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>where <italic>m</italic> is the number of tokens in the answer.</p>
</sec>
<sec>
<title>3.2.3. InterviewBERT</title>
<p>To improve the pre-trained BERT language model for interview language understanding, we first extended it with a large interview answer corpus of over 3 million answers. Each answer is first tokenized using WordPiece tokenizer<xref ref-type="fn" rid="fn0006"><sup>6</sup></xref> that also adds a special token [<italic>CLS</italic>] to the start of each answer to enable an answer level representation. InterviewBERT is then pre-trained using the same tasks detailed in Section 2.4. The training process updates the word embeddings based on interview context, while not losing the prior knowledge in general domains. That is, after the pre-training, InterviewBERT contains the general language knowledge already encoded in the initial BERT with the addition of job interview specific knowledge learnt from interview answers.</p>
<p>With pre-trained InterviewBERT, we can either fetch answer representations by passing each answer through its encoder and then training a personality predictor based on those representations (i.e., train an independent regressor, as shown in <xref ref-type="fig" rid="F3">Figure 3A</xref>), or fine-tune the pre-trained model itself for personality prediction task (i.e., add a regression layer to InterviewBERT itself, as shown in <xref ref-type="fig" rid="F3">Figure 3B</xref>). A hybrid approach is to get contextualized answer embeddings after fine-tuning InterviewBERT for a regression task and then training a personality predictor separately (as shown in <xref ref-type="fig" rid="F3">Figure 3C</xref>). The advantage of the hybrid approach is that the representation of the same answer could be optimized individually for each regression task (e.g., predicting the trait Extraversion vs. Agreeableness) to obtain better results for the given task than using a generic representation. To make a fair comparison with other methods based on context-free representations, in which the personality predictors are separately trained after obtaining answer representations, we used the hybrid approach in building the InterviewBERT based models.</p>
<fig id="F3" position="float">
<label>Figure 3</label>
<caption><p>Illustration of different ways of using InterviewBERT. <bold>(A)</bold> Train an independent regressor without fine-tuning. <bold>(B)</bold> Fine-tune InterviewBERT. <bold>(C)</bold> Hybrid approach.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-865841-g0003.tif"/>
</fig>
<p>In order to obtain task and context specific representations we explored two approaches; (a) fine-tuning the model using the learned context-specific [<italic>CLS</italic>] representations (InterviewBERT-CLS) or (b) average all the learned context-specific word embeddings in an answer (InterviewBERT-AVE) as the answer representation. We briefly describe each method below and then report the performance of each approach in Section 4.</p>
<p><bold>InterviewBERT-CLS</bold> we first input an answer to the encoder of InterviewBERT and get the last hidden representation <inline-formula><mml:math id="M8"><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>v</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>c</mml:mi><mml:mi>l</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup><mml:mo>&#x02208;</mml:mo><mml:msup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>R</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> for <italic>[CLS]</italic> as the answer representation, where <italic>d</italic> &#x0003D; 768. We then passed it through a regression layer to get the personality score <inline-formula><mml:math id="M9"><mml:msubsup><mml:mrow><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup><mml:mo>&#x02208;</mml:mo><mml:mstyle mathvariant="bold"><mml:mi>R</mml:mi></mml:mstyle></mml:math></inline-formula>. The model is fine-tuned by minimizing the mean squared error loss <italic>L</italic> between prediction scores <inline-formula><mml:math id="M10"><mml:msubsup><mml:mrow><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> and the ground truth scores <inline-formula><mml:math id="M11"><mml:msubsup><mml:mrow><mml:mi>&#x00177;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> for trait <italic>T</italic>:</p>
<disp-formula id="E3"><label>(3)</label><mml:math id="M12"><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:msup><mml:mi>L</mml:mi><mml:mi>T</mml:mi></mml:msup><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:msup><mml:mi>X</mml:mi><mml:mi>T</mml:mi></mml:msup></mml:mrow></mml:mfrac><mml:mstyle displaystyle='true'><mml:munderover><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:msup><mml:mi>X</mml:mi><mml:mi>T</mml:mi></mml:msup></mml:mrow></mml:munderover><mml:mrow><mml:msup><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:msubsup><mml:mi>y</mml:mi><mml:mi>i</mml:mi><mml:mi>T</mml:mi></mml:msubsup><mml:mo>&#x02212;</mml:mo><mml:msubsup><mml:mover accent='true'><mml:mi>y</mml:mi><mml:mo>&#x0005E;</mml:mo></mml:mover><mml:mi>i</mml:mi><mml:mi>T</mml:mi></mml:msubsup><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mn>2</mml:mn></mml:msup></mml:mrow></mml:mstyle></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>where X is the total number of answers to fine-tune the models for trait <italic>T</italic>.</p>
<p>With the fine-tuned model, we can get context-specific representations <inline-formula><mml:math id="M13"><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>V</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> by passing answers through the InterviewBERT encoder:</p>
<disp-formula id="E4"><label>(4)</label><mml:math id="M14"><mml:mtable class="eqnarray" columnalign="right center left"><mml:mtr><mml:mtd><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>V</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup><mml:mo>=</mml:mo><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>v</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>c</mml:mi><mml:mi>l</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p><bold>InterviewBERT-AVE</bold> instead of using an aggregated answer representation as above, word level representations are averaged to get an answer level representation. We obtained all the hidden word representations from the model except for the [<italic>CLS</italic>] token from the last layer, and then averaged across all word representations to obtain the answer representation <inline-formula><mml:math id="M15"><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>V</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup><mml:mo>&#x02208;</mml:mo><mml:msup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>R</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula>. Then, similar to [<italic>CLS</italic>] based representation, we pass <inline-formula><mml:math id="M16"><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>V</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> through a regression layer to get personality score <inline-formula><mml:math id="M17"><mml:msubsup><mml:mrow><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula>. The model is fine-tuned by minimizing the mean squared error loss <italic>L</italic> between prediction scores and the ground truth scores.</p>
<p>With the fine-tuned model, we can get a context-specific answer representation <inline-formula><mml:math id="M18"><mml:msubsup><mml:mrow><mml:mstyle mathvariant="bold"><mml:mi>V</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> by passing the answer through the InterviewBERT encoder, and averaging the output word embeddings:</p>
<disp-formula id="E5"><label>(5)</label><mml:math id="M19"><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:msubsup><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>V</mml:mi></mml:mstyle><mml:mi>i</mml:mi><mml:mi>T</mml:mi></mml:msubsup><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mi>m</mml:mi><mml:mo>&#x02212;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:mfrac><mml:mstyle displaystyle='true'><mml:munderover><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>j</mml:mi><mml:mo>=</mml:mo><mml:mn>2</mml:mn></mml:mrow><mml:mi>m</mml:mi></mml:munderover><mml:mrow><mml:msubsup><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>v</mml:mi></mml:mstyle><mml:mrow><mml:msub><mml:mi>i</mml:mi><mml:mi>j</mml:mi></mml:msub></mml:mrow><mml:mi>T</mml:mi></mml:msubsup></mml:mrow></mml:mstyle></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>where <italic>m</italic> is the number of tokens in answer <italic>a</italic><sub><italic>i</italic></sub>.</p>
<p>The models are implemented using the Huggingface transformers package<xref ref-type="fn" rid="fn0007"><sup>7</sup></xref> and optimized on Nvidia Tesla T4 using the Adam optimizer (Kingma and Ba, <xref ref-type="bibr" rid="B36">2014</xref>) with &#x003B2;<sub>1</sub> &#x0003D; 0.9, &#x003B2;<sub>2</sub> &#x0003D; 0.999, &#x003F5; &#x0003D; 1<italic>e</italic>&#x02212;6, weight decay tuned among [0.001, <italic>0.01</italic>]. The learning rate, which is warmed up over the first 500 steps, is tuned between [<italic>1e-4</italic>, 1e-3], and then linearly decayed. The model is trained with a dropout, which is tuned between [<italic>0.1</italic>, 0.2], on all layers and attention weights to avoid overfitting. The batch size is tuned between [<italic>8</italic>, 16] and the maximum sequence length is tuned between [256, <italic>512</italic>]. We use the development datasets to select the best hyperparameters and the optimal hyperparameters are highlighted above in <italic>italics</italic>. The model is trained for a maximum of 5 epochs with evaluation for every 2,000 steps. Early stopping is set once convergence is determined, i.e., when the loss <italic>L</italic> on the development set does not decrease after 10,000 steps.</p>
</sec>
</sec>
<sec>
<title>3.3. Personality Inference</title>
<p>Once an answer representation is obtained using the context-free and context-specific methods discussed above, the inference task involves building a regressor for each HEXACO trait using the text representation as the independent variable and the self-rating score as the dependent (target) variable. We used the Random Forest algorithm (Breiman, <xref ref-type="bibr" rid="B10">2001</xref>) implemented using sklearn<xref ref-type="fn" rid="fn0008"><sup>8</sup></xref> to train regression models for each trait with a maximum tree depth set to 50 and the number of trees in the forest set to 100. Given each participant responded to 5&#x02013;7 interview questions but only had a single trait score from the self-report items, two methods were explored to aggregate the answer representations in building the regression model. One method was to train a regression model using each individual answer representation to predict the trait score of a participant and then average scores across answers to get the final individual trait score. Second method was to average all answer representations for a candidate and use the averaged answer representation to train a regression model to predict the trait score. We found that method one provided higher accuracy than method two and hence report results from the first method in Section 4.</p>
</sec>
</sec>
<sec sec-type="results" id="s4">
<title>4. Results</title>
<p>We evaluated the trained models on the 10% of the data set left out for testing using the Pearson correlation coefficient, <italic>r</italic>, between the ground truth personality scores, &#x00177;, and the predicted personality scores, <italic>p</italic>.</p>
<p><xref ref-type="table" rid="T2">Table 2</xref> presents the performance of the models trained on different answer representation methods. All representation methods produced predictive models with varying levels of positive correlations (<italic>p</italic> &#x0003C; 0.001) for all six HEXACO traits. This demonstrates that language used in responding to interview questions are predictive of one&#x00027;s personality. Further the higher average correlations in context-specific InterviewBERT models over the context-free approaches highlight the superiority of InterviewBERT.</p>
<table-wrap position="float" id="T2">
<label>Table 2</label>
<caption><p>Correlation coefficient <italic>r</italic> of different methods which aggregate the scores after prediction.</p></caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th valign="top" align="left" colspan="2"><bold>Methods</bold></th>
<th valign="top" align="left"><bold>H</bold></th>
<th valign="top" align="center"><bold>E</bold></th>
<th valign="top" align="center"><bold>X</bold></th>
<th valign="top" align="center"><bold>A</bold></th>
<th valign="top" align="center"><bold>C</bold></th>
<th valign="top" align="center"><bold>O</bold></th>
<th valign="top" align="center"><bold>Ave. <italic>r</italic></bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Context-free</td>
<td valign="top" align="left">TF-IDF &#x0002B; LDA</td>
<td valign="top" align="center">0.31</td>
<td valign="top" align="center">0.28</td>
<td valign="top" align="center">0.41</td>
<td valign="top" align="center">0.23</td>
<td valign="top" align="center">0.38</td>
<td valign="top" align="center">0.39</td>
<td valign="top" align="center">0.333</td>
</tr>
<tr>
<td/>
<td valign="top" align="left">GloVe</td>
<td valign="top" align="center">0.34</td>
<td valign="top" align="center"><bold>0.30</bold></td>
<td valign="top" align="center">0.41</td>
<td valign="top" align="center">0.24</td>
<td valign="top" align="center">0.38</td>
<td valign="top" align="center">0.39</td>
<td valign="top" align="center">0.343</td>
</tr>
<tr>
<td valign="top" align="left">Context-specific</td>
<td valign="top" align="left">InterviewBERT-AVE</td>
<td valign="top" align="center"><bold>0.37</bold></td>
<td valign="top" align="center"><bold>0.30</bold></td>
<td valign="top" align="center"><bold>0.44</bold></td>
<td valign="top" align="center"><bold>0.28</bold></td>
<td valign="top" align="center">0.40</td>
<td valign="top" align="center"><bold>0.45</bold></td>
<td valign="top" align="center"><bold>0.373</bold></td>
</tr>
<tr>
<td/>
<td valign="top" align="left">InterviewBERT-CLS</td>
<td valign="top" align="center"><bold>0.37</bold></td>
<td valign="top" align="center"><bold>0.30</bold></td>
<td valign="top" align="center"><bold>0.44</bold></td>
<td valign="top" align="center"><bold>0.28</bold></td>
<td valign="top" align="center"><bold>0.41</bold></td>
<td valign="top" align="center">0.44</td>
<td valign="top" align="center"><bold>0.373</bold></td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>All correlations are significant with p &#x0003C; 0.001. Bold numbers indicate the best correlation for each trait</italic>.</p>
</table-wrap-foot>
</table-wrap>
<p><xref ref-type="table" rid="T3">Table 3</xref> presents the correlation between answer word length and the predicted trait scores of different models. This demonstrates that the answer length to interview questions have a weaker correlation with one&#x00027;s personality compared with language use.</p>
<table-wrap position="float" id="T3">
<label>Table 3</label>
<caption><p>Correlation coefficient <italic>r</italic> between answer length and trait scores predicted by different methods.</p></caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th valign="top" align="left"><bold>Methods</bold></th>
<th valign="top" align="center"><bold>H</bold></th>
<th valign="top" align="center"><bold>E</bold></th>
<th valign="top" align="center"><bold>X</bold></th>
<th valign="top" align="center"><bold>A</bold></th>
<th valign="top" align="center"><bold>C</bold></th>
<th valign="top" align="center"><bold>O</bold></th>
<th valign="top" align="center"><bold>Ave. |<italic>r</italic>|</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Ground truth</td>
<td valign="top" align="center">&#x02212;0.04</td>
<td valign="top" align="center">0.02</td>
<td valign="top" align="center">0.08</td>
<td valign="top" align="center">0.01</td>
<td valign="top" align="center">0.08</td>
<td valign="top" align="center">&#x02212;0.04</td>
<td valign="top" align="center">0.045</td>
</tr>
<tr>
<td valign="top" align="left">TF-IDF &#x0002B; LDA</td>
<td valign="top" align="center">&#x02212;0.07</td>
<td valign="top" align="center">0.17</td>
<td valign="top" align="center">0.10</td>
<td valign="top" align="center">0.01</td>
<td valign="top" align="center">0.10</td>
<td valign="top" align="center">&#x02212;0.13</td>
<td valign="top" align="center">0.096</td>
</tr>
<tr>
<td valign="top" align="left">GloVe</td>
<td valign="top" align="center">-0.13</td>
<td valign="top" align="center">0.15</td>
<td valign="top" align="center">0.07</td>
<td valign="top" align="center">0.01</td>
<td valign="top" align="center">0.07</td>
<td valign="top" align="center">&#x02212;0.12</td>
<td valign="top" align="center">0.092</td>
</tr>
<tr>
<td valign="top" align="left">InterviewBERT-AVE</td>
<td valign="top" align="center">&#x02212;0.06</td>
<td valign="top" align="center">0.11</td>
<td valign="top" align="center">0.16</td>
<td valign="top" align="center">0.04</td>
<td valign="top" align="center">0.11</td>
<td valign="top" align="center">&#x02212;0.02</td>
<td valign="top" align="center">0.083</td>
</tr>
<tr>
<td valign="top" align="left">InterviewBERT-CLS</td>
<td valign="top" align="center">&#x02212;0.05</td>
<td valign="top" align="center">0.13</td>
<td valign="top" align="center">0.15</td>
<td valign="top" align="center">0.06</td>
<td valign="top" align="center">0.12</td>
<td valign="top" align="center">&#x02212;0.01</td>
<td valign="top" align="center">0.087</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>All correlations are significant with p &#x0003C; 0.001</italic>.</p>
</table-wrap-foot>
</table-wrap>
<p><xref ref-type="table" rid="T4">Table 4</xref> presents the inter-correlations between the personality scores inferred using the InterviewBERT-CLS model on an independent group of <italic>N</italic> = 11,433 candidates. This demonstrates that there are strong inter-correlations (|<italic>r</italic>|&#x0003E;0.20) between some personality traits, and these correlations are consistent with previous findings in literature (Ashton and Lee, <xref ref-type="bibr" rid="B5">2009</xref>; Lee and Ashton, <xref ref-type="bibr" rid="B38">2018</xref>; Moshagen et al., <xref ref-type="bibr" rid="B56">2019</xref>; Skimina et al., <xref ref-type="bibr" rid="B76">2020</xref>).</p>
<table-wrap position="float" id="T4">
<label>Table 4</label>
<caption><p>Intercorrelations between personality scores inferred using InterviewBERT on an independent group of 11,433 candidates.</p></caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th valign="top" align="center"><bold>H</bold></th>
<th valign="top" align="center"><bold>E</bold></th>
<th valign="top" align="center"><bold>X</bold></th>
<th valign="top" align="center"><bold>A</bold></th>
<th valign="top" align="center"><bold>C</bold></th>
<th valign="top" align="center"><bold>O</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Honesty-humility (H)</td>
<td valign="top" align="center">1.00</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
</tr>
<tr>
<td valign="top" align="left">Emotionality (E)</td>
<td valign="top" align="center">0.18</td>
<td valign="top" align="center">1.00</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
</tr>
<tr>
<td valign="top" align="left">Extraversion (X)</td>
<td valign="top" align="center">&#x02212;0.18</td>
<td valign="top" align="center">&#x02212;0.12</td>
<td valign="top" align="center">1.00</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
</tr>
<tr>
<td valign="top" align="left">Agreeableness (A)</td>
<td valign="top" align="center">0.31</td>
<td valign="top" align="center">0.14</td>
<td valign="top" align="center">0.05</td>
<td valign="top" align="center">1.00</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
</tr>
<tr>
<td valign="top" align="left">Conscientiousness (C)</td>
<td valign="top" align="center">0.29</td>
<td valign="top" align="center">0.20</td>
<td valign="top" align="center">0.24</td>
<td valign="top" align="center">0.19</td>
<td valign="top" align="center">1.00</td>
<td valign="top" align="center">&#x02013;</td>
</tr>
<tr>
<td valign="top" align="left">Openness (O)</td>
<td valign="top" align="center">&#x02212;0.10</td>
<td valign="top" align="center">0.11</td>
<td valign="top" align="center">0.43</td>
<td valign="top" align="center">&#x02212;0.11</td>
<td valign="top" align="center">0.33</td>
<td valign="top" align="center">1.00</td>
</tr>
</tbody>
</table>
</table-wrap>
<p><xref ref-type="fig" rid="F4">Figure 4</xref> shows an example of four attention heatmaps from InterviewBERT for answers from two candidates with corresponding tokens related to Openness (O) and extraversion (X); tokens with higher attention are in darker color. This demonstrates how InterviewBERT can provide us with reasonable language-level explanations of personality inference results allowing further analysis of language patterns related to personality.</p>
<fig id="F4" position="float">
<label>Figure 4</label>
<caption><p>Visualization of attentions from InterveiwBERT along with corresponding tokens for openness (O) and extraversion (X). Tokens with higher attention are in darker color.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-865841-g0004.tif"/>
</fig>
<p><xref ref-type="table" rid="T5">Table 5</xref> presents <italic>t</italic>-test results for mean difference between female (<italic>N</italic> = 5,673) and male (<italic>N</italic> = 5,760) on predicted trait scores. These findings of gender differences are similar to that reported by others using standard self-rated item inventories (Feingold, <xref ref-type="bibr" rid="B19">1994</xref>; Ashton and Lee, <xref ref-type="bibr" rid="B5">2009</xref>; Wakabayashi, <xref ref-type="bibr" rid="B86">2014</xref>; Lee and Ashton, <xref ref-type="bibr" rid="B39">2020</xref>; Skimina et al., <xref ref-type="bibr" rid="B76">2020</xref>).</p>
<table-wrap position="float" id="T5">
<label>Table 5</label>
<caption><p>Mean differences between predicted female (<italic>N</italic> = 5,673) and male (<italic>N</italic> = 5,760) trait scores (<italic>p</italic> &#x0003C; 0.001).</p></caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th valign="top" align="left"><bold>Traits</bold></th>
<th valign="top" align="center"><bold>Diff</bold></th>
<th valign="top" align="center"><bold><italic>t</italic></bold></th>
<th valign="top" align="center"><bold>Cohen <italic>d</italic></bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">H</td>
<td valign="top" align="center">0.04</td>
<td valign="top" align="center">4.61</td>
<td valign="top" align="center">0.26</td>
</tr>
<tr>
<td valign="top" align="left">E</td>
<td valign="top" align="center">0.02</td>
<td valign="top" align="center">5.71</td>
<td valign="top" align="center">0.23</td>
</tr>
<tr>
<td valign="top" align="left">X</td>
<td valign="top" align="center">&#x02212;0.02</td>
<td valign="top" align="center">&#x02212;5.49</td>
<td valign="top" align="center">&#x02212;0.16</td>
</tr>
<tr>
<td valign="top" align="left">A</td>
<td valign="top" align="center">0.02</td>
<td valign="top" align="center">4.74</td>
<td valign="top" align="center">0.17</td>
</tr>
<tr>
<td valign="top" align="left">C</td>
<td valign="top" align="center">0.01</td>
<td valign="top" align="center">2.0</td>
<td valign="top" align="center">0.09</td>
</tr>
<tr>
<td valign="top" align="left">O</td>
<td valign="top" align="center">&#x02212;0.04</td>
<td valign="top" align="center">&#x02212;8.87</td>
<td valign="top" align="center">&#x02212;0.31</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>Positive differences indicate higher females means</italic>.</p>
</table-wrap-foot>
</table-wrap>
</sec>
<sec sec-type="discussion" id="s5">
<title>5. Discussion</title>
<p>The job interview is one of the most widely used assessment tools in the selection process. Personality perception through verbal and non-verbal signals is a common practice used by interviewers in employment interviews. Perceived personality traits of candidates, especially the traits Openness to experience and Conscientiousness, have been found to be positively correlated to interview outcomes (Caldwell and Burger, <xref ref-type="bibr" rid="B11">1998</xref>; Van Dam, <xref ref-type="bibr" rid="B82">2003</xref>). Personality perception is related to the notions of Spontaneous Trait Inference (STI) (Uleman, <xref ref-type="bibr" rid="B80">1989</xref>) and Intentional Trait Inference (ITI) (Uleman, <xref ref-type="bibr" rid="B81">1999</xref>) found in social psychology. Spontaneous trait inferences require little mental effort and are difficult to suppress or modify (closer to being unconscious or automatic), while intentional trait inferences (ITI) require deliberate effort to make a relevant social judgement. In an interview setting the interviewers are expected to get rid of STI, which may unintentionally bring in subjective biases and unavoidably affect the interview results. However, recent research in neuroimaging (Van Duynslaeger et al., <xref ref-type="bibr" rid="B83">2007</xref>; Ma et al., <xref ref-type="bibr" rid="B47">2011</xref>) and social experiments (Ferreira et al., <xref ref-type="bibr" rid="B20">2012</xref>) have shown that STI and ITI often run in synchrony without our awareness. Structured interviews on the other hand attempt to mitigate the impact of interviewer biases by asking the same questions from all candidates with limited interviewer probing and by using a clear scoring rubric to evaluate the candidates based on their responses (Levashina et al., <xref ref-type="bibr" rid="B42">2013</xref>). However, in in-person or video-based structured job interviews where a candidate&#x00027;s appearance or verbal signals are available, it is difficult to avoid interviewers unintentionally forming impressions of candidates influenced by attributes such as race, gender, age, and appearance (Purkiss et al., <xref ref-type="bibr" rid="B68">2006</xref>). Especially when faced with a large number of interviews, human interviewers can hardly infer traits accurately and efficiently.</p>
<p>Our work shows that textual content of interview responses can offer interviewers with rich and deep understanding of a candidate&#x00027;s personality. When used with a structured interview, algorithmic inference of personality from interview responses can help reduce errors in spontaneous trait inference as discussed above. As shown in <xref ref-type="table" rid="T2">Table 2</xref> the content and context of answers are strongly correlated with self-reported personality scores. The results also highlight the superiority of context-specific answer representation approaches over context-free approaches by producing the most accurate models. The InterviewBERT based models reached the highest average accuracy of <italic>r</italic> = 0.373 (<italic>p</italic> &#x0003C; 0.001) across the six HEXACO traits while models for Extraversion, Conscientiousness, and Openness exceeded 0.4 correlation, a value typically considered a &#x0201C;correlation upper-limit&#x0201D; for predicting personality with behavior (Meyer et al., <xref ref-type="bibr" rid="B52">2001</xref>; Roberts et al., <xref ref-type="bibr" rid="B71">2007</xref>). Among the different answer representation methods, context-specific methods achieved better results than context-free ones. This is reasonable since context-free methods assume the meaning of a word or a sentence to be relatively stable and unlikely to change across different contexts (see discussion in Section 2.3). On the contrary, context-specific methods are closer to a human in understanding language that consider the context of the words and inter-word correlations in a sentence to better understand the answers. Similar results are also reported on personality prediction using social media data, like tweets and Facebook posts, where context-specific methods (Christian et al., <xref ref-type="bibr" rid="B12">2021</xref>; Lucky and Suhartono, <xref ref-type="bibr" rid="B46">2021</xref>) achieved better performance than context-free methods, such as TF-IDF (Pratama and Sarno, <xref ref-type="bibr" rid="B67">2015</xref>), LDA (Ong et al., <xref ref-type="bibr" rid="B59">2017</xref>), and GloVe (Tandera et al., <xref ref-type="bibr" rid="B79">2017</xref>).</p>
<p>With context-specific methods, traits Extraversion, Conscientiousness, and Openness achieved higher accuracies with correlations exceeding 0.4. These three traits also achieved the highest correlations for context-free methods, albeit only Extraversion exceeding 0.4. Jayaratne and Jayatilleke (<xref ref-type="bibr" rid="B32">2020</xref>), also predicting personality from interview responses, report similar results with Conscientiousness and Openness exceeding 0.4 and Extraversion achieving a correlation of 0.34. In a large study using social media data from over 75,000 volunteers, Schwartz et al. (<xref ref-type="bibr" rid="B75">2013</xref>) also report their highest correlations for Big5 dimensions Openness, Extraversion, and Conscientiousness at 0.42, 0.38, and 0.35, respectively. On the other hand, we observed that Emotionality and Agreeableness are harder to predict from textual responses with correlations &#x0003C; &#x0003D; 0.3. This is in line with Jayaratne and Jayatilleke (<xref ref-type="bibr" rid="B32">2020</xref>) and Schwartz et al. (<xref ref-type="bibr" rid="B75">2013</xref>), where they reported their lowest correlations for Agreeableness and Emotionality (Neuroticism in the case of Schwartz et al., <xref ref-type="bibr" rid="B75">2013</xref> using Big5). The above highlights the different degrees to which language encodes personality signals for different personality traits. Exploring which characteristics of language-use lead to these differences is a useful future direction that is out of scope for this paper.</p>
<p>It is important to highlight here that previous work by Jayaratne and Jayatilleke (<xref ref-type="bibr" rid="B32">2020</xref>) using only context-free approaches reached an average correlation of <italic>r</italic> = 0.387 on a similar study with interview responses based personality inference. There are two fundamental differences between this previous study and the current one that we see as improvements, apart from the use of context-specific InterviewBERT. Firstly, the previous study used a concatenated string combining all 5&#x02013;7 answers per candidate as input to the regressions model while the current study used individual answers to predict personality and then averaged the predicted scores to obtain a final score. Use of individual answers to build InterviewBERT and the proceeding trait prediction models allow the retention of individual answer context compared to combining with other answers. Further the models are less susceptible to the variance in text length due to the varying number of questions in different interviews. Secondly, the TF-IDF &#x0002B; LDA approach used in the previous study lacks the ability to provide explainability as enabled by the InterviewBERT approach.</p>
<p>As shown in <xref ref-type="table" rid="T3">Table 3</xref>, the average correlations between answer length and personality scores are low (Ave.|<italic>r</italic>| &#x0003C; 0.10). Only Extraversion (X), Conscientiousness (C), and Emotionality (E) have relatively higher correlations with answer length for predicted results (|<italic>r</italic>|&#x0003E;0.10). While further work is required in explaining these higher correlations, some hypotheses can be formed based on the reported characteristics of the traits. For example Extraversion is associated with being sociable and more confident in expressing themselves (McCabe and Fleeson, <xref ref-type="bibr" rid="B50">2016</xref>; Diener and Lucas, <xref ref-type="bibr" rid="B17">2019</xref>), and a long answer may indicate these tendencies. Conscientiousness is associated with striving for accuracy and perfection (McCabe and Fleeson, <xref ref-type="bibr" rid="B50">2016</xref>; Pletzer et al., <xref ref-type="bibr" rid="B66">2019</xref>) and it is reasonable that they tend to answer questions with longer and more elaborate responses. It is interesting to note that the predicted Emotionality (E) scores of all four methods showed correlations of |<italic>r</italic>|&#x0003E;0.10 with the answer length while the correlation with the ground truth remained at 0.02. Further analysis using the explainability features available in InterviewBERT can help explain these correlations by identifying the patterns in language that lead to higher vs. lower scores.</p>
<p>While our context-specific models are trained to predict each personality trait individually, there are inherent inter-correlations among the different personality traits. As shown in <xref ref-type="table" rid="T4">Table 4</xref>, the HEXACO personality scores inferred by InterviewBERT were weakly correlated overall (Ave. |<italic>r</italic>| &#x02264; 0.20). However, their are high correlations (<italic>r</italic>&#x0003E;0.20) between Honesty-humility and Agreeableness (H-A), Honesty-humility and Conscientiousness (H-C), Extraversion and Conscientiousness (X-C), Extraversion and Openness (X-O), and Conscientiousness and Openness (C-O). These high correlations have also been reported elsewhere in self-report studies (Ashton and Lee, <xref ref-type="bibr" rid="B5">2009</xref>; Lee and Ashton, <xref ref-type="bibr" rid="B38">2018</xref>; Moshagen et al., <xref ref-type="bibr" rid="B56">2019</xref>; Skimina et al., <xref ref-type="bibr" rid="B76">2020</xref>). (Skimina et al., <xref ref-type="bibr" rid="B76">2020</xref>) report a high inter-correlation for H-A (<italic>r</italic> &#x0003D; 0.44), H-C (<italic>r</italic> &#x0003D; 0.28), X-C (<italic>r</italic> &#x0003D; 0.24), X-O (<italic>r</italic> &#x0003D; 0.22), C-O (<italic>r</italic> &#x0003D; 0.21) based on HEXACO-60 and HEXACO-100 self-rated inventories, which is consistent with the correlations we found based on textual answers to interview questions. Lee and Ashton (<xref ref-type="bibr" rid="B38">2018</xref>) report a high H-A correlation in different test groups (0.28 &#x0003C; <italic>r</italic> &#x0003C; 0.42). Moshagen et al. (<xref ref-type="bibr" rid="B56">2019</xref>) also found H-A to have the highest correlation and the correlation between X-C to be the second highest. Ashton and Lee (<xref ref-type="bibr" rid="B5">2009</xref>) report a high correlation for H-A (<italic>r</italic> &#x0003D; 0.25) and X-O (<italic>r</italic> &#x0003D; 0.26) on a community sample of 734 candidates. These results indicate that our findings of language inferred trait inter-correlations are in-line with other previous studies.</p>
<p>Gender related differences are another aspect where our findings are in line with some of the previous findings (Ashton and Lee, <xref ref-type="bibr" rid="B5">2009</xref>; Lee and Ashton, <xref ref-type="bibr" rid="B38">2018</xref>; Moshagen et al., <xref ref-type="bibr" rid="B56">2019</xref>; Skimina et al., <xref ref-type="bibr" rid="B76">2020</xref>). As shown in <xref ref-type="table" rid="T5">Table 5</xref> female candidates show a higher mean difference in Honesty-humility (H) (<italic>d</italic> &#x0003D; 0.26) and Emotionality (E) (<italic>d</italic> &#x0003D; 0.23) compared to males, which is consistent with a study in 48 countries with 347,192 participants (Lee and Ashton, <xref ref-type="bibr" rid="B39">2020</xref>), a study of 522 participants aged 16&#x02013;75 with 56.3% female (Skimina et al., <xref ref-type="bibr" rid="B76">2020</xref>), and a study of 734 participants with 413 females (Ashton and Lee, <xref ref-type="bibr" rid="B5">2009</xref>). Further, our results show that male candidates on average are higher in Openness to experience (O) than females candidates (<italic>d</italic> &#x0003D; 0.31). While some previous studies have also reported similar results, especially for the Inquisitiveness facet in O with <italic>d</italic> &#x0003D; 0.44 (Skimina et al., <xref ref-type="bibr" rid="B76">2020</xref>) tested on Polish participants, large scale studies such as (Lee and Ashton, <xref ref-type="bibr" rid="B39">2020</xref>) found otherwise. As for Extraversion (X), Agreeableness (A) and Conscientiousness (C), there are no significant differences between female and male candidates (|<italic>d</italic>| &#x0003C; 0.20) and this is consistent with Lee and Ashton (<xref ref-type="bibr" rid="B39">2020</xref>), who also tested personality using the HEXACO model.</p>
<p>Explainability is one of the key attributes of ethical use of algorithms together with aspects such as accountability and fairness (Hagendorff, <xref ref-type="bibr" rid="B26">2020</xref>). It addresses the &#x0201C;black-box&#x0201D; problem raised by users of machine learning related to the lack of transparency on how the algorithm works and explaining the outcomes. The ability to see into the &#x0201C;black-box&#x0201D; of the algorithm to get at least a high level understanding of how the outcome is derived increases the user&#x00027;s trust. Using the self-attention mechanism in InterviewBERT (ref. Sections 2.4, 3.2.3) we are able to visualize the attention weights of different words on real interview answers to better examine and understand how various language patterns influence trait outcomes. <xref ref-type="fig" rid="F4">Figure 4</xref> shows an example of four attention heat-maps for answers from two candidates with corresponding tokens for Openness (O) and Extraversion (X); tokens with higher attention are in darker color. As can be seen, the attention-based methods provide us with reasonable language-level details to analyze the associations learnt by the machine learning models between language patterns and personality. While further work is required in analyzing these associations to discover general patterns (e.g., which words or phrase co-occurrences are more likely to make someone high in Agreeableness), the attention weights in InterviewBERT provide us the data to conduct such a study.</p>
</sec>
<sec sec-type="conclusions" id="s6">
<title>6. Conclusion</title>
<p>In this work, we demonstrate how textual content from answers to interview questions related to situational judgement and past behavior can be used to infer personality traits based on the HEXACO model. We extend the Bidirectional Encoder Representations from Transformers (BERT) with a large interview answer corpus of over 3 million answers (over 330 million words) to build InterviewBERT, and use it as the underlying model for personality trait inference from interview responses. The InterviewBERT model is able to better contextualize interview responses based on the interview specific knowledge learnt from the answer corpus in addition to the general language knowledge already encoded in the initial pre-trained BERT. Moreover, we show how &#x0201C;Attention-based&#x0201D; learning approaches in deep neural networks can be used to develop more explainable personality inference models. With regard to gender difference in personality, we show that mean differences in inferred trait scores between male and female groups are similar to those reported by others using standard self-rated item inventories.</p>
<p>Our results show the potential of algorithms to objectively infer a candidate&#x00027;s personality in an explainable manner using only the textual content of interview responses, presenting significant opportunities to remove the subjective biases involved in human interviewer judgement of candidate personality.</p>
<p>For future work, we plan to explore the words and terms discovered through attentions, and analyze how the language usage and the related context are correlated with different personality traits. Since our methodology shows promising results on predicting personality based on individual answers, we are interested in unearthing interview questions that lead to more accurate personality scores from their answers, and explore the effectiveness of different questions. In terms of the underlying algorithms, we are interested in exploring the applicability of other large-scale pre-trained models and different deep regression layers that are jointly fine-tuned with the pre-trained models. Given the inherent inter-correlations among different personality traits, training a multi-task model that could jointly predict different traits is also a direction worth exploring.</p>
</sec>
<sec sec-type="data-availability" id="s7">
<title>Data Availability Statement</title>
<p>The use of data for this study was permissible under the terms of the Sapia Candidate Privacy Policy to which consent is given by all candidates. Due to legal and privacy restrictions, the authors are not able to make the data publicly available. Requests to access these datasets should be directed to <email>buddhi&#x00040;sapia.ai</email>.</p>
</sec>
<sec id="s8">
<title>Ethics Statement</title>
<p>Ethical review and approval are not required for the study of human participants in accordance with local legislation and institutional requirements. The use of data for this study was permissible under the terms of the Sapia Candidate Privacy Policy to which consent is given by all candidates. All candidate data used for this research has been in de-identified form. No potentially identifiable human images or data are presented in this study.</p>
</sec>
<sec id="s9">
<title>Author Contributions</title>
<p>YD, MJ, and BJ contributed to the conception and design of the study. MJ and YD organized the dataset. YD developed the computer code for machine learning model training, performed the statistical analysis, and wrote the first draft of the manuscript. MJ and BJ reviewed the early results of the analysis, provided feedback, and wrote sections of the manuscript. All authors contributed to manuscript revision, read, and approved the submitted version.</p>
</sec>
<sec sec-type="COI-statement" id="conf1">
<title>Conflict of Interest</title>
<p>YD, MJ, and BJ were employees of Sapia&#x00026;Co Pty Ltd. The authors declare that this study received funding from Sapia&#x00026;Co Pty Ltd. The funder had the following involvement with the study: funded the research and provided the data set on which the research is based.</p>
</sec>
<sec sec-type="disclaimer" id="s10">
<title>Publisher&#x00027;s Note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
</body>
<back>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Anglim</surname> <given-names>J.</given-names></name> <name><surname>Lievens</surname> <given-names>F.</given-names></name> <name><surname>Everton</surname> <given-names>L.</given-names></name> <name><surname>Grant</surname> <given-names>S. L.</given-names></name> <name><surname>Marty</surname> <given-names>A.</given-names></name></person-group> (<year>2018</year>). <article-title>HEXACO personality predicts counterproductive work behavior and organizational citizenship behavior in low-stakes and job applicant contexts</article-title>. <source>J. Res. Pers</source>. <volume>77</volume>, <fpage>11</fpage>&#x02013;<lpage>20</lpage>. <pub-id pub-id-type="doi">10.1016/j.jrp.2018.09.003</pub-id></citation>
</ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ariyabuddhiphongs</surname> <given-names>V.</given-names></name> <name><surname>Marican</surname> <given-names>S.</given-names></name></person-group> (<year>2015</year>). <article-title>Big five personality traits and turnover intention among Thai hotel employees</article-title>. <source>Int. J. Hosp. Tour. Administr</source>. <volume>16</volume>, <fpage>355</fpage>&#x02013;<lpage>374</lpage>. <pub-id pub-id-type="doi">10.1080/15256480.2015.1090257</pub-id></citation>
</ref>
<ref id="B3">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Arnoux</surname> <given-names>P.-H.</given-names></name> <name><surname>Xu</surname> <given-names>A.</given-names></name> <name><surname>Boyette</surname> <given-names>N.</given-names></name> <name><surname>Mahmud</surname> <given-names>J.</given-names></name> <name><surname>Akkiraju</surname> <given-names>R.</given-names></name> <name><surname>Sinha</surname> <given-names>V.</given-names></name></person-group> (<year>2017</year>). <article-title>&#x0201C;25 tweets to know you: a new model to predict personality with social media,&#x0201D;</article-title> in <source>Proceedings of the International AAAI Conference on Web and Social Media, Vol. 11</source> (<publisher-loc>Montr&#x000E9;al, QC</publisher-loc>).</citation>
</ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ashton</surname> <given-names>M. C.</given-names></name> <name><surname>Lee</surname> <given-names>K.</given-names></name></person-group> (<year>2007</year>). <article-title>Empirical, theoretical, and practical advantages of the HEXACO model of personality structure</article-title>. <source>Pers. Soc. Psychol. Rev</source>. <volume>11</volume>, <fpage>150</fpage>&#x02013;<lpage>166</lpage>. <pub-id pub-id-type="doi">10.1177/1088868306294907</pub-id><pub-id pub-id-type="pmid">18453460</pub-id></citation></ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ashton</surname> <given-names>M. C.</given-names></name> <name><surname>Lee</surname> <given-names>K.</given-names></name></person-group> (<year>2009</year>). <article-title>The Hexaco-60: a short measure of the major dimensions of personality</article-title>. <source>J. Pers. Assess</source>. <volume>91</volume>, <fpage>340</fpage>&#x02013;<lpage>345</lpage>. <pub-id pub-id-type="doi">10.1080/00223890902935878</pub-id><pub-id pub-id-type="pmid">20017063</pub-id></citation></ref>
<ref id="B6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Barrick</surname> <given-names>M. R.</given-names></name> <name><surname>Mount</surname> <given-names>M. K.</given-names></name></person-group> (<year>1991</year>). <article-title>The big five personality dimensions and job performance: a meta-analysis</article-title>. <source>Pers. Psychol</source>. <volume>44</volume>, <fpage>1</fpage>&#x02013;<lpage>26</lpage>. <pub-id pub-id-type="doi">10.1111/j.1744-6570.1991.tb00688.x</pub-id></citation>
</ref>
<ref id="B7">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Beltagy</surname> <given-names>I.</given-names></name> <name><surname>Lo</surname> <given-names>K.</given-names></name> <name><surname>Cohan</surname> <given-names>A.</given-names></name></person-group> (<year>2019</year>). <article-title>&#x0201C;Scibert: a pretrained language model for scientific text,&#x0201D;</article-title> in <source>Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP/IJCNLP 2019)</source> (<publisher-loc>Hong Kong</publisher-loc>: <publisher-name>Association for Computational Linguistics</publisher-name>), <fpage>3606</fpage>&#x02013;<lpage>3611</lpage>. <pub-id pub-id-type="doi">10.18653/v1/D19-1371</pub-id><pub-id pub-id-type="pmid">35062081</pub-id></citation></ref>
<ref id="B8">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Blei</surname> <given-names>D. M.</given-names></name> <name><surname>Ng</surname> <given-names>A. Y.</given-names></name> <name><surname>Jordan</surname> <given-names>M. I.</given-names></name></person-group> (<year>2003</year>). <article-title>Latent dirichlet allocation</article-title>. <source>J. Mach. Learn. Res</source>. <volume>3</volume>, <fpage>993</fpage>&#x02013;<lpage>1022</lpage>. <pub-id pub-id-type="doi">10.5555/944919.944937</pub-id></citation>
</ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Boyd</surname> <given-names>R. L.</given-names></name> <name><surname>Pennebaker</surname> <given-names>J. W.</given-names></name></person-group> (<year>2017</year>). <article-title>Language-based personality: a new approach to personality in a digital world</article-title>. <source>Curr. Opin. Behav. Sci</source>. <volume>18</volume>, <fpage>63</fpage>&#x02013;<lpage>68</lpage>. <pub-id pub-id-type="doi">10.1016/j.cobeha.2017.07.017</pub-id></citation>
</ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Breiman</surname> <given-names>L.</given-names></name></person-group> (<year>2001</year>). <article-title>Random forests</article-title>. <source>Mach. Learn</source>. <volume>45</volume>, <fpage>5</fpage>&#x02013;<lpage>32</lpage>. <pub-id pub-id-type="doi">10.1023/A:1010933404324</pub-id></citation>
</ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Caldwell</surname> <given-names>D. F.</given-names></name> <name><surname>Burger</surname> <given-names>J. M.</given-names></name></person-group> (<year>1998</year>). <article-title>Personality characteristics of job applicants and success in screening interviews</article-title>. <source>Pers. Psychol</source>. <volume>51</volume>, <fpage>119</fpage>&#x02013;<lpage>136</lpage>. <pub-id pub-id-type="doi">10.1111/j.1744-6570.1998.tb00718.x</pub-id><pub-id pub-id-type="pmid">35511822</pub-id></citation></ref>
<ref id="B12">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Christian</surname> <given-names>H.</given-names></name> <name><surname>Suhartono</surname> <given-names>D.</given-names></name> <name><surname>Chowanda</surname> <given-names>A.</given-names></name> <name><surname>Zamli</surname> <given-names>K. Z.</given-names></name></person-group> (<year>2021</year>). <article-title>Text based personality prediction from multiple social media data sources using pre-trained language model and model averaging</article-title>. <source>J. Big Data</source> <volume>8</volume>, <fpage>1</fpage>&#x02013;<lpage>20</lpage>. <pub-id pub-id-type="doi">10.1186/s40537-021-00459-1</pub-id></citation>
</ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Christopher</surname> <given-names>D. M.</given-names></name> <name><surname>Prabhakar</surname> <given-names>R.</given-names></name> <name><surname>Hinrich</surname> <given-names>S.</given-names></name></person-group> (<year>2008</year>). <source>Introduction to Information Retrieval</source>.</citation>
</ref>
<ref id="B14">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Costa</surname> <given-names>P. T.</given-names> <suffix>Jr</suffix></name>  <name><surname>McCrae</surname> <given-names>R. R</given-names></name></person-group>. (<year>2008</year>). <article-title>&#x0201C;The revised NEO personality inventory (NEO-PI-R),&#x0201D;</article-title> in <source>The SAGE Handbook of Personality Theory and Assessment, Vol. 2: Personality Measurement and Testing</source>, (Thousand Oaks, CA: Sage Publications, Inc.), <fpage>179</fpage>&#x02013;<lpage>198</lpage>. <pub-id pub-id-type="doi">10.4135/9781849200479.n9</pub-id></citation>
</ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>de Vries</surname> <given-names>R. E.</given-names></name> <name><surname>van Gelder</surname> <given-names>J.-L.</given-names></name></person-group> (<year>2015</year>). <article-title>Explaining workplace delinquency: the role of honesty-humility, ethical culture, and employee surveillance</article-title>. <source>Pers. Individ. Differ</source>. <volume>86</volume>, <fpage>112</fpage>&#x02013;<lpage>116</lpage>. <pub-id pub-id-type="doi">10.1016/j.paid.2015.06.008</pub-id></citation>
</ref>
<ref id="B16">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Devlin</surname> <given-names>J.</given-names></name> <name><surname>Chang</surname> <given-names>M.-W.</given-names></name> <name><surname>Lee</surname> <given-names>K.</given-names></name> <name><surname>Toutanova</surname> <given-names>K.</given-names></name></person-group> (<year>2019</year>). <article-title>&#x0201C;Bert: pre-training of deep bidirectional transformers for language understanding,&#x0201D;</article-title> in <source>Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL/HLT 2019)</source> (<publisher-loc>Minneapolis, MN</publisher-loc>: <publisher-name>Association for Computational Linguistics</publisher-name>), <fpage>4171</fpage>&#x02013;<lpage>4186</lpage>.</citation>
</ref>
<ref id="B17">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Diener</surname> <given-names>E.</given-names></name> <name><surname>Lucas</surname> <given-names>R. E.</given-names></name></person-group> (<year>2019</year>). <article-title>&#x0201C;Personality traits,&#x0201D;</article-title> in <source>General Psychology: Required Reading</source>, eds J. A. Cummings and L. Sanders (<publisher-loc>Saskatoon, SK</publisher-loc>: <publisher-name>University of Saskatchewan Open Press</publisher-name>), <fpage>278</fpage>.</citation>
</ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fast</surname> <given-names>L. A.</given-names></name> <name><surname>Funder</surname> <given-names>D. C.</given-names></name></person-group> (<year>2008</year>). <article-title>Personality as manifest in word use: correlations with self-report, acquaintance report, and behavior</article-title>. <source>J. Pers. Soc. Psychol</source>. <volume>94</volume>, <fpage>334</fpage>&#x02013;<lpage>46</lpage>. <pub-id pub-id-type="doi">10.1037/0022-3514.94.2.334</pub-id><pub-id pub-id-type="pmid">18211181</pub-id></citation></ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Feingold</surname> <given-names>A.</given-names></name></person-group> (<year>1994</year>). <article-title>Gender differences in personality: a meta-analysis</article-title>. <source>Psychol. Bull</source>. <volume>116</volume>:<fpage>429</fpage>. <pub-id pub-id-type="doi">10.1037/0033-2909.116.3.429</pub-id><pub-id pub-id-type="pmid">7809307</pub-id></citation></ref>
<ref id="B20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ferreira</surname> <given-names>M. B.</given-names></name> <name><surname>Garcia-Marques</surname> <given-names>L.</given-names></name> <name><surname>Hamilton</surname> <given-names>D.</given-names></name> <name><surname>Ramos</surname> <given-names>T.</given-names></name> <name><surname>Uleman</surname> <given-names>J. S.</given-names></name> <name><surname>Jer&#x000F3;nimo</surname> <given-names>R.</given-names></name></person-group> (<year>2012</year>). <article-title>On the relation between spontaneous trait inferences and intentional inferences: an inference monitoring hypothesis</article-title>. <source>J. Exp. Soc. Psychol</source>. <volume>48</volume>, <fpage>1</fpage>&#x02013;<lpage>12</lpage>. <pub-id pub-id-type="doi">10.1016/j.jesp.2011.06.013</pub-id></citation>
</ref>
<ref id="B21">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Gill</surname> <given-names>A. J.</given-names></name> <name><surname>Nowson</surname> <given-names>S.</given-names></name> <name><surname>Oberlander</surname> <given-names>J.</given-names></name></person-group> (<year>2009</year>). <article-title>&#x0201C;What are they blogging about? Personality, topic and motivation in blogs,&#x0201D;</article-title> in <source>Third International AAAI Conference on Weblogs and Social Media</source> (<publisher-loc>San Jose, CA</publisher-loc>).</citation>
</ref>
<ref id="B22">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Gjurkovi&#x00107;</surname> <given-names>M.</given-names></name> <name><surname>&#x00160;najder</surname> <given-names>J.</given-names></name></person-group> (<year>2018</year>). <article-title>&#x0201C;Reddit: a gold mine for personality prediction,&#x0201D;</article-title> in <source>Proceedings of the Second Workshop on Computational Modeling of People&#x00027;s Opinions, Personality, and Emotions in Social Media</source> (<publisher-loc>New Orleans, LA</publisher-loc>), <fpage>87</fpage>&#x02013;<lpage>97</lpage>. <pub-id pub-id-type="doi">10.18653/v1/W18-1112</pub-id></citation>
</ref>
<ref id="B23">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Gleitman</surname> <given-names>L.</given-names></name> <name><surname>Papafragou</surname> <given-names>A.</given-names></name></person-group> (<year>2012</year>). <article-title>&#x0201C;New perspectives on language and thought,&#x0201D;</article-title> in <source>The Oxford Handbook of Thinking and Reasoning</source>, 2nd Edn, eds K. J. Holyoak and R. G. Morrison (<publisher-loc>New York, NY</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>). <pub-id pub-id-type="doi">10.1093/oxfordhb/9780199734689.013.0028</pub-id></citation>
</ref>
<ref id="B24">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Golbeck</surname> <given-names>J.</given-names></name> <name><surname>Robles</surname> <given-names>C.</given-names></name> <name><surname>Edmondson</surname> <given-names>M.</given-names></name> <name><surname>Turner</surname> <given-names>K.</given-names></name></person-group> (<year>2011</year>). <article-title>&#x0201C;Predicting personality from twitter,&#x0201D;</article-title> in <source>2011 IEEE Third International Conference on Privacy, Security, Risk and Trust and 2011 IEEE Third International Conference on Social Computing</source> (<publisher-loc>Boston, MA</publisher-loc>), <fpage>149</fpage>&#x02013;<lpage>156</lpage>. <pub-id pub-id-type="doi">10.1109/PASSAT/SocialCom.2011.33</pub-id></citation>
</ref>
<ref id="B25">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Goldberg</surname> <given-names>L. R.</given-names></name></person-group> (<year>1993</year>). <article-title>The structure of phenotypic personality traits</article-title>. <source>Am. Psychol</source>. <volume>48</volume>, <fpage>26</fpage>&#x02013;<lpage>34</lpage>. <pub-id pub-id-type="doi">10.1037/0003-066X.48.1.26</pub-id></citation>
</ref>
<ref id="B26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hagendorff</surname> <given-names>T.</given-names></name></person-group> (<year>2020</year>). <article-title>The ethics of AI ethics: an evaluation of guidelines</article-title>. <source>Minds Mach</source>. <volume>30</volume>, <fpage>99</fpage>&#x02013;<lpage>120</lpage>. <pub-id pub-id-type="doi">10.1007/s11023-020-09517-8</pub-id></citation>
</ref>
<ref id="B27">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hall</surname> <given-names>A. N.</given-names></name> <name><surname>Matz</surname> <given-names>S. C.</given-names></name></person-group> (<year>2020</year>). <article-title>Targeting item-level nuances leads to small but robust improvements in personality prediction from digital footprints</article-title>. <source>Eur. J. Pers</source>. <volume>34</volume>, <fpage>873</fpage>&#x02013;<lpage>884</lpage>. <pub-id pub-id-type="doi">10.1002/per.2253</pub-id></citation>
</ref>
<ref id="B28">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ham</surname> <given-names>J.</given-names></name> <name><surname>Vonk</surname> <given-names>R.</given-names></name></person-group> (<year>2003</year>). <article-title>Smart and easy: co-occurring activation of spontaneous trait inferences and spontaneous situational inferences</article-title>. <source>J. Exp. Soc. Psychol</source>. <volume>39</volume>, <fpage>434</fpage>&#x02013;<lpage>447</lpage>. <pub-id pub-id-type="doi">10.1016/S0022-1031(03)00033-7</pub-id></citation>
</ref>
<ref id="B29">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hirsh</surname> <given-names>J. B.</given-names></name> <name><surname>Peterson</surname> <given-names>J. B.</given-names></name></person-group> (<year>2009</year>). <article-title>Personality and language use in self-narratives</article-title>. <source>J. Res. Pers</source>. <volume>43</volume>, <fpage>524</fpage>&#x02013;<lpage>527</lpage>. <pub-id pub-id-type="doi">10.1016/j.jrp.2009.01.006</pub-id></citation>
</ref>
<ref id="B30">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Huang</surname> <given-names>K.</given-names></name> <name><surname>Altosaar</surname> <given-names>J.</given-names></name> <name><surname>Ranganath</surname> <given-names>R.</given-names></name></person-group> (<year>2019</year>). <article-title>Clinicalbert: modeling clinical notes and predicting hospital readmission</article-title>. <source>arXiv preprint arXiv:1904.05342</source>. <pub-id pub-id-type="doi">10.48550/arXiv.1904.05342</pub-id></citation>
</ref>
<ref id="B31">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Iacobelli</surname> <given-names>F.</given-names></name> <name><surname>Gill</surname> <given-names>A. J.</given-names></name> <name><surname>Nowson</surname> <given-names>S.</given-names></name> <name><surname>Oberlander</surname> <given-names>J.</given-names></name></person-group> (<year>2011</year>). <article-title>&#x0201C;Large scale personality classification of bloggers,&#x0201D;</article-title> in <source>Affective Computing and Intelligent Interaction, Lecture Notes in Computer Science</source>, eds S. D&#x00027;Mello, A. Graesser, B. Schuller, and J. C. Martin (<publisher-loc>Memphis, TN</publisher-loc>: <publisher-name>Springer</publisher-name>), <fpage>568</fpage>&#x02013;<lpage>577</lpage>. <pub-id pub-id-type="doi">10.1007/978-3-642-24571-8_71</pub-id></citation>
</ref>
<ref id="B32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jayaratne</surname> <given-names>M.</given-names></name> <name><surname>Jayatilleke</surname> <given-names>B.</given-names></name></person-group> (<year>2020</year>). <article-title>Predicting personality using answers to open-ended interview questions</article-title>. <source>IEEE Access</source> <volume>8</volume>, <fpage>115345</fpage>&#x02013;<lpage>115355</lpage>. <pub-id pub-id-type="doi">10.1109/ACCESS.2020.3004002</pub-id></citation>
</ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>John</surname> <given-names>O. P.</given-names></name> <name><surname>Angleitner</surname> <given-names>A.</given-names></name> <name><surname>Ostendorf</surname> <given-names>F.</given-names></name></person-group> (<year>1988</year>). <article-title>The lexical approach to personality: a historical review of trait taxonomic research</article-title>. <source>Eur. J. Pers</source>. <volume>2</volume>, <fpage>171</fpage>&#x02013;<lpage>203</lpage>. <pub-id pub-id-type="doi">10.1002/per.2410020302</pub-id></citation>
</ref>
<ref id="B34">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Johnson</surname> <given-names>M. K.</given-names></name> <name><surname>Rowatt</surname> <given-names>W. C.</given-names></name> <name><surname>Petrini</surname> <given-names>L.</given-names></name></person-group> (<year>2011</year>). <article-title>A new trait on the market: Honesty-humility as a unique predictor of job performance ratings</article-title>. <source>Pers. Individ. Differ</source>. <volume>50</volume>, <fpage>857</fpage>&#x02013;<lpage>862</lpage>. <pub-id pub-id-type="doi">10.1016/j.paid.2011.01.011</pub-id></citation>
</ref>
<ref id="B35">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Kamijo</surname> <given-names>K.</given-names></name> <name><surname>Nasukawa</surname> <given-names>T.</given-names></name> <name><surname>Kitamura</surname> <given-names>H.</given-names></name></person-group> (<year>2016</year>). <article-title>&#x0201C;Personality estimation from Japanese text,&#x0201D;</article-title> in <source>Proceedings of the Workshop on Computational Modeling of People&#x00027;s Opinions, Personality, and Emotions in Social Media (PEOPLES)</source> (<publisher-loc>Osaka</publisher-loc>), <fpage>101</fpage>&#x02013;<lpage>109</lpage>.</citation>
</ref>
<ref id="B36">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kingma</surname> <given-names>D. P.</given-names></name> <name><surname>Ba</surname> <given-names>J.</given-names></name></person-group> (<year>2014</year>). <article-title>Adam: a method for stochastic optimization</article-title>. <source>arXiv preprint arXiv:1412.6980</source>. <pub-id pub-id-type="doi">10.48550/arXiv.1412.698</pub-id></citation>
</ref>
<ref id="B37">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>J.</given-names></name> <name><surname>Yoon</surname> <given-names>W.</given-names></name> <name><surname>Kim</surname> <given-names>S.</given-names></name> <name><surname>Kim</surname> <given-names>D.</given-names></name> <name><surname>Kim</surname> <given-names>S.</given-names></name> <name><surname>So</surname> <given-names>C. H.</given-names></name> <etal/></person-group>. (<year>2020</year>). <article-title>Biobert: a pre-trained biomedical language representation model for biomedical text mining</article-title>. <source>Bioinformatics</source> <volume>36</volume>, <fpage>1234</fpage>&#x02013;<lpage>1240</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btz682</pub-id><pub-id pub-id-type="pmid">31501885</pub-id></citation></ref>
<ref id="B38">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>K.</given-names></name> <name><surname>Ashton</surname> <given-names>M. C.</given-names></name></person-group> (<year>2018</year>). <article-title>Psychometric properties of the HEXACO-100</article-title>. <source>Assessment</source> <volume>25</volume>, <fpage>543</fpage>&#x02013;<lpage>556</lpage>. <pub-id pub-id-type="doi">10.1177/1073191116659134</pub-id><pub-id pub-id-type="pmid">27411678</pub-id></citation></ref>
<ref id="B39">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>K.</given-names></name> <name><surname>Ashton</surname> <given-names>M. C.</given-names></name></person-group> (<year>2020</year>). <article-title>Sex differences in Hexaco personality characteristics across countries and ethnicities</article-title>. <source>J. Pers</source>. <volume>88</volume>, <fpage>1075</fpage>&#x02013;<lpage>1090</lpage>. <pub-id pub-id-type="doi">10.1111/jopy.12551</pub-id><pub-id pub-id-type="pmid">32394462</pub-id></citation></ref>
<ref id="B40">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>K.</given-names></name> <name><surname>Ashton</surname> <given-names>M. C.</given-names></name> <name><surname>Morrison</surname> <given-names>D. L.</given-names></name> <name><surname>Cordery</surname> <given-names>J.</given-names></name> <name><surname>Dunlop</surname> <given-names>P. D.</given-names></name></person-group> (<year>2008</year>). <article-title>Predicting integrity with the HEXACO personality model: Use of self- and observer reports</article-title>. <source>J. Occup. Organ. Psychol</source>. <volume>81</volume>, <fpage>147</fpage>&#x02013;<lpage>167</lpage>. <pub-id pub-id-type="doi">10.1348/096317907X195175</pub-id></citation>
</ref>
<ref id="B41">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>K.</given-names></name> <name><surname>Ashton</surname> <given-names>M. C.</given-names></name> <name><surname>Vries</surname> <given-names>R. E. d.</given-names></name></person-group> (<year>2005</year>). <article-title>Predicting workplace delinquency and integrity with the HEXACO and five-factor models of personality structure</article-title>. <source>Hum. Perform</source>. <volume>18</volume>, <fpage>179</fpage>&#x02013;<lpage>197</lpage>. <pub-id pub-id-type="doi">10.1207/s15327043hup1802_4</pub-id></citation>
</ref>
<ref id="B42">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Levashina</surname> <given-names>J.</given-names></name> <name><surname>Hartwell</surname> <given-names>C. J.</given-names></name> <name><surname>Morgeson</surname> <given-names>F.</given-names></name> <name><surname>Campion</surname> <given-names>M.</given-names></name></person-group> (<year>2013</year>). <article-title>The structured employment interview: narrative and quantitative review of the recent literature</article-title>. <source>Pers. Psychol</source>. <volume>67</volume>:<fpage>241</fpage>. <pub-id pub-id-type="doi">10.1111/peps.12052</pub-id></citation>
</ref>
<ref id="B43">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>F.</given-names></name> <name><surname>Perez</surname> <given-names>J.</given-names></name> <name><surname>Nowson</surname> <given-names>S.</given-names></name></person-group> (<year>2016</year>). <article-title>&#x0201C;A recurrent and compositional model for personality trait recognition from short texts,&#x0201D;</article-title> in <source>Proceedings of the Workshop on Computational Modeling of People&#x00027;s Opinions, Personality, and Emotions in Social Media (PEOPLES)</source> (<publisher-loc>Osaka</publisher-loc>), <fpage>20</fpage>&#x02013;<lpage>29</lpage>.</citation>
</ref>
<ref id="B44">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lounsbury</surname> <given-names>J. W.</given-names></name> <name><surname>Foster</surname> <given-names>N.</given-names></name> <name><surname>Patel</surname> <given-names>H.</given-names></name> <name><surname>Carmody</surname> <given-names>P.</given-names></name> <name><surname>Gibson</surname> <given-names>L. W.</given-names></name> <name><surname>Stairs</surname> <given-names>D. R.</given-names></name></person-group> (<year>2012</year>). <article-title>An investigation of the personality traits of scientists versus nonscientists and their relationship with career satisfaction</article-title>. <source>R&#x00026;D Manage</source>. <volume>42</volume>, <fpage>47</fpage>&#x02013;<lpage>59</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-9310.2011.00665.x</pub-id></citation>
</ref>
<ref id="B45">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lounsbury</surname> <given-names>J. W.</given-names></name> <name><surname>Steel</surname> <given-names>R. P.</given-names></name> <name><surname>Gibson</surname> <given-names>L. W.</given-names></name> <name><surname>Drost</surname> <given-names>A. W.</given-names></name></person-group> (<year>2008</year>). <article-title>Personality traits and career satisfaction of human resource professionals</article-title>. <source>Hum. Resour. Dev. Int</source>. <volume>11</volume>, <fpage>351</fpage>&#x02013;<lpage>366</lpage>. <pub-id pub-id-type="doi">10.1080/13678860802261215</pub-id></citation>
</ref>
<ref id="B46">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Lucky</surname> <given-names>H.</given-names></name> <name><surname>Suhartono</surname> <given-names>D.</given-names></name></person-group>. (<year>2021</year>). <article-title>&#x0201C;Towards classification of personality prediction model: a combination of BERT word embedding and mlsmote,&#x0201D;</article-title> in <source>2021 1st International Conference on Computer Science and Artificial Intelligence (ICCSAI), Vol. 1</source> (<publisher-loc>Jakarta</publisher-loc>), <fpage>346</fpage>&#x02013;<lpage>350</lpage>. <pub-id pub-id-type="doi">10.1109/ICCSAI53272.2021.9609750</pub-id><pub-id pub-id-type="pmid">27295638</pub-id></citation></ref>
<ref id="B47">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ma</surname> <given-names>N.</given-names></name> <name><surname>Vandekerckhove</surname> <given-names>M.</given-names></name> <name><surname>Van Overwalle</surname> <given-names>F.</given-names></name> <name><surname>Seurinck</surname> <given-names>R.</given-names></name> <name><surname>Fias</surname> <given-names>W.</given-names></name></person-group> (<year>2011</year>). <article-title>Spontaneous and intentional trait inferences recruit a common mentalizing network to a different degree: spontaneous inferences activate only its core areas</article-title>. <source>Soc. Neurosci</source>. <volume>6</volume>, <fpage>123</fpage>&#x02013;<lpage>138</lpage>. <pub-id pub-id-type="doi">10.1080/17470919.2010.485884</pub-id><pub-id pub-id-type="pmid">20661837</pub-id></citation></ref>
<ref id="B48">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Macan</surname> <given-names>T.</given-names></name></person-group> (<year>2009</year>). <article-title>The employment interview: a review of current studies and directions for future research</article-title>. <source>Hum. Resour. Manage. Rev</source>. <volume>19</volume>, <fpage>203</fpage>&#x02013;<lpage>218</lpage>. <pub-id pub-id-type="doi">10.1016/j.hrmr.2009.03.006</pub-id></citation>
</ref>
<ref id="B49">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Majumder</surname> <given-names>N.</given-names></name> <name><surname>Poria</surname> <given-names>S.</given-names></name> <name><surname>Gelbukh</surname> <given-names>A.</given-names></name> <name><surname>Cambria</surname> <given-names>E.</given-names></name></person-group> (<year>2017</year>). <article-title>Deep learning-based document modeling for personality detection from text</article-title>. <source>IEEE Intell. Syst</source>. <volume>32</volume>, <fpage>74</fpage>&#x02013;<lpage>79</lpage>. <pub-id pub-id-type="doi">10.1109/MIS.2017.23</pub-id></citation>
</ref>
<ref id="B50">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>McCabe</surname> <given-names>K. O.</given-names></name> <name><surname>Fleeson</surname> <given-names>W.</given-names></name></person-group> (<year>2016</year>). <article-title>Are traits useful? Explaining trait manifestations as tools in the pursuit of goals</article-title>. <source>J. Pers. Soc. Psychol</source>. <volume>110</volume>:<fpage>287</fpage>. <pub-id pub-id-type="doi">10.1037/a0039490</pub-id><pub-id pub-id-type="pmid">26280839</pub-id></citation></ref>
<ref id="B51">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mcdaniel</surname> <given-names>M.</given-names></name> <name><surname>Whetzel</surname> <given-names>D.</given-names></name> <name><surname>Schmidt</surname> <given-names>F.</given-names></name> <name><surname>Maurer</surname> <given-names>S.</given-names></name></person-group> (<year>1994</year>). <article-title>The validity of employment interviews: a comprehensive review and meta-analysis</article-title>. <source>J. Appl. Psychol</source>. <volume>79</volume>, <fpage>599</fpage>&#x02013;<lpage>616</lpage>. <pub-id pub-id-type="doi">10.1037/0021-9010.79.4.599</pub-id></citation>
</ref>
<ref id="B52">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Meyer</surname> <given-names>G. J.</given-names></name> <name><surname>Finn</surname> <given-names>S. E.</given-names></name> <name><surname>Eyde</surname> <given-names>L. D.</given-names></name> <name><surname>Kay</surname> <given-names>G. G.</given-names></name> <name><surname>Moreland</surname> <given-names>K. L.</given-names></name> <name><surname>Dies</surname> <given-names>R. R.</given-names></name> <etal/></person-group>. (<year>2001</year>). <article-title>Psychological testing and psychological assessment. A review of evidence and issues</article-title>. <source>Am. Psychol</source>. <volume>56</volume>, <fpage>128</fpage>&#x02013;<lpage>165</lpage>. <pub-id pub-id-type="doi">10.1037/0003-066X.56.2.128</pub-id><pub-id pub-id-type="pmid">11279806</pub-id></citation></ref>
<ref id="B53">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Mikolov</surname> <given-names>T.</given-names></name> <name><surname>Sutskever</surname> <given-names>I.</given-names></name> <name><surname>Chen</surname> <given-names>K.</given-names></name> <name><surname>Corrado</surname> <given-names>G. S.</given-names></name> <name><surname>Dean</surname> <given-names>J.</given-names></name></person-group> (<year>2013</year>). <article-title>&#x0201C;Distributed representations of words and phrases and their compositionality,&#x0201D;</article-title> in <source>Proceedings of the 2013 Advances in Neural Information Processing Systems (NIPS 2013)</source> (<publisher-loc>Lake Tahoe, Nevada</publisher-loc>: <publisher-name>Association for Computing Machinery</publisher-name>), <fpage>3111</fpage>&#x02013;<lpage>3119</lpage>.<pub-id pub-id-type="pmid">31840584</pub-id></citation></ref>
<ref id="B54">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Morgeson</surname> <given-names>F. P.</given-names></name> <name><surname>Campion</surname> <given-names>M. A.</given-names></name> <name><surname>Dipboye</surname> <given-names>R. L.</given-names></name> <name><surname>Hollenbeck</surname> <given-names>J. R.</given-names></name> <name><surname>Murphy</surname> <given-names>K.</given-names></name> <name><surname>Schmitt</surname> <given-names>N.</given-names></name></person-group> (<year>2007a</year>). <article-title>Are we getting fooled again? Coming to terms with limitations in the use of personality tests for personnel selection</article-title>. <source>Pers. Psychol</source>. <volume>60</volume>, <fpage>1029</fpage>&#x02013;<lpage>1049</lpage>. <pub-id pub-id-type="doi">10.1111/j.1744-6570.2007.00100.x</pub-id></citation>
</ref>
<ref id="B55">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Morgeson</surname> <given-names>F. P.</given-names></name> <name><surname>Campion</surname> <given-names>M. A.</given-names></name> <name><surname>Dipboye</surname> <given-names>R. L.</given-names></name> <name><surname>Hollenbeck</surname> <given-names>J. R.</given-names></name> <name><surname>Murphy</surname> <given-names>K.</given-names></name> <name><surname>Schmitt</surname> <given-names>N.</given-names></name></person-group> (<year>2007b</year>). <article-title>Reconsidering the use of personality tests in personnel selection contexts</article-title>. <source>Pers. Psychol</source>. <volume>60</volume>, <fpage>683</fpage>&#x02013;<lpage>729</lpage>. <pub-id pub-id-type="doi">10.1111/j.1744-6570.2007.00089.x</pub-id></citation>
</ref>
<ref id="B56">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Moshagen</surname> <given-names>M.</given-names></name> <name><surname>Thielmann</surname> <given-names>I.</given-names></name> <name><surname>Hilbig</surname> <given-names>B. E.</given-names></name> <name><surname>Zettler</surname> <given-names>I.</given-names></name></person-group> (<year>2019</year>). <article-title>Meta-analytic investigations of the HEXACO personality inventory(-revised)</article-title>. <source>Zeitsch. Psychol</source>. <volume>227</volume>, <fpage>186</fpage>&#x02013;<lpage>194</lpage>. <pub-id pub-id-type="doi">10.1027/2151-2604/a000377</pub-id></citation>
</ref>
<ref id="B57">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Neuman</surname> <given-names>Y.</given-names></name> <name><surname>Cohen</surname> <given-names>Y.</given-names></name></person-group> (<year>2014</year>). <article-title>A vectorial semantics approach to personality assessment</article-title>. <source>Nat. Sci. Rep</source>. <volume>4</volume>, <fpage>1</fpage>&#x02013;<lpage>6</lpage>. <pub-id pub-id-type="doi">10.1038/srep04761</pub-id><pub-id pub-id-type="pmid">24755833</pub-id></citation></ref>
<ref id="B58">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ones</surname> <given-names>D. S.</given-names></name> <name><surname>Dilchert</surname> <given-names>S.</given-names></name> <name><surname>Viswesvaran</surname> <given-names>C.</given-names></name> <name><surname>Judge</surname> <given-names>T. A.</given-names></name></person-group> (<year>2007</year>). <article-title>In support of personality assessment in organizational settings</article-title>. <source>Pers. Psychol</source>. <volume>60</volume>, <fpage>995</fpage>&#x02013;<lpage>1027</lpage>. <pub-id pub-id-type="doi">10.1111/j.1744-6570.2007.00099.x</pub-id></citation>
</ref>
<ref id="B59">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Ong</surname> <given-names>V.</given-names></name> <name><surname>Rahmanto</surname> <given-names>A. D.</given-names></name> <name><surname>Suhartono</surname> <given-names>D.</given-names></name> <name><surname>Nugroho</surname> <given-names>A. E.</given-names></name> <name><surname>Andangsari</surname> <given-names>E. W.</given-names></name> <name><surname>Suprayogi</surname> <given-names>M. N.</given-names></name> <etal/></person-group>. (<year>2017</year>). <article-title>&#x0201C;Personality prediction based on twitter information in Bahasa Indonesia,&#x0201D;</article-title> in <source>2017 Federated Conference on Computer Science and Information Systems (FedCSIS)</source> (<publisher-loc>Prague</publisher-loc>), <fpage>367</fpage>&#x02013;<lpage>372</lpage>.</citation>
</ref>
<ref id="B60">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Park</surname> <given-names>G.</given-names></name> <name><surname>Schwartz</surname> <given-names>H. A.</given-names></name> <name><surname>Eichstaedt</surname> <given-names>J. C.</given-names></name> <name><surname>Kern</surname> <given-names>M. L.</given-names></name> <name><surname>Kosinski</surname> <given-names>M.</given-names></name> <name><surname>Stillwell</surname> <given-names>D. J.</given-names></name> <etal/></person-group>. (<year>2015</year>). <article-title>Automatic personality assessment through social media language</article-title>. <source>J. Pers. Soc. Psychol</source>. <volume>108</volume>:<fpage>934</fpage>. <pub-id pub-id-type="doi">10.1037/pspp0000020</pub-id><pub-id pub-id-type="pmid">25365036</pub-id></citation></ref>
<ref id="B61">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Pennebaker</surname> <given-names>J. W.</given-names></name> <name><surname>Boyd</surname> <given-names>R.</given-names></name> <name><surname>Jordan</surname> <given-names>K.</given-names></name> <name><surname>Blackburn</surname> <given-names>K.</given-names></name></person-group> (<year>2015</year>). <source>The Development and Psychometric Properties of LIWC2015</source>. <publisher-loc>Austin, TX</publisher-loc>: <publisher-name>University of Texas at Austin</publisher-name>. <pub-id pub-id-type="doi">10.15781/T29G6Z</pub-id><pub-id pub-id-type="pmid">35330723</pub-id></citation></ref>
<ref id="B62">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Pennington</surname> <given-names>J.</given-names></name> <name><surname>Socher</surname> <given-names>R.</given-names></name> <name><surname>Manning</surname> <given-names>C. D.</given-names></name></person-group> (<year>2014</year>). <article-title>&#x0201C;Glove: global vectors for word representation,&#x0201D;</article-title> in <source>Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP 2014)</source> (<publisher-loc>Doha</publisher-loc>: <publisher-name>Association for Computational Linguistics</publisher-name>), <fpage>1532</fpage>&#x02013;<lpage>1543</lpage>. <pub-id pub-id-type="doi">10.3115/v1/D14-1162</pub-id></citation>
</ref>
<ref id="B63">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Peters</surname> <given-names>M.</given-names></name> <name><surname>Neumann</surname> <given-names>M.</given-names></name> <name><surname>Iyyer</surname> <given-names>M.</given-names></name> <name><surname>Gardner</surname> <given-names>M.</given-names></name> <name><surname>Clark</surname> <given-names>C.</given-names></name> <name><surname>Lee</surname> <given-names>K.</given-names></name> <etal/></person-group>. (<year>2018</year>). <article-title>&#x0201C;Deep contextualized word representations,&#x0201D;</article-title> in <source>Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL/HLT 2018)</source> (<publisher-loc>New Orleans, LA</publisher-loc>: <publisher-name>Association for Computational Linguistics</publisher-name>), <fpage>2227</fpage>&#x02013;<lpage>2237</lpage>. <pub-id pub-id-type="doi">10.18653/v1/N18-1202</pub-id><pub-id pub-id-type="pmid">34343877</pub-id></citation></ref>
<ref id="B64">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Pinker</surname> <given-names>S.</given-names></name></person-group> (<year>2007</year>). <source>The Stuff of Thought: Language As a Window Into Human Nature</source>. <publisher-loc>New York, NY</publisher-loc>: <publisher-name>Penguin Group (Viking Press)</publisher-name>.<pub-id pub-id-type="pmid">17896455</pub-id></citation></ref>
<ref id="B65">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Plank</surname> <given-names>B.</given-names></name> <name><surname>Hovy</surname> <given-names>D.</given-names></name></person-group> (<year>2015</year>). <article-title>&#x0201C;Personality traits on twitter-or-how to get 1,500 personality tests in a week,&#x0201D;</article-title> in <source>Proceedings of the 6th Workshop on Computational Approaches to Subjectivity, Sentiment and Social Media Analysis</source> (<publisher-loc>Lisbon</publisher-loc>), <fpage>92</fpage>&#x02013;<lpage>98</lpage>. <pub-id pub-id-type="doi">10.18653/v1/W15-2913</pub-id></citation>
</ref>
<ref id="B66">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pletzer</surname> <given-names>J. L.</given-names></name> <name><surname>Bentvelzen</surname> <given-names>M.</given-names></name> <name><surname>Oostrom</surname> <given-names>J. K.</given-names></name> <name><surname>de Vries</surname> <given-names>R. E.</given-names></name></person-group> (<year>2019</year>). <article-title>A meta-analysis of the relations between personality and workplace deviance: big five versus HEXACO</article-title>. <source>J. Vocat. Behav</source>. <volume>112</volume>, <fpage>369</fpage>&#x02013;<lpage>383</lpage>. <pub-id pub-id-type="doi">10.1016/j.jvb.2019.04.004</pub-id></citation>
</ref>
<ref id="B67">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Pratama</surname> <given-names>B. Y.</given-names></name> <name><surname>Sarno</surname> <given-names>R.</given-names></name></person-group> (<year>2015</year>). <article-title>&#x0201C;Personality classification based on twitter text using naive Bayes, KNN and SVM,&#x0201D;</article-title> in <source>2015 International Conference on Data and Software Engineering (ICoDSE)</source> (<publisher-loc>Yogyakarta</publisher-loc>), <fpage>170</fpage>&#x02013;<lpage>174</lpage>. <pub-id pub-id-type="doi">10.1109/ICODSE.2015.7436992</pub-id></citation>
</ref>
<ref id="B68">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Purkiss</surname> <given-names>S. L. S.</given-names></name> <name><surname>Perrew&#x000E9;</surname> <given-names>P. L.</given-names></name> <name><surname>Gillespie</surname> <given-names>T. L.</given-names></name> <name><surname>Mayes</surname> <given-names>B. T.</given-names></name> <name><surname>Ferris</surname> <given-names>G. R.</given-names></name></person-group> (<year>2006</year>). <article-title>Implicit sources of bias in employment interview judgments and decisions</article-title>. <source>Organ. Behav. Hum. Decis. Process</source>. <volume>101</volume>, <fpage>152</fpage>&#x02013;<lpage>167</lpage>. <pub-id pub-id-type="doi">10.1016/j.obhdp.2006.06.005</pub-id></citation>
</ref>
<ref id="B69">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Qiu</surname> <given-names>L.</given-names></name> <name><surname>Lin</surname> <given-names>H.</given-names></name> <name><surname>Ramsay</surname> <given-names>J.</given-names></name> <name><surname>Yang</surname> <given-names>F.</given-names></name></person-group> (<year>2012</year>). <article-title>You are what you tweet: personality expression and perception on twitter</article-title>. <source>J. Res. Pers</source>. <volume>46</volume>, <fpage>710</fpage>&#x02013;<lpage>718</lpage>. <pub-id pub-id-type="doi">10.1016/j.jrp.2012.08.008</pub-id></citation>
</ref>
<ref id="B70">
<citation citation-type="web"><person-group person-group-type="author"><name><surname>Radford</surname> <given-names>A.</given-names></name> <name><surname>Narasimhan</surname> <given-names>K.</given-names></name> <name><surname>Salimans</surname> <given-names>T.</given-names></name> <name><surname>Sutskever</surname> <given-names>I.</given-names></name></person-group> (<year>2018</year>). <source>Improving Language Understanding by Generative Pre-training</source>. Available online at: <ext-link ext-link-type="uri" xlink:href="https://openai.com/blog/language-unsupervised/">https://openai.com/blog/language-unsupervised/</ext-link></citation>
</ref>
<ref id="B71">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Roberts</surname> <given-names>B. W.</given-names></name> <name><surname>Kuncel</surname> <given-names>N. R.</given-names></name> <name><surname>Shiner</surname> <given-names>R.</given-names></name> <name><surname>Caspi</surname> <given-names>A.</given-names></name> <name><surname>Goldberg</surname> <given-names>L. R.</given-names></name></person-group> (<year>2007</year>). <article-title>The power of personality: The comparative validity of personality traits, socioeconomic status, and cognitive ability for predicting important life outcomes</article-title>. <source>Perspect. Psychol. Sci</source>. <volume>2</volume>, <fpage>313</fpage>&#x02013;<lpage>345</lpage>. <pub-id pub-id-type="doi">10.1111/j.1745-6916.2007.00047.x</pub-id><pub-id pub-id-type="pmid">26151971</pub-id></citation></ref>
<ref id="B72">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rothmann</surname> <given-names>S.</given-names></name> <name><surname>Coetzer</surname> <given-names>E. P.</given-names></name></person-group> (<year>2003</year>). <article-title>The big five personality dimensions and job performance</article-title>. <source>SA J. Indus. Psychol</source>. <volume>29</volume>, <fpage>68</fpage>&#x02013;<lpage>74</lpage>. <pub-id pub-id-type="doi">10.4102/sajip.v29i1.88</pub-id></citation>
</ref>
<ref id="B73">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Salgado</surname> <given-names>J. F.</given-names></name></person-group> (<year>2002</year>). <article-title>The big five personality dimensions and counterproductive behaviors</article-title>. <source>Int. J. Select. Assess</source>. <volume>10</volume>, <fpage>117</fpage>&#x02013;<lpage>125</lpage>. <pub-id pub-id-type="doi">10.1111/1468-2389.00198</pub-id></citation>
</ref>
<ref id="B74">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Saucier</surname> <given-names>G.</given-names></name> <name><surname>Goldberg</surname> <given-names>L. R.</given-names></name></person-group> (<year>1996</year>). <article-title>&#x0201C;The language of personality: lexical perspectives on the five-factor model,&#x0201D;</article-title> in <source>The Five-Factor Model of Personality: Theoretical Perspectives</source>, ed J. S. Wiggins (<publisher-loc>New York, NY</publisher-loc>: <publisher-name>Guilford Press</publisher-name>), <fpage>21</fpage>&#x02013;<lpage>50</lpage>.</citation>
</ref>
<ref id="B75">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schwartz</surname> <given-names>H. A.</given-names></name> <name><surname>Eichstaedt</surname> <given-names>J. C.</given-names></name> <name><surname>Kern</surname> <given-names>M. L.</given-names></name> <name><surname>Dziurzynski</surname> <given-names>L.</given-names></name> <name><surname>Ramones</surname> <given-names>S. M.</given-names></name> <name><surname>Agrawal</surname> <given-names>M.</given-names></name> <etal/></person-group>. (<year>2013</year>). <article-title>Personality, gender, and age in the language of social media: the open-vocabulary approach</article-title>. <source>PLoS ONE</source> <volume>8</volume>:<fpage>e73791</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0073791</pub-id><pub-id pub-id-type="pmid">24086296</pub-id></citation></ref>
<ref id="B76">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Skimina</surname> <given-names>E.</given-names></name> <name><surname>Strus</surname> <given-names>W.</given-names></name> <name><surname>Cieciuch</surname> <given-names>J.</given-names></name> <name><surname>Szarota</surname> <given-names>P.</given-names></name> <name><surname>Izdebski</surname> <given-names>P. K.</given-names></name></person-group> (<year>2020</year>). <article-title>Psychometric properties of the polish versions of the hexaco-60 and the hexaco-100 personality inventories</article-title>. <source>Curr. Issues Pers. Psychol</source>. <volume>8</volume>, <fpage>259</fpage>&#x02013;<lpage>278</lpage>. <pub-id pub-id-type="doi">10.5114/cipp.2020.98693</pub-id></citation>
</ref>
<ref id="B77">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Sumner</surname> <given-names>C.</given-names></name> <name><surname>Byers</surname> <given-names>A.</given-names></name> <name><surname>Boochever</surname> <given-names>R.</given-names></name> <name><surname>Park</surname> <given-names>G. J.</given-names></name></person-group> (<year>2012</year>). <article-title>&#x0201C;Predicting dark triad personality traits from twitter usage and a linguistic analysis of tweets,&#x0201D;</article-title> in <source>2012 11th International Conference on Machine Learning and Applications, Vol. 2</source> (<publisher-loc>Boca Raton, FL</publisher-loc>), <fpage>386</fpage>&#x02013;<lpage>393</lpage>. <pub-id pub-id-type="doi">10.1109/ICMLA.2012.218</pub-id></citation>
</ref>
<ref id="B78">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tadesse</surname> <given-names>M. M.</given-names></name> <name><surname>Lin</surname> <given-names>H.</given-names></name> <name><surname>Xu</surname> <given-names>B.</given-names></name> <name><surname>Yang</surname> <given-names>L.</given-names></name></person-group> (<year>2018</year>). <article-title>Personality predictions based on user behavior on the Facebook social media platform</article-title>. <source>IEEE Access</source> <volume>6</volume>, <fpage>61959</fpage>&#x02013;<lpage>61969</lpage>. <pub-id pub-id-type="doi">10.1109/ACCESS.2018.2876502</pub-id></citation>
</ref>
<ref id="B79">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tandera</surname> <given-names>T.</given-names></name> <name><surname>Suhartono</surname> <given-names>D.</given-names></name> <name><surname>Wongso</surname> <given-names>R.</given-names></name> <name><surname>Prasetio</surname> <given-names>Y. L.</given-names></name> <etal/></person-group>. (<year>2017</year>). <article-title>Personality prediction system from Facebook users</article-title>. <source>Proc. Comput. Sci</source>. <volume>116</volume>, <fpage>604</fpage>&#x02013;<lpage>611</lpage>. <pub-id pub-id-type="doi">10.1016/j.procs.2017.10.016</pub-id></citation>
</ref>
<ref id="B80">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Uleman</surname> <given-names>J. S.</given-names></name></person-group> (<year>1989</year>). <article-title>&#x0201C;Spontaneous trait inference,&#x0201D;</article-title> in <source>Unintended Thought</source>, eds J. S. Uleman and J. A. Bargh (<publisher-loc>New York, NY</publisher-loc>: <publisher-name>Guilford Press</publisher-name>), <fpage>155</fpage>&#x02013;<lpage>188</lpage>.</citation>
</ref>
<ref id="B81">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Uleman</surname> <given-names>J. S.</given-names></name></person-group> (<year>1999</year>). <article-title>&#x0201C;Spontaneous versus intentional inferences in impression formation,&#x0201D;</article-title> in <source>Dual-Process Theories in Social Psychology</source>, eds S. Chaiken and Y. Trope (<publisher-loc>New York, NY</publisher-loc>: <publisher-name>Guilford Press</publisher-name>), <fpage>141</fpage>&#x02013;<lpage>160</lpage>.</citation>
</ref>
<ref id="B82">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Van Dam</surname> <given-names>K.</given-names></name></person-group> (<year>2003</year>). <article-title>Trait perception in the employment interview: a five-factor model perspective. International <italic>J. Select. Assess</italic></article-title>. <volume>11</volume>, <fpage>43</fpage>&#x02013;<lpage>55</lpage>. <pub-id pub-id-type="doi">10.1111/1468-2389.00225</pub-id></citation>
</ref>
<ref id="B83">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Van Duynslaeger</surname> <given-names>M.</given-names></name> <name><surname>Van Overwalle</surname> <given-names>F.</given-names></name> <name><surname>Verstraeten</surname> <given-names>E.</given-names></name></person-group> (<year>2007</year>). <article-title>Electrophysiological time course and brain areas of spontaneous and intentional trait inferences</article-title>. <source>Soc. Cogn. Affect. Neurosci</source>. <volume>2</volume>, <fpage>174</fpage>&#x02013;<lpage>188</lpage>. <pub-id pub-id-type="doi">10.1093/scan/nsm016</pub-id><pub-id pub-id-type="pmid">18985139</pub-id></citation></ref>
<ref id="B84">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Vaswani</surname> <given-names>A.</given-names></name> <name><surname>Shazeer</surname> <given-names>N.</given-names></name> <name><surname>Parmar</surname> <given-names>N.</given-names></name> <name><surname>Uszkoreit</surname> <given-names>J.</given-names></name> <name><surname>Jones</surname> <given-names>L.</given-names></name> <name><surname>Gomez</surname> <given-names>A. N.</given-names></name> <etal/></person-group>. (<year>2017</year>). <article-title>&#x0201C;Attention is all you need,&#x0201D;</article-title> in <source>Proceedings of the 31st Annual Conference on Neural Information Processing Systems (NIPS 2017)</source> (<publisher-loc>Long Beach, CA</publisher-loc>: <publisher-name>Association for Computing Machinery</publisher-name>), <fpage>5998</fpage>&#x02013;<lpage>6008</lpage>.</citation>
</ref>
<ref id="B85">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Verhoeven</surname> <given-names>B.</given-names></name> <name><surname>Daelemans</surname> <given-names>W.</given-names></name> <name><surname>Plank</surname> <given-names>B.</given-names></name></person-group> (<year>2016</year>). <article-title>&#x0201C;Twisty: a multilingual twitter stylometry corpus for gender and personality profiling,&#x0201D;</article-title> in <source>Proceedings of the Tenth International Conference on Language Resources and Evaluation (LREC&#x00027;16)</source> (<publisher-loc>Portoro&#x0017E;</publisher-loc>), <fpage>1632</fpage>&#x02013;<lpage>1637</lpage>.</citation>
</ref>
<ref id="B86">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wakabayashi</surname> <given-names>A.</given-names></name></person-group> (<year>2014</year>). <article-title>A sixth personality domain that is independent of the big five domains: the psychometric properties of the Hexaco personality inventory in a Japanese sample</article-title>. <source>Jpn. Psychol. Res</source>. <volume>56</volume>, <fpage>211</fpage>&#x02013;<lpage>223</lpage>. <pub-id pub-id-type="doi">10.1111/jpr.12045</pub-id></citation>
</ref>
<ref id="B87">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>Z.</given-names></name> <name><surname>Wu</surname> <given-names>C.</given-names></name> <name><surname>Zheng</surname> <given-names>K.</given-names></name> <name><surname>Niu</surname> <given-names>X.</given-names></name> <name><surname>Wang</surname> <given-names>X.</given-names></name></person-group> (<year>2019</year>). <article-title>SMOTETomek-based resampling for personality recognition</article-title>. <source>IEEE Access</source> <volume>7</volume>, <fpage>129678</fpage>&#x02013;<lpage>129689</lpage>. <pub-id pub-id-type="doi">10.1109/ACCESS.2019.2940061</pub-id></citation>
</ref>
<ref id="B88">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xue</surname> <given-names>D.</given-names></name> <name><surname>Hong</surname> <given-names>Z.</given-names></name> <name><surname>Guo</surname> <given-names>S.</given-names></name> <name><surname>Gao</surname> <given-names>L.</given-names></name> <name><surname>Wu</surname> <given-names>L.</given-names></name> <name><surname>Zheng</surname> <given-names>J.</given-names></name> <name><surname>Zhao</surname> <given-names>N.</given-names></name></person-group> (<year>2017</year>). <article-title>Personality recognition on social media with label distribution learning</article-title>. <source>IEEE Access</source> <volume>5</volume>, <fpage>13478</fpage>&#x02013;<lpage>13488</lpage>. <pub-id pub-id-type="doi">10.1109/ACCESS.2017.2719018</pub-id></citation>
</ref>
</ref-list>
<fn-group>
<fn id="fn0001"><p><sup>1</sup><ext-link ext-link-type="uri" xlink:href="https://www.sapia.ai/">https://www.sapia.ai/</ext-link></p></fn>
<fn id="fn0002"><p><sup>2</sup><ext-link ext-link-type="uri" xlink:href="https://scikit-learn.org/stable/modules/generated/sklearn.feature_extraction.text.TfidfVectorizer.html">https://scikit-learn.org/stable/modules/generated/sklearn.feature_extraction.text.TfidfVectorizer.html</ext-link></p></fn>
<fn id="fn0003"><p><sup>3</sup><ext-link ext-link-type="uri" xlink:href="https://radimrehurek.com/gensim/">https://radimrehurek.com/gensim/</ext-link></p></fn>
<fn id="fn0004"><p><sup>4</sup><ext-link ext-link-type="uri" xlink:href="https://nlp.stanford.edu/projects/glove/">https://nlp.stanford.edu/projects/glove/</ext-link></p></fn>
<fn id="fn0005"><p><sup>5</sup><ext-link ext-link-type="uri" xlink:href="https://torchtext.readthedocs.io/en/latest/vocab.html&#x00023;glove">https://torchtext.readthedocs.io/en/latest/vocab.html&#x00023;glove</ext-link></p></fn>
<fn id="fn0006"><p><sup>6</sup><ext-link ext-link-type="uri" xlink:href="https://huggingface.co/docs/transformers/v4.15.0/en/main_classes/tokenizer">https://huggingface.co/docs/transformers/v4.15.0/en/main_classes/tokenizer</ext-link></p></fn>
<fn id="fn0007"><p><sup>7</sup><ext-link ext-link-type="uri" xlink:href="https://huggingface.co/docs/transformers/v4.15.0/en/index">https://huggingface.co/docs/transformers/v4.15.0/en/index</ext-link></p></fn>
<fn id="fn0008"><p><sup>8</sup><ext-link ext-link-type="uri" xlink:href="https://scikit-learn.org/stable/modules/generated/sklearn.ensemble.RandomForestRegressor.html">https://scikit-learn.org/stable/modules/generated/sklearn.ensemble.RandomForestRegressor.html</ext-link></p></fn>
</fn-group>
</back>
</article>