<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2017.00378</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Conducting Online Behavioral Research Using Crowdsourcing Services in Japan</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Majima</surname> <given-names>Yoshimasa</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="author-notes" rid="fn001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/379541/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Nishiyama</surname> <given-names>Kaoru</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/384413/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Nishihara</surname> <given-names>Aki</given-names></name>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
</contrib>
<contrib contrib-type="author">
<name><surname>Hata</surname> <given-names>Ryosuke</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/419286/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Department of Psychology for Well-Being, Hokusei Gakuen University</institution> <country>Sapporo, Japan</country></aff>
<aff id="aff2"><sup>2</sup><institution>Department of Foreign Language Education, Hokusei Gakuen University</institution> <country>Sapporo, Japan</country></aff>
<aff id="aff3"><sup>3</sup><institution>Department of Social Work, Hokusei Gakuen University</institution> <country>Sapporo, Japan</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Lynne D. Roberts, Curtin University, Australia</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Emma Buchtel, Hong Kong Institute of Education, Hong Kong; Ilka H Gleibs, London School of Economics and Political Science, UK</p></fn>
<fn fn-type="corresp" id="fn001"><p>&#x0002A;Correspondence: Yoshimasa Majima <email>majima.y&#x00040;hokusei.ac.jp</email></p></fn>
<fn fn-type="other" id="fn002"><p>This article was submitted to Quantitative Psychology and Measurement, a section of the journal Frontiers in Psychology</p></fn></author-notes>
<pub-date pub-type="epub">
<day>14</day>
<month>03</month>
<year>2017</year>
</pub-date>
<pub-date pub-type="collection">
<year>2017</year>
</pub-date>
<volume>8</volume>
<elocation-id>378</elocation-id>
<history>
<date date-type="received">
<day>12</day>
<month>10</month>
<year>2016</year>
</date>
<date date-type="accepted">
<day>27</day>
<month>02</month>
<year>2017</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2017 Majima, Nishiyama, Nishihara and Hata.</copyright-statement>
<copyright-year>2017</copyright-year>
<copyright-holder>Majima, Nishiyama, Nishihara and Hata</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) or licensor are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>Recent research on human behavior has often collected empirical data from the online labor market, through a process known as crowdsourcing. As well as the United States and the major European countries, there are several crowdsourcing services in Japan. For research purpose, Amazon&#x00027;s Mechanical Turk (MTurk) is the widely used platform among those services. Previous validation studies have shown many commonalities between MTurk workers and participants from traditional samples based on not only personality but also performance on reasoning tasks. The present study aims to extend these findings to non-MTurk (i.e., Japanese) crowdsourcing samples in which workers have different ethnic backgrounds from those of MTurk. We conducted three surveys (<italic>N</italic> &#x0003D; 426, 453, 167, respectively) designed to compare Japanese crowdsourcing workers and university students in terms of their demographics, personality traits, reasoning skills, and attention to instructions. The results generally align with previous studies and suggest that non-MTurk participants are also eligible for behavioral research. Furthermore, small screen devices are found to impair participants&#x00027; attention to instructions. Several recommendations concerning this sample are presented.</p>
</abstract>
<kwd-group>
<kwd>online study</kwd>
<kwd>non-MTurk crowdsourcing</kwd>
<kwd>personality</kwd>
<kwd>reasoning</kwd>
<kwd>instructional manipulation check</kwd>
</kwd-group>
<counts>
<fig-count count="0"/>
<table-count count="6"/>
<equation-count count="0"/>
<ref-count count="39"/>
<page-count count="13"/>
<word-count count="11220"/>
</counts>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>Introduction</title>
<p>Online survey research is becoming increasingly popular in psychology and other social sciences on human behavior. Researchers often collect data from participants in online labor markets, through a process known as <italic>crowdsourcing</italic>. Recruiting participants from a crowdsourcing service is attractive to researchers because of its advantages over using <italic>traditional</italic> samples.</p>
<p>Estelles-Arolas and Gonzalez-Ladron-De-Guevara (<xref ref-type="bibr" rid="B7">2012</xref>) define crowdsourcing as an online activity in which a group of diverse individuals (users) voluntarily undertake a task proposed by an individual or a profit/non-profit organization (crowdsourcer) and in which the users receive monetary and other forms of compensation in exchange for their contributions, while the crowdsourcer benefits from the work performed by the users. In behavioral research, crowdsourcing websites offer researchers a useful platform that provides convenient access to a large set of people who are willing to undertake tasks, including research studies, at a relatively low cost. One of the most well-known crowdsourcing sites is Amazon&#x00027;s Mechanical Turk, which is often abbreviated as MTurk.</p>
<p>MTurk specializes in recruiting users, who are referred to as <italic>workers</italic>, to complete small tasks that are known as <italic>HIT</italic>s (human intelligence tasks). For research purposes, researchers (<italic>requesters</italic>) post a HIT that contain surveys and/or experiments that can be completed on a computer using supplied templates. Sometimes, requesters post a link to external survey tools, such as SurveyMonkey and Qualtrics. Workers will browse or search tasks, and they are paid in exchange for their successful contribution to a task.</p>
<p>Mason and Suri (<xref ref-type="bibr" rid="B20">2012</xref>) identified four advantages of MTurk: access to a large, stable pool of participants; participant diversity; a low cost and built-in payment system; and faster theory and/or experiment cycle. Because of these benefits, MTurk is becoming popular as a potential participant pool for psychology and other social sciences.</p>
<p>Along with increasing usage in behavioral research, the validity of the data obtained from MTurk participants has been examined in several studies (for a recent review, see Paolacci and Chandler, <xref ref-type="bibr" rid="B25">2014</xref>). These investigations have typically compared MTurk data with those from traditional samples, such as university students and other community members. First, demographic surveys have shown that MTurk workers are mostly residents of the United States and India and that they are in about their thirties, which is older than typical students who are in their late teens and twenties (e.g., Paolacci et al., <xref ref-type="bibr" rid="B26">2010</xref>; Behrend et al., <xref ref-type="bibr" rid="B1">2011</xref>; Goodman et al., <xref ref-type="bibr" rid="B11">2013</xref>). In addition, workers and participants from traditional samples differ in terms of their personality traits. For example, MTurk workers are less extraverted and emotionally stable and show lower self-esteem than students. They also value money more than time and exhibit higher materialism than an age-matched community sample (Goodman et al., <xref ref-type="bibr" rid="B11">2013</xref>).</p>
<p>The two samples were also different in their performances of reasoning and attention to instructions. For example, Goodman et al. (<xref ref-type="bibr" rid="B11">2013</xref>) found that MTurk workers show lower cognitive capabilities than students on the Cognitive Reflection Test (Frederick, <xref ref-type="bibr" rid="B9">2005</xref>), which requires effortful system-2 thinking and on a &#x0201C;trap&#x0201D; task that involves what is known as instructional manipulation checks (IMCs), which reflect participants&#x00027; inattentive response to the survey questions. However, Goodman et al. also pointed out that failures in IMC were mainly found in ESL and non-US participants. Therefore, the language proficiency, as well as careful reading of study materials, is essential for the successful solution to IMC. Furthermore, Hauser and Schwarz (<xref ref-type="bibr" rid="B13">2015</xref>) showed that answering IMCs before &#x0201C;tricky&#x0201D; reasoning tasks improves performance of subsequent tasks; the authors explained that the IMC itself alters participants&#x00027; attention to subsequent tasks and prompts participants to adopt a more deliberative thinking strategy, which results in improved performance on these tasks.</p>
<p>The MTurk and traditional samples also have several commonalities. For example, MTurk workers and students show similar performance on classical heuristic-bias judgment tasks, such as the <italic>Linda problem</italic> (Tversky and Kahneman, <xref ref-type="bibr" rid="B36">1983</xref>) and the <italic>Asian disease problem</italic> (Tversky and Kahneman, <xref ref-type="bibr" rid="B35">1981</xref>). Paolacci et al. (<xref ref-type="bibr" rid="B26">2010</xref>) showed that both students and MTurk workers exhibited a significant framing effect, conjunction fallacy, and outcome bias. They also exhibit a significant anchoring-and-adjustment effect; however, the anchoring bias is mainly shown in the community sample, and the MTurk workers do not show anchoring bias, partly because they might &#x0201C;search&#x0201D; the correct answer on the Internet. In addition, MTurk workers perform similarly in traditional experimental psychology tasks, such as the Stroop, Flanker, attentional blink, and categorical learning tasks (Crump et al., <xref ref-type="bibr" rid="B5">2013</xref>).</p>
<p>In sum, although MTurk participants and traditional participants differ in terms of a few features, they share many common properties. Therefore, crowdsourcing is considered to be a fruitful data collection tool for psychology and other social sciences (Goodman et al., <xref ref-type="bibr" rid="B11">2013</xref>; Paolacci and Chandler, <xref ref-type="bibr" rid="B25">2014</xref>).</p>
<p>MTurk appears to provide a promising approach to behavioral studies owing to its advantages over traditional offline data collection. Despite these advantages, there are some limitations of MTurk as a participant pool for empirical studies. First, there are issues with sample diversity. Demographic surveys have repeatedly shown that the majority of MTurk workers are Caucasian residents of the United States, followed by Asian workers who live in India (Paolacci et al., <xref ref-type="bibr" rid="B26">2010</xref>; Behrend et al., <xref ref-type="bibr" rid="B1">2011</xref>; Goodman et al., <xref ref-type="bibr" rid="B11">2013</xref>). Currently, MTurk requires their workers to provide valid US taxpayer identification information when they get paid (either in US dollars or Indian Rupees), otherwise they can only transfer their earnings to Amazon&#x00027;s gift card. This restriction on monetary compensation may substantially reduce the number of non-US workers. Because of its biased population, it is difficult for researchers in other countries to collect data from residents of their own cultures. Of course, there are other crowdsourcing services, such as Prolific Academic and CrowdFlower, although it seems that Caucasian residents of the USA, the UK, and other European countries are also the predominant participants of these pools. Therefore, researchers who wish to collect data from samples of other ethnicities or nationalities should utilize other crowdsourcing services. This is exactly the case with Japanese researchers.</p>
<p>The second issue is of a technical nature. At this time, a US bank account is required to be a <italic>requester</italic> in MTurk. This requirement also constitutes an obstacle to adopt MTurk as a participant pool for researchers outside the Unites States<xref ref-type="fn" rid="fn0001"><sup>1</sup></xref>. On these grounds, MTurk is considered to provide limited access to participant pools for behavioral researchers around the world.</p>
<p>Several studies also pointed out potential pitfalls of online studies with MTurk. First, Zhou and Fishbach (<xref ref-type="bibr" rid="B39">2016</xref>) claimed that researchers should pay attention to attrition rate that poses a threat to internal validity of the study. They also recommended that researchers not only implement dropout-reduction strategies, but also explore causes of, increase the visibility of, and report participant attrition. Second, Chandler et al. (<xref ref-type="bibr" rid="B3">2014</xref>, <xref ref-type="bibr" rid="B4">2015</xref>) suggested that MTurk workers are likely to participate in multiple surveys, hence workers might be less na&#x000EF;ve than participants from other (e.g., student) samples. They also pointed out that the prior experience with commonly used survey question (e.g., Cognitive Reflection Test, Frederick, <xref ref-type="bibr" rid="B9">2005</xref>) inflates performances on the task, and suggested that the repeated participation of workers may threaten the predictive accuracy of the task and reduce effect sizes of research findings. Furthermore, Stewart et al. (<xref ref-type="bibr" rid="B32">2015</xref>) estimated the size of the population of active MTurk workers and suggested that the average laboratory can collect data from the relatively smaller numbers of active workers (about 7,300 compared to 500,000 registered MTurk workers). Thus, multiple participations to similar surveys are likely to happen than expected. These pitfalls, the high rate of non-naivety of participants in particular, can be resolved if researchers recruit participants from alternative crowdsourcing services. In addition, conducting online surveys and experiments with multiple crowdsourcing platforms will be beneficial for researchers who look for a more diverse sample.</p>
<p>As noted previously, the quality of data collected from MTurk participants have been verified. It is also shown that the other crowdsourcing pools, such as Clickworker and Prolific Academic, are practical alternatives to MTurk (e.g., Lutz, <xref ref-type="bibr" rid="B17">2016</xref>; Peer et al., <xref ref-type="bibr" rid="B27">2017</xref>). However, data from other crowdsourcing samples, particularly from non-Caucasian samples, have not as yet been fully investigated. To promote research using other crowdsourcing services, we must examine whether the data obtained from other crowdsourcing pools are as reliable as those from MTurk.</p>
</sec>
<sec id="s2">
<title>Research objectives and general research method</title>
<p>The primary goal of the present study is to extend existing findings of previous validation studies of MTurk to other non-MTurk crowdsourcing samples. Specifically, we investigated the following questions.</p>
<list list-type="simple">
<list-item><p>Question 1: Do the demographic properties of workers from the other (i.e., non-MTurk) crowdsourcing samples differ from those of students? If so, how are they different?</p></list-item>
<list-item><p>Question 2: Do psychometric properties, such as personality traits or those of consumer behavior, differ across non-MTurk workers and students?</p></list-item>
<list-item><p>Question 3: Is the quality of non-MTurk workers&#x00027; performance on reasoning and judgment tasks relevant to effortful System-2 thinking in comparison with that of students?</p></list-item>
<list-item><p>Question 4: How do non-MTurk workers respond to &#x0201C;trap&#x0201D; questions? Are they more (or less) attentive to the instructions for these tasks?</p></list-item>
</list>
<p>The present study compared crowdsourcing participants with university students in terms of their personality, psychometric properties regarding decision making, and consumer behavior (Survey 1), thinking disposition, reasoning performance, and attention to the study materials (Surveys 2 and 3). In all of the surveys, the crowdsourcing participants were recruited from CrowdWorks (a Japanese crowdsourcing service, which is abbreviated as CW hereafter; <ext-link ext-link-type="uri" xlink:href="https://crowdworks.jp">https://crowdworks.jp</ext-link>). We adopted CW as a participant pool for the following reasons. Firstly, CW has a growing and sufficiently large pool of registered workers for validation studies (a total of more than 1 million workers as of August 2016). Second, because it offers a user interface that is written in Japanese, the majority of workers are native Japanese speakers, and as a result, it enables data collection from participants of different ethnic groups than MTurk. Third, it offers a similar payment system as MTurk, and it does not charge a commission fee for micro tasks. In addition, it accepts several payment methods, such as bank transfer, credit cards, and PayPal. The student sample was collected from two middle-sized universities that are located in Sapporo, which is a large northern city of Japan. The CW participants received monetary compensation in exchange for their participation in the survey. However, the students received extra course credit or voluntarily participated in the survey.</p>
<p>All of the participants answered web-based questionnaires that were administered by SurveyMonkey (Surveys 1 and 2) or Qualtrics (Survey 3). For the CW sample, we posted a link to the survey site to the CW task. When the participants reached the site, they were presented with general instructions, and they were asked to provide their consent to participate in the survey by clicking an &#x0201C;agree&#x0201D; button. If they agreed to take the survey, the online questionnaires were presented in a designed sequence. After they completed the questionnaires, they received a randomized completion code, and they were asked to enter it into the CW task page to receive payment. Because CW allocates a unique ID per person, it is possible to restrict the same worker to a single task more than once. In addition, we also enabled SurveyMonkey and Qualtrics restriction features to prohibit multiple participations. After the correct completion code had been entered, the experimenter approved the compensation to be sent to the participants&#x00027; accounts. The CW participants were completely anonymous throughout the entire survey process.</p>
<p>The university students were recruited from introductory psychology, statistics, English, or social welfare classes, and they were provided with a leaflet that described a link to the equivalent web-based survey site. When they reached the site, they received the same general instructions and the same request for their consent to take the survey as the CW participants. After they completed the survey, they were provided with a randomly generated completion code that was required for them to receive credit.</p>
<p>The present study was approved and conducted in compliance with the guidelines of the Hokusei Gakuen University Ethics Committee. All of the participants gave their web-based informed consent instead of written consent.</p>
</sec>
<sec id="s3">
<title>Survey 1: personality and psychometric properties</title>
<p>Survey 1 compared the CW and university samples in terms of their demographic status, personality traits and psychometric properties, which included the so-called Big Five traits, as well as self-esteem, goal orientation, and materialism as an aspect of consumer behavior. These scales were adopted from previous validation studies of MTurk (e.g., Behrend et al., <xref ref-type="bibr" rid="B1">2011</xref>; Goodman et al., <xref ref-type="bibr" rid="B11">2013</xref>).</p>
<sec>
<title>Method</title>
<sec>
<title>Participants</title>
<p>A total of 319 crowdsourcing workers agreed to participate in the survey; however, we excluded 7 participants because they did not complete the questionnaire. We also excluded 17 responses because of IP address duplication, which left 295 in the final sample. The participants received 50 JPY for completing a 10 min survey.</p>
<p>In addition, we collected 144 students, but we excluded 12 participants from the analyses for the following reasons: incomplete responses (11 participants) and IP address duplication (1 participant). We also excluded one participant from the analyses because of a failure to indicate that he or she was currently a university student in the demographic question. A final sample of 131 undergraduate students participated in the survey.</p>
<p>The sample size of the present survey was decided in reference to previous validation studies of MTurk and other practical reasons. For example, Behrend et al. (<xref ref-type="bibr" rid="B1">2011</xref>) collected 270 MTurk and 270 undergraduate students, and Goodman et al. (<xref ref-type="bibr" rid="B11">2013</xref>) sampled 207 MTurk and 131 student participants. In addition, based on our previous experience in using CW, we estimated that a growth in the number of CW participants slowed down if we recruited more than 300 participants. Furthermore, the size of the student sample was determined by rather practical reason, i.e., class attendance. However, as shown above, the present survey collected as many student participants as those of Goodman et al. (<xref ref-type="bibr" rid="B11">2013</xref>)&#x00027;s study. Although Goodman et al. (<xref ref-type="bibr" rid="B11">2013</xref>) did not mention effect sizes, Behrend et al. (<xref ref-type="bibr" rid="B1">2011</xref>) reported that effect sizes on the difference in personality traits between MTurk and student samples ranged from <italic>d</italic> &#x0003D; 0.31&#x02013;0.86. We conducted power analysis in G-Power to determine sufficient sample size using an alpha of 0.05, a power of 0.8, effect size (<italic>d</italic> &#x0003D; 0.3), and two tails. Based on the aforementioned assumptions, the desired sizes for the first and the second sample were 285 and 127. The result indicated that sample size of the present survey was sufficiently large.</p>
</sec>
<sec>
<title>Materials and procedure</title>
<p>As the measures for personality traits, we administered two widely used personality inventories: a brief measure of the Big-Five personality dimensions (10-Item Personality Inventory, TIPI; Gosling et al., <xref ref-type="bibr" rid="B12">2003</xref>) and Rosenberg&#x00027;s self-esteem scale (RSE; Rosenberg, <xref ref-type="bibr" rid="B30">1965</xref>). In this survey, we adopted the Japanese version of the TIPI (TIPI-J; Oshio et al., <xref ref-type="bibr" rid="B24">2012</xref>) and the RSE, which was translated into Japanese by Yamamoto et al. (<xref ref-type="bibr" rid="B38">1982</xref>). Furthermore, we administered two additional scales that were also used in the previous validation studies: the performance prove/avoid goal orientation scale (PPGO and PAGO; Vandewalle, <xref ref-type="bibr" rid="B37">1997</xref>) and the Material Value Scale (MVS; Richins, <xref ref-type="bibr" rid="B28">2004</xref>). Finally, we asked for participants&#x00027; demographic status: age, gender, ethnicity, nationality, educational level, and employment status.</p>
<p>All of the participants completed identical measures in an identical order. In the first step, they answered each of TIPI-J items on a 7-point Likert scale (1 &#x0003D; <italic>Disagree Strongly</italic> to 7 &#x0003D; <italic>Agree Strongly</italic>). Next, the participants answered the PAGO and PPGO in mixed order (6-point scale, 1 &#x0003D; <italic>Strongly disagree</italic> to 6 &#x0003D; <italic>Strongly agree</italic>). Subsequently, the participants were presented with 10 items of the RSE followed by a nine-item version of the MVS and provided answers to each item on a 5-point scale that ranged from 1 &#x0003D; <italic>Strongly disagree</italic> to 5 &#x0003D; <italic>Strongly agree</italic>. Finally, they answered demographic questions before the end of the survey.</p>
</sec>
</sec>
<sec>
<title>Results</title>
<p>All of the statistical analyses of the present study were performed using SPSS 21.0. In addition, when we report &#x003B7;<sup>2</sup> as an index of effect size of ANOVA, where the value designates partial &#x003B7;<sup>2</sup>.</p>
<sec>
<title>Demographics</title>
<p>Table <xref ref-type="table" rid="T1">1</xref> summarizes the demographic status of both samples. The CW workers were significantly higher in age than the students (UNIV), <italic>M</italic><sub>C</sub> &#x0003D; 36.9 vs. <italic>M</italic><sub>U</sub> &#x0003D; 19.6 years, <italic>t</italic><sub>(423)</sub> &#x0003D; 22.4, <italic>p</italic> &#x0003C; 0.001, <italic>d</italic> &#x0003D; 2.35; percentage of female, CW &#x0003D; 63.7% vs. UNIV &#x0003D; 43.5%, <inline-formula><mml:math id="M1"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 15.2, <italic>p</italic> &#x0003C; 0.001; and median level of education, Mdn<sub>C</sub> &#x0003D; &#x0201C;associate degree,&#x0201D; Mdn<sub>U</sub> &#x0003D; &#x0201C;high school,&#x0201D; Wilcoxon&#x00027;s <italic>Z</italic> &#x0003D; 10.3, <italic>p</italic> &#x0003C; 0.001. The two samples were also different in their years of work experience, <italic>M</italic><sub>C</sub> &#x0003D; 12.4 vs. <italic>M</italic><sub>U</sub> &#x0003D; 2.5, <italic>t</italic><sub>(263)</sub> &#x0003D; 4.7, <italic>p</italic> &#x0003C; 0.001, <italic>d</italic> &#x0003D; 1.17; and employment status, <inline-formula><mml:math id="M2"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>5</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 77.4, <italic>p</italic> &#x0003C; 0.001. On the one hand, 74.8% of the students were not currently employed, and 17.6% were part-time workers. On the other hand, 40.3% of the CW workers were not employed, 21.4% were full-time employees, 19.7% were self-employed, and 13.2% were part-time workers.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p><bold>Demographic results of Survey 1 and 2</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Survey</bold></th>
<th valign="top" align="center"><bold>Sample</bold></th>
<th valign="top" align="center" colspan="2" style="border-bottom: thin solid #000000;"><bold>Age</bold></th>
<th valign="top" align="center"><bold>Female %</bold></th>
<th valign="top" align="center" colspan="2" style="border-bottom: thin solid #000000;"><bold>Work experience (years)</bold></th>
<th valign="top" align="center" style="border-bottom: thin solid #000000;"><bold>Ethnicity</bold></th>
<th valign="top" align="center" style="border-bottom: thin solid #000000;"><bold>Nationality</bold></th>
<th/>
</tr>
<tr>
<th/>
<th/>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th/>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th valign="top" align="center"><bold>%JP</bold></th>
<th valign="top" align="center"><bold>%JP</bold></th>
<th valign="top" align="center"><bold><italic>N</italic></bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">1</td>
<td valign="top" align="center">UNIV</td>
<td valign="top" align="center">19.56</td>
<td valign="top" align="center">1.12</td>
<td valign="top" align="center">43.5</td>
<td valign="top" align="center">2.47</td>
<td valign="top" align="center">4.12</td>
<td valign="top" align="center">98</td>
<td valign="top" align="center">100</td>
<td valign="top" align="center">131</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">CW</td>
<td valign="top" align="center">36.89</td>
<td valign="top" align="center">8.81</td>
<td valign="top" align="center">63.7</td>
<td valign="top" align="center">12.38</td>
<td valign="top" align="center">8.68</td>
<td valign="top" align="center">98</td>
<td valign="top" align="center">99</td>
<td valign="top" align="center">295</td>
</tr>
<tr>
<td valign="top" align="left">2</td>
<td valign="top" align="center">UNIV</td>
<td valign="top" align="center">19.70</td>
<td valign="top" align="center">1.34</td>
<td valign="top" align="center">46.2</td>
<td valign="top" align="center">3.04</td>
<td valign="top" align="center">4.37</td>
<td valign="top" align="center">98</td>
<td valign="top" align="center">100</td>
<td valign="top" align="center">156</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">CW</td>
<td valign="top" align="center">36.59</td>
<td valign="top" align="center">9.19</td>
<td valign="top" align="center">62.3</td>
<td valign="top" align="center">12.43</td>
<td valign="top" align="center">9.16</td>
<td valign="top" align="center">98</td>
<td valign="top" align="center">100</td>
<td valign="top" align="center">297</td>
</tr>
<tr style="border-top: thin solid #000000;">
<td/>
<td/>
<td/>
<td valign="top" align="center"><bold>Middle</bold></td>
<td valign="top" align="center"><bold>High</bold></td>
<td valign="top" align="center"><bold>Associate</bold></td>
<td valign="top" align="center"><bold>Bachelor</bold></td>
<td valign="top" align="center"><bold>Graduate</bold></td>
<td valign="top" align="center"><bold>NA</bold></td>
<td/>
</tr>
<tr style="border-top: thin solid #000000;">
<td valign="top" align="left" colspan="10" style="background-color:#bbbdc0"><bold>HIGHEST EDUCATIONAL LEVEL</bold></td>
</tr>
<tr>
<td valign="top" align="left">1</td>
<td valign="top" align="center">UNIV</td>
<td/>
<td valign="top" align="center">0</td>
<td valign="top" align="center">117</td>
<td valign="top" align="center">3</td>
<td valign="top" align="center">10</td>
<td valign="top" align="center">0</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">131</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">CW</td>
<td/>
<td valign="top" align="center">1</td>
<td valign="top" align="center">90</td>
<td valign="top" align="center">76</td>
<td valign="top" align="center">118</td>
<td valign="top" align="center">8</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">295</td>
</tr>
<tr>
<td valign="top" align="left">2</td>
<td valign="top" align="center">UNIV</td>
<td/>
<td valign="top" align="center">0</td>
<td valign="top" align="center">141</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">9</td>
<td valign="top" align="center">0</td>
<td valign="top" align="center">5</td>
<td valign="top" align="center">156</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">CW</td>
<td/>
<td valign="top" align="center">9</td>
<td valign="top" align="center">94</td>
<td valign="top" align="center">40</td>
<td valign="top" align="center">114</td>
<td valign="top" align="center">9</td>
<td valign="top" align="center">31</td>
<td valign="top" align="center">297</td>
</tr>
<tr style="border-top: thin solid #000000;">
<td/>
<td/>
<td valign="top" align="center"><bold>Full-time</bold></td>
<td valign="top" align="center"><bold>Part-time</bold></td>
<td valign="top" align="center"><bold>Self</bold></td>
<td valign="top" align="center"><bold>Employer</bold></td>
<td valign="top" align="center"><bold>Retired</bold></td>
<td valign="top" align="center"><bold>Not-employed</bold></td>
<td valign="top" align="center"><bold>NA</bold></td>
<td/>
</tr>
<tr style="border-top: thin solid #000000;">
<td valign="top" align="left" colspan="10" style="background-color:#bbbdc0"><bold>EMPLOYMENT STATUS</bold></td>
</tr>
<tr>
<td valign="top" align="left">1</td>
<td valign="top" align="center">UNIV</td>
<td valign="top" align="center">0</td>
<td valign="top" align="center">23</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">0</td>
<td valign="top" align="center">98</td>
<td valign="top" align="center">8</td>
<td valign="top" align="center">131</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">CW</td>
<td valign="top" align="center">63</td>
<td valign="top" align="center">39</td>
<td valign="top" align="center">58</td>
<td valign="top" align="center">7</td>
<td valign="top" align="center">3</td>
<td valign="top" align="center">119</td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">295</td>
</tr>
<tr>
<td valign="top" align="left">2</td>
<td valign="top" align="center">UNIV</td>
<td valign="top" align="center">0</td>
<td valign="top" align="center">26</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">0</td>
<td valign="top" align="center">111</td>
<td valign="top" align="center">16</td>
<td valign="top" align="center">156</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">CW</td>
<td valign="top" align="center">69</td>
<td valign="top" align="center">48</td>
<td valign="top" align="center">50</td>
<td valign="top" align="center">5</td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">109</td>
<td valign="top" align="center">10</td>
<td valign="top" align="center">297</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>UNIV, student sample; CW, CrowdWorks sample; %JP, percentage of Japanese; Self, self-employed; NA, either &#x0201C;no answer&#x0201D; or &#x0201C;not applicable.&#x0201D;</italic></p>
</table-wrap-foot>
</table-wrap>
</sec>
<sec>
<title>Personality traits</title>
<p>Table <xref ref-type="table" rid="T2">2</xref> summarizes the result of the Big Five personality and self-esteem scale. In the following analyses, we considered sample and gender as independent variables (age was excluded because of a strong point-biserial correlation with sample, <italic>r</italic><sub>pb</sub> &#x0003D; 0.74). The gender was included because several previous studies with Japanese participants have shown gender differences in these personality traits (e.g., Kawamoto et al., <xref ref-type="bibr" rid="B15">2015</xref>; Okada et al., <xref ref-type="bibr" rid="B22">2015</xref>; the gender issue was discussed in the Discussion Section). The TIPI-J scores were submitted to a sample &#x000D7; gender MANOVA, and they showed a significant multivariate effect of the sample, <italic>F</italic><sub>(5, 418)</sub> &#x0003D; 5.95, Wilk&#x00027;s &#x0039B; &#x0003D; 0.93, <italic>p</italic> &#x0003C; 0.001, &#x003B7;<sup>2</sup> &#x0003D; 0.07. The multivariate effect was also significant for gender, <italic>F</italic><sub>(5, 418)</sub> &#x0003D; 6.63, &#x0039B; &#x0003D; 0.93, <italic>p</italic> &#x0003C; 0.001, &#x003B7;<sup>2</sup> &#x0003D; 0.08. The univariate <italic>F</italic>-tests revealed significant differences of the samples in Extraversion, <italic>F</italic><sub>(1, 422)</sub> &#x0003D; 8.57, <italic>MSE</italic> &#x0003D; 8.55, <italic>p</italic> &#x0003D; 0.004, &#x003B7;<sup>2</sup> &#x0003D; 0.02; Agreeableness, <italic>F</italic><sub>(1, 422)</sub> &#x0003D; 5.22, <italic>MSE</italic> &#x0003D; 5.38, <italic>p</italic> &#x0003D; 0.023, &#x003B7;<sup>2</sup> &#x0003D; 0.01; and Conscientiousness, <italic>F</italic><sub>(1, 422)</sub> &#x0003D; 6.07, <italic>MSE</italic> &#x0003D; 6.97, <italic>p</italic> &#x0003D; 0.014, &#x003B7;<sup>2</sup> &#x0003D; 0.01. The two samples were not different in Emotional Stability and Openness (<italic>F</italic>s &#x0003C; 1). The results also showed that males were significantly higher than females in Emotional Stability, <italic>F</italic><sub>(1, 422)</sub> &#x0003D; 14.08, <italic>MSE</italic> &#x0003D; 6.39, <italic>p</italic> &#x0003C; 0.001, &#x003B7;<sup>2</sup> &#x0003D; 0.03; and Openness, <italic>F</italic><sub>(1, 422)</sub> &#x0003D; 9.11, <italic>MSE</italic> &#x0003D; 6.34, <italic>p</italic> &#x0003D; 0.003, &#x003B7;<sup>2</sup> &#x0003D; 0.02. The gender differences were not found in Extraversion, Agreeableness, and Conscientiousness, <italic>F</italic>s<sub>(1, 422)</sub> &#x0003C; 1.44, <italic>p</italic>s &#x0003E; 0.23. Although the multivariate sample &#x000D7; gender interaction was not significant, <italic>F</italic><sub>(5, 418)</sub> &#x0003D; 1.29, &#x0039B; &#x0003D; 0.98, <italic>p</italic> &#x0003D; 0.267, &#x003B7;<sup>2</sup> &#x0003D; 0.02, the univariate analysis showed a significant sample &#x000D7; gender interaction on Openness, <italic>F</italic><sub>(1, 422)</sub> &#x0003D; 5.25, <italic>MSE</italic> &#x0003D; 6.34, <italic>p</italic> &#x0003D; 0.022, &#x003B7;<sup>2</sup> &#x0003D; 0.01. The analysis of the simple main effect indicated that male students were more open than female students, <italic>F</italic><sub>(1, 422)</sub> &#x0003D; 10.38, <italic>MSE</italic> &#x0003D; 6.34, <italic>p</italic> &#x0003D; 0.001, &#x003B7;<sup>2</sup> &#x0003D; 0.02; however, no such difference was found for CW workers, <italic>F</italic> &#x0003C; 1.</p>
<table-wrap position="float" id="T2">
<label>Table 2</label>
<caption><p><bold>Means, standard deviations, and Cronbach&#x00027;s alpha coefficients of personality traits and psychometric properties as functions of the sample (UNIV, student; CW, CrowdWorks) and gender (Survey 1)</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th/>
<th valign="top" align="center" colspan="7" style="border-bottom: thin solid #000000;"><bold>UNIV (</bold><italic><bold>N</bold></italic> &#x0003D; <bold>131)</bold></th>
<th valign="top" align="center" colspan="7" style="border-bottom: thin solid #000000;"><bold>CW (</bold><italic><bold>N</bold></italic> &#x0003D; <bold>295)</bold></th>
</tr>
<tr>
<th/>
<th/>
<th valign="top" align="center" colspan="3" style="border-bottom: thin solid #000000;"><bold>Male (</bold><italic><bold>n</bold></italic> &#x0003D; <bold>74)</bold></th>
<th valign="top" align="center" colspan="3" style="border-bottom: thin solid #000000;"><bold>Female (</bold><italic><bold>n</bold></italic> &#x0003D; <bold>57)</bold></th>
<th/>
<th valign="top" align="center" colspan="3" style="border-bottom: thin solid #000000;"><bold>Male (</bold><italic><bold>n</bold></italic> &#x0003D; <bold>107)</bold></th>
<th valign="top" align="center" colspan="3" style="border-bottom: thin solid #000000;"><bold>Female (</bold><italic><bold>n</bold></italic> &#x0003D; <bold>188)</bold></th>
</tr>
<tr>
<th/>
<th valign="top" align="center"><bold>&#x003B1;</bold></th>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th valign="top" align="center"><bold>[95% CI]</bold></th>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th valign="top" align="center"><bold>[95% CI]</bold></th>
<th valign="top" align="center"><bold>&#x003B1;</bold></th>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th valign="top" align="center"><bold>[95%CI]</bold></th>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th valign="top" align="center"><bold>[95% CI]</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">EX<xref ref-type="table-fn" rid="TN1"><sup>a</sup></xref></td>
<td valign="top" align="center">0.706</td>
<td valign="top" align="center">7.39</td>
<td valign="top" align="center">3.17</td>
<td valign="top" align="center">[6.72, 8.06]</td>
<td valign="top" align="center">7.28</td>
<td valign="top" align="center">3.30</td>
<td valign="top" align="center">[6.52, 8.04]</td>
<td valign="top" align="center">0.746</td>
<td valign="top" align="center">5.99</td>
<td valign="top" align="center">2.53</td>
<td valign="top" align="center">[5.43, 6.55]</td>
<td valign="top" align="center">6.85</td>
<td valign="top" align="center">2.91</td>
<td valign="top" align="center">[6.43, 7.27]</td>
</tr>
<tr>
<td valign="top" align="left">A<xref ref-type="table-fn" rid="TN1"><sup>a</sup></xref></td>
<td valign="top" align="center">0.411</td>
<td valign="top" align="center">9.82</td>
<td valign="top" align="center">2.51</td>
<td valign="top" align="center">[9.29, 10.35]</td>
<td valign="top" align="center">9.74</td>
<td valign="top" align="center">2.25</td>
<td valign="top" align="center">[9.13, 10.34]</td>
<td valign="top" align="center">0.447</td>
<td valign="top" align="center">9.09</td>
<td valign="top" align="center">2.48</td>
<td valign="top" align="center">[8.65, 9.53]</td>
<td valign="top" align="center">9.34</td>
<td valign="top" align="center">2.16</td>
<td valign="top" align="center">[9.00, 9.67]</td>
</tr>
<tr>
<td valign="top" align="left">C<xref ref-type="table-fn" rid="TN1"><sup>a</sup></xref></td>
<td valign="top" align="center">0.479</td>
<td valign="top" align="center">6.58</td>
<td valign="top" align="center">3.03</td>
<td valign="top" align="center">[5.98, 7.18]</td>
<td valign="top" align="center">6.26</td>
<td valign="top" align="center">2.40</td>
<td valign="top" align="center">[5.58, 6.95]</td>
<td valign="top" align="center">0.569</td>
<td valign="top" align="center">6.88</td>
<td valign="top" align="center">2.31</td>
<td valign="top" align="center">[6.38, 7.38]</td>
<td valign="top" align="center">7.36</td>
<td valign="top" align="center">2.72</td>
<td valign="top" align="center">[6.98, 7.73]</td>
</tr>
<tr>
<td valign="top" align="left">ES<xref ref-type="table-fn" rid="TN2"><sup>b</sup></xref></td>
<td valign="top" align="center">0.309</td>
<td valign="top" align="center">7.28</td>
<td valign="top" align="center">2.61</td>
<td valign="top" align="center">[6.71, 7.86]</td>
<td valign="top" align="center">6.04</td>
<td valign="top" align="center">2.28</td>
<td valign="top" align="center">[5.38, 6.69]</td>
<td valign="top" align="center">0.601</td>
<td valign="top" align="center">7.21</td>
<td valign="top" align="center">2.29</td>
<td valign="top" align="center">[6.73, 7.69]</td>
<td valign="top" align="center">6.43</td>
<td valign="top" align="center">2.69</td>
<td valign="top" align="center">[6.06, 6.79]</td>
</tr>
<tr>
<td valign="top" align="left">O<xref ref-type="table-fn" rid="TN2"><sup>b</sup></xref><sup>,</sup> <xref ref-type="table-fn" rid="TN3"><sup>c</sup></xref></td>
<td valign="top" align="center">0.245</td>
<td valign="top" align="center">8.50</td>
<td valign="top" align="center">2.53</td>
<td valign="top" align="center">[7.92, 9.08]</td>
<td valign="top" align="center">7.07</td>
<td valign="top" align="center">2.65</td>
<td valign="top" align="center">[6.41, 7.73]</td>
<td valign="top" align="center">0.533</td>
<td valign="top" align="center">7.81</td>
<td valign="top" align="center">2.54</td>
<td valign="top" align="center">[7.33, 8.29]</td>
<td valign="top" align="center">7.62</td>
<td valign="top" align="center">2.46</td>
<td valign="top" align="center">[7.26, 7.98]</td>
</tr>
<tr>
<td valign="top" align="left">RSE<xref ref-type="table-fn" rid="TN3"><sup>c</sup></xref></td>
<td valign="top" align="center">0.814</td>
<td valign="top" align="center">29.39</td>
<td valign="top" align="center">7.88</td>
<td valign="top" align="center">[27.63, 31.15]</td>
<td valign="top" align="center">26.35</td>
<td valign="top" align="center">6.14</td>
<td valign="top" align="center">[24.34, 28.36]</td>
<td valign="top" align="center">0.889</td>
<td valign="top" align="center">27.59</td>
<td valign="top" align="center">7.58</td>
<td valign="top" align="center">[26.12, 29.05]</td>
<td valign="top" align="center">28.24</td>
<td valign="top" align="center">8.13</td>
<td valign="top" align="center">[27.14, 29.35]</td>
</tr>
<tr>
<td valign="top" align="left">PAGO<xref ref-type="table-fn" rid="TN1"><sup>a</sup></xref></td>
<td valign="top" align="center">0.671</td>
<td valign="top" align="center">16.31</td>
<td valign="top" align="center">4.11</td>
<td valign="top" align="center">[15.56, 17.07]</td>
<td valign="top" align="center">16.16</td>
<td valign="top" align="center">2.90</td>
<td valign="top" align="center">[15.30, 17.02]</td>
<td valign="top" align="center">0.719</td>
<td valign="top" align="center">15.19</td>
<td valign="top" align="center">3.26</td>
<td valign="top" align="center">[14.56, 15.82]</td>
<td valign="top" align="center">15.23</td>
<td valign="top" align="center">3.08</td>
<td valign="top" align="center">[14.75, 15.70]</td>
</tr>
<tr>
<td valign="top" align="left">PPGO<xref ref-type="table-fn" rid="TN1"><sup>a</sup></xref></td>
<td valign="top" align="center">0.737</td>
<td valign="top" align="center">16.23</td>
<td valign="top" align="center">4.28</td>
<td valign="top" align="center">[15.46, 17.00]</td>
<td valign="top" align="center">15.42</td>
<td valign="top" align="center">3.45</td>
<td valign="top" align="center">[14.54, 16.30]</td>
<td valign="top" align="center">0.707</td>
<td valign="top" align="center">14.94</td>
<td valign="top" align="center">2.97</td>
<td valign="top" align="center">[14.30, 15.59]</td>
<td valign="top" align="center">15.02</td>
<td valign="top" align="center">3.16</td>
<td valign="top" align="center">[14.54, 15.51]</td>
</tr>
<tr>
<td valign="top" align="left">MVS<xref ref-type="table-fn" rid="TN1"><sup>a</sup></xref></td>
<td valign="top" align="center">0.744</td>
<td valign="top" align="center">27.97</td>
<td valign="top" align="center">5.99</td>
<td valign="top" align="center">[26.63, 29.31]</td>
<td valign="top" align="center">26.37</td>
<td valign="top" align="center">6.09</td>
<td valign="top" align="center">[24.84, 27.89]</td>
<td valign="top" align="center">0.782</td>
<td valign="top" align="center">26.16</td>
<td valign="top" align="center">5.46</td>
<td valign="top" align="center">[25.05, 27.27]</td>
<td valign="top" align="center">25.70</td>
<td valign="top" align="center">5.96</td>
<td valign="top" align="center">[24.86, 26.54]</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>EX, Extraversion; A, Agreeableness; C, Conscientiousness; ES, Emotional Stability; O, Openness to Experience; RSE, Rosenberg&#x00027;s Self-Esteem scale; PAGO, Performance Avoid Goal Orientation; PPGO, Performance Prove Goal Orientation; MVS, Material Value scale</italic>.</p>
<fn id="TN1">
<label>a</label>
<p><italic>CW participants are significantly different from students (p &#x0003C; 0.05)</italic>,</p></fn>
<fn id="TN2">
<label>b</label>
<p><italic>significant gender difference (p &#x0003C; 0.05)</italic>,</p></fn>
<fn id="TN3">
<label>c</label>
<p><italic>significant sample &#x000D7; gender interaction (p &#x0003C; 0.05)</italic>.</p></fn>
</table-wrap-foot>
</table-wrap>
<p>A similar ANOVA on the RSE scale failed to show significant sample and gender differences, <italic>F</italic>s<sub>(1, 422)</sub> &#x0003C; 2.1, <italic>p</italic>s &#x0003E; 0.148. However, we found a significant sample &#x000D7; gender interaction, <italic>F</italic><sub>(1, 422)</sub> &#x0003D; 5.02, <italic>MSE</italic> &#x0003D; 59.51, <italic>p</italic> &#x0003D; 0.026, &#x003B7;<sup>2</sup> &#x0003D; 0.01. The analysis of simple effects revealed that male students scored slightly higher than female students, <italic>F</italic><sub>(1, 422)</sub> &#x0003D; 5.0, <italic>p</italic> &#x0003D; 0.026, <italic>MSE</italic> &#x0003D; 59.51, &#x003B7;<sup>2</sup> &#x0003D; 0.01; however, no gender difference was found for the CW sample, <italic>F</italic> &#x0003C; 1.</p>
</sec>
<sec>
<title>Goal orientation and material value</title>
<p>Table <xref ref-type="table" rid="T2">2</xref> shows the results of goal orientation and materialism. A MANOVA on two goal orientations indicated the multivariate effect of the sample, <italic>F</italic><sub>(2, 421)</sub> &#x0003D; 5.11, &#x0039B; &#x0003D; 0.98, <italic>p</italic> &#x0003D; 0.006, &#x003B7;<sup>2</sup> &#x0003D; 0.02. Subsequent univariate <italic>F</italic>-tests revealed that the students were higher than the workers in both PAGO and PPGO, <italic>F</italic>s<sub>(1, 422)</sub> &#x0003D; 8.43, 5.45, <italic>MSE</italic>s &#x0003D; 10.93, 11.40, <italic>p</italic>s &#x0003D; 0.004, 0.020, &#x003B7;<sup>2</sup>s &#x0003D; 0.02, 0.01, respectively. However, neither the effect of gender nor the interaction effect were significant, <italic>F</italic>s &#x0003C; 1.</p>
<p>Then, a sample &#x000D7; gender ANOVA was conducted, and the result showed that the students were more materialistic than the crowdsourcing sample, <italic>F</italic><sub>(1, 422)</sub> &#x0003D; 3.94, <italic>MSE</italic> &#x0003D; 34.34, <italic>p</italic> &#x0003D; 0.048, &#x003B7;<sup>2</sup> &#x0003D; 0.01. However, gender main effect and interaction were not significant, <italic>F</italic>s<sub>(1, 422)</sub> &#x0003D; 2.72, 0.83, <italic>p</italic> &#x0003D; 0.100, 0.362.</p>
</sec>
</sec>
<sec>
<title>Discussion</title>
<p>In Survey 1, we found a significant, but not surprising, difference between the students and the CW workers in terms of their demographic status. The findings also showed that some personality characteristics differed between the two samples. For example, the CW participants were less extraverted and agreeable, although they were more conscientious than the students. In addition, the CW participants were less materialistic and their pursuit performance-avoid or prove goals were lower than those of the students.</p>
<p>Some of these results, such as demographics, extraversion, openness, and performance-avoid goal orientation, were compatible with the previous validation studies using MTurk (Paolacci et al., <xref ref-type="bibr" rid="B26">2010</xref>; Behrend et al., <xref ref-type="bibr" rid="B1">2011</xref>; Goodman et al., <xref ref-type="bibr" rid="B11">2013</xref>). There were also several inconsistent results on the difference between the two samples compared to the previous studies. For example, Goodman et al. (<xref ref-type="bibr" rid="B11">2013</xref>) showed that MTurk workers were more emotionally unstable, i.e., neurotic, than the students and community sample; however, we did not find any such difference between the samples, but we did find a gender difference. Recently, Kawamoto et al. (<xref ref-type="bibr" rid="B15">2015</xref>) showed that Japanese females scored higher in neuroticism than males, particularly in their younger adulthood. Our result is compatible with this finding if we consider the distribution of age in both of the samples (UNIV &#x0003D; the majority of the participants were in their late teens or early twenties, CW &#x0003D; 40% were in their thirties, 30% were in their forties, and 19% were in their twenties). Goodman et al. (<xref ref-type="bibr" rid="B11">2013</xref>) also found that MTurk workers were less conscientious than students. However, we found an opposite direction of results; our results were consistent with the findings of Big Five personality and showed that conscientiousness was likely to develop during adulthood (e.g., McCrae et al., <xref ref-type="bibr" rid="B21">2000</xref>; Srivastava et al., <xref ref-type="bibr" rid="B31">2003</xref>; Kawamoto et al., <xref ref-type="bibr" rid="B15">2015</xref>). Furthermore, we found that the male students were higher in self-esteem than the female students; however, no gender difference was found in the CW sample. Our findings were compatible with the previous offline investigations, which showed that males had higher self-esteem than females, and this gender difference decreased throughout adulthood (Kling et al., <xref ref-type="bibr" rid="B16">1999</xref>; Robins et al., <xref ref-type="bibr" rid="B29">2002</xref>; Okada et al., <xref ref-type="bibr" rid="B22">2015</xref>).</p>
<p>To summarize, our results indicated both similarities and differences between the CW workers and the students, which is generally consistent with existing findings. It is also important to note that the effect sizes of the sample differences were relatively small, as has been shown in previous studies.</p>
</sec>
</sec>
<sec id="s4">
<title>Survey 2: attentional check and system-2 thinking</title>
<p>Survey 2 aimed to compare the crowdsourcing workers and students in terms of their thinking disposition, as well as their reasoning and judgment biases related to systematic System-2 thinking.</p>
<p>As a measure of thinking disposition, we administered the Cognitive Reflection Test (CRT; Frederick, <xref ref-type="bibr" rid="B9">2005</xref>), which is a set of widely used tasks to measure individual differences in dual process thought, particularly in effortful System 2 thinking. We also administered the following three tasks to measure the participants&#x00027; biases in reasoning and judgment. The first task was the probabilistic reasoning task (Toplak et al., <xref ref-type="bibr" rid="B33">2011</xref>), which aimed to measure denominator neglect bias in a hypothetical scenario. The second task was the logical reasoning task, which consisted of eight syllogisms (Markovits and Nantel, <xref ref-type="bibr" rid="B19">1989</xref>; Majima, <xref ref-type="bibr" rid="B18">2015</xref>) in which the validity of the conclusion always conflicted with common belief. These syllogisms were designed to measure the strength of the belief bias effect (Evans et al., <xref ref-type="bibr" rid="B8">1983</xref>). The third task was a classical anchoring-and-adjustment task (Tversky and Kahneman, <xref ref-type="bibr" rid="B34">1974</xref>).</p>
<p>We also investigated sample differences in their attention to instructions by using instructional manipulation checks (IMCs; Oppenheimer et al., <xref ref-type="bibr" rid="B23">2009</xref>). In addition, we examined whether answering to the IMCs promoted successful solutions to the other &#x0201C;tricky&#x0201D; reasoning tasks, as shown in Hauser and Schwarz (<xref ref-type="bibr" rid="B13">2015</xref>). To investigate whether the interventional effects of an IMC on the subsequent tasks were replicated in the Japanese sample, two questionnaire orders were introduced: IMC-first, in which IMC was administered before the CRT and other reasoning tasks, and IMC-last, in which IMC was administered after those tasks.</p>
<sec>
<title>Method</title>
<sec>
<title>Participants</title>
<p>We collected data from 338 CW workers; however, data from 27 of the participants were excluded due to incomplete responses, and data from 14 participants were excluded because of IP address duplication, which left 297 in the final sample. The participants received 80 JPY for the 15 min survey.</p>
<p>We also collected 166 undergraduate students from the same university as in Study 1 as the student sample. However, 10 of the participants were excluded from the analysis for the following reasons: incomplete response &#x0003D; 5 participants, IP address duplication &#x0003D; 1 participant, and failure to choose &#x0201C;student&#x0201D; as the current status at demographic question &#x0003D; 4 participants.</p>
<p>The sample size was decided based on the same rationale as Survey 1. We also conducted power analysis to determine sufficient sample size using an alpha of 0.05, a power of 0.8, effect size (<italic>d</italic> &#x0003D; 0.28), and two tails. The effect size was calculated based on the difference in performance of CRT score between the MTurk and the student participants that was reported by Goodman et al. (<xref ref-type="bibr" rid="B11">2013</xref>, Study2). Based on the aforementioned assumptions, the desired sizes for the first and the second sample were 292 and 154. Therefore, the present survey collected the sufficient number of participants.</p>
</sec>
<sec>
<title>Materials and procedure</title>
<p>In this survey, the participants were presented with five tasks that measured their thinking disposition, reasoning and judgment biases, and attention to instructions: CRT, probabilistic reasoning, syllogism, anchoring-and-adjustment, and IMC. A 2 (sample; CW and UNIV) &#x000D7; 2 (IMC order; first and last) factorial design was adopted.</p>
<p>After the participants read general instructions and provided their consent, those who were assigned to the IMC-first order (<italic>N</italic> &#x0003D; 77 for UNIV, and <italic>N</italic> &#x0003D; 155 for CW) answered IMC questions (sports participation task derived from Oppenheimer et al., <xref ref-type="bibr" rid="B23">2009</xref>). The participants were presented with 11 alternatives that consisted of 10 sports and 1 &#x0201C;other&#x0201D; option, and they were asked to choose activities in which they engaged regularly. However, the instructions also asked the participants to &#x0201C;select other&#x0201D; and enter &#x0201C;I read the instructions&#x0201D; to the text box at the end. If the participants carefully read and followed the instructions, they were scored as &#x0201C;correct.&#x0201D; The participants who were assigned to the IMC-last order (<italic>N</italic> &#x0003D; 79 for student, <italic>N</italic> &#x0003D; 142 for CW) answered IMC questions after the other reasoning tasks.</p>
<p>Next, we administered a three-item version of the CRT (Frederick, <xref ref-type="bibr" rid="B9">2005</xref>). The participants were asked to enter their response into a text entry box. Subsequently, the participants responded to a probabilistic reasoning task (Toplak et al., <xref ref-type="bibr" rid="B33">2011</xref>). In this task, the participants were asked to imagine that they were presented with two trays of black and white <italic>go</italic> stones<xref ref-type="fn" rid="fn0002"><sup>2</sup></xref>: a large tray with 100 <italic>go</italic> stones (8 black and 92 white) and a small tray with 10 stones (1 black and 9 white). The participants were also told to imagine that if they drew a black stone, they would win 300 JPY. The participants showed their preference by clicking one of two radio buttons that were labeled &#x0201C;small tray&#x0201D; or &#x0201C;large tray.&#x0201D; The rational choice of this task was the small tray because the chance of winning a prize was higher in the small (1/10) rather than in the large (8/100) tray. However, people often neglect the denominator and prefer the large number of black stones in the large tray.</p>
<p>Then, the participants were presented with the logical reasoning task. They were presented with eight syllogisms one at a time, and they answered by clicking either &#x0201C;True&#x0201D; or &#x0201C;False&#x0201D; on each conclusion. Following the syllogisms, an anchoring-and-adjustment task that was adopted from Goodman et al. (<xref ref-type="bibr" rid="B11">2013</xref>) was administered. The participants entered the last two digits of their phone number, they show whether the number of countries in Africa is larger or smaller than that number, and they estimated the exact number of African countries. Finally, we probed whether the participants had previously experienced each of the six tasks. The participants also answered the same demographic questions that were used in Survey 1.</p>
</sec>
</sec>
<sec>
<title>Results</title>
<p>In the following analysis, the participants who answered &#x0201C;yes&#x0201D; to the probe question to each task were excluded from the analyses.</p>
<sec>
<title>Demographics</title>
<p>Table <xref ref-type="table" rid="T1">1</xref> summarizes the demographic properties of both samples. Similar to Survey 1, the CW participants were significantly different from the students in their age, <italic>M</italic><sub>C</sub> &#x0003D; 36.6 vs. <italic>M</italic><sub>U</sub> &#x0003D; 19.7, <italic>t</italic><sub>(451)</sub> &#x0003D; 22.8, <italic>p</italic> &#x0003C; 0.001, <italic>d</italic> &#x0003D; 2.26; percentage of females, CW &#x0003D; 62.3% vs. UNIV &#x0003D; 46.2%, <inline-formula><mml:math id="M3"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 10.8, <italic>p</italic> &#x0003C; 0.001, and median level of education, Mdn<sub>C</sub> &#x0003D; &#x0201C;associate degree,&#x0201D; Mdn<sub>U</sub> &#x0003D; &#x0201C;high school,&#x0201D; Wilcoxon&#x00027;s <italic>Z</italic> &#x0003D; 9.7, <italic>p</italic> &#x0003C; 0.001. Furthermore, the CW participants had longer work experience than the students, <italic>M</italic><sub>C</sub> &#x0003D; 12.4 vs. <italic>M</italic><sub>U</sub> &#x0003D; 3.0, <italic>t</italic><sub>(271)</sub> &#x0003D; 5.1, <italic>p</italic> &#x0003C; 0.001, <italic>d</italic> &#x0003D; 1.06.</p>
</sec>
<sec>
<title>IMC performance</title>
<p>Table <xref ref-type="table" rid="T3">3</xref> summarizes the performance of attentional check and the other reasoning tasks. Three of the students and six of the workers were excluded from the following analysis because they answered yes to the probe question. The percentages of the participants who successfully passed the IMC are shown in Table <xref ref-type="table" rid="T3">3</xref>. We conducted a logistic regression analysis to ascertain the effects of sample and presentation order on the pass rate of the IMC inquiry. In this analysis, the independent variables were simultaneously introduced into the model. The logistic regression model was statistically significant, <inline-formula><mml:math id="M4"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>3</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 22.83, <italic>p</italic> &#x0003C; 0.001, Negelkerke&#x00027;s pseudo-<italic>R</italic><sup>2</sup> &#x0003D; 0.07 (Table <xref ref-type="table" rid="T4">4</xref>). The results showed that more CW participants successfully passed the IMC than students, UNIV &#x0003D; 35.3%, CW &#x0003D; 53.6%, odds ratio &#x0003D; 3.44, 95% CI &#x0003D; [1.91, 6.21], <inline-formula><mml:math id="M5"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 16.8, <italic>p</italic> &#x0003C; 0.001. We also found a significant sample &#x000D7; order interaction, odds ratio &#x0003D; 0.40, 95% CI &#x0003D; [0.18, 0.90], <inline-formula><mml:math id="M6"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 4.85, <italic>p</italic> &#x0003D; 0.028. Then, we conducted follow-up logistic regression analyses stratified by sample. The order did not affect performance in the student sample; however, the CW participants performed worse if they performed the IMC at the beginning of the survey than at the end of the survey, IMC-first &#x0003D; 45.5%, IMC-last &#x0003D; 62.8%, odds ratio &#x0003D; 0.49, 95% CI &#x0003D; [0.31, 0.79], <inline-formula><mml:math id="M7"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 8.65, <italic>p</italic> &#x0003D; 0.003.</p>
<table-wrap position="float" id="T3">
<label>Table 3</label>
<caption><p><bold>Performance of attentional check and reasoning tasks (Survey 2)</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th/>
<th/>
<th valign="top" align="center" colspan="4" style="border-bottom: thin solid #000000;"><bold>UNIV (</bold><italic><bold>N</bold></italic> &#x0003D; <bold>156)</bold></th>
<th valign="top" align="center" colspan="4" style="border-bottom: thin solid #000000;"><bold>CW (</bold><italic><bold>N</bold></italic> &#x0003D; <bold>297)</bold></th>
</tr>
<tr>
<th valign="top" align="left"><bold>Task</bold></th>
<th valign="top" align="left"><bold>IMC performance</bold></th>
<th valign="top" align="center"><bold><italic>n</italic></bold></th>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th valign="top" align="center"><bold>[95% CI]</bold></th>
<th valign="top" align="center"><bold><italic>n</italic></bold></th>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th valign="top" align="center"><bold>[95% CI]</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left" colspan="10" style="background-color:#bbbdc0"><bold>INSTRUCTIONAL MANIPULATION CHECK % CORRECT<xref ref-type="table-fn" rid="TN4"><sup>a</sup></xref><sup>,</sup> <xref ref-type="table-fn" rid="TN7"><sup>d</sup></xref></bold></td>
</tr>
<tr>
<td valign="top" align="left">IMC first</td>
<td/>
<td valign="top" align="center">77</td>
<td valign="top" align="center">37.7%</td>
<td/>
<td/>
<td valign="top" align="center">154</td>
<td valign="top" align="center">45.5%</td>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">IMC last</td>
<td/>
<td valign="top" align="center">76</td>
<td valign="top" align="center">32.9%</td>
<td/>
<td/>
<td valign="top" align="center">137</td>
<td valign="top" align="center">62.8%</td>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left" colspan="10" style="background-color:#bbbdc0"><bold>COGNITIVE REFLECTION TEST<xref ref-type="table-fn" rid="TN5"><sup>b</sup></xref><sup>,</sup> <xref ref-type="table-fn" rid="TN6"><sup>c</sup></xref></bold></td>
</tr>
<tr>
<td valign="top" align="left">IMC first</td>
<td valign="top" align="left">Pass</td>
<td valign="top" align="center">22</td>
<td valign="top" align="center">1.91</td>
<td valign="top" align="center">1.02</td>
<td valign="top" align="center">[1.47, 2.35]</td>
<td valign="top" align="center">62</td>
<td valign="top" align="center">1.53</td>
<td valign="top" align="center">1.00</td>
<td valign="top" align="center">[1.27, 1.80]</td>
</tr>
<tr>
<td/>
<td valign="top" align="left">Failure</td>
<td valign="top" align="center">46</td>
<td valign="top" align="center">1.17</td>
<td valign="top" align="center">1.16</td>
<td valign="top" align="center">[0.87, 1.48]</td>
<td valign="top" align="center">77</td>
<td valign="top" align="center">1.09</td>
<td valign="top" align="center">1.08</td>
<td valign="top" align="center">[0.85, 1.33]</td>
</tr>
<tr>
<td valign="top" align="left">IMC last</td>
<td valign="top" align="left">Pass</td>
<td valign="top" align="center">19</td>
<td valign="top" align="center">0.84</td>
<td valign="top" align="center">1.01</td>
<td valign="top" align="center">[0.36, 1.32]</td>
<td valign="top" align="center">79</td>
<td valign="top" align="center">1.48</td>
<td valign="top" align="center">1.00</td>
<td valign="top" align="center">[1.25, 1.72]</td>
</tr>
<tr>
<td/>
<td valign="top" align="left">Failure</td>
<td valign="top" align="center">40</td>
<td valign="top" align="center">1.20</td>
<td valign="top" align="center">1.07</td>
<td valign="top" align="center">[0.87, 1.53]</td>
<td valign="top" align="center">48</td>
<td valign="top" align="center">0.96</td>
<td valign="top" align="center">1.11</td>
<td valign="top" align="center">[0.66, 1.26]</td>
</tr>
<tr>
<td valign="top" align="left">DN % rational<xref ref-type="table-fn" rid="TN6"><sup>c</sup></xref></td>
<td valign="top" align="left">Pass</td>
<td valign="top" align="center">53</td>
<td valign="top" align="center">73.6%</td>
<td/>
<td/>
<td valign="top" align="center">154</td>
<td valign="top" align="center">73.4%</td>
<td/>
<td/>
</tr>
<tr>
<td/>
<td valign="top" align="left">Failure</td>
<td valign="top" align="center">93</td>
<td valign="top" align="center">64.5%</td>
<td/>
<td/>
<td valign="top" align="center">125</td>
<td valign="top" align="center">59.2%</td>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">Syllogism<xref ref-type="table-fn" rid="TN6"><sup>c</sup></xref></td>
<td valign="top" align="left">Pass</td>
<td valign="top" align="center">47</td>
<td valign="top" align="center">4.49</td>
<td valign="top" align="center">2.58</td>
<td valign="top" align="center">[3.73, 5.13]</td>
<td valign="top" align="center">151</td>
<td valign="top" align="center">3.51</td>
<td valign="top" align="center">2.55</td>
<td valign="top" align="center">[3.14, 3.91]</td>
</tr>
<tr>
<td/>
<td valign="top" align="left">Failure</td>
<td valign="top" align="center">82</td>
<td valign="top" align="center">3.21</td>
<td valign="top" align="center">2.08</td>
<td valign="top" align="center">[2.68, 3.73]</td>
<td valign="top" align="center">118</td>
<td valign="top" align="center">3.31</td>
<td valign="top" align="center">2.37</td>
<td valign="top" align="center">[2.81, 3.73]</td>
</tr>
<tr>
<td valign="top" align="left" colspan="10" style="background-color:#bbbdc0"><bold>ANCHORING</bold></td>
</tr>
<tr>
<td valign="top" align="left">Mean estimation</td>
<td/>
<td valign="top" align="center">147</td>
<td valign="top" align="center">36.51</td>
<td valign="top" align="center">17.60</td>
<td/>
<td valign="top" align="center">276</td>
<td valign="top" align="center">38.71</td>
<td valign="top" align="center">20.26</td>
<td/>
</tr>
<tr>
<td valign="top" align="left"><italic>r</italic><xref ref-type="table-fn" rid="TN8"><sup>e</sup></xref></td>
<td/>
<td/>
<td valign="top" align="center">0.171<xref ref-type="table-fn" rid="TN9"><sup>&#x0002A;</sup></xref></td>
<td/>
<td/>
<td/>
<td valign="top" align="center">0.064</td>
<td/>
<td/>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>Sum of ns may not be equal to total number of sample because the number of participants reporting they have experienced the question was different by means of task. UNIV, student sample; CW, CrowdWorks sample; IMC first, IMC was presented before other tasks; IMC last, IMC was presented after other tasks. DN % rational, percentages of participants who chose small, i.e., high probability of win, tray in denominator neglect bias task. IMC order was pooled for results of other reasoning tasks except for CRT</italic>.</p>
<fn id="TN4">
<label>a</label>
<p><italic>CW participants are statistically different (p &#x0003C; 0.05) from students</italic>,</p></fn>
<fn id="TN5">
<label>b</label>
<p><italic>significant effect of presentation order (p &#x0003C; 0.05)</italic>,</p></fn>
<fn id="TN6">
<label>c</label>
<p><italic>significant effect of IMC performance (p &#x0003C; 0.05)</italic>,</p></fn>
<fn id="TN7">
<label>d</label>
<p><italic>significant sample &#x000D7; order interaction (p &#x0003C; 0.05)</italic>,</p></fn>
<fn id="TN8">
<label>e</label>
<p><italic>Pearson product moment correlation coefficients between the estimated number of African countries and anchor, i.e., last two digits of phone number</italic>.</p></fn>
<fn id="TN9">
<label>&#x0002A;</label>
<p><italic>p &#x0003C; 0.05</italic>.</p></fn>
</table-wrap-foot>
</table-wrap>
<table-wrap position="float" id="T4">
<label>Table 4</label>
<caption><p><bold>Logistic regression analysis predicting the likelihood of passing IMC question by sample and task order, and separate analyses stratified by sample (Survey 2)</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Variables</bold></th>
<th/>
<th/>
<th/>
<th/>
<th/>
<th/>
<th valign="top" align="center" colspan="4" style="border-bottom: thin solid #000000;"><bold>Model evaluation</bold></th>
</tr>
<tr>
<th/>
<th valign="top" align="center"><bold>&#x003B2;</bold></th>
<th valign="top" align="center"><bold><italic>SE</italic> (&#x003B2;)</bold></th>
<th valign="top" align="center"><bold>Wald&#x00027;s &#x003C7;<sup>2</sup></bold></th>
<th valign="top" align="center"><bold><italic>p</italic></bold></th>
<th valign="top" align="center"><bold>Odds ratio</bold></th>
<th valign="top" align="center"><bold>[95% CI]</bold></th>
<th valign="top" align="center"><bold>&#x003C7;<sup>2</sup></bold></th>
<th valign="top" align="center"><bold><italic>df</italic></bold></th>
<th valign="top" align="center"><bold><italic>p</italic></bold></th>
<th valign="top" align="center"><bold>pseudo-<italic>R</italic><sup>2</sup></bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">[Overall model]</td>
<td/>
<td/>
<td/>
<td/>
<td/>
<td/>
<td valign="top" align="center">22.83</td>
<td valign="top" align="center">3</td>
<td valign="top" align="center">&#x0003C;0.001</td>
<td valign="top" align="center">0.067</td>
</tr>
<tr>
<td valign="top" align="left">&#x000A0;&#x000A0;&#x000A0;Constant</td>
<td valign="top" align="char" char=".">&#x02212;0.71</td>
<td valign="top" align="center">0.24</td>
<td valign="top" align="center">8.53</td>
<td valign="top" align="center">0.003</td>
<td valign="top" align="center">0.49</td>
<td/>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">&#x000A0;&#x000A0;&#x000A0;Sample (UNIV &#x0003D; 0, CW &#x0003D; 1)</td>
<td valign="top" align="char" char=".">1.24</td>
<td valign="top" align="center">0.30</td>
<td valign="top" align="center">16.80</td>
<td valign="top" align="center">&#x0003C;0.001</td>
<td valign="top" align="center">3.44</td>
<td valign="top" align="center">[1.91, 6.21]</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">&#x000A0;&#x000A0;&#x000A0;IMC Order (Last &#x0003D; 0, First &#x0003D; 1)</td>
<td valign="top" align="char" char=".">0.21</td>
<td valign="top" align="center">0.34</td>
<td valign="top" align="center">0.38</td>
<td valign="top" align="center">0.537</td>
<td valign="top" align="center">1.23</td>
<td valign="top" align="center">[0.63, 2.40]</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">&#x000A0;&#x000A0;&#x000A0;Sample &#x000D7; Order</td>
<td valign="top" align="char" char=".">&#x02212;0.91</td>
<td valign="top" align="center">0.42</td>
<td valign="top" align="center">4.85</td>
<td valign="top" align="center">0.028</td>
<td valign="top" align="center">0.40</td>
<td valign="top" align="center">[0.18, 0.90]</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">[UNIV]</td>
<td/>
<td/>
<td/>
<td/>
<td/>
<td/>
<td valign="top" align="center">0.38</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">0.537</td>
<td valign="top" align="center">0.003</td>
</tr>
<tr>
<td valign="top" align="left">&#x000A0;&#x000A0;&#x000A0;Constant</td>
<td valign="top" align="char" char=".">&#x02212;0.71</td>
<td valign="top" align="center">0.24</td>
<td valign="top" align="center">8.53</td>
<td valign="top" align="center">0.003</td>
<td valign="top" align="center">0.49</td>
<td/>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">&#x000A0;&#x000A0;&#x000A0;IMC Order</td>
<td valign="top" align="char" char=".">0.21</td>
<td valign="top" align="center">0.34</td>
<td valign="top" align="center">0.38</td>
<td valign="top" align="center">0.537</td>
<td valign="top" align="center">1.23</td>
<td valign="top" align="center">[0.63, 2.40]</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">[CW]</td>
<td/>
<td/>
<td/>
<td/>
<td/>
<td/>
<td valign="top" align="center">8.80</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">0.003</td>
<td valign="top" align="center">0.040</td>
</tr>
<tr>
<td valign="top" align="left">&#x000A0;&#x000A0;&#x000A0;Constant</td>
<td valign="top" align="char" char=".">0.52</td>
<td valign="top" align="center">0.18</td>
<td valign="top" align="center">8.74</td>
<td valign="top" align="center">0.003</td>
<td valign="top" align="center">1.69</td>
<td/>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">&#x000A0;&#x000A0;&#x000A0;IMC Order</td>
<td valign="top" align="char" char=".">&#x02212;0.70</td>
<td valign="top" align="center">0.24</td>
<td valign="top" align="center">8.65</td>
<td valign="top" align="center">0.003</td>
<td valign="top" align="center">0.49</td>
<td valign="top" align="center">[0.31, 0.79]</td>
<td/>
<td/>
<td/>
<td/>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>UNIV, student sample; CW, CrowdWorks sample; pseudo-R<sup>2</sup>, Negelkerke&#x00027;s R<sup>2</sup></italic>.</p>
</table-wrap-foot>
</table-wrap>
</sec>
<sec>
<title>Cognitive reflection test</title>
<p>We excluded 29 students and 31 workers from the following analysis owing to their previous experience with the task. A 2 (sample) &#x000D7; 2 (IMC order) &#x000D7; 2 (IMC performance; pass vs. failure) ANOVA revealed significant main effects of the IMC performance and task order, <italic>F</italic>s<sub>(1, 385)</sub> &#x0003D; 7.8, 6.5, <italic>MSE</italic> &#x0003D; 1.12, <italic>p</italic>s &#x0003D; 0.006, 0.011, &#x003B7;<sup>2</sup>s &#x0003D; 0.02, respectively. An order &#x000D7; performance interaction was also significant, <italic>F</italic><sub>(1, 385)</sub> &#x0003D; 4.4, <italic>p</italic> &#x0003D; 0.036, &#x003B7;<sup>2</sup> &#x0003D; 0.01. The analyses of simple main effects revealed that the simple main effect of IMC order was significant among the participants who passed the IMC, <italic>F</italic><sub>(1, 385)</sub> &#x0003D; 8.8, <italic>p</italic> &#x0003D; 0.003, <italic>MSE</italic> &#x0003D; 1.12, &#x003B7;<sup>2</sup> &#x0003D; 0.02; however, the effect of order was not found among those who failed the IMC (<italic>F</italic> &#x0003C; 1). This result indicates that the participants performed better at the CRT only if they successfully solved the IMC question before the CRT. Furthermore, we found a significant three-way interaction, <italic>F</italic><sub>(1, 385)</sub> &#x0003D; 5.9, <italic>p</italic> &#x0003D; 0.015, <italic>MSE</italic> &#x0003D; 1.12, &#x003B7;<sup>2</sup> &#x0003D; 0.02. An analysis of the simple interaction effects showed a significant order &#x000D7; performance interaction for UNIV, <italic>F</italic><sub>(1, 385)</sub> &#x0003D; 7.37, <italic>p</italic> &#x0003D; 0.007, &#x003B7;<sup>2</sup> &#x0003D; 0.02; however, no effect was found for CW (<italic>F</italic> &#x0003C; 1).</p>
</sec>
<sec>
<title>Denominator neglect and belief bias</title>
<p>The probe analysis excluded 10 students and 18 workers. We conducted logistic regression analyses that predicted the likelihood of high-probability choice by IMC order and performance (see Table <xref ref-type="table" rid="T5">5</xref>). The sample was excluded from the model, as the preliminary analysis failed to show any effect of and interactions with the sample. In the first analysis, IMC order, performance, and an order &#x000D7; performance interaction were simultaneously introduced to the model (Model 1), <inline-formula><mml:math id="M8"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>3</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 8.36, <italic>p</italic> &#x0003D; 0.039, pseudo-<italic>R</italic><sup>2</sup> &#x0003D; 0.03, AIC &#x0003D; 27.84. However, we failed to find any significant effects of the predictors. Then, we introduced IMC performance solely into the model (Model 2). This model showed a slightly good fit, <inline-formula><mml:math id="M9"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 6.94, <italic>p</italic> &#x0003D; 0.008, pseudo-<italic>R</italic><sup>2</sup> &#x0003D; 0.02, AIC &#x0003D; 15.32. As is shown in Table <xref ref-type="table" rid="T5">5</xref>, the participants who successfully passed the IMC tended to choose a higher probability alternative than those who failed at the IMC, 73.4% vs. 61.5%, odds ratio &#x0003D; 1.73, 95% CI &#x0003D; [1.15, 2.61], <inline-formula><mml:math id="M10"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 6.84, <italic>p</italic> &#x0003D; 0.009.</p>
<table-wrap position="float" id="T5">
<label>Table 5</label>
<caption><p><bold>Logistic regression analyses predicting the denominator neglect bias (Survey 2)</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th/>
<th valign="top" align="center" colspan="5" style="border-bottom: thin solid #000000;"><bold>Model 1</bold></th>
<th valign="top" align="center" colspan="5" style="border-bottom: thin solid #000000;"><bold>Model 2</bold></th>
</tr>
<tr>
<th valign="top" align="left"><bold>Variables</bold></th>
<th valign="top" align="center"><bold>&#x003B2;</bold></th>
<th valign="top" align="center"><bold><italic>SE</italic> (&#x003B2;)</bold></th>
<th valign="top" align="center"><bold>Wald&#x00027;s &#x003C7;<sup>2</sup></bold></th>
<th valign="top" align="center"><bold><italic>p</italic></bold></th>
<th valign="top" align="center"><bold>Odds ratio [95% CI]</bold></th>
<th valign="top" align="center"><bold>&#x003B2;</bold></th>
<th valign="top" align="center"><bold><italic>SE</italic> (&#x003B2;)</bold></th>
<th valign="top" align="center"><bold>Wald&#x00027;s &#x003C7;<sup>2</sup></bold></th>
<th valign="top" align="center"><bold><italic>p</italic></bold></th>
<th valign="top" align="center"><bold>Odds ratio [95% CI]</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Constant</td>
<td valign="top" align="center">0.66</td>
<td valign="top" align="center">0.22</td>
<td valign="top" align="center">9.23</td>
<td valign="top" align="center">0.002</td>
<td valign="top" align="center">1.94</td>
<td valign="top" align="center">0.47</td>
<td valign="top" align="center">0.14</td>
<td valign="top" align="center">11.26</td>
<td valign="top" align="center">&#x0003C;0.001</td>
<td valign="top" align="center">1.60</td>
</tr>
<tr>
<td valign="top" align="left">IMC Order (Last &#x0003D; 0, First &#x0003D; 1)</td>
<td valign="top" align="center">&#x02212;0.34</td>
<td valign="top" align="center">0.28</td>
<td valign="top" align="center">1.40</td>
<td valign="top" align="center">0.236</td>
<td valign="top" align="center">0.71 [0.41, 1.25]</td>
<td/>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">IMC Performance (Failure &#x0003D; 0, Pass &#x0003D; 1)</td>
<td valign="top" align="center">0.37</td>
<td valign="top" align="center">0.31</td>
<td valign="top" align="center">1.42</td>
<td valign="top" align="center">0.233</td>
<td valign="top" align="center">1.44 [0.79, 2.63]</td>
<td valign="top" align="center">0.55</td>
<td valign="top" align="center">0.21</td>
<td valign="top" align="center">6.84</td>
<td valign="top" align="center">0.009</td>
<td valign="top" align="center">1.73 [1.15, 2.61]</td>
</tr>
<tr>
<td valign="top" align="left">Order &#x000D7; Performance</td>
<td valign="top" align="center">0.31</td>
<td valign="top" align="center">0.42</td>
<td valign="top" align="center">0.55</td>
<td valign="top" align="center">0.460</td>
<td valign="top" align="center">1.37 [0.60, 3.14]</td>
<td/>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left" colspan="11" style="background-color:#bbbdc0"><bold>MODEL EVALUATION</bold></td>
</tr>
<tr>
<td valign="top" align="left">&#x003C7;<sup>2</sup></td>
<td valign="top" align="center">8.36</td>
<td/>
<td/>
<td/>
<td/>
<td valign="top" align="center">6.94</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left"><italic>df</italic></td>
<td valign="top" align="center">3</td>
<td/>
<td/>
<td/>
<td/>
<td valign="top" align="center">1</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left"><italic>p</italic></td>
<td valign="top" align="center">0.039</td>
<td/>
<td/>
<td/>
<td/>
<td valign="top" align="center">0.008</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">Negelkerke&#x00027;s pseudo-<italic>R</italic><sup>2</sup></td>
<td valign="top" align="center">0.027</td>
<td/>
<td/>
<td/>
<td/>
<td valign="top" align="center">0.023</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">AIC</td>
<td valign="top" align="center">27.84</td>
<td/>
<td/>
<td/>
<td/>
<td valign="top" align="center">15.32</td>
<td/>
<td/>
<td/>
<td/>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>Low probability choice was coded as 0, high probability (rational) choice as 1</italic>.</p>
</table-wrap-foot>
</table-wrap>
<p>Next, we conducted a three-way ANOVA on the number of correctly solved syllogisms. In this analysis, 27 students and 28 workers were excluded because of their previous experience with the task. The results showed a significant main effect of IMC performance, <italic>M</italic><sub>PASSED</sub> &#x0003D; 3.7 vs. <italic>M</italic><sub>FAILED</sub> &#x0003D; 3.3, <italic>F</italic><sub>(1, 390)</sub> &#x0003D; 7.6, <italic>p</italic> &#x0003D; 0.006, <italic>MSE</italic> &#x0003D; 5.82, &#x003B7;<sup>2</sup> &#x0003D; 0.02. We also found marginal main effects of order and sample &#x000D7; performance interaction, <italic>F</italic>s<sub>(1, 390)</sub> &#x0003D; 2.8, 3.2, <italic>p</italic>s &#x0003D; 0.094, 0.072, &#x003B7;<sup>2</sup>s &#x0003D; 0.01, respectively. <italic>Post-hoc</italic> analyses indicated that the students who successfully passed the IMC scored higher in the syllogism task than those who failed, <italic>F</italic><sub>(1, 390)</sub> &#x0003D; 7.7, <italic>p</italic> &#x0003D; 0.006, <italic>MSE</italic> &#x0003D; 5.82, &#x003B7;<sup>2</sup> &#x0003D; 0.02; on the other hand, the performance of attentional check was not associated with the solution of the syllogisms for the CW workers.</p>
</sec>
<sec>
<title>Anchoring and adjustment</title>
<p>We excluded 9 students and 18 workers from the following analyses owing to their previous experience. We also excluded three workers who estimated extremely large numbers (&#x0003E;mean &#x0002B;3 <italic>SD</italic>) of African countries, such as 350. To examine the anchoring-and-adjustment effect, we regressed the participants&#x00027; estimates on their mean-centered phone number (i.e., <italic>anchor</italic>), sample, and an anchor &#x000D7; sample interaction, and we found a marginal positive association between estimates and anchors, &#x003B2; &#x0003D; 0.16, <italic>p</italic> &#x0003D; 0.061; however, this initial model failed to show a good fit, adjusted <italic>R</italic><sup>2</sup> &#x0003D; 0.007, <italic>F</italic><sub>(3, 419)</sub> &#x0003D; 2.0, <italic>p</italic> &#x0003D; 0.112. Then, we conducted additional separate regression analyses for each of the two samples. On the one hand, we found that the anchor was a significant predictor for the students, &#x003B2; &#x0003D; 0.17, <italic>p</italic> &#x0003D; 0.038, adjusted <italic>R</italic><sup>2</sup> &#x0003D; 0.023. On the other hand, this was not the case in the CW sample, &#x003B2; &#x0003D; 0.06, <italic>p</italic> &#x0003D; 0.288, adjusted <italic>R</italic><sup>2</sup> &#x0003C; 0.001. These results suggest that students are more prone to the anchoring-and-adjustment bias than CW participants.</p>
<p>The present study allowed the participants to answer the survey using their PC or mobile device at their convenience; therefore, they might have searched for accurate answers on the Internet. Seven students (4.8%) and 26 workers (9.4%) &#x0201C;estimated&#x0201D; the correct number of African countries that could be found on a Wikipedia query (56 countries) or a document by the Ministry of Foreign Affairs of Japan (54 countries). The percentage of correctly &#x0201C;estimated&#x0201D; participants was slightly higher in CW, but the difference was relatively small, <inline-formula><mml:math id="M11"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 2.89, <italic>p</italic> &#x0003D; 0.089. In addition, the overall percentage of this type of cheating was similar to that found in a previous study (10%; Goodman et al., <xref ref-type="bibr" rid="B11">2013</xref>).</p>
</sec>
<sec>
<title>Previous experience with the commonly used tasks</title>
<p>The number of excluded participants owing to previous experience was different across tasks. We compared the proportion of participants who answered &#x0201C;Yes&#x0201D; to the probe question between student and CW participants. A series of Chi-square tests revealed that students were more likely to have the experience of participating CRT (% of excluded, UNIV &#x0003D; 18.6 vs. CW &#x0003D; 10.4) and syllogism task (UNIV &#x0003D; 17.3% vs. CW &#x0003D; 9.4%), &#x003C7;<sup>2</sup>s<sub>(1)</sub> &#x0003D; 5.92, 5.95, <italic>p</italic>s &#x0003C; 0.02, respectively. However, the proportion of excluded participants was not different in IMC, denominator neglect, and anchoring-and-adjustment tasks, &#x003C7;<sup>2</sup>s &#x0003C; 1.</p>
</sec>
</sec>
<sec>
<title>Discussion</title>
<p>Survey 2 showed that the workers and students did not differ in their overall performance in the CRT and probabilistic and logical reasoning. In addition, the CW participants were less prone to anchoring-and-adjustment bias than the non-crowdsourced sample (similar results were reported by Goodman et al., <xref ref-type="bibr" rid="B11">2013</xref>), and this may be partly because the CW participants obtained precisely correct responses from Internet searches. These results are generally consistent with the previous validation studies. However, the present results also show that the students seem not to read instructions carefully in comparison with CW participants and that presentation order has a limited impact on the overall performance of IMC. Contrary to Hauser and Schwarz (<xref ref-type="bibr" rid="B13">2015</xref>), mere exposure to IMC did little to improve subsequent reasoning tasks that required systematic thinking. A rather successful solution to the trap question was associated with the successful solution of the other reasoning tasks. Notably, the present participants showed poorer performance on the IMC than those in previous research using a MTurk sample. We suspect that this was partly because the majority of the students accessed the survey site using their mobile device (i.e., smartphones and tablets). We discuss this issue in the section on Survey 3.</p>
<p>The number of participants with prior experience in the task was different between two samples only for CRT and syllogism task. Furthermore, the percentages of previously exposed participants were somewhat lower than MTurk workers. For example, Chandler et al. (<xref ref-type="bibr" rid="B3">2014</xref>) indicated that the proportion of workers who reported participating in commonly used paradigms, such as Trolley problem and Prisoner&#x00027;s dilemma was ranged from approximately 10 to 60% (Trolley problem &#x0003D; 30%, Prisoner&#x00027;s dilemma &#x0003D; 56%) except for Dictator&#x00027;s game (0%). The percentages of prior exposure among the present online participants were ranged from 2.0% (IMC) to 10.4% (CRT). Therefore, CW participants seem to be more na&#x000EF;ve compared with MTurk workers.</p>
</sec>
</sec>
<sec id="s5">
<title>Survey 3: effect of device type on attentional checks</title>
<p>Survey 2 indicated that the participants were less attentive to the instructions of the task. We suspected that this may have partly been caused by the fact that many of the participants, particularly the students, reached the survey site using small screen devices, such as smartphones. However, because we did not collect information regarding device or browser type in Survey 2, whether small screen devices compared to larger screen devices lead to less attentive responses remains unclear. In Survey 3, we examined whether the use of small screen devices facilitated failure in attentional checks and poorer performance on the other reasoning tasks. As in Survey 2, we also investigated whether the order of the IMC question affected performance in the subsequent reasoning tasks.</p>
<sec>
<title>Method</title>
<sec>
<title>Participants</title>
<p>Similar to Surveys 1 and 2, we recruited participants from CrowdWorks; however, in this survey, we decided to hide the task from the workers if their acceptance rate was less than 95%. We collected 205 participants from CrowdWorks; however, 38 of the participants were excluded from analysis for the following reasons: providing incomplete response (3 participants), searching for correct answers or responding randomly (see Materials and Procedure Section; 34 participants), and participating both in mobile and PC surveys (1 participant). One participant in the mobile condition was also excluded because the device type information indicated that he or she had participated in the survey using a PC. Consequently, 167 participants remained in the final sample. (Mobile group, <italic>N</italic> &#x0003D; 81, mean age &#x0003D; 33.4, female &#x0003D; 66.7%; PC group, <italic>N</italic> &#x0003D; 85, mean age &#x0003D; 37.1, female &#x0003D; 44.7%). The participants received 80 JPY for their participation in the survey.</p>
</sec>
<sec>
<title>Materials and procedure</title>
<p>The tasks and the procedure were almost identical to those of Survey 2 except that the anchoring-and-adjustment task was omitted from this survey. We posted two different CW tasks that were designed for two experimental conditions (device type: Mobile and PC). The participants were asked to choose one of two tasks appropriate for their device. And they were also asked not to participate twice. Both of the tasks consisted of the same general instructions, and the link to the online survey was administered by Qualtrics. The two tasks differed in terms of the following device-specific instructions. The instructions for the mobile condition asked the participants to use their mobile devices (smartphone or tablet) and not to use a PC. However, the participants in the PC condition were asked to take the survey using their PC. In addition, we collected device-type information (e.g., OS, Browser and its version, and screen resolution; these data were collected by Meta Info question of Qualtrics) to prevent those who accessed with inappropriate devices from participating in the survey. At the end of the survey, the participants were presented with two probe questions that asked whether they searched for any correct answers or responded randomly during the survey. Those who answered yes to at least one probe question were excluded from the following analyses.</p>
</sec>
</sec>
<sec>
<title>Results and discussion</title>
<sec>
<title>Demographics</title>
<p>Table <xref ref-type="table" rid="T6">6</xref> shows the demographic status and performance of the four tasks. The participants in the mobile group were younger than those in the PC group, <italic>t</italic><sub>(164)</sub> &#x0003D; 2.6, <italic>p</italic> &#x0003D; 0.009, <italic>d</italic> &#x0003D; 0.41; and the percentage of females was higher in the mobile group than in the PC group, <inline-formula><mml:math id="M12"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 8.1, <italic>p</italic> &#x0003D; 0.004.</p>
<table-wrap position="float" id="T6">
<label>Table 6</label>
<caption><p><bold>Demographic results and performance of reasoning tasks as a function of device type and IMC performance (Survey 3)</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Variable</bold></th>
<th/>
<th valign="top" align="center" colspan="4" style="border-bottom: thin solid #000000;"><bold>Mobile (</bold><italic><bold>N</bold></italic> &#x0003D; <bold>81)</bold></th>
<th valign="top" align="center" colspan="4" style="border-bottom: thin solid #000000;"><bold>PC (</bold><italic><bold>N</bold></italic> &#x0003D; <bold>85)</bold></th>
</tr>
<tr>
<th/>
<th/>
<th valign="top" align="center"><bold><italic>n</italic></bold></th>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th valign="top" align="center"><bold>[95% CI]</bold></th>
<th valign="top" align="center"><bold><italic>n</italic></bold></th>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th valign="top" align="center"><bold>[95% CI]</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Age<xref ref-type="table-fn" rid="TN10"><sup>a</sup></xref></td>
<td/>
<td/>
<td valign="top" align="center">33.4</td>
<td valign="top" align="center">8.99</td>
<td/>
<td/>
<td valign="top" align="center">37.1</td>
<td valign="top" align="center">9.07</td>
<td/>
</tr>
<tr>
<td valign="top" align="left">Female %<xref ref-type="table-fn" rid="TN10"><sup>a</sup></xref></td>
<td/>
<td/>
<td valign="top" align="center">66.7%</td>
<td/>
<td/>
<td/>
<td valign="top" align="center">44.7%</td>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">% Passed IMC<xref ref-type="table-fn" rid="TN10"><sup>a</sup></xref></td>
<td valign="top" align="left">IMC first</td>
<td valign="top" align="center">44</td>
<td valign="top" align="center">72.7%</td>
<td/>
<td/>
<td valign="top" align="center">40</td>
<td valign="top" align="center">92.5%</td>
<td/>
<td/>
</tr>
<tr>
<td/>
<td valign="top" align="left">IMC last</td>
<td valign="top" align="center">37</td>
<td valign="top" align="center">64.9%</td>
<td/>
<td/>
<td valign="top" align="center">45</td>
<td valign="top" align="center">93.3%</td>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">CRT</td>
<td valign="top" align="left">Passed IMC</td>
<td valign="top" align="center">52</td>
<td valign="top" align="center">1.54</td>
<td valign="top" align="center">1.11</td>
<td valign="top" align="center">[1.24, 1.83]</td>
<td valign="top" align="center">64</td>
<td valign="top" align="center">1.63</td>
<td valign="top" align="center">1.02</td>
<td valign="top" align="center">[1.35, 1.88]</td>
</tr>
<tr>
<td/>
<td valign="top" align="left">Failed IMC</td>
<td valign="top" align="center">21</td>
<td valign="top" align="center">1.05</td>
<td valign="top" align="center">1.07</td>
<td valign="top" align="center">[0.59, 1.51]</td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">1.17</td>
<td valign="top" align="center">1.17</td>
<td valign="top" align="center">[0.31, 2.03]</td>
</tr>
<tr>
<td valign="top" align="left">DN % rational<xref ref-type="table-fn" rid="TN11"><sup>b</sup></xref></td>
<td valign="top" align="left">Passed IMC</td>
<td valign="top" align="center">54</td>
<td valign="top" align="center">64.8%</td>
<td/>
<td/>
<td valign="top" align="center">70</td>
<td valign="top" align="center">78.6%</td>
<td/>
<td/>
</tr>
<tr>
<td/>
<td valign="top" align="left">Failed IMC</td>
<td valign="top" align="center">24</td>
<td valign="top" align="center">62.5%</td>
<td/>
<td/>
<td valign="top" align="center">6</td>
<td valign="top" align="center">66.7%</td>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">Syllogism</td>
<td valign="top" align="left">Passed IMC</td>
<td valign="top" align="center">56</td>
<td valign="top" align="center">4.34</td>
<td valign="top" align="center">1.94</td>
<td valign="top" align="center">[3.73, 4.94]</td>
<td valign="top" align="center">79</td>
<td valign="top" align="center">4.32</td>
<td valign="top" align="center">2.40</td>
<td valign="top" align="center">[3.79, 4.80]</td>
</tr>
<tr>
<td/>
<td valign="top" align="left">Failed IMC</td>
<td valign="top" align="center">25</td>
<td valign="top" align="center">4.12</td>
<td valign="top" align="center">2.55</td>
<td valign="top" align="center">[3.23, 5.03]</td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">4.33</td>
<td valign="top" align="center">1.63</td>
<td valign="top" align="center">[2.50, 6.16]</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>Sum of ns may not be equal to total number of sample because the number of participants reporting they have experienced the question was different by means of task. IMC first, IMC question was administered at the beginning; IMC last, IMC question were administered after other tasks. DN % rational, percentages of participants who choosing the high probability of win option in denominator neglect bias task. IMC order was pooled for results of the other reasoning tasks</italic>.</p>
<fn id="TN10">
<label>a</label>
<p><italic>mobile participants are significantly different from PC participants</italic>,</p></fn>
<fn id="TN11">
<label>b</label>
<p><italic>mobile participants are marginally different from PC</italic>.</p></fn>
</table-wrap-foot>
</table-wrap>
</sec>
<sec>
<title>Attentional check</title>
<p>The likelihood of passing IMC instructions is shown in Table <xref ref-type="table" rid="T6">6</xref>. We conducted a logistic regression analysis to ascertain the effects of device type and task order on the IMC solution, and we found that the logistic regression model was statistically significant, <inline-formula><mml:math id="M13"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>3</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 17.0, <italic>p</italic> &#x0003C; 0.001, pseudo-<italic>R</italic><sup>2</sup> &#x0003D; 0.16. The participants in the mobile group were less attentive than the participants in the PC group, 69.1 vs. 92.9%, odds ratio &#x0003D; 0.22, 95% CI &#x0003D; [0.06, 0.83], <inline-formula><mml:math id="M14"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 4.94, <italic>p</italic> &#x0003D; 0.026; however, presentation order did not affect performance, first &#x0003D; 82.1% vs. last &#x0003D; 80.5%, odds ratio &#x0003D; 1.14, <inline-formula><mml:math id="M15"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 0.02, <italic>p</italic> &#x0003D; 0.881. Similarly, the device &#x000D7; order interaction was not significant, odds ratio &#x0003D; 0.61, <inline-formula><mml:math id="M16"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 0.26, <italic>p</italic> &#x0003D; 0.612.</p>
</sec>
<sec>
<title>Systematic reasoning tasks</title>
<p>A three-way (device type &#x000D7; IMC order &#x000D7; IMC performance) ANOVA on CRT score was conducted; however, 24 participants (8 in mobile 16 in PC group) were excluded from the analysis because they declared that they had experienced with the CRT before the survey. The results showed a marginally significant effect of IMC performance, <italic>M</italic><sub>PASSED_IMC</sub> &#x0003D; 1.58 vs. <italic>M</italic><sub>FAILED_IMC</sub> &#x0003D; 1.11, <italic>F</italic><sub>(1, 135)</sub> &#x0003D; 3.14, <italic>p</italic> &#x0003D; 0.079, <italic>MSE</italic> &#x0003D; 1.13, &#x003B7;<sup>2</sup> &#x0003D; 0.02; however, none of the other effects were significant, <italic>F</italic>s<sub>(1, 135)</sub> &#x0003C; 2.2, <italic>p</italic>s &#x0003E; 0.14.</p>
<p>Then, we conducted a logistic regression analysis to predict the likelihood of high-probability choice in a probabilistic reasoning task by device type, IMC order and IMC performance. Three participants in the mobile condition and nine participants in the PC condition were excluded from the following analysis due to previous experience with the task. However, this model failed to show a good fit, <inline-formula><mml:math id="M17"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>7</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 4.77, <italic>p</italic> &#x0003D; 0.688, pseudo-<italic>R</italic><sup>2</sup> &#x0003D; 0.04. Instead, the model including device type solely (odds ratio &#x0003D; 0.51, 95% CI &#x0003D; [0.25, 1.05]) showed a slightly good fit, <inline-formula><mml:math id="M18"><mml:msubsup><mml:mrow><mml:mo>&#x003C7;</mml:mo></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> &#x0003D; 3.43, <italic>p</italic> &#x0003D; 0.064, pseudo-<italic>R</italic><sup>2</sup> &#x0003D; 0.03. This result implies that the participants in the mobile condition tended to neglect the denominator.</p>
<p>Finally, the number of correct responses to eight syllogism tasks was submitted to a similar three-way ANOVA; however, neither the main effects for device type, IMC order, IMC performance nor their interactions were significant, <italic>F</italic>s<sub>(1, 158)</sub> &#x0003C; 1.3, <italic>p</italic>s &#x0003E; 0.256.</p>
<p>To summarize, Survey 3 indicated that the participants were less attentive to instructions when they used their mobile devices, i.e., small screen devices. They were also prone to denominator neglect bias. However, type of device does not affect other reasoning tasks that are associated with analytic System 2 thinking. Furthermore, the participants were likely to answer reflectively if they read the instructions carefully. These results indicate that small screen devices hinder the careful reading of instructions; however, this might not necessarily spoil the performance of reasoning tasks.</p>
</sec>
</sec>
</sec>
<sec id="s6">
<title>General discussion</title>
<sec>
<title>The characteristics of crowdworks as a participant pool</title>
<p>In the present study, we compared participants from a Japanese crowdsourcing service with a Japanese student sample in terms of their demography, personality traits, reasoning skills, and attention to instructions. In general, the results were compatible with the existing findings of MTurk validation studies. The present results showed many similarities between the CW workers and the students; however, we also found interesting differences between the two samples.</p>
<p>First, but not surprisingly, the CW workers were older and hence had longer work experience than the students. Second, the CW workers and students were different in some of the personality traits, such as extraversion, conscientiousness, and performance-avoid goal orientation; however, these differences were relatively small and compatible with previous MTurk validation studies (Paolacci et al., <xref ref-type="bibr" rid="B26">2010</xref>; Behrend et al., <xref ref-type="bibr" rid="B1">2011</xref>; Goodman et al., <xref ref-type="bibr" rid="B11">2013</xref>) and other studies on personality (e.g., Kling et al., <xref ref-type="bibr" rid="B16">1999</xref>; Srivastava et al., <xref ref-type="bibr" rid="B31">2003</xref>; Kawamoto et al., <xref ref-type="bibr" rid="B15">2015</xref>; Okada et al., <xref ref-type="bibr" rid="B22">2015</xref>). Third, the CW participants performed better at attentional checks; however, they showed similar responses in the other reasoning tasks. These findings suggest that the Japanese crowdsourcing sample was as reliable a pool as that which included MTurk workers.</p>
<p>We also identified a few important dimensions that differed from previous validation studies. First, the present participants, particularly the students, showed poorer IMC performance than participants in previous studies. The failure rate of the present CW participants (46%) was equal to the failure rate in Oppenheimer et al. (<xref ref-type="bibr" rid="B23">2009</xref>, Study 1); however, it was remarkably higher than that for recent MTurk workers (e.g., Hauser and Schwarz, <xref ref-type="bibr" rid="B13">2015</xref>, <xref ref-type="bibr" rid="B14">2016</xref>). This might be partly due to the device that was used by the participants to answer the survey. If the participants reached the site using a mobile device (i.e., small screen), they were more likely to miss important instructions than those who used larger screen devices. Second, previous exposure to IMC had a limited impact on the improvement of subsequent tricky-seeming tasks. Hauser and Schwarz (<xref ref-type="bibr" rid="B13">2015</xref>) found that answering IMCs prior to a task improved performance on both the CRT and probability reasoning; however, the present results indicated that correctly passing an attentional check, rather than IMC presentation order, was associated with better performance on the other reasoning tasks. Poor performance on the IMC, particularly for the participants who accessed the site with their mobile devices, raises an important methodological issue regarding online data collection. Recently, the penetration rate of smartphone user over population has continued to grow worldwide (e.g., eMarketer, <xref ref-type="bibr" rid="B6">2014</xref>). Moreover, smartphone penetration is much stronger in the younger than the older population. Our findings give the following suggestions. First, IMC could be a useful tool for the elimination of inattentive responses in online studies. Second, it might be wise to ask workers to participate in online studies using relatively large screen devices or to prohibit mobile users from taking the survey. On the other hand, the present findings indicated that IMC performance is moderately associated with other reasoning tasks such as CRT, however the prior exposure to IMC does not necessarily improve performance of subsequent task. Therefore, as one reviewer pointed out, IMC itself might reflect certain personality traits, such as conscientiousness or thoughtfulness, rather than an indicator of a tendency to respond dishonestly. Further, studies would be needed to explore what performances of IMC and other tasks assess.</p>
</sec>
<sec>
<title>Recommendations regarding the use of a non-MTurk participant pool</title>
<p>It is important to note that the present study showed both commonalities and differences between CW workers and the Japanese student sample, which was compatible with the existing literature comprising MTurk validation studies. Despite a few inconsistencies, the present study suggested that online data collection using non-MTurk crowdsourcing services remains a promising approach for behavioral research.</p>
<p>However, at the same time, we recommend that researchers consider the following issues if they collect empirical data from non-MTurk crowdsourcing studies. First, the language that is used in CrowdWorks is limited to Japanese. Therefore, a solid level of language skill is required for both the researchers and the participants to conduct or participate in online surveys with this platform. It may be an obstacle for researchers who are not literate in the Japanese language, and this may also be the case for other crowdsourcing services in which the majority of potential workers are not literate in English. However, in other words, it is also a good opportunity to encourage researchers with different cultural backgrounds to conduct cooperative studies.</p>
<p>Second, MTurk provides a useful command-line interface and API that are designed to control HITs including the ability to obtain a worker&#x00027;s ID. Conversely, CrowdWorks provides only a web-based graphical interface to requesters. This may not necessarily be a disadvantage, since researchers can download the data that includes workers&#x00027; ID and the survey completion code entered by individual workers from the CrowdWorks website. Therefore, if researchers allocate the unique completion code to each participant, they can examine whether a certain participant has participated in their own surveys before when the naivety in sample is essential. However, it is still impossible to identify whether a certain participant already took similar surveys or experiments that have been administered by other researchers. If the survey includes widely used tasks, such as the CRT, it is helpful to ask participants whether they have already answered previous versions of such tasks. As suggested by several previous studies (e.g., Chandler et al., <xref ref-type="bibr" rid="B3">2014</xref>, <xref ref-type="bibr" rid="B4">2015</xref>; Stewart et al., <xref ref-type="bibr" rid="B32">2015</xref>), data from nonna&#x000EF;ve online participants may threaten the quality of data. Although the present study suggested that the CrowdWorks workers are more na&#x000EF;ve compared to MTurk workers for now, a growing usage of this sample will lead some active workers to be <italic>professional</italic> participants, as were MTurk workers. Future investigations should explore whether and how multiple participation across similar surveys might endanger the quality of data from CrowdWorks (and other crowdsourcing) participants.</p>
<p>Third, MTurk workers sometimes receive very little compensation to complete HITs (e.g., $.10 for a 5 min survey) compared to that received in traditional laboratory research (for recent ethical questions concerning online studies, see Gleibs, <xref ref-type="bibr" rid="B10">2016</xref>). In this study, we paid 50 or 80 JPY (approximately 0.40 to 0.70$) for 10&#x02013;15 min surveys. This rate was relatively higher than that for a typical MTurk study but was still less than the lowest wage of paid workers in the local city (764JPY per hour). It has been shown that data quality does not seem to be impaired by the amount of payment (e.g., Buhrmester et al., <xref ref-type="bibr" rid="B2">2011</xref>; Paolacci and Chandler, <xref ref-type="bibr" rid="B25">2014</xref>); however, the research community may be called on to establish guidelines for ethically valid compensation for participation in surveys.</p>
</sec>
</sec>
<sec id="s7">
<title>Author contributions</title>
<p>YM designed the study, lead the data collection and analysis, and was a main author. KN co-lead the design, assisted in the analysis and interpretation, and was a contributing author. AN assisted in the data collection and contributed with the manuscript drafting. RH co-lead the data collection and assisted in the design of the study.</p>
<sec>
<title>Conflict of interest statement</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
</sec>
</body>
<back>
<ack><p>This research was financially supported by the Special Group Research Grant (Year 2015) from Hokusei Gakuen University.</p>
</ack>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Behrend</surname> <given-names>T. S.</given-names></name> <name><surname>Sharek</surname> <given-names>D. J.</given-names></name> <name><surname>Meade</surname> <given-names>A. W.</given-names></name> <name><surname>Wiebe</surname> <given-names>E. N.</given-names></name></person-group> (<year>2011</year>). <article-title>The viability of crowdsourcing for survey research</article-title>. <source>Behav. Res. Methods</source> <volume>43</volume>, <fpage>800</fpage>&#x02013;<lpage>813</lpage>. <pub-id pub-id-type="doi">10.3758/s13428-011-0081-0</pub-id><pub-id pub-id-type="pmid">21437749</pub-id></citation>
</ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Buhrmester</surname> <given-names>M.</given-names></name> <name><surname>Kwang</surname> <given-names>T.</given-names></name> <name><surname>Gosling</surname> <given-names>S. D.</given-names></name></person-group> (<year>2011</year>). <article-title>Amazon&#x00027;s mechanical turk: a new source of inexpensive, yet high-quality, data?</article-title> <source>Perspect. Psychol. Sci.</source> <volume>6</volume>, <fpage>3</fpage>&#x02013;<lpage>5</lpage>. <pub-id pub-id-type="doi">10.1177/1745691610393980</pub-id><pub-id pub-id-type="pmid">26162106</pub-id></citation>
</ref>
<ref id="B3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chandler</surname> <given-names>J.</given-names></name> <name><surname>Mueller</surname> <given-names>P.</given-names></name> <name><surname>Paolacci</surname> <given-names>G.</given-names></name></person-group> (<year>2014</year>). <article-title>Nonna&#x000EF;vet&#x000E9; among Amazon Mechanical Turk workers: consequences and solutions for behavioral researchers</article-title>. <source>Behav. Res. Methods</source> <volume>46</volume>, <fpage>112</fpage>&#x02013;<lpage>130</lpage>. <pub-id pub-id-type="doi">10.3758/s13428-013-0365-7</pub-id><pub-id pub-id-type="pmid">23835650</pub-id></citation>
</ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chandler</surname> <given-names>J.</given-names></name> <name><surname>Paolacci</surname> <given-names>G.</given-names></name> <name><surname>Peer</surname> <given-names>E.</given-names></name> <name><surname>Mueller</surname> <given-names>P.</given-names></name> <name><surname>Ratliff</surname> <given-names>K. A.</given-names></name></person-group> (<year>2015</year>). <article-title>Using nonnaive participants can reduce effect sizes</article-title>. <source>Psychol. Sci.</source> <volume>26</volume>, <fpage>1131</fpage>&#x02013;<lpage>1139</lpage>. <pub-id pub-id-type="doi">10.1177/0956797615585115</pub-id><pub-id pub-id-type="pmid">26063440</pub-id></citation>
</ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crump</surname> <given-names>M. J.</given-names></name> <name><surname>McDonnell</surname> <given-names>J. V.</given-names></name> <name><surname>Gureckis</surname> <given-names>T. M.</given-names></name></person-group> (<year>2013</year>). <article-title>Evaluating amazon&#x00027;s mechanical turk as a tool for experimental behavioral research</article-title>. <source>PLoS ONE</source> <volume>8</volume>:<fpage>e57410</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0057410</pub-id><pub-id pub-id-type="pmid">23516406</pub-id></citation>
</ref>
<ref id="B6">
<citation citation-type="web"><person-group person-group-type="author"><collab>eMarketer</collab></person-group> (<year>2014</year>). <source>Worldwide Smartphone Usage to Grow 25% in 2014</source>. Available online at: <ext-link ext-link-type="uri" xlink:href="http://www.emarketer.com/Article/Worldwide-Smartphone-Usage-Grow-25-2014/1010920">http://www.emarketer.com/Article/Worldwide-Smartphone-Usage-Grow-25-2014/1010920</ext-link> (Accessed).</citation>
</ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Estelles-Arolas</surname> <given-names>E.</given-names></name> <name><surname>Gonzalez-Ladron-De-Guevara</surname> <given-names>F.</given-names></name></person-group> (<year>2012</year>). <article-title>Towards an integrated crowdsourcing definition</article-title>. <source>J. Inf. Sci.</source> <volume>38</volume>, <fpage>189</fpage>&#x02013;<lpage>200</lpage>. <pub-id pub-id-type="doi">10.1177/0165551512437638</pub-id></citation>
</ref>
<ref id="B8">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Evans</surname> <given-names>J. S.</given-names></name> <name><surname>Barston</surname> <given-names>J. L.</given-names></name> <name><surname>Pollard</surname> <given-names>P.</given-names></name></person-group> (<year>1983</year>). <article-title>On the conflict between logic and belief in syllogistic reasoning</article-title>. <source>Mem. Cognit.</source> <volume>11</volume>, <fpage>295</fpage>&#x02013;<lpage>306</lpage>. <pub-id pub-id-type="doi">10.3758/BF03196976</pub-id><pub-id pub-id-type="pmid">6621345</pub-id></citation>
</ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Frederick</surname> <given-names>S.</given-names></name></person-group> (<year>2005</year>). <article-title>Cognitive reflection and decision making</article-title>. <source>J. Econ. Perspect.</source> <volume>19</volume>, <fpage>25</fpage>&#x02013;<lpage>42</lpage>. <pub-id pub-id-type="doi">10.1257/089533005775196732</pub-id></citation>
</ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gleibs</surname> <given-names>I. H.</given-names></name></person-group> (<year>2016</year>). <article-title>Are all &#x0201C;research fields&#x0201D; equal? Rethinking practice for the use of data from crowdsourcing market places</article-title>. <source>Behav. Res. Methods</source>. [Epub ahead of print]. <pub-id pub-id-type="doi">10.3758/s13428-016-0789-y</pub-id><pub-id pub-id-type="pmid">27515317</pub-id></citation>
</ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Goodman</surname> <given-names>J. K.</given-names></name> <name><surname>Cryder</surname> <given-names>C. E.</given-names></name> <name><surname>Cheema</surname> <given-names>A.</given-names></name></person-group> (<year>2013</year>). <article-title>Data collection in a flat world: the strengths and weaknesses of mechanical turk samples</article-title>. <source>J. Behav. Decis. Making</source> <volume>26</volume>, <fpage>213</fpage>&#x02013;<lpage>224</lpage>. <pub-id pub-id-type="doi">10.1002/bdm.1753</pub-id></citation>
</ref>
<ref id="B12">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gosling</surname> <given-names>S. D.</given-names></name> <name><surname>Rentfrow</surname> <given-names>P. J.</given-names></name> <name><surname>Swann</surname> <given-names>W. B.</given-names> <suffix>Jr.</suffix></name></person-group> (<year>2003</year>). <article-title>A very brief measure of the big-five personality domains</article-title>. <source>J. Res. Pers.</source> <volume>37</volume>, <fpage>504</fpage>&#x02013;<lpage>528</lpage>. <pub-id pub-id-type="doi">10.1016/S0092-6566(03)00046-1</pub-id></citation>
</ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hauser</surname> <given-names>D. J.</given-names></name> <name><surname>Schwarz</surname> <given-names>N.</given-names></name></person-group> (<year>2015</year>). <article-title>It&#x00027;s a trap! Instructional manipulation checks prompt systematic thinking on &#x0201C;Tricky&#x0201D; tasks</article-title>. <source>SAGE Open</source> <volume>5</volume>:<fpage>2158244015584617</fpage>. <pub-id pub-id-type="doi">10.1177/2158244015584617</pub-id></citation>
</ref>
<ref id="B14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hauser</surname> <given-names>D. J.</given-names></name> <name><surname>Schwarz</surname> <given-names>N.</given-names></name></person-group> (<year>2016</year>). <article-title>Attentive Turkers: MTurk participants perform better on online attention checks than do subject pool participants</article-title>. <source>Behav. Res. Methods</source> <volume>48</volume>, <fpage>400</fpage>&#x02013;<lpage>407</lpage>. <pub-id pub-id-type="doi">10.3758/s13428-015-0578-z</pub-id><pub-id pub-id-type="pmid">25761395</pub-id></citation>
</ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kawamoto</surname> <given-names>T.</given-names></name> <name><surname>Oshio</surname> <given-names>A.</given-names></name> <name><surname>Abe</surname> <given-names>S.</given-names></name> <name><surname>Tsubota</surname> <given-names>Y.</given-names></name> <name><surname>Hirashima</surname> <given-names>T.</given-names></name> <name><surname>Ito</surname> <given-names>H.</given-names></name> <etal/></person-group>. (<year>2015</year>). <article-title>Big Five personality tokusei no nenreisa to seisa: daikibo oudan chousa ni yoru kentou [Age and gender differences of Big Five personality traits in a cross-sectional Japanese sample]</article-title>. <source>Jpn. J. Develop. Psychol.</source> <volume>26</volume>, <fpage>107</fpage>&#x02013;<lpage>122</lpage>.</citation>
</ref>
<ref id="B16">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kling</surname> <given-names>K. C.</given-names></name> <name><surname>Hyde</surname> <given-names>J. S.</given-names></name> <name><surname>Showers</surname> <given-names>C. J.</given-names></name> <name><surname>Buswell</surname> <given-names>B. N.</given-names></name></person-group> (<year>1999</year>). <article-title>Gender differences in self-esteem: a meta-analysis</article-title>. <source>Psychol. Bull.</source> <volume>125</volume>, <fpage>470</fpage>&#x02013;<lpage>500</lpage>. <pub-id pub-id-type="doi">10.1037/0033-2909.125.4.470</pub-id><pub-id pub-id-type="pmid">10414226</pub-id></citation>
</ref>
<ref id="B17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lutz</surname> <given-names>J.</given-names></name></person-group> (<year>2016</year>). <article-title>The validity of crowdsourcing data in studying anger and aggressive behavior</article-title>. <source>Soc. Psychol.</source> <volume>47</volume>, <fpage>38</fpage>&#x02013;<lpage>51</lpage>. <pub-id pub-id-type="doi">10.1027/1864-9335/a000256</pub-id></citation>
</ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Majima</surname> <given-names>Y.</given-names></name></person-group> (<year>2015</year>). <article-title>Belief in pseudoscience, cognitive style and science literacy</article-title>. <source>Appl. Cognit. Psychol.</source> <volume>29</volume>, <fpage>552</fpage>&#x02013;<lpage>559</lpage>. <pub-id pub-id-type="doi">10.1002/acp.3136</pub-id></citation>
</ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Markovits</surname> <given-names>H.</given-names></name> <name><surname>Nantel</surname> <given-names>G.</given-names></name></person-group> (<year>1989</year>). <article-title>The belief-bias effect in the production and evaluation of logical conclusions</article-title>. <source>Mem. Cognit.</source> <volume>17</volume>, <fpage>11</fpage>&#x02013;<lpage>17</lpage>. <pub-id pub-id-type="doi">10.3758/BF03199552</pub-id><pub-id pub-id-type="pmid">2913452</pub-id></citation>
</ref>
<ref id="B20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mason</surname> <given-names>W.</given-names></name> <name><surname>Suri</surname> <given-names>S.</given-names></name></person-group> (<year>2012</year>). <article-title>Conducting behavioral research on Amazon&#x00027;s mechanical Turk</article-title>. <source>Behav. Res. Methods</source> <volume>44</volume>, <fpage>1</fpage>&#x02013;<lpage>23</lpage>. <pub-id pub-id-type="doi">10.3758/s13428-011-0124-6</pub-id><pub-id pub-id-type="pmid">21717266</pub-id></citation>
</ref>
<ref id="B21">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>McCrae</surname> <given-names>R. R.</given-names></name> <name><surname>Costa</surname> <given-names>P. T.</given-names> <suffix>Jr.</suffix></name> <name><surname>Ostendorf</surname> <given-names>F.</given-names></name> <name><surname>Angleitner</surname> <given-names>A.</given-names></name> <name><surname>Hreb&#x000ED;ckov&#x000E1;</surname> <given-names>M.</given-names></name> <name><surname>Avia</surname> <given-names>M. D.</given-names></name> <etal/></person-group>. (<year>2000</year>). <article-title>Nature over nurture: temperament, personality, and life span development</article-title>. <source>J. Pers. Soc. Psychol.</source> <volume>78</volume>, <fpage>173</fpage>&#x02013;<lpage>186</lpage>. <pub-id pub-id-type="doi">10.1037/0022-3514.78.1.173</pub-id><pub-id pub-id-type="pmid">10653513</pub-id></citation>
</ref>
<ref id="B22">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Okada</surname> <given-names>R.</given-names></name> <name><surname>Oshio</surname> <given-names>A.</given-names></name> <name><surname>Mogaki</surname> <given-names>M.</given-names></name> <name><surname>Wakita</surname> <given-names>T.</given-names></name> <name><surname>Namikawa</surname> <given-names>T.</given-names></name></person-group> (<year>2015</year>). <article-title>Nihon-jin ni okeru jison kanjou no seisa ni kansuru meta bunseki [A meta-analysis of gender differences in self-esteem in Japanese]</article-title>. <source>Jpn. J. Pers.</source> <volume>24</volume>, <fpage>49</fpage>&#x02013;<lpage>60</lpage>. <pub-id pub-id-type="doi">10.2132/personality.24.49</pub-id></citation>
</ref>
<ref id="B23">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Oppenheimer</surname> <given-names>D. M.</given-names></name> <name><surname>Meyvis</surname> <given-names>T.</given-names></name> <name><surname>Davidenko</surname> <given-names>N.</given-names></name></person-group> (<year>2009</year>). <article-title>Instructional manipulation checks: detecting satisficing to increase statistical power</article-title>. <source>J. Exp. Soc. Psychol.</source> <volume>45</volume>, <fpage>867</fpage>&#x02013;<lpage>872</lpage>. <pub-id pub-id-type="doi">10.1016/j.jesp.2009.03.009</pub-id></citation>
</ref>
<ref id="B24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Oshio</surname> <given-names>A.</given-names></name> <name><surname>Abe</surname> <given-names>S.</given-names></name> <name><surname>Cutrone</surname> <given-names>P.</given-names></name></person-group> (<year>2012</year>). <article-title>Nihongoban ten item personality inventory (TIPI-J) sakusei no kokoromi [Development, reliability, and validity of the Japanese version of ten item personality inventory (TIPI-J)]</article-title>. <source>Jpn. J. Pers.</source> <volume>21</volume>, <fpage>40</fpage>&#x02013;<lpage>52</lpage>. <pub-id pub-id-type="doi">10.2132/personality.21.40</pub-id></citation>
</ref>
<ref id="B25">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Paolacci</surname> <given-names>G.</given-names></name> <name><surname>Chandler</surname> <given-names>J.</given-names></name></person-group> (<year>2014</year>). <article-title>Inside the turk: understanding mechanical turk as a participant pool</article-title>. <source>Curr. Dir. Psychol. Sci.</source> <volume>23</volume>, <fpage>184</fpage>&#x02013;<lpage>188</lpage>. <pub-id pub-id-type="doi">10.1177/0963721414531598</pub-id></citation>
</ref>
<ref id="B26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Paolacci</surname> <given-names>G.</given-names></name> <name><surname>Chandler</surname> <given-names>J.</given-names></name> <name><surname>Ipeirotis</surname> <given-names>P.</given-names></name></person-group> (<year>2010</year>). <article-title>Running experiments on Amazon mechanical turk</article-title>. <source>Judge. Decis. Making</source> <volume>5</volume>, <fpage>411</fpage>&#x02013;<lpage>419</lpage>. Retrieved from: <ext-link ext-link-type="uri" xlink:href="http://journal.sjdm.org/10/10630a/jdm10630a.pdf">http://journal.sjdm.org/10/10630a/jdm10630a.pdf</ext-link></citation>
</ref>
<ref id="B27">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Peer</surname> <given-names>E.</given-names></name> <name><surname>Brandimarte</surname> <given-names>L.</given-names></name> <name><surname>Samat</surname> <given-names>S.</given-names></name> <name><surname>Acquisti</surname> <given-names>A.</given-names></name></person-group> (<year>2017</year>). <article-title>Beyond the Turk: alternative platforms for crowdsourcing behavioral research</article-title>. <source>J. Exp. Soc. Psychol.</source> <volume>70</volume>, <fpage>153</fpage>&#x02013;<lpage>163</lpage>. <pub-id pub-id-type="doi">10.1016/j.jesp.2017.01.006</pub-id></citation>
</ref>
<ref id="B28">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Richins</surname> <given-names>M. L.</given-names></name></person-group> (<year>2004</year>). <article-title>The material values scale: measurement properties and development of a short form</article-title>. <source>J. Consum. Res.</source> <volume>31</volume>, <fpage>209</fpage>&#x02013;<lpage>219</lpage>. <pub-id pub-id-type="doi">10.1086/383436</pub-id></citation>
</ref>
<ref id="B29">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Robins</surname> <given-names>R. W.</given-names></name> <name><surname>Trzesniewski</surname> <given-names>K. H.</given-names></name> <name><surname>Tracy</surname> <given-names>J. L.</given-names></name> <name><surname>Gosling</surname> <given-names>S. D.</given-names></name> <name><surname>Potter</surname> <given-names>J.</given-names></name></person-group> (<year>2002</year>). <article-title>Global self-esteem across the life span</article-title>. <source>Psychol. Aging</source> <volume>17</volume>, <fpage>423</fpage>&#x02013;<lpage>434</lpage>. <pub-id pub-id-type="doi">10.1037/0882-7974.17.3.423</pub-id><pub-id pub-id-type="pmid">12243384</pub-id></citation>
</ref>
<ref id="B30">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Rosenberg</surname> <given-names>M.</given-names></name></person-group> (<year>1965</year>). <source>Society and the Adolescent Self-Image</source>. <publisher-loc>Princeton, NJ</publisher-loc>: <publisher-name>Princeton University Press</publisher-name>.</citation>
</ref>
<ref id="B31">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Srivastava</surname> <given-names>S.</given-names></name> <name><surname>John</surname> <given-names>O. P.</given-names></name> <name><surname>Gosling</surname> <given-names>S. D.</given-names></name> <name><surname>Potter</surname> <given-names>J.</given-names></name></person-group> (<year>2003</year>). <article-title>Development of personality in early and middle adulthood: set like plaster or persistent change?</article-title> <source>J. Pers. Soc. Psychol.</source> <volume>84</volume>, <fpage>1041</fpage>&#x02013;<lpage>1053</lpage>. <pub-id pub-id-type="doi">10.1037/0022-3514.84.5.1041</pub-id><pub-id pub-id-type="pmid">12757147</pub-id></citation>
</ref>
<ref id="B32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Stewart</surname> <given-names>N.</given-names></name> <name><surname>Ungemach</surname> <given-names>C.</given-names></name> <name><surname>Harris</surname> <given-names>A. J. L.</given-names></name> <name><surname>Bartels</surname> <given-names>D. M.</given-names></name> <name><surname>Newell</surname> <given-names>B. R.</given-names></name> <name><surname>Paolacci</surname> <given-names>G.</given-names></name> <etal/></person-group>. (<year>2015</year>). <article-title>The average laboratory samples a population of 7,300 Amazon Mechanical Turk workers</article-title>. <source>Judge. Decis. Making</source> <volume>10</volume>, <fpage>479</fpage>&#x02013;<lpage>491</lpage>. Retrieved from: <ext-link ext-link-type="uri" xlink:href="http://journal.sjdm.org/14/14725/jdm14725.pdf">http://journal.sjdm.org/14/14725/jdm14725.pdf</ext-link></citation>
</ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Toplak</surname> <given-names>M. E.</given-names></name> <name><surname>West</surname> <given-names>R. F.</given-names></name> <name><surname>Stanovich</surname> <given-names>K. E.</given-names></name></person-group> (<year>2011</year>). <article-title>The cognitive reflection test as a predictor of performance on heuristics-and-biases tasks</article-title>. <source>Mem. Cogn.</source> <volume>39</volume>, <fpage>1275</fpage>&#x02013;<lpage>1289</lpage>. <pub-id pub-id-type="doi">10.3758/s13421-011-0104-1</pub-id><pub-id pub-id-type="pmid">21541821</pub-id></citation>
</ref>
<ref id="B34">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tversky</surname> <given-names>A.</given-names></name> <name><surname>Kahneman</surname> <given-names>D.</given-names></name></person-group> (<year>1974</year>). <article-title>Judgment under uncertainty: heuristics and biases</article-title>. <source>Science</source> <volume>185</volume>, <fpage>1124</fpage>&#x02013;<lpage>1131</lpage>. <pub-id pub-id-type="doi">10.1126/science.185.4157.1124</pub-id><pub-id pub-id-type="pmid">17835457</pub-id></citation>
</ref>
<ref id="B35">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tversky</surname> <given-names>A.</given-names></name> <name><surname>Kahneman</surname> <given-names>D.</given-names></name></person-group> (<year>1981</year>). <article-title>The framing of decisions and the psychology of choice</article-title>. <source>Science</source> <volume>211</volume>, <fpage>453</fpage>. <pub-id pub-id-type="doi">10.1126/science.7455683</pub-id><pub-id pub-id-type="pmid">7455683</pub-id></citation>
</ref>
<ref id="B36">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tversky</surname> <given-names>A.</given-names></name> <name><surname>Kahneman</surname> <given-names>D.</given-names></name></person-group> (<year>1983</year>). <article-title>Extensional versus intuitive reasoning: the conjunction fallacy in probability judgment</article-title>. <source>Psychol. Rev.</source> <volume>90</volume>, <fpage>293</fpage>&#x02013;<lpage>315</lpage>. <pub-id pub-id-type="doi">10.1037/0033-295X.90.4.293</pub-id></citation>
</ref>
<ref id="B37">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Vandewalle</surname> <given-names>D.</given-names></name></person-group> (<year>1997</year>). <article-title>Development and validation of a work domain goal orientation instrument</article-title>. <source>Educ. Psychol. Meas.</source> <volume>57</volume>, <fpage>995</fpage>&#x02013;<lpage>1015</lpage>. <pub-id pub-id-type="doi">10.1177/0013164497057006009</pub-id></citation>
</ref>
<ref id="B38">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yamamoto</surname> <given-names>M.</given-names></name> <name><surname>Matsui</surname> <given-names>Y.</given-names></name> <name><surname>Yamanari</surname> <given-names>Y.</given-names></name></person-group> (<year>1982</year>). <article-title>Ninchi sareta jiko no shosokumen no kouzou [The structure of perceived aspects of self]</article-title>. <source>Jpn. J. Educ. Psychol.</source> <volume>30</volume>, <fpage>64</fpage>&#x02013;<lpage>68</lpage>. <pub-id pub-id-type="doi">10.5926/jjep1953.30.1_64</pub-id></citation>
</ref>
<ref id="B39">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhou</surname> <given-names>H.</given-names></name> <name><surname>Fishbach</surname> <given-names>A.</given-names></name></person-group> (<year>2016</year>). <article-title>The pitfall of experimenting on the web: How unattended selective attrition leads to surprising (yet false) research conclusions</article-title>. <source>J. Pers. Soc. Psychol.</source> <volume>111</volume>, <fpage>493</fpage>&#x02013;<lpage>504</lpage>. <pub-id pub-id-type="doi">10.1037/pspa0000056</pub-id><pub-id pub-id-type="pmid">27295328</pub-id></citation>
</ref>
</ref-list>
<fn-group>
<fn id="fn0001"><p><sup>1</sup>Non-US researchers can post HITs on MTurk, if they use outside service.</p></fn>
<fn id="fn0002"><p><sup>2</sup>In Toplak et al. (<xref ref-type="bibr" rid="B33">2011</xref>)&#x00027;s task, black and white marble were used, although we replaced the marble with the stone of the <italic>go</italic> game, a popular board game in East Asia.</p></fn>
</fn-group>
</back>
</article>