<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="editorial">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Big Data</journal-id>
<journal-title>Frontiers in Big Data</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Big Data</abbrev-journal-title>
<issn pub-type="epub">2624-909X</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fdata.2023.1193412</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Big Data</subject>
<subj-group>
<subject>Editorial</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Editorial: Critical data and algorithm studies</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Mayer</surname> <given-names>Katja</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/660454/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Pfeffer</surname> <given-names>J&#x000FC;rgen</given-names></name>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Science and Technology Studies, University of Vienna</institution>, <addr-line>Vienna</addr-line>, <country>Austria</country></aff>
<aff id="aff2"><sup>2</sup><institution>Computational Social Science and Big Data, Technical University of Munich</institution>, <addr-line>Munich</addr-line>, <country>Germany</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited and reviewed by: Xintao Wu, University of Arkansas, United States</p></fn>
<corresp id="c001">&#x0002A;Correspondence: Katja Mayer <email>katja.mayer&#x00040;univie.ac.at</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>10</day>
<month>05</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>6</volume>
<elocation-id>1193412</elocation-id>
<history>
<date date-type="received">
<day>24</day>
<month>03</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>18</day>
<month>04</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2023 Mayer and Pfeffer.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Mayer and Pfeffer</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license></permissions>
<related-article id="RA1" related-article-type="commentary-article" xlink:href="https://www.frontiersin.org/research-topics/9570/critical-data-and-algorithm-studies" ext-link-type="uri">Editorial on the Research Topic <article-title>Critical data and algorithm studies</article-title></related-article>
<kwd-group>
<kwd>critical data studies</kwd>
<kwd>critical data science</kwd>
<kwd>critical algorithm studies</kwd>
<kwd>data work</kwd>
<kwd>social theory</kwd>
<kwd>big data</kwd>
<kwd>computational social science</kwd>
</kwd-group>
<counts>
<fig-count count="0"/>
<table-count count="0"/>
<equation-count count="0"/>
<ref-count count="5"/>
<page-count count="3"/>
<word-count count="1679"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Data Mining and Management</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec id="s1">
<title>1. Introduction</title>
<p>As digitalisation, mobile computing and social media platforms have become ubiquitous in modern life, there has been a surge in the use of big data analytics and data-driven research approaches. Scholars in human-computer interaction, critical data studies, and critical algorithm studies have long been concerned with the social challenges posed by data science, including issues of bias, opaque access to data and infrastructure, and issues of representation (Agre, <xref ref-type="bibr" rid="B1">1997</xref>; Boyd and Crawford, <xref ref-type="bibr" rid="B2">2012</xref>; Kitchin, <xref ref-type="bibr" rid="B3">2014</xref>; Zook et al., <xref ref-type="bibr" rid="B5">2017</xref>; Moats and Seaver, <xref ref-type="bibr" rid="B4">2019</xref>). To bridge the gap between cultures of critique and those of practice, this Research Topic is dedicated to bringing together critical expertise from data-driven research fields to reflect on the social impact of their research. Furthermore, the papers in this Research Topic address a range of issues, from understanding the limitations of social data and methods, the need to consider social theories for better outcomes and research integrity in the datafication of social behavior, to pressing issues of research infrastructures and the complexities of social media research.</p></sec>
<sec id="s2">
<title>2. The Research Topic of articles</title>
<p>The first paper in our Research Topic, &#x0201C;<italic>Social data: biases, methodological pitfalls, and ethical boundaries</italic>&#x0201D; by <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fdata.2019.00013">Olteanu et al.</ext-link>, lists and discusses the biases and inaccuracies in social data, as well as methodological limitations and ethical concerns. The article argues for the importance of auditing social data and algorithms to address potential biases and makes four recommendations: detailed documentation of datasets and models; expanding social data research to different platforms, topics, timings, and subpopulations; enabling transparency mechanisms to facilitate auditing of social software; and broadening research on guidelines, standards, methodologies, and protocols to address the limitations of social data.</p>
<p>In this vein, <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fdata.2020.00018">Radford and Joseph&#x00027;s</ext-link> paper argues that social theory is crucial to addressing problems with machine learning models used to analyse social data. Technical solutions alone are not enough. Social theory provides insights into methodological and interpretive questions that cannot be answered by technical fixes. The paper recommends further research into how social theory can be applied to address bias and inequality in machine learning models. By developing a systematic theoretical framework for social data, researchers can identify and address various challenges and limitations associated with the use of social data, such as biases in data collection and processing, methodological limitations, and ethical concerns. A structured approach to social data analysis can lead to more accurate and reliable results, improving decision-making and policy development.</p>
<p>The paper by <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fdata.2020.519957">Poechhacker and Kacianka</ext-link> applies a theoretical perspective to the ongoing debate on algorithmic accountability in automated decision making and machine learning. It focuses on the use of structural causal models (SCMs) as a means of establishing accountability by describing the causal relationships between different factors in an algorithmic system. As such, SCMs provide transparency that allow for public scrutiny, as people can review the rules and decisions to ensure that they are fair and ethical. However, the authors argue that the concept of causality needs further exploration. They bring insights from social theory, particularly pragmatism, and suggest that formal expressions of causality need to be considered within the social system in which they are applied.</p>
<p>In &#x0201C;<italic>Decentralized but globally coordinated biodiversity data</italic>&#x0201D; by <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fdata.2020.519133">Sterner et al.</ext-link>, the authors draw our attention to the central role of research infrastructures and their governance. They argue that centralized biodiversity data aggregation is failing to meet societal needs due to data quality shortcomings and propose a decentralized approach to data coordination. The authors suggest that this approach will lead to sustained expert engagement, higher quality data products and greater societal impact. The decentralized approach encourages the emergence and evolution of multiple self-identifying communities of practice that can control the social and informational design of their local data infrastructures.</p>
<p>The paper &#x0201C;<italic>The datafication of hate: expectations and challenges in automated hate speech monitoring</italic>&#x0201D; by <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fdata.2020.00003">Laaksonen et al.</ext-link> reflects on an action research setting aimed at monitoring social media updates for hate speech during the Finnish local elections in 2017. The study examines how hate speech emerged as a technical problem, and how an algorithmic solution was developed using supervised machine learning. The paper highlights the oversimplification of the automated approach and research design, and suggests practical implications for hate speech detection.</p>
<p>The need to reflect on one&#x00027;s own research and data handling is also considered in the paper by <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fdata.2020.509954">Kinder-Kurlanda and Weller</ext-link>. The authors argue that the details of scientific practice in data science and computational social science are not well-described in the literature. They propose a perspective that recognizes the everyday &#x0201C;data work&#x0201D; required to conduct social media research at different stages of a data lifecycle. The authors highlight the complexity faced by social media researchers and suggest that documenting research decisions is necessary to better understand what drives social media research and to address structural challenges in the research ecosystem. The overall outcome of such documentation is improved research rigor and transparency, leading to more accurate and trustworthy results.</p>
<p>The paper by <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fdata.2020.00005">Allhutter et al.</ext-link> examines the societal impact of algorithms in decision making. They discuss the algorithmic profiling used by the Austrian Public Employment Service (AMS) to categorize job seekers based on their prospects in the labor market. Drawing on their interdisciplinary collaboration, the authors highlight the tensions, challenges and biases inherent in the AMS algorithm and question the objectivity and neutrality of data claims and evidence-based decision making. The paper thus sheds light on (semi)automated management practices in employment agencies and the framing of unemployment under austerity policies, providing insights not only for critical data studies but also for public service policies.</p>
<p>In &#x0201C;<italic>Staying with the trouble of networks</italic>,&#x0201D; <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fdata.2022.510310">van Geenen et al.</ext-link> examine the use of network visualizations in different fields, such as scientific research or journalism, and argue that problems with such visualization practices can provide opportunities to reflect on knowledge practices in social contexts. While network visualizations are both discovery and narrative tools, the authors make explicit the epistemic assumptions built into network tools and emphasize the need to pay attention to the cultural practices of interpreting and understanding social relations with network analysis. By attending to the different settings and situations in which network graphs and maps are created and used, the authors suggest that we can develop a more nuanced understanding of the role of networks in collective forms of inquiry and sense-making. The article contributes to the concept of critical data practice by highlighting the need to consider the social and ethical implications of data.</p></sec>
<sec id="s3">
<title>3. Conclusion</title>
<p>In summary, this Research Topic has brought together a diverse range of articles that highlight this need for critical technical practice in data-driven research areas. The articles contribute to the growing fields of Critical Data Studies and Critical Algorithm Studies by addressing scientific, ethical, and social challenges, and by promoting a cross- and transdisciplinary exchange between practical and critical positions, thus encouraging collaborations between computer scientists, social scientists, humanities scholars, journalists, policy makers, and activists, and providing a productive intersection between cultures of critique and those of practice. Overall, the Research Topic offers a valuable contribution to ongoing critical reflection on scientific methods, data sources, modeling, validation, replication and review procedures, while highlighting the performative and normative aspects of data science practices.</p></sec>
<sec sec-type="author-contributions" id="s4">
<title>Author contributions</title>
<p>All authors listed have made a substantial, direct, and intellectual contribution to the work and approved it for publication.</p></sec>
</body>
<back>
<sec sec-type="funding-information" id="s5">
<title>Funding</title>
<p>The work of KM was supported by the Austrian Science Fund FWF (V699-G29).</p>
</sec>
<sec sec-type="COI-statement" id="conf1">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s6">
<title>Publisher&#x00027;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Agre</surname> <given-names>P.</given-names></name></person-group> (<year>1997</year>). <source>Computation and Human Experience</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation>
</ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Boyd</surname> <given-names>D.</given-names></name> <name><surname>Crawford</surname> <given-names>K.</given-names></name></person-group> (<year>2012</year>). <article-title>Critical questions for big data: provocations for a cultural, technological, and scholarly phenomenon</article-title>. <source>Inform. Commun. Soc.</source> <volume>15</volume>, <fpage>662</fpage>&#x02013;<lpage>679</lpage>. <pub-id pub-id-type="doi">10.1080/1369118X.2012.678878</pub-id></citation>
</ref>
<ref id="B3">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Kitchin</surname> <given-names>R.</given-names></name></person-group> (<year>2014</year>). <source>The Data Revolution: Big Data, Open Data, Data Infrastructures and Their Consequences</source>. <publisher-loc>London</publisher-loc>: <publisher-name>SAGE</publisher-name>.<pub-id pub-id-type="pmid">29453548</pub-id></citation></ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Moats</surname> <given-names>D.</given-names></name> <name><surname>Seaver</surname> <given-names>N.</given-names></name></person-group> (<year>2019</year>). <article-title>&#x0201C;You social scientists love mind games&#x0201D;: Experimenting in the &#x0201C;divide&#x0201D; between data science and critical algorithm studies</article-title>. <source>Big Data Soc.</source> <volume>6</volume>, <fpage>2053951719833404</fpage>. <pub-id pub-id-type="doi">10.1177/2053951719833404</pub-id></citation>
</ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zook</surname> <given-names>M.</given-names></name> <name><surname>Barocas</surname> <given-names>S.</given-names></name> <name><surname>Boyd</surname> <given-names>D.</given-names></name> <name><surname>Crawford</surname> <given-names>K.</given-names></name> <name><surname>Keller</surname> <given-names>E.</given-names></name> <name><surname>Gangadharan</surname> <given-names>S. P.</given-names></name> <etal/></person-group>. (<year>2017</year>). <article-title>Ten simple rules for responsible big data research</article-title>. <source>PLoS Comput. Biol.</source> <volume>13</volume>, <fpage>e1005399</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pcbi.1005399</pub-id><pub-id pub-id-type="pmid">28358831</pub-id></citation></ref>
</ref-list>
</back>
</article>