<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article article-type="editorial" dtd-version="2.3" xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Pharmacol.</journal-id>
<journal-title>Frontiers in Pharmacology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Pharmacol.</abbrev-journal-title>
<issn pub-type="epub">1663-9812</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">1226756</article-id>
<article-id pub-id-type="doi">10.3389/fphar.2023.1226756</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Pharmacology</subject>
<subj-group>
<subject>Editorial</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Editorial: Opportunities and challenges in reusing public genomics data</article-title>
<alt-title alt-title-type="left-running-head">Ahmed and Kim</alt-title>
<alt-title alt-title-type="right-running-head">
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fphar.2023.1226756">10.3389/fphar.2023.1226756</ext-link>
</alt-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Ahmed</surname>
<given-names>Mahmoud</given-names>
</name>
<uri xlink:href="https://loop.frontiersin.org/people/1472223/overview"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Kim</surname>
<given-names>Deok Ryong</given-names>
</name>
<xref ref-type="corresp" rid="c001">&#x2a;</xref>
<uri xlink:href="https://loop.frontiersin.org/people/1430421/overview"/>
</contrib>
</contrib-group>
<aff>
<institution>Department of Biochemistry and Convergence Medical Sciences</institution>, <institution>Institute of Health Sciences</institution>, <institution>College of Medicine</institution>, <institution>Gyeongsang National University</institution>, <addr-line>Jinju</addr-line>, <country>Republic of Korea</country>
</aff>
<author-notes>
<fn fn-type="edited-by">
<p>
<bold>Edited and reviewed by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/46314/overview">Dov Greenbaum</ext-link>, Yale University, United States</p>
</fn>
<corresp id="c001">&#x2a;Correspondence: Deok Ryong Kim, <email>drkim@gnu.ac.kr</email>
</corresp>
</author-notes>
<pub-date pub-type="epub">
<day>12</day>
<month>06</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>14</volume>
<elocation-id>1226756</elocation-id>
<history>
<date date-type="received">
<day>22</day>
<month>05</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>05</day>
<month>06</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#xa9; 2023 Ahmed and Kim.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Ahmed and Kim</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<related-article id="RA1" related-article-type="commentary-article" journal-id="Front. Pharmacol." xlink:href="https://www.frontiersin.org/researchtopic/37301" ext-link-type="uri">Editorial on the Research Topic <article-title>Opportunities and challenges in reusing public genomics data</article-title> </related-article>
<kwd-group>
<kwd>reusing public data</kwd>
<kwd>genomics</kwd>
<kwd>data sharing</kwd>
<kwd>metadata</kwd>
<kwd>data curation and integration</kwd>
</kwd-group>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>ELSI in Science and Genetics</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<p>Genomics data is accumulating in public repositories at an ever-increasing rate. Large consortia and individual labs continue to probe animal and plant tissue and cell cultures, generating vast amounts of data using established and novel technologies. The human genome project kick started the era of systems biology (<xref ref-type="bibr" rid="B5">Lander et al., 2001</xref>; <xref ref-type="bibr" rid="B4">Gates et al., 2021</xref>). Ambitious projects followed to characterize non-coding regions, variations across species, and between populations (<xref ref-type="bibr" rid="B3">Feingold et al., 2004</xref>; <xref ref-type="bibr" rid="B9">Sabeti et al., 2007</xref>; <xref ref-type="bibr" rid="B1">Auton et al., 2015</xref>). The cost reduction allowed individual labs to generate numerous smaller high-throughput datasets (<xref ref-type="bibr" rid="B2">Edgar et al., 2002</xref>; <xref ref-type="bibr" rid="B8">Parkinson et al., 2007</xref>; <xref ref-type="bibr" rid="B7">Metzker, 2010</xref>; <xref ref-type="bibr" rid="B6">Leinonen et al., 2011</xref>). As a result, the scientific community should consider strategies to overcome the challenges and maximize the opportunities to use these resources for research and the public good. In this Research Topic, we have elicited opinions and perspectives from researchers in the field on the opportunities and challenges of reusing public genomics data. The articles in this Research Topic converge on the need for data sharing while acknowledging the challenges that come with it. Two articles defined and highlighted the distinction between data and metadata. The characteristic of each should be considered when designing optimal sharing strategies. One article focuses on the specific issues surrounding the sharing of genomics interval data, and another on balancing the need for protecting pediatric rights and the sharing benefits.</p>
<p>The definition of what counts as data is itself a moving target. As technology advances, data can be produced in more ways and from novel sources. Events of recent years have highlighted this fact. &#x201c;The pandemic has underscored the urgent need to recognize health data as a global public good with mechanisms to facilitate rapid data sharing and governance,&#x201d; <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fdgth.2020.612339/full">Schwalbe et al</ext-link>. The challenges facing these mechanisms could be technical, economic, legal, or political. Defining what data is and its type, therefore, is necessary to overcome these barriers because &#x201c;the mechanisms to facilitate data sharing are often specific to data types.&#x201d; Unlike genomics data, which has established platforms, sharing clinical data &#x201c;remains in a nascent phase.&#x201d; The article by <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fgene.2022.872586/full">Patrinos et al.</ext-link> considers the strong ethical imperative for protecting pediatric data while acknowledging the need to avoid over protections. The authors discuss a model of consent for pediatric research that can balance the need to protect participants and generate health benefits.</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fgene.2023.1155809/full">Xue et al.</ext-link> focus on reusing genomic interval data. Identifying and retrieving the relevant data can be difficult, given the state of the repositories and the size of these data. Similarly, integrating interval data in reference genomes can be hard. The author calls for standardized formats for the data and the metadata to facilitate reuse.</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fgene.2023.1154198/full">Sheffield et al.</ext-link> highlight the distinction between data and metadata. Metadata describes the characteristics of the sample, experiment, and analysis. The nature of this information differs from that of the primary data in size, source, and ways of use. Therefore, an optimal strategy should consider these specific attributes for sharing metadata. Challenges specifics to sharing metadata include the need for standardized terms and formats, making it portable and easier to find.</p>
<p>We go beyond the reuse issue to highlight two other aspects that might increase the utility of available public data in <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fgene.2023.1106631/full">Ahmed et al</ext-link>. These are curation and integration. Despite being generated using different protocols, combining the datasets from separate groups could help to fill the gaps in the design and increase the statistical power of the analysis. Integrating data types can be beneficial to either verify or complement the observations made based on a single data type. We also emphasize the critical requirements for these strategies to be successful. We draw on our experience and others in using publicly available datasets to support, develop, and extend our research interest.</p>
<p>The articles in this Research Topic converge on the importance of data sharing. In addition, the articles present the challenges facing data sharing and reuse and propose models to increase the utility of public data.</p>
</body>
<back>
<sec id="s1">
<title>Author contributions</title>
<p>MA and DK wrote and revised the manuscript. All authors contributed to the article and approved the submitted version.</p>
</sec>
<sec id="s2">
<title>Funding</title>
<p>This study was supported by the National Research Foundation of Korea (NRF) grant funded by the Ministry of Science and ICT (MSIT) of the Korea government (2020R1A2C2011416) and by the Commercializations Promotion Agency for R&#x26;D Outcomes (COMPA) grant funded by the Korea government (MSIT) (1711173796).</p>
</sec>
<ack>
<p>We thank all the lab members for their thoughtful feedback on this study.</p>
</ack>
<sec sec-type="COI-statement" id="s3">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s4">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Auton</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Abecasis</surname>
<given-names>G. R.</given-names>
</name>
<name>
<surname>Altshuler</surname>
<given-names>D. M.</given-names>
</name>
<name>
<surname>Durbin</surname>
<given-names>R. M.</given-names>
</name>
<name>
<surname>Bentley</surname>
<given-names>D. R.</given-names>
</name>
<name>
<surname>Chakravarti</surname>
<given-names>A.</given-names>
</name>
<etal/>
</person-group> (<year>2015</year>). <article-title>A global reference for human genetic variation</article-title>. <source>Nature</source> <volume>526</volume>, <fpage>68</fpage>&#x2013;<lpage>74</lpage>. <pub-id pub-id-type="doi">10.1038/nature15393</pub-id>
</citation>
</ref>
<ref id="B2">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Edgar</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Domrachev</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Lash</surname>
<given-names>A. E.</given-names>
</name>
</person-group> (<year>2002</year>). <article-title>Gene expression omnibus: NCBI gene expression and hybridization array data repository</article-title>. <source>Nucleic acids Res.</source> <volume>30</volume>, <fpage>207</fpage>&#x2013;<lpage>210</lpage>. <pub-id pub-id-type="doi">10.1093/nar/30.1.207</pub-id>
</citation>
</ref>
<ref id="B3">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Feingold</surname>
<given-names>E. A.</given-names>
</name>
<name>
<surname>Good</surname>
<given-names>P. J.</given-names>
</name>
<name>
<surname>Guyer</surname>
<given-names>M. S.</given-names>
</name>
<name>
<surname>Kamholz</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Liefer</surname>
<given-names>L.</given-names>
</name>
<etal/>
</person-group> (<year>2004</year>). <article-title>The ENCODE (ENCyclopedia of DNA elements) project</article-title>. <source>Science</source> <volume>306</volume>, <fpage>636</fpage>&#x2013;<lpage>640</lpage>. <pub-id pub-id-type="doi">10.1126/science.1105136</pub-id>
</citation>
</ref>
<ref id="B4">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Gates</surname>
<given-names>A. J.</given-names>
</name>
<name>
<surname>Gysi</surname>
<given-names>D. M.</given-names>
</name>
<name>
<surname>Kellis</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Barab&#xe1;si</surname>
<given-names>A. L.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>A wealth of discovery built on the human genome project &#x2014; By the numbers</article-title>. <source>Nature</source> <volume>590</volume>, <fpage>212</fpage>&#x2013;<lpage>215</lpage>. <pub-id pub-id-type="doi">10.1038/d41586-021-00314-6</pub-id>
</citation>
</ref>
<ref id="B5">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lander</surname>
<given-names>E. S.</given-names>
</name>
<name>
<surname>Linton</surname>
<given-names>L. M.</given-names>
</name>
<name>
<surname>Birren</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Nusbaum</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Zody</surname>
<given-names>M. C.</given-names>
</name>
<name>
<surname>Baldwin</surname>
<given-names>J.</given-names>
</name>
<etal/>
</person-group> (<year>2001</year>). <article-title>Initial sequencing and analysis of the human genome</article-title>. <source>Nature</source> <volume>409</volume>, <fpage>860</fpage>&#x2013;<lpage>921</lpage>. <pub-id pub-id-type="doi">10.1038/35057062</pub-id>
</citation>
</ref>
<ref id="B6">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Leinonen</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Sugawara</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Shumway</surname>
<given-names>M.</given-names>
</name>
</person-group> (<year>2011</year>). <article-title>The sequence read archive</article-title>. <source>Nucleic Acids Res.</source> <volume>39</volume>, <fpage>D19</fpage>&#x2013;<lpage>D21</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkq1019</pub-id>
</citation>
</ref>
<ref id="B7">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Metzker</surname>
<given-names>M. L.</given-names>
</name>
</person-group> (<year>2010</year>). <article-title>Sequencing technologies the next generation</article-title>. <source>Nat. Rev. Genet.</source> <volume>11</volume>, <fpage>31</fpage>&#x2013;<lpage>46</lpage>. <pub-id pub-id-type="doi">10.1038/nrg2626</pub-id>
</citation>
</ref>
<ref id="B8">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Parkinson</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Kapushesky</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Shojatalab</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Abeygunawardena</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Coulson</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Farne</surname>
<given-names>A.</given-names>
</name>
<etal/>
</person-group> (<year>2007</year>). <article-title>ArrayExpress - a public database of microarray experiments and gene expression profiles</article-title>. <source>Nucleic Acids Res.</source> <volume>35</volume>, <fpage>D747</fpage>&#x2013;<lpage>D750</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkl995</pub-id>
</citation>
</ref>
<ref id="B9">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Sabeti</surname>
<given-names>P. C.</given-names>
</name>
<name>
<surname>Varilly</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Fry</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Lohmueller</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Hostetter</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Cotsapas</surname>
<given-names>C.</given-names>
</name>
<etal/>
</person-group> (<year>2007</year>). <article-title>Genome-wide detection and characterization of positive selection in human populations</article-title>. <source>Nature</source> <volume>449</volume>, <fpage>913</fpage>&#x2013;<lpage>918</lpage>. <pub-id pub-id-type="doi">10.1038/nature06250</pub-id>
</citation>
</ref>
</ref-list>
</back>
</article>