<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Plant Sci.</journal-id>
<journal-title>Frontiers in Plant Science</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Plant Sci.</abbrev-journal-title>
<issn pub-type="epub">1664-462X</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpls.2017.01034</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Plant Science</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Expressed Centromere Specific Histone 3 (<italic>CENH3</italic>) Variants in Cultivated Triploid and Wild Diploid Bananas (<italic>Musa</italic> spp.)</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name><surname>Muiruri</surname> <given-names>Kariuki S.</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/418670/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Britt</surname> <given-names>Anne</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/22468/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Amugune</surname> <given-names>Nelson O.</given-names></name>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/431601/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Nguu</surname> <given-names>Edward K.</given-names></name>
<xref ref-type="aff" rid="aff4"><sup>4</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/128274/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Chan</surname> <given-names>Simon</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<xref ref-type="author-notes" rid="fn002"><sup>&#x2020;</sup></xref>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name><surname>Tripathi</surname> <given-names>Leena</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="author-notes" rid="fn001"><sup>&#x002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/115555/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>International Institute of Tropical Agriculture</institution> <country>Nairobi, Kenya</country></aff>
<aff id="aff2"><sup>2</sup><institution>School of Biological Sciences, University of Nairobi</institution> <country>Nairobi, Kenya</country></aff>
<aff id="aff3"><sup>3</sup><institution>Department of Plant Biology, University of California, Davis, Davis</institution> <country>CA, United States</country></aff>
<aff id="aff4"><sup>4</sup><institution>Department of Biochemistry, University of Nairobi</institution> <country>Nairobi, Kenya</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: <italic>Junhua Peng, Center for Life Sci&#x0026;Tech of China National Seed Group Co. Ltd., China</italic></p></fn>
<fn fn-type="edited-by"><p>Reviewed by: <italic>Liang Chen, University of Chinese Academy of Sciences (UCAS), China; Xiaoli Jin, Zhejiang University, China; Guangxiao Yang, Huazhong University of Science and Technology, China</italic></p></fn>
<fn fn-type="corresp" id="fn001"><p>&#x002A;Correspondence: <italic>Leena Tripathi, <email>l.tripathi@cgiar.org</email></italic></p></fn>
<fn fn-type="other" id="fn002"><p><sup>&#x2020;</sup><italic>Deceased</italic></p></fn>
<fn fn-type="other" id="fn003"><p>This article was submitted to Plant Biotechnology, a section of the journal Frontiers in Plant Science</p></fn>
</author-notes>
<pub-date pub-type="epub">
<day>29</day>
<month>06</month>
<year>2017</year>
</pub-date>
<pub-date pub-type="collection">
<year>2017</year>
</pub-date>
<volume>8</volume>
<elocation-id>1034</elocation-id>
<history>
<date date-type="received">
<day>24</day>
<month>02</month>
<year>2017</year>
</date>
<date date-type="accepted">
<day>30</day>
<month>05</month>
<year>2017</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2017 Muiruri, Britt, Amugune, Nguu, Chan and Tripathi.</copyright-statement>
<copyright-year>2017</copyright-year>
<copyright-holder>Muiruri, Britt, Amugune, Nguu, Chan and Tripathi</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) or licensor are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>Centromeres are specified by a centromere specific histone 3 (CENH3) protein, which exists in a complex environment, interacting with conserved proteins and rapidly evolving satellite DNA sequences. The interactions may become more challenging if multiple CENH3 versions are introduced into the zygote as this can affect post-zygotic mitosis and ultimately sexual reproduction. Here, we characterize <italic>CENH3</italic> variant transcripts expressed in cultivated triploid and wild diploid progenitor bananas. We describe both splice- and allelic-[Single Nucleotide Polymorphisms (SNP)] variants and their effects on the predicted secondary structures of protein. Expressed <italic>CENH3</italic> transcripts from six banana genotypes were characterized and clustered into three groups (<italic>MusaCENH</italic>-1A, <italic>MusaCENH</italic>-1B, and <italic>MusaCENH</italic>-2) based on similarity. The <italic>CENH3</italic> groups differed with SNPs as well as presence of indels resulting from retained and/or skipped exons. The <italic>CENH3</italic> transcripts from different banana genotypes were spliced in either 7/6, 5/4 or 6/5 exons/introns. The 7/6 and the 5/4 exon/intron structures were found in both diploids and triploids, however, 7/6 was most predominant. The 6/5 exon/introns structure was a result of failure of the 7/6 to splice correctly. The various transcripts obtained were predicted to encode highly variable N-terminal tails and a relatively conserved C-terminal histone fold domain (HFD). The SNPs were predicted in some cases to affect the secondary structure of protein by lengthening or shorting the affected domains. Sequencing of banana <italic>CENH3</italic> transcripts predicts SNP variations that affect amino acid sequences and alternatively spliced transcripts. Most of these changes affect the N-terminal tail of CENH3.</p>
</abstract>
<kwd-group>
<kwd><italic>CENH3</italic></kwd>
<kwd>splice variants</kwd>
<kwd>genotype</kwd>
<kwd>centromere</kwd>
<kwd>histones</kwd>
<kwd>banana</kwd>
</kwd-group>
<contract-num rid="cn001">Basic Research Enabling Agricultural Development (BREAD) project number 1109882</contract-num>
<contract-num rid="cn002">Basic Research Enabling Agricultural Development (BREAD) project number 1109882</contract-num>
<contract-sponsor id="cn001">National Science Foundation<named-content content-type="fundref-id">10.13039/100000865</named-content></contract-sponsor>
<contract-sponsor id="cn002">Bill and Melinda Gates Foundation<named-content content-type="fundref-id">10.13039/100000865</named-content></contract-sponsor>
<counts>
<fig-count count="5"/>
<table-count count="2"/>
<equation-count count="0"/>
<ref-count count="44"/>
<page-count count="12"/>
<word-count count="0"/>
</counts>
</article-meta>
</front>
<body>
<sec><title>Introduction</title>
<p>Centromeres are assembly sites for the kinetochore, a protein complex that connects chromosomes to spindle fibers during meiosis and mitosis. The structure, size, and distribution of centromeres differ with species in spite of their common function (<xref ref-type="bibr" rid="B37">Talbert et al., 2004</xref>). Centromeres in both plants and animals often contain arrays of rapidly evolving tandemly repeated DNA sequences (<xref ref-type="bibr" rid="B11">Gent et al., 2011</xref>; <xref ref-type="bibr" rid="B40">Verdaasdonk and Bloom, 2011</xref>). The high rate of evolution in these repeats is remarkable given the fact that the function of centromeres is highly conserved. The role of the repeats is a subject of debate with the most common proposition being that they maintain the large heterochromatic domains associated with centromeres (<xref ref-type="bibr" rid="B25">Malik and Henikoff, 2009</xref>; <xref ref-type="bibr" rid="B2">Black and Cleveland, 2011</xref>). It is reported that CENH3 [aka Centromere Protein A (CENP-A) in humans and CID in drosophila] epigenetically determines and maintains centromeres (<xref ref-type="bibr" rid="B24">Malik and Henikoff, 2001</xref>; <xref ref-type="bibr" rid="B6">Dawe and Henikoff, 2006</xref>; <xref ref-type="bibr" rid="B9">Ekwall, 2007</xref>; <xref ref-type="bibr" rid="B1">Allshire and Karpen, 2008</xref>; <xref ref-type="bibr" rid="B10">Fachinetti et al., 2013</xref>). CENH3 contains a highly variable N-terminal tail and a relatively conserved histone fold domain (HFD) (<xref ref-type="bibr" rid="B32">Ravi et al., 2010</xref>; <xref ref-type="bibr" rid="B22">Lermontova et al., 2014</xref>). The majority of diploid plant species have been shown to encode a single <italic>CENH3</italic> gene (<xref ref-type="bibr" rid="B44">Zhong et al., 2002</xref>). However, more than one copy (alpha and beta) of the gene per genome are present in some species like wheat, barley, <italic>Arabidopsis halleri</italic> and <italic>A. lyrata</italic> (<xref ref-type="bibr" rid="B17">Kawabe et al., 2006</xref>; <xref ref-type="bibr" rid="B34">Sanei et al., 2011</xref>; <xref ref-type="bibr" rid="B43">Yuan et al., 2015</xref>).</p>
<p>The majority of cultivated bananas exist as allo- or autopolyploids and a variety of <italic>CENH3</italic> isoforms are presumed to coexist in the nucleus. Polyploidization brings together multiple gene copies within the same background and can result in additive or non-additive gene expression leading to biased or unbiased homeolog expression (<xref ref-type="bibr" rid="B28">Pignatta and Comai, 2009</xref>; <xref ref-type="bibr" rid="B15">Hui et al., 2010</xref>; <xref ref-type="bibr" rid="B30">Rapp et al., 2010</xref>; <xref ref-type="bibr" rid="B42">Yoo et al., 2013</xref>). Unlike many diploid species where a single copy of <italic>CENH3</italic> gene is encoded, multiple copies have been observed in newly synthesized allopolyploids of rice, wheat, brassica, and pea (<xref ref-type="bibr" rid="B13">Hirsch et al., 2009</xref>; <xref ref-type="bibr" rid="B15">Hui et al., 2010</xref>; <xref ref-type="bibr" rid="B41">Wang et al., 2011</xref>; <xref ref-type="bibr" rid="B27">Neumann et al., 2012</xref>; <xref ref-type="bibr" rid="B43">Yuan et al., 2015</xref>). <italic>CENH3</italic> variants have also been characterized in wild and cultivated carrots (<xref ref-type="bibr" rid="B8">Dunemann et al., 2014</xref>) and in stable polyploids of different angiosperms (<xref ref-type="bibr" rid="B26">Masonbrink et al., 2014</xref>). Multiple <italic>CENH3</italic> copies observed in polyploids might result from coming together of single-<italic>CENH3</italic>-expressing genomes or multiple-<italic>CENH3</italic> expressing progenitor genomes. Crosses of diploid parents encoding multiple <italic>CENH3</italic> transcripts have resulted in stable hybrids. For example, in stable hybrids from <italic>Hordeum vulgare</italic> &#x00D7; <italic>H. bulbosum</italic> crosses, both alpha and beta <italic>CENH3</italic> variants from <italic>H. vulgare</italic> were incorporated into the centromeric nucleosomes of the hybrid. In contrast, a hybrid of <italic>H. bulbosum</italic> &#x00D7; <italic>Triticum aestivum</italic> incorporated the <italic>H. bulbosum CENH3</italic> variant Hb&#x03B1;CENH3 only (<xref ref-type="bibr" rid="B34">Sanei et al., 2011</xref>).</p>
<p>Unlike stable hybrids, embryos derived from unstable crosses have been observed to undergo uniparental genome elimination, resulting in haploids carrying genetic material from only one parent (<xref ref-type="bibr" rid="B31">Ravi and Chan, 2010</xref>; <xref ref-type="bibr" rid="B34">Sanei et al., 2011</xref>; <xref ref-type="bibr" rid="B35">Seymour et al., 2012</xref>; <xref ref-type="bibr" rid="B23">Maheshwari et al., 2015</xref>). The genome of <italic>H. bulbosum</italic> in embryos from <italic>H. vulgare</italic> &#x00D7; <italic>H. bulbosum</italic> crosses for example was completely lost within 5&#x2013;9 days post-fertilization. Despite elimination of the <italic>H. bulbosum</italic> genome later in post-zygotic mitosis, <italic>H. vulgare</italic> &#x00D7; <italic>H. bulbosum</italic> unstable crosses have been observed to transcribe <italic>CENH3</italic> transcript variants from both parents (<xref ref-type="bibr" rid="B34">Sanei et al., 2011</xref>). In <italic>A. thaliana</italic>, uniparental genome elimination was also observed in offspring from crosses between mutant &#x2018;haploid inducer&#x2019; (parent with modified <italic>CENH3</italic>) and wild-type (carrying wild-type <italic>CENH3</italic> version) (<xref ref-type="bibr" rid="B31">Ravi and Chan, 2010</xref>). The modification of <italic>CENH</italic>3 in this case was generated by replacing the N-terminal tail with that of the variant H3.3 and tagging it with GFP. Apart from obtaining haploids in these crosses, novel genetic rearrangements were observed (<xref ref-type="bibr" rid="B23">Maheshwari et al., 2015</xref>). Currently, there are efforts undergoing to transfer this technology to many crops including banana (<xref ref-type="bibr" rid="B4">Comai, 2014</xref>). Crosses of <italic>A. thaliana</italic> null-mutants carrying gene constructs expressing <italic>CENH3</italic> from distant species to plants wild-type for <italic>CENH3</italic> have also resulted in haploids (<xref ref-type="bibr" rid="B23">Maheshwari et al., 2015</xref>). Furthermore, uniparental genome elimination has been observed in crosses of wild-type <italic>A. thaliana</italic> plants to null mutants complemented with <italic>CENH3</italic> carrying missense point mutations in conserved regions of the HFD (<xref ref-type="bibr" rid="B20">Kuppu et al., 2015</xref>).</p>
<p>Banana breeding involves crossing of tetraploids to diploids to give triploids and this may add into the complexity of the space CENH3 exists. Therefore, it would be interesting and useful to understand <italic>CENH3</italic> dynamics in cultivated polyploids and their diploid progenitors. Furthermore, a clear understanding of <italic>CENH3</italic> behavior in cultivated crops like banana is essential if breeding tools such as <italic>CENH3</italic>-based haploid technology are to be effectively applied (<xref ref-type="bibr" rid="B3">Britt and Kuppu, 2016</xref>). Therefore, in this study the expression of <italic>CENH</italic>3 was characterized in cultivated triploid and wild-type diploid progenitor bananas. The existence and evolutionary relationships of <italic>CENH3</italic> SNPs and/or splice variants as well as their predicted secondary folding of protein were analyzed.</p>
</sec>
<sec id="s1" sec-type="materials|methods">
<title>Materials and Methods</title>
<sec><title>Plant Materials</title>
<p>Six banana genotypes including wild diploids &#x2018;Calcutta 4&#x2019; (AA) and &#x2018;Zebrina GF&#x2019; (AA) both from the species <italic>Musa acuminata</italic>, the species <italic>M. balbisiana</italic> (BB) and cultivated triploids &#x2018;Sukali Ndiizi&#x2019; (AAB), &#x2018;Pisang Awak&#x2019; (ABB) and &#x2018;Gros Michel&#x2019; (AAA) were used in this study. All plant materials used were obtained from <italic>in vitro</italic> collection at IITA Kenya.</p>
</sec>
<sec><title>Identification of Genomic Sequence of Banana <italic>CENH3</italic></title>
<p>To identify putative genomic sequence of banana <italic>CENH3</italic>, a nucleotide BLAST (BLASTN) analysis was performed using genomic sequence of <italic>A. thaliana CENH3</italic> (At1g01030) against the whole-genome shotgun contigs (wgs) of <italic>M. acuminata</italic> (tax id: 4641) for &#x201C;somewhat similar sequences&#x201D;. In order to identify the exact genomic region of <italic>CENH3</italic>, consensus sequences from conserved regions at the beginning and end of selected monocot <italic>CENH3</italic> CDSs were mapped to the BLASTN hit results. The conserved consensus, which we considered as representative <italic>CENH3</italic> &#x2018;landmark&#x2019; regions for monocots, were obtained by aligning sequences of <italic>CENH3</italic> from the monocots <italic>Zea mays</italic> (NM_001112050), <italic>H. vulgare</italic> (JF419328), <italic>T. aestivum</italic> (JF969285.1) and <italic>Oryza sativa</italic> (AY438639.1). To identify the genomic regions of the <italic>CENH3</italic> from <italic>M. acuminata</italic>, BLASTN hits, we mapped the <italic>CENH3</italic> &#x2018;landmarks&#x2019; and regions with >75% nucleotide identities were selected. The primers CENH3_END_F (GGCGAGAACGAAGCATC) and CENH3_END_R (TCACCAATGTCTTCTTCCTCC) were designed to amplify the CDS (from the beginning to the end of the coding region) derived from <italic>in silico</italic> analysis of the putative banana genomic sequence (Accession: CAIC01023700).</p>
</sec>
<sec><title>RNA Extraction and RT-PCR</title>
<p>Total RNA was extracted from 100 mg of young incompletely open leaves. Extraction was performed using RNeasy<sup>&#x00AE;</sup> plant mini kit (Hilden, Germany) as per the manufacturer&#x2019;s protocol except for the elution volume which was reduced to 40 &#x03BC;l. Genomic DNA contamination was removed from the extracted RNA through DNase I (Thermo Scientific, Waltham, MA, United States) treatment by incubating at 37&#x00B0;C for 30 min and then terminating the reaction by adding 1 mM EDTA and heating at 70&#x00B0;C for 5 min. RNA quality and quantity were checked using a NanoDrop<sup>TM</sup> 2000 (Thermo Scientific, Waltham, MA, United States) spectrophotometer.</p>
<p>First strand cDNA was synthesized from 1 &#x03BC;g of DNA-free total RNA with random hexamer primers using maxima first strand reverse transcriptase kit (Thermo Scientific, Waltham, MA, United States). Two independent cDNA synthesis reactions were performed for each of the genotype.</p>
<p>The <italic>CENH3</italic> transcripts were amplified from cDNA in a total of six PCR reactions (three reactions for each of the two cDNA synthesis) per genotype. Each PCR reaction was performed in a 20 &#x03BC;l volume, which contained 50 ng of cDNA template, 1x Q5 reaction buffer containing 2.5 mM MgCl<sub>2</sub>, 500 &#x03BC;M of each dNTP, 10 &#x03BC;M each of <italic>CENH3</italic> primers (CENH3_END_F and CENH3_END_R) and 1unit of Q5 high fidelity DNA polymerase (New England Biolabs, MA). The reactions were performed in an ABI 9700 PCR machine with the conditions set at initial denaturation of 98&#x00B0;C for 4 min, 35 cycles of 98&#x00B0;C for 15 s, 66&#x00B0;C for 30 s and 72&#x00B0;C for 45 s and a final extension at 72&#x00B0;C for 10 min. An aliquot of PCR product (2 &#x03BC;l) was run on a 1.5% agarose gel stained with GelRed (Biotium, CA) to confirm amplification. For PCR reactions in each genotype that had observable band(s) on agarose gel, the remainder (18 &#x03BC;l) PCR product was purified using Bioneer PCR purification kit (Daeongeon, South Korea) and eluted in 15 &#x03BC;l water.</p>
</sec>
<sec><title>Cloning and Sequencing of <italic>CENH3</italic> Genes</title>
<p>Purified PCR products were cloned into pJET 1.2 cloning vector (Thermo Scientific, MA) and transformed into competent <italic>Escherichia coli</italic> (DH5&#x03B1;) cells using heat shock method. The transformed <italic>E. coli</italic> colonies were selected on Luria Bertani (LB) agar (10 g/l Tryptone, 5 g/l Yeast extract, 10 g/l NaCl, 15 g/l Agar, pH 7.5) containing 50 mg/L ampicillin. One to 10 transformed colonies from each PCR reaction were screened for presence of the insert by colony-PCR. A maximum of 60 colonies were screened for each genotype. The primer pairs pJET 1.2_F: CGACTCACTATAGGGAGAGCGGC and pJET 1.2_R: AAGAACATCGATTTTCCATGGCAG were used for colony PCR. Colonies with amplicon sizes >200 bp were cultured in LB broth medium overnight at 37&#x00B0;C and plasmid DNA extracted using Qiagen plasmid miniprep kit. Each clone with product >200 bp was sequenced bi-directionally in three replicates using the primers pJET 1.2_F and Pjet 1.2_R. Sequencing was performed on ABI 3130 analyzer (Applied Biosystems, Foster City, CA, United States) using BigDye Terminator Kit version 3.1.</p>
</sec>
<sec><title>Sequence Analysis and Multiple Alignments</title>
<p>Sequences were analyzed in Geneious version 7.1 (Biomatter, NZ) (<xref ref-type="bibr" rid="B18">Kearse et al., 2012</xref>) by manually checking the quality of the chromatograms. Sequences with quality above 50% (based on Phred values) across the entire sequence length were used for analysis. Sequences were further screened and &#x2018;dirty&#x2019; sections at the ends were manually trimmed to retain only high quality regions. Sequences within any of the six genotypes that were independently derived (those obtained from amplification of independently synthesized cDNA transcripts) and had 100% similarity were considered to represent the same transcript.</p>
<p>Multiple alignments of amino acids were conducted among translated banana CENH3 sequences and monocots (<italic>Z. mays, T. aestivum, O. sativa</italic>, and <italic>H. vulgare</italic>) and dicots (<italic>A. thaliana and Brassica rapa</italic>) in MUSCLE as implemented in Geneious version 7.1 using default parameters. Phylogenetic trees comparing transcript sequences were drawn in the software &#x201C;Molecular and Evolutionary Genetic Analysis&#x201D; (MEGA) version 6.0 (<xref ref-type="bibr" rid="B38">Tamura et al., 2013</xref>) based on only the conserved tail sections and entire HFD region.</p>
</sec>
<sec><title>Identification of Exon/Intron Structures</title>
<p>Since the banana <italic>CENH3</italic> from the genotypes used in this study had not been sequenced previously, splicing patterns for the transcript sequences were predicted by aligning them to the then available banana genomic sequence (accession number: <ext-link ext-link-type="DDBJ/EMBL/GenBank" xlink:href="CAIC01023700">CAIC01023700</ext-link> positions 70772 to 76310) from <italic>M. acuminata</italic> genotype &#x2018;DH Pahang&#x2019; using the program Splign (<xref ref-type="bibr" rid="B16">Kapustin et al., 2008</xref>).</p>
</sec>
<sec><title>Protein Structure Modeling</title>
<p>Secondary structures of proteins were predicted using the original Garnier Osguthorpe Robson algorithm (GOR I) provided by the European Molecular Biology Open Software Suite (EMBOSS) 6.5.7 (<xref ref-type="bibr" rid="B33">Rice et al., 2000</xref>) and implemented in Geneious version 7.1.9 as garnier tool (<xref ref-type="bibr" rid="B18">Kearse et al., 2012</xref>). Predicted protein structures from transcripts of different length, SNP and splice were visually compared to determine any variation in their secondary folding.</p>
</sec>
</sec>
<sec><title>Results</title>
<sec><title>Identification of Genomic Sequence of Banana <italic>CENH3</italic></title>
<p>To identify <italic>CENH3</italic> genomic sequence from completely sequenced banana genome [doubled haploid (DH) genotype &#x2018;DH Pahang&#x2019; (&#x2018;Malaccensis&#x2019; group)] (<xref ref-type="bibr" rid="B14">Hont et al., 2012</xref>), a BLASTN was performed for &#x2018;somewhat&#x2019; similar targets using <italic>A. thaliana CENH3</italic> to query <italic>M. acuminata</italic> whole-genome contigs. This search resulted in a total of 46 hits (<bold>Additional File <xref ref-type="supplementary-material" rid="SM1">S1</xref></bold>). To identify the exact banana <italic>CENH3</italic> genomic region(s), conserved consensus sequences at the beginning (ATGGCSMGMACSAAGCAYCCGGCSGTGMGSAARAGC) and end (GCAAGGCGWATMGGAGGRAGRAGRCATTGGTGATGA) of <italic>CENH3</italic> CDSs from four monocotyledonous plants (rice, maize, barley, and millet), referred as monocot <italic>CENH3 &#x2018;</italic>landmarks&#x2019;, were searched within the 46 BLASTN hits. A search for these consensus sequences within the 46 BLASTN hit revealed an 82 Kb contiguous sequence (GenBank accession number: <ext-link ext-link-type="DDBJ/EMBL/GenBank" xlink:href="CAIC01023700">CAIC01023700</ext-link>) as containing the putative banana <italic>CENH3</italic> genomic region. The exact location of the sequence within the 82 Kb contig CAIC01023700 was from positions 70772 to 76310 resulting in a 5538 bp long sequence.</p>
</sec>
<sec><title>Banana <italic>CENH3</italic> Sequences and Expressed Variants</title>
<p>In an effort to identify banana <italic>CENH3</italic> transcripts in each of the six banana genotypes, PCR products from amplification of cDNA template obtained from two independent synthesis reactions were cloned and sequenced. One to seven unique transcripts were obtained per genotype by sequencing of the multiple clones. The multiple clones sequenced were derived from three independent PCR amplifications of the two cDNA templates for a maximum of six reactions per genotype (<bold>Table <xref ref-type="table" rid="T1">1</xref></bold>). The genotype &#x2018;Calcutta 4&#x2019; and &#x2018;<italic>M. balbisiana</italic>&#x2019; had only one unique sequence each, where as &#x2018;Gros Michel&#x2019; and &#x2018;Pisang Awak&#x2019; had two unique sequences, &#x2018;Zebrina GF&#x2019; had four and &#x2018;Sukali Ndiizi&#x2019; had seven unique sequences (<bold>Table <xref ref-type="table" rid="T1">1</xref></bold>). All unique cDNA sequences from this study were deposited in GenBank (<bold>Table <xref ref-type="table" rid="T1">1</xref></bold>). The transcripts obtained were of variable lengths (471, 477, 504, 591, and 760 bp). The open reading frames of the cDNA sequences encoded proteins of about 156&#x2013;167 amino acids. The <italic>CENH3</italic> sequence from &#x2018;Calcutta 4&#x2019; (KT600803) was used as a reference as it had 100% identity to the exons of the publicly available genomic sequence of banana genotype &#x2018;DH Pahang&#x2019; (GenBank accession number <ext-link ext-link-type="DDBJ/EMBL/GenBank" xlink:href="CAIC01023700">CAIC01023700</ext-link> position 70772 to 76310). The protein translation of the &#x2018;Calcutta 4&#x2019; cDNA sequence resulted in a 167 amino acid long protein.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p>Description of <italic>CENH3</italic> transcripts from different genotypes of banana.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<th valign="top" align="left">Genotype</th>
<th valign="top" align="center">Genomic group</th>
<th valign="top" align="center">Banana <italic>CENH3</italic> group</th>
<th valign="top" align="center">Unique sequence identifier</th>
<th valign="top" align="center">Total number of clones</th>
<th valign="top" align="center">CDS length</th>
<th valign="top" align="center">Exon/Intron Structure</th>
<th valign="top" align="center">Functional status</th>
<th valign="top" align="center">Genbank Accession Number</th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Gros Michel</td>
<td valign="top" align="center">AAA</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">591</td>
<td valign="top" align="center">6/5</td>
<td valign="top" align="center">Non-functional</td>
<td valign="top" align="center">KP878227</td>
</tr>
<tr>
<td valign="top" align="left">Gros Michel</td>
<td valign="top" align="center">AAA</td>
<td valign="top" align="center"><italic>MusaCENH3-1A</italic></td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">7</td>
<td valign="top" align="center">504</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878231</td>
</tr>
<tr>
<td valign="top" align="left">Pisang Awak</td>
<td valign="top" align="center">ABB</td>
<td valign="top" align="center"><italic>MusaCENH3-1A</italic></td>
<td valign="top" align="center">4</td>
<td valign="top" align="center">5</td>
<td valign="top" align="center">504</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878229</td>
</tr>
<tr>
<td valign="top" align="left">Pisang Awak</td>
<td valign="top" align="center">ABB</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
<td valign="top" align="center">5</td>
<td valign="top" align="center">4</td>
<td valign="top" align="center">760</td>
<td valign="top" align="center">5/4</td>
<td valign="top" align="center">Non-functional</td>
<td valign="top" align="center">KP878228</td>
</tr>
<tr>
<td valign="top" align="left">Sukali Ndiizi</td>
<td valign="top" align="center">AAB</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
<td valign="top" align="center">G</td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">471</td>
<td valign="top" align="center">6/5</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878221</td>
</tr>
<tr>
<td valign="top" align="left">Sukali Ndiizi</td>
<td valign="top" align="center">AAB</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
<td valign="top" align="center">A</td>
<td valign="top" align="center">4</td>
<td valign="top" align="center">504</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878225</td>
</tr>
<tr>
<td valign="top" align="left">Sukali Ndiizi</td>
<td valign="top" align="center">AAB</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
<td valign="top" align="center">F</td>
<td valign="top" align="center">7</td>
<td valign="top" align="center">477</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878222</td>
</tr>
<tr>
<td valign="top" align="left">Sukali Ndiizi</td>
<td valign="top" align="center">AAB</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
<td valign="top" align="center">H</td>
<td valign="top" align="center">3</td>
<td valign="top" align="center">504</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878226</td>
</tr>
<tr>
<td valign="top" align="left">Sukali Ndiizi</td>
<td valign="top" align="center">AAB</td>
<td valign="top" align="center"><italic>MusaCENH3-2</italic></td>
<td valign="top" align="center">B</td>
<td valign="top" align="center">5</td>
<td valign="top" align="center">471</td>
<td valign="top" align="center">5/4</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878238</td>
</tr>
<tr>
<td valign="top" align="left">Sukali Ndiizi</td>
<td valign="top" align="center">AAB</td>
<td valign="top" align="center"><italic>MusaCENH3-2</italic></td>
<td valign="top" align="center">C</td>
<td valign="top" align="center">4</td>
<td valign="top" align="center">471</td>
<td valign="top" align="center">5/4</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878236</td>
</tr>
<tr>
<td valign="top" align="left">Sukali Ndiizi</td>
<td valign="top" align="center">AAB</td>
<td valign="top" align="center"><italic>MusaCENH3-2</italic></td>
<td valign="top" align="center">E</td>
<td valign="top" align="center">4</td>
<td valign="top" align="center">471</td>
<td valign="top" align="center">5/4</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878239</td>
</tr>
<tr>
<td valign="top" align="left">Zebrina GF</td>
<td valign="top" align="center">AA</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">7</td>
<td valign="top" align="center">504</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878223</td>
</tr>
<tr>
<td valign="top" align="left">Zebrina GF</td>
<td valign="top" align="center">AA</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
<td valign="top" align="center">7</td>
<td valign="top" align="center">9</td>
<td valign="top" align="center">504</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878224</td>
</tr>
<tr>
<td valign="top" align="left">Zebrina GF</td>
<td valign="top" align="center">AA</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
<td valign="top" align="center">8</td>
<td valign="top" align="center">5</td>
<td valign="top" align="center">504</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878220</td>
</tr>
<tr>
<td valign="top" align="left">Zebrina GF</td>
<td valign="top" align="center">AA</td>
<td valign="top" align="center"><italic>MusaCENH3-2</italic></td>
<td valign="top" align="center">9</td>
<td valign="top" align="center">3</td>
<td valign="top" align="center">471</td>
<td valign="top" align="center">5/4</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KP878237</td>
</tr>
<tr>
<td valign="top" align="left"><italic>Musa balbisiana</italic></td>
<td valign="top" align="center">BB</td>
<td valign="top" align="center"><italic>MusaCENH3-1A</italic></td>
<td valign="top" align="center">10</td>
<td valign="top" align="center">13</td>
<td valign="top" align="center">504</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KT600804</td>
</tr>
<tr>
<td valign="top" align="left">Calcutta 4</td>
<td valign="top" align="center">AA</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
<td valign="top" align="center">11</td>
<td valign="top" align="center">13</td>
<td valign="top" align="center">504</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center">Functional</td>
<td valign="top" align="center">KT600803</td>
</tr>
<tr>
<td valign="top" align="left"></td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Based on similarity of conserved cDNA regions (partially in the tail and entire HFD region), the <italic>CENH3</italic> sequences were clustered into three major groups denoted as <italic>MusaCENH3-1A</italic> (transcripts of <italic>M. balbisiana-10</italic>, Gros Michel-2, Pisang Awak-4) <italic>MusaCENH3-1B (</italic>Gros Michel-1, Zebrina GF-6, 7, and 8, Calcutta 4-11, Pisang Awak-5, and Sukali Ndiizi-A, F, G and H) and <italic>MusaCENH3-2</italic> (Sukali Ndiizi-B, C and E and Zebrina GF-9) (<bold>Figure <xref ref-type="fig" rid="F1">1</xref></bold> and <bold>Table <xref ref-type="table" rid="T1">1</xref></bold>). The transcripts within each group had slight variations mainly less than two SNPs. The first two groups (<italic>MusaCENH3-1A</italic> and <italic>MusaCENH3-1B</italic>) differed from <italic>MusaCENH3-2</italic> with a C to G substitution within the HFD &#x03B1;-2 helix region that resulted in alanine (A) to proline (P) substitution in the later. In addition to this HFD SNP, transcripts in <italic>MusaCENH3-2</italic> group consistently had a 46 bp longer exon 1 than <italic>MusaCENH3-1A</italic> and <italic>MusaCENH3-1B</italic> and also lacked extra two exons (exons 2 and 3), which were otherwise present in <italic>MusaCENH3-1A</italic> and <italic>MusaCENH3-1B</italic> groups. The 46 bp extra length in exon 1 as well as lack of exons 2 and 3 in <italic>MusaCENH3-2</italic> suggests that this is a different type of <italic>CENH3</italic> in bananas. However, since we did not sequence the whole genome of the genotypes used in this study, we cannot definitively prove that the missing exons or 46 bp extension are indeed different genes or splice variants (as suggested by alignment to the published sequence), but this seems likely and we will refer to them as such. There were also multiple SNPs within transcripts of each <italic>CENH3</italic> group, majority of which were within the HFD (<bold>Figures <xref ref-type="fig" rid="F2">2</xref></bold>, <bold><xref ref-type="fig" rid="F3">3</xref></bold>). In comparison to CENH3s from other monocots and dicot species, banana sequences were observed to be highly variable within the tail region and conserved only in the loop 2 of the HDF (<bold>Figure <xref ref-type="fig" rid="F2">2</xref></bold>).</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption><p>Phylogenetic tree of banana <italic>CENH3s</italic>. Unrooted Phylogenetic tree based on histone fold domain (HFD) and conserved <italic>CENH3</italic> tail sections of six banana genotypes showing the <italic>MusaCENH3-1A, MusaCENH3-1B</italic>, and <italic>MusaCENH3-2</italic> groups. Values at the root are bootstrap support values at 1000 replicates. The tree was drawn in MEGA 6 (<xref ref-type="bibr" rid="B38">Tamura et al., 2013</xref>).</p></caption>
<graphic xlink:href="fpls-08-01034-g001.tif"/>
</fig>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption><p>Multiple alignment of banana CENH3s. The blue lines separate the different alignments: Block 1 is a <italic>MusaCENH3-1A</italic> alignment, block 2 is a <italic>MusaCENH3-1B</italic>, block 3 is <italic>MusaCENH3-2</italic> and block 4 is an alignment to other monocots and dicots. The red highlights in the alignment are some amino acids substitutions observed in banana alleles within the HFD. Inset red box is the similarity index. Alignments were conducted in ClustalW (<xref ref-type="bibr" rid="B21">Larkin et al., 2007</xref>) as implemented in Geneious version 7.1 (<xref ref-type="bibr" rid="B18">Kearse et al., 2012</xref>).</p></caption>
<graphic xlink:href="fpls-08-01034-g002.tif"/>
</fig>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption><p>Alignment of transcript variants to the reference transcript from diploid banana genotype &#x2018;Calcutta 4&#x2019;. Blocks <bold>(A&#x2013;D)</bold> are alignments of genotypes &#x2018;Gros Michel&#x2019;, &#x2018;Pisang Awak&#x2019;, &#x2018;Sukali Ndiizi&#x2019;, and a combination of &#x2018;Zebrina GF&#x2019; and species &#x2018;<italic>Musa balbisiana</italic>&#x2019; to &#x2018;Calcutta 4&#x2019;, respectively. Inset in red is the nucleotide alignment similarity index.</p></caption>
<graphic xlink:href="fpls-08-01034-g003.tif"/>
</fig>
<p>The <italic>MusaCENH3-1A</italic> and <italic>MusaCENH3-1B</italic> groups were more similar to each other in both sequence and splicing in comparison to transcripts in group <italic>MusaCENH3-2</italic>. The <italic>MusaCENH3-1A</italic> and <italic>MusaCENH3-1B</italic> transcripts differed at five SNP sites (<bold>Figure <xref ref-type="fig" rid="F3">3</xref></bold>), which resulted in one non-synonymous amino acid substitution (<bold>Figure <xref ref-type="fig" rid="F2">2</xref></bold>). The <italic>MusaCENH3-1A</italic> was observed in both A and B genomes. The genotypes &#x2018;Zebrina GF&#x2019; and &#x2018;Sukali Ndiizi&#x2019; had the highest number of SNP variants observed (<bold>Table <xref ref-type="table" rid="T2">2</xref></bold>).</p>
<table-wrap position="float" id="T2">
<label>Table 2</label>
<caption><p>Minimum banana <italic>CENH3</italic> allele and splice variants.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<th valign="top" align="left">Banana Genotype</th>
<th valign="top" align="center">Minimum number of SNP-allele variants</th>
<th valign="top" align="center">Minimum number of splice variants</th>
<th valign="top" align="center">Splicing mechanism (s)</th>
<th valign="top" align="center">Splice variant <italic>CENH3</italic> group(s)</th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Gros Michel</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">7/6, 6/5</td>
<td valign="top" align="center"><italic>MusaCENH3-1A</italic> and <italic>-1B</italic></td>
</tr>
<tr>
<td valign="top" align="left">Pisang Awak</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">7/6, 5/4</td>
<td valign="top" align="center"><italic>MusaCENH3-1A</italic> and -<italic>1B</italic></td>
</tr>
<tr>
<td valign="top" align="left">Sukali Ndiizi</td>
<td valign="top" align="center">4</td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">7/6, 5/4, and 6/5</td>
<td valign="top" align="center"><italic>MusaCENH3-1A, -1B</italic> and <italic>-2</italic></td>
</tr>
<tr>
<td valign="top" align="left">Zebrina GF</td>
<td valign="top" align="center">4</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">7/6 and 5/4</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic> and <italic>-2</italic></td>
</tr>
<tr>
<td valign="top" align="left"><italic>Musa balbisiana</italic></td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">0</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center"><italic>MusaCENH3-1A</italic></td>
</tr>
<tr>
<td valign="top" align="left">Calcutta 4</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">0</td>
<td valign="top" align="center">7/6</td>
<td valign="top" align="center"><italic>MusaCENH3-1B</italic></td>
</tr>
<tr>
<td valign="top" align="left"></td>
</tr>
</tbody>
</table>
</table-wrap>
<p>These three banana <italic>CENH3</italic> groups differed in the number of exons as identified by alignment to the genomic sequence obtained through BLASTN analysis (CAIC01023700 position 70772&#x2013;76310). The alignment confirmed that <italic>MusaCENH3-1A</italic> and <italic>MusaCENH3-1B</italic> have seven exons whereas <italic>MusaCENH3-2</italic> had five exons with exemptions of specific cases that differed due to exon skipping or intron retention. The <italic>MusaCENH3-1A</italic> and <italic>MusaCENH3-1B</italic> transcripts were 471 bp to 760 bp long while those in the <italic>MusaCENH3-2</italic> group were 471 bp long. The three <italic>CENH3</italic> groups had few SNPs among them that were observed mainly in transcripts from different genotypes.</p>
<p>To check the homology of banana CENH3 proteins to those of other plant species, banana translated protein sequences were aligned to monocot (<italic>T. aestivum</italic>, <italic>O. sativa</italic>, and <italic>Z. mays</italic>) and dicot species (<italic>A. thaliana</italic> and <italic>B. rapa</italic>). This alignment resulted in conserved &#x03B1;N-helix, &#x03B1;1-helix, &#x03B1;2-helix and &#x03B1;3-helix of the C-terminal, a specific loop 1 and a highly variable N-terminal tail (<bold>Figure <xref ref-type="fig" rid="F2">2</xref></bold>). The loop 1 and &#x03B1;2-helix of the C-terminal constitute the CENP-A targeting domain (CATD) and these two domains were found to be conserved in banana sequences except for two amino acid substitutions within the &#x03B1;2-helix in the sequences <italic>M. balbisiana-10</italic> (alignment position 144) and Sukali Ndiizi-B, C, and E and in the Zebrina GF-9 (alignment position 134) (<bold>Figure <xref ref-type="fig" rid="F2">2</xref></bold>).</p>
</sec>
<sec><title>Variants in Autotriploid Genotype &#x2018;Gros Michel&#x2019;</title>
<p>&#x2018;Gros Michel&#x2019; had two variable and unique sequences as grouped in <italic>MusaCENH3-1A</italic> (Gros Michel-2) and <italic>MusaCENH3-1B</italic> (Gros Michel-1). Over and above having the five SNPs that differentiated <italic>MusaCENH3-1A</italic> from <italic>MusaCENH3-1B</italic>, the transcript Gros Michel-1 had an 87 bp indel that resulted from retention of intron 2 and spanned alignment positions 133 to 219 (<bold>Figure <xref ref-type="fig" rid="F3">3A</xref></bold>). This retained intron resulted in a frame shift and introduced a premature stop codon in the tail region (nucleotide position 219) rendering it non-functional. The transcript Gros Michel-2 differed to Calcutta 4 at nine SNPs and out of these, six were in the tail region. Two (alignment positions 105 and 276) out of six SNPs in the tail were synonymous substitutions. The other four SNPs were aligned at positions 244, 246, 247, and 277 resulted in a total of three amino acid substitutions.</p>
</sec>
<sec><title>Variants in Allotriploid Genotype &#x2018;Pisang Awak&#x2019;</title>
<p>The allotriploid cultivated genotype &#x2018;Pisang Awak&#x2019; had two unique transcripts that fell into the <italic>CENH3</italic> groups <italic>MusaCENH3-1A</italic> (Pisang Awak-5) and <italic>MusaCENH3-1B</italic> (Pisang Awak-4). Despite being in the <italic>MusaCENH3-1B</italic> group, the transcript Pisang Awak-5 had retained two introns (introns 2 and 3) (<bold>Figure <xref ref-type="fig" rid="F3">3B</xref></bold>). These retained introns resulted in a non-functional protein by introducing multiple premature stop codons the first one at nucleotide position 219 in the tail region. The transcript Pisang Awak-4 carried two additional SNPs [one in the tail and one in HFD (alignment positions 105 and 622)] in addition to the five that allowed it to be grouped into the <italic>MusaCENH3-1B.</italic> Both SNPs were silent and did not result in any amino acid substitution.</p>
</sec>
<sec><title>Variants in Allotriploid Genotype &#x2018;Sukali Ndiizi&#x2019;</title>
<p>The allotriploid genotype &#x2018;Sukali Ndiizi&#x2019; had seven variants, four within <italic>MusaCENH3-1B</italic> group (transcripts Sukali Ndiizi-A, F, G, and H) and three within <italic>MusaCENH3-2</italic> (Sukali Ndiizi-B, C, and E). Despite being in the same <italic>MusaCENH3-1B</italic> group, the sequences of Sukali Ndiizi-A and H differed to the Calcutta 4 at one SNP position each; positions 117 (A to C) and 461 (A to G) in Sukali Ndiizi-A and H, respectively, with the latter resulting in a aspartic acid (D) to glycine (G) substitution in the protein sequence (<bold>Figures <xref ref-type="fig" rid="F2">2</xref></bold>, <bold><xref ref-type="fig" rid="F3">3C</xref></bold>). The transcripts Sukali Ndiizi-F and G varied from each other with indels; Sukali Ndiizi-F had 27 bp indel (alignment position 63 &#x2013; 89) as well as substitution from T to C at position 184 which resulted in a serine (S) to proline (P) substitution in the protein translation. The 27 bp indel resulted in a shortened protein sequence with 158 amino acids due to a deletion in exon 1. The transcript Sukali Ndiizi-G on the other hand had a 37 bp indel from alignment positions 133&#x2013;169 (from skipping of exon 3), which resulted in alternative 3&#x2032; and 5&#x2032; splice sites. The two splice variations did not cause any shift in the reading frames and therefore resulted in functional proteins.</p>
<p>The transcripts falling within the <italic>MusaCENH3-2</italic> group (Sukali Ndiizi-C, B, and E) were observed to be 471 bp long, which is 33 bp shorter than those in <italic>MusaCENH3-1A</italic> and <italic>MusaCENH3-1B</italic> and especially with Calcutta 4 (<bold>Figure <xref ref-type="fig" rid="F3">3C</xref></bold>). The resultant proteins were all 156 amino acids long and functional. Despite all the three transcripts being in the same group (<italic>MusaCENH3-2</italic>) they differed among themselves at six nucleotide positions, four of which resulted in amino acid substitutions at positions 94 and 117, 179 and 181 in Sukali Ndiizi-C (<bold>Figure <xref ref-type="fig" rid="F2">2</xref></bold>).</p>
</sec>
<sec><title>Variants in Diploid Banana &#x2018;Zebrina GF&#x2019; and &#x2018;<italic>Musa balbisiana</italic>&#x2019;</title>
<p>The diploid banana genotype &#x2018;Zebrina GF&#x2019; expressed transcripts that fell into both the <italic>MusaCENH3-1B</italic> (transcripts Zebrina GF-6, 7, and 8) and <italic>MusaCENH3-2</italic> (Zebrina GF-9) categories. The three transcripts in <italic>MusaCENH3-1B</italic> differed amongst themselves with four SNPs, three of which were non-synonymous substitutions at alignment positions 34 (T to C) and 400 (A to G) in Zebrina GF-8 and position 494 in Zebrina GF-7 (G to A) (<bold>Figure <xref ref-type="fig" rid="F3">3D</xref></bold>). The only <italic>MusaCENH3-2</italic> representative sequence in this genotype was transcript Zebrina GF-9, which was 471 bp encoding a 156 amino acid long protein. This transcript, like others in the same group from other cultivars, had an exon 1 that was 45 bp longer, the C to G substitution in the &#x03B1;-2 helix and in addition an A to G non-synonymous substitution at alignment position 80 (<bold>Figure <xref ref-type="fig" rid="F3">3D</xref></bold>) that resulted in glycine (Q) to argenine (R) substitution at protein alignment position 28 (<bold>Figure <xref ref-type="fig" rid="F2">2</xref></bold>).</p>
<p>The diploid species &#x2018;<italic>M. balbisiana&#x2019;</italic> had one unique 504 bp long sequence that encoded a 167 amino acid long protein. This transcript fell into the <italic>MusaCENH3-1A</italic> group and differed from others in the same group with one major non-synonymous SNP site in the HFD that resulted in the substitution of the amino acid threonine (T) to isoleucine (I) at alignment position 144.</p>
</sec>
<sec><title>Exon/Intron Structures in Banana <italic>CENH3</italic></title>
<p>To get an insight into the splicing approaches and the intron/exon structures of the transcripts obtained and to also know if the differences in lengths of the transcripts were due to splicing variations, the unique banana <italic>CENH3</italic> transcripts were mapped to genomic sequence of putative <italic>CENH3</italic> from &#x2018;DH Pahang&#x2019; (<bold>Figure <xref ref-type="fig" rid="F4">4</xref></bold> and <bold>Additional File <xref ref-type="supplementary-material" rid="SM2">S2</xref></bold>). Three exon/intron structures (7/6, 6/5, and 5/4) were observed, which were probably as a result of differences in splicing patterns (<bold>Figure <xref ref-type="fig" rid="F4">4</xref></bold>). The 7 exon/6 intron structure was most frequently observed (10 transcripts out of 17 unique clones). This structure was observed in both diploid and triploid genotypes with three of the four transcripts from the diploid &#x2018;Zebrina GF&#x2019; (Zebrina GF-6, 7, and 8), diploid &#x2018;Calcutta 4&#x2019; and &#x2018;<italic>M. balbisiana&#x2019;</italic>, triploid genotype &#x2018;Pisang Awak&#x2019; (Pisang Awak-4), in three of the seven sequences in the genotype &#x2018;Sukali Ndiizi&#x2019; (Sukali Ndiizi-A, F, and H) and in one transcript from the autopolyploid &#x2018;Gros Michel&#x2019; (Gros Michel-2) (<bold>Table <xref ref-type="table" rid="T1">1</xref></bold>).</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption><p>Alignment of the <italic>CENH3</italic> transcripts to genomic sequence of genotype &#x2018;DH Pahang&#x2019; to identify splice mechanisms. The 7/6, 6/5, and 5/4 structures are represented intron/exon structures are represented. Alignment was performed using Splign (<xref ref-type="bibr" rid="B16">Kapustin et al., 2008</xref>).</p></caption>
<graphic xlink:href="fpls-08-01034-g004.tif"/>
</fig>
<p>The 5 exon/4 intron structure was observed in genotypes &#x2018;Pisang Awak&#x2019; (Pisang Awak-5), &#x2018;Sukali Ndiizi&#x2019; (Sukali Ndiizi-B, C, and E) and the diploid &#x2018;Zebrina GF&#x2019; (Zebrina GF-9) (<bold>Figure <xref ref-type="fig" rid="F4">4</xref></bold>). This structure resulted from skipping of the second and the third exons in all sequences apart from Pisang Awak-5, which had this structure due to retention of introns two and three.</p>
<p>The 6 exon/5 intron pattern was observed in two transcripts of Gros Michel-1 and Sukali Ndiizi-G from the genotype &#x2018;Gros Michel&#x2019; and &#x2018;Sukali Ndiizi&#x2019;, respectively (<bold>Figure <xref ref-type="fig" rid="F4">4</xref></bold>). This structure was as a result of skipping of exon 2 for Sukali Ndiizi-G and retention of intron 2 in the transcript Gros Michel-1.</p>
</sec>
<sec><title>Alternative Splicing of <italic>CENH3</italic> in Banana</title>
<p>Alternative splicing achieves diversity and novelty of proteins. Alternatively spliced variants were obtained based on deviations from splicing in their respective banana <italic>CENH3</italic> groups. Four out of the seventeen unique transcripts were alternatively spliced with two of these resulting in unique proteins while the rest introduced premature stop codons. Some of the alternatively spliced transcripts also had SNP variations. Three alternative splicing approaches were observed: exon skipping, intron retention and alternate 3&#x2032; and 5&#x2032; splice site (<bold>Figure <xref ref-type="fig" rid="F4">4</xref></bold> and <bold>Additional File <xref ref-type="supplementary-material" rid="SM2">S2</xref></bold>).</p>
</sec>
<sec><title>Alternative Splicing by Exon Skipping</title>
<p>Exon skipping was observed in only one transcript (Sukali Ndiizi-G) of the triploid cultivar &#x2018;Sukali Ndiizi&#x2019;. This alternate splicing mechanism resulted from skipping of exon 3 and resulted in a functional transcript (<bold>Figure <xref ref-type="fig" rid="F4">4</xref></bold> and <bold>Additional File <xref ref-type="supplementary-material" rid="SM2">S2</xref></bold>). Exon skipping resulted in shorter transcript length, where the transcript affected (Sukali Ndiizi-G) had a reduced length of 471 bp instead of the 504 bp in transcripts from the same <italic>CENH3</italic> group.</p>
</sec>
<sec><title>Alternative Splicing by Intron Retention</title>
<p>Intron retention as an alternative splicing mechanism was observed in two transcripts (Gros Michel-1 and Pisang Awak-5), which retained one and two introns, respectively (<bold>Figure <xref ref-type="fig" rid="F4">4</xref></bold>). The intron retention resulted in non-functional proteins due to introduction of at least one stop codon in either of the two transcripts. The transcript Pisang Awak-5 had five stop codons introduced, four in the tail and one in the HFD, whereas Gros Michel-1 only had one stop codon in the tail region.</p>
</sec>
<sec><title>Splice Variation by Alternative Splice Site Selection</title>
<p>The alternate 3&#x2032; and 5&#x2032; splice site selection resulted in variation in the length of exon 1 (<bold>Figure <xref ref-type="fig" rid="F4">4</xref></bold>). Partial deletion of a 27 bp segment from positions 63&#x2013;89 of exon 1 was observed in Sukali Ndiizi-F. This deletion resulted in a change of the splice junction from CCCC/GGTC to TTTC/GGTC resulting in a change in splice sites. The transcript Sukali Ndiizi-G was also observed to have a different splice site selection by retaining the nucleotide G from intron 2 (Exon 3 was skipped) and retaining the nucleotide C of intron 3.</p>
</sec>
<sec><title>Secondary Structure Prediction</title>
<p>There was general conservation in predicted secondary structure within each of the banana <italic>CENH3</italic> groups, although slight variations at specific sections were observed (<bold>Figure <xref ref-type="fig" rid="F5">5</xref></bold>). The <italic>MusaCENH3-1A</italic> and <italic>MusaCENH3-1B</italic> had similar predicted secondary folding and varied in the second and the third last turns of the N-terminal tail where they were merged into one due to the lack of the predicted intervening beta sheet in the <italic>MusaCENH3-1A</italic> group. The secondary structures in the <italic>MusaCENH3-2</italic> had more structural variation in comparison to the <italic>MusaCENH3-1A</italic> and <italic>MusaCENH3-1B</italic>. The structural modifications included addition, loss, elongation or shortening of coils, turns, &#x03B1;-helices and &#x03B2;-sheets. Splice variations affecting the tail region resulted in loss of &#x03B1;-helices and beta strands, coils and turns in Sukali Ndiizi-G and F (<bold>Figure <xref ref-type="fig" rid="F5">5</xref></bold>). The CENH3 proteins for Sukali Ndiizi-C, B, E, and Zebrina GF-9 also gained and lost domains within the tail region. The major form of variation observed within the HFD was point mutations some of which resulted in non-synonymous substitution. These substitutions resulted in elongations and/or shortening of some predicted secondary structures. The Proline (P) to Alanine (A) substitution within the HFD in Sukali Ndiizi-C, B, E, and Zebrina GF-9 resulted in an elongated loop 1 and a shortened &#x03B1;2 helix a clear structural feature unique to the <italic>MusaCENH3-2</italic> CENH3 groups (<bold>Figure <xref ref-type="fig" rid="F5">5</xref></bold>). The substitutions of glutamine (E) with glycine (G) at two different positions in Sukali Ndiizi-C resulted in structural changes within the &#x03B1;N- and &#x03B1;1-helices (<bold>Figure <xref ref-type="fig" rid="F5">5</xref></bold>).</p>
<fig id="F5" position="float">
<label>FIGURE 5</label>
<caption><p>Predicted secondary structures of banana CENH3. 1- <italic>MusaCENH3-1A</italic>, 2<italic>- MusaCENH3-1B</italic>, and 3- <italic>MusaCENH3-2.</italic> At the bottom are the different CENH3 domains, above the structures are the amino acids in logo format. Inset is the key to the secondary structures.</p></caption>
<graphic xlink:href="fpls-08-01034-g005.tif"/>
</fig>
</sec>
</sec>
<sec><title>Discussion</title>
<p>In this study, we observed that <italic>CENH3</italic> in diploid and triploid bananas exists as single or multiple allele variants depending on the genotype. The variants were SNPs in both the tail and the HFD region of CENH3. The non-synonymous SNPs resulted in modification of the predicted secondary structures of proteins. The majority of splice variants (apart from two) were predicted to translate in-frame. These splice variations only affected the tail region of CENH3. The presence of multiple <italic>CENH3</italic> SNP-alleles in a diploid genotype like &#x2018;Zebrina GF&#x2019; suggests that bananas maybe carrying more than one <italic>CENH3</italic> gene per genome.</p>
<p>Cultivated bananas are mainly triploids. Banana breeding involves crossing fertile triploids with diploids to get tetraploids which are then crossed to diploid accessions to give triploid cultivars (<xref ref-type="bibr" rid="B29">Pillay et al., 2004</xref>). In other plant species, the presence of different <italic>CENH3s</italic> from different parents in embryos has been observed to result in uniparental genome elimination, aneuploids, or stable hybrids (<xref ref-type="bibr" rid="B31">Ravi and Chan, 2010</xref>; <xref ref-type="bibr" rid="B20">Kuppu et al., 2015</xref>; <xref ref-type="bibr" rid="B23">Maheshwari et al., 2015</xref>; <xref ref-type="bibr" rid="B39">Tan et al., 2015</xref>; <xref ref-type="bibr" rid="B19">Kelliher et al., 2016</xref>). The focus in banana breeding programs is to first establish tetraploids and then use them to develop triploids. It has been suggested that crosses between diploids and triploids result in viable diploids (<xref ref-type="bibr" rid="B7">De Langhe et al., 2010</xref>). Crosses between <italic>A. thaliana</italic> wild-type parents and pollen donors carrying specific point mutations within the HFD resulted in uniparental genome elimination- with loss of the genome derived from the mutant line (<xref ref-type="bibr" rid="B20">Kuppu et al., 2015</xref>). Two out of the five point mutations that resulted in uniparental genome elimination in Arabidopsis were within the centromere targeting domain (CATD). Some of the SNPs in our study were observed to be within the conserved HFD domains including &#x03B1;2- and &#x03B1;3-helices. Furthermore, mutations of CENP-A (CENH3 of humans) residues resulted in reduced retention of CENP-A in centromere of human cells and this was due to the effect on &#x03B1;2-helix length which plays a key role of maintaining orientation at nucleosome entry and exit (<xref ref-type="bibr" rid="B36">Tachiwana et al., 2011</xref>). The SNPs were within the CATD and these resulted in a predicted shortening of the &#x03B1;2-helix. These CATD SNP variations can result in CENH3 nucleosome instability and may affect crosses of bananas having <italic>CENH3s</italic> with variation at these positions. It would be interesting to know if multiple <italic>CENH3s</italic> affect banana breeding and if they do, then it may be important to consider <italic>CENH3</italic> type when choosing parents for crossing.</p>
<p>We observed three <italic>CENH3</italic> variants in both diploid and triploid bananas, which differed at the tail region and have SNPs in the HFD region. The presence of these three variants in a diploid line indicates presence of more than one <italic>CENH3</italic> in a single banana genome. Alpha and beta <italic>CENH3</italic> variants in wheat were observed to have different functional roles. Reduced expression of alpha version resulted in extreme dwarfing and weakened root system whereas reduced expression of the beta version resulted in reduced plant height and reproductive fitness leading to the conclusion that the two versions are involved in plant development and reproductive development, respectively (<xref ref-type="bibr" rid="B43">Yuan et al., 2015</xref>). Although this study did not explore the functions of the banana <italic>CENH3</italic> variants, it would be worth conducting such studies in future to verify if the variants differ in functionality.</p>
<p>The presence of multiple <italic>CENH3</italic> allele variants in a wild diploid banana (four in diploid &#x2018;Zebrina GF&#x2019;) corroborate the hypothesis that domestication of cultivated hybrids passed through intermediate hybrids (<xref ref-type="bibr" rid="B7">De Langhe et al., 2010</xref>). &#x2018;Zebrina GF&#x2019; is a wild diploid that has been shown to segregate during crosses, an indication that it has a high degree of heterozygozity (personal communication from Professor Rony Swennen, Banana breeder at IITA and collector of this genotype).</p>
<p>The observation that alternative splicing of <italic>CENH3</italic> in bananas only affected the N-terminal tail is consistent with those made in the angiosperms <italic>Oryza</italic> spp., <italic>Brassica</italic> spp., and <italic>Gossypium</italic> spp. (<xref ref-type="bibr" rid="B41">Wang et al., 2011</xref>; <xref ref-type="bibr" rid="B26">Masonbrink et al., 2014</xref>). It was interesting to observe that some of the splice variations resulted in transcripts that were translatable into proteins as these could further add into the diversity of banana <italic>CENH3s.</italic> However, it is not clear if the in-frame splice variants translate into proteins <italic>in vivo</italic> and whether they are loaded into the centromere. The role of the out-of-frame variants is also not clear and future studies targeting the <italic>CENH3</italic> splice variants and their proteins (if translated) are required to identify their role(s) and fate.</p>
<p>The variations observed in the three main banana <italic>CENH3</italic> groups were observed to affect the predicted secondary structures of the respective proteins. This is interesting considering that crosses of Arabidopsis null mutant lines complemented with a <italic>CENH3</italic> version in which tail was replaced with that of histone H3.3 and GFP-tagged to wild-type resulted in uniparental genome elimination (<xref ref-type="bibr" rid="B31">Ravi and Chan, 2010</xref>). The highly variable N-terminal tail of the <italic>CENH3</italic> indicates its role in the evolving centromeric satellites (<xref ref-type="bibr" rid="B15">Hui et al., 2010</xref>; <xref ref-type="bibr" rid="B32">Ravi et al., 2010</xref>; <xref ref-type="bibr" rid="B12">Hayden and Willard, 2012</xref>) or affecting the targeting of centromeres that might be a mode of bringing in new CENH3 proteins in response to increased centromere size (<xref ref-type="bibr" rid="B26">Masonbrink et al., 2014</xref>).</p>
<p>The frequency of non-synonymous SNPs within each of the banana <italic>CENH3</italic> groups was observed to be higher within the HFD region, while the frequency of both synonymous and non-synonymous SNPs between different <italic>CENH3</italic> groups was higher in the tail region. A study on evolution of <italic>CENH3</italic> in drosophila observed that the frequency of interspecific <italic>CENH3</italic> polymorphisms were higher in the tail than the HFD although the ratios of such changes were lower within the same species (<xref ref-type="bibr" rid="B24">Malik and Henikoff, 2001</xref>). One of the <italic>CENH3</italic> groups (<italic>MusaCENH3-1A</italic>) was observed to be specific to the diploid <italic>Musa</italic> species &#x2018;<italic>M. balbisiana</italic>&#x2019; and differed from other <italic>CENH3</italic> groups with non-synonymous substitutions, majority of which were in the tail region. Majority of the non-synonymous SNPs in transcripts within banana <italic>CENH3</italic> group were observed to be within the CATD, which may affect CENH3 targeting to the centromere because the CATD specifically the loop 1 has been shown to be involved in localization (<xref ref-type="bibr" rid="B5">Dalal et al., 2007</xref>).</p>
<p>The number of <italic>CENH3</italic> exons and introns in the respective exon/intron structures has been found to vary in different plant species. Seven <italic>CENH3</italic> transcripts obtained from five rice species were observed to have 7 exons and 6 introns despite having different CDS lengths (<xref ref-type="bibr" rid="B13">Hirsch et al., 2009</xref>). In carrots, a similar structure of 7 exons and 6 introns was observed while in brassica two different structures were observed in <italic>CENH3s</italic> of varying lengths, one with 7 exons/6 introns and second with 9 exons/8 introns structure. In this study, three exons/introns structures (7/6, 6/5, and 5/4) were observed in bananas. The 7/6 and the 5/4 exon/intron structures were found in both diploids and triploids, however, 7/6 was most predominant. The 6 exons/5 introns structure was only observed in triploid bananas and this mechanism resulted in functional and non-functional transcripts. In this analysis, it is clear that there was more bias toward having a 7 exons/6 introns structure whereas the 5 exons/4 introns structure was minor and the 6/5 structure was a result of failure of the 7/6 to splice correctly.</p>
<p>This study provided insight into how <italic>CENH3</italic> is expressed in diploid and triploid bananas. Additional genotypes including tetraploids should be included in future studies. Due to the emergence of CENH3-based breeding techniques, the knowledge obtained here indicates that checking the <italic>CENH3</italic> type may be used as a criterion in selection of parents for banana breeding.</p>
</sec>
<sec><title>Availability of Supporting Data</title>
<p>Gene sequences for banana <italic>CENH3</italic> obtained in this study were deposited in the GenBank and the accession numbers are provided in this manuscript, all other supporting data is providing as additional files.</p>
</sec>
<sec><title>Author Contributions</title>
<p>KM, LT, AB, and SC conceived the idea and designed the experiments. KM performed the experiments and wrote the manuscript. LT and AB supervised the experimentation. All authors contributed in interpreting the data. KM, AB, LT, NA, and EN contributed to reviewing and editing the manuscript.</p>
</sec>
<sec><title>Conflict of Interest Statement</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
</body>
<back>
<ack>
<p>This work was supported by funding from National Science Foundation (NSF) and the Bill &#x0026; Melinda Gates Foundation (BMGF) under the program Basic Research Enabling Agricultural Development (BREAD) project number 1109882: Fast Breeding for Slow Cycling Crops: Doubled Haploids in Cassava and Banana/Plantain.</p>
</ack>
<sec sec-type="supplementary material">
<title>Supplementary Material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="http://journal.frontiersin.org/article/10.3389/fpls.2017.01034/full#supplementary-material">http://journal.frontiersin.org/article/10.3389/fpls.2017.01034/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Data_Sheet_1.PDF" id="SM1" mimetype="application/pdf" xmlns:xlink="http://www.w3.org/1999/xlink">
<p><bold>ADDITIONAL FILE S1 &#x007C;</bold> BLAST hits on <italic>Musa acuminata</italic> whole genome shot gun contigs using <italic>A. thaliana CENH3</italic> genomic sequence as the query.</p>
</supplementary-material>
<supplementary-material xlink:href="Data_Sheet_1.PDF" id="S1" mimetype="application/pdf" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Data_Sheet_2.pdf" id="SM2" mimetype="application/pdf" xmlns:xlink="http://www.w3.org/1999/xlink">
<p><bold>ADDITIONAL FILE S2 &#x007C;</bold> Alignment of banana <italic>CENH3</italic> transcripts to the putative genomic sequence obtained using <italic>A. thaliana CENH3</italic> genomic sequence to query <italic>M. acuminata</italic> whole genome contig.</p>
</supplementary-material>
<supplementary-material xlink:href="Data_Sheet_2.pdf" id="S2" mimetype="application/pdf" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<ref-list>
<title>References</title>
<ref id="B1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Allshire</surname> <given-names>R. C.</given-names></name> <name><surname>Karpen</surname> <given-names>G. H.</given-names></name></person-group> (<year>2008</year>). <article-title>Epigenetic regulation of centromeric chromatin: old dogs, new tricks?</article-title> <source><italic>Nat. Rev. Genet.</italic></source> <volume>9</volume> <fpage>923</fpage>&#x2013;<lpage>937</lpage>. <pub-id pub-id-type="doi">10.1038/nrg2466</pub-id></citation></ref>
<ref id="B2"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Black</surname> <given-names>B. E.</given-names></name> <name><surname>Cleveland</surname> <given-names>D. W.</given-names></name></person-group> (<year>2011</year>). <article-title>Epigenetic centromere propagation and the nature of CENP-A nucleosomes.</article-title> <source><italic>Cell</italic></source> <volume>144</volume> <fpage>471</fpage>&#x2013;<lpage>479</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2011.02.002</pub-id></citation></ref>
<ref id="B3"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Britt</surname> <given-names>A. B.</given-names></name> <name><surname>Kuppu</surname> <given-names>S.</given-names></name></person-group> (<year>2016</year>). <article-title>Cenh3: an emerging player in haploid induction technology.</article-title> <source><italic>Front. Plant Sci.</italic></source> <volume>7</volume>:<issue>357</issue>. <pub-id pub-id-type="doi">10.3389/fpls.2016.00357</pub-id></citation></ref>
<ref id="B4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Comai</surname> <given-names>L.</given-names></name></person-group> (<year>2014</year>). <article-title>Genome elimination: translating basic research into a future tool for plant breeding.</article-title> <source><italic>PLoS Biol.</italic></source> <volume>12</volume>:<issue>e1001876</issue>. <pub-id pub-id-type="doi">10.1371/journal.pbio.1001876</pub-id></citation></ref>
<ref id="B5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dalal</surname> <given-names>Y.</given-names></name> <name><surname>Furuyama</surname> <given-names>T.</given-names></name> <name><surname>Vermaak</surname> <given-names>D.</given-names></name> <name><surname>Henikoff</surname> <given-names>S.</given-names></name></person-group> (<year>2007</year>). <article-title>Structure, dynamics, and evolution of centromeric nucleosomes.</article-title> <source><italic>Proc. Natl. Acad. Sci. U.S.A.</italic></source> <volume>104</volume> <fpage>15974</fpage>&#x2013;<lpage>15981</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.0707648104</pub-id></citation></ref>
<ref id="B6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dawe</surname> <given-names>R. K.</given-names></name> <name><surname>Henikoff</surname> <given-names>S.</given-names></name></person-group> (<year>2006</year>). <article-title>Centromeres put epigenetics in the driver&#x2019;s seat.</article-title> <source><italic>Trends Biochem. Sci.</italic></source> <volume>31</volume> <fpage>662</fpage>&#x2013;<lpage>669</lpage>. <pub-id pub-id-type="doi">10.1016/j.tibs.2006.10.004</pub-id></citation></ref>
<ref id="B7"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>De Langhe</surname> <given-names>E.</given-names></name> <name><surname>H&#x0159;ibov&#x00E1;</surname> <given-names>E.</given-names></name> <name><surname>Carpentier</surname> <given-names>S.</given-names></name> <name><surname>Doleel</surname> <given-names>J.</given-names></name> <name><surname>Swennen</surname> <given-names>R.</given-names></name></person-group> (<year>2010</year>). <article-title>Did backcrossing contribute to the origin of hybrid edible bananas?</article-title> <source><italic>Ann. Bot.</italic></source> <volume>106</volume> <fpage>849</fpage>&#x2013;<lpage>857</lpage>. <pub-id pub-id-type="doi">10.1093/aob/mcq187</pub-id></citation></ref>
<ref id="B8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dunemann</surname> <given-names>F.</given-names></name> <name><surname>Schrader</surname> <given-names>O.</given-names></name> <name><surname>Budahn</surname> <given-names>H.</given-names></name> <name><surname>Houben</surname> <given-names>A.</given-names></name></person-group> (<year>2014</year>). <article-title>Characterization of centromeric histone H3 (CENH3) variants in cultivated and wild carrots (<italic>Daucus</italic> sp.).</article-title> <source><italic>PLoS ONE</italic></source> <volume>9</volume>:<issue>e98504</issue>. <pub-id pub-id-type="doi">10.1371/journal.pone.0098504</pub-id></citation></ref>
<ref id="B9"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ekwall</surname> <given-names>K.</given-names></name></person-group> (<year>2007</year>). <article-title>Epigenetic control of centromere behavior.</article-title> <source><italic>Annu. Rev. Genet.</italic></source> <volume>41</volume> <fpage>63</fpage>&#x2013;<lpage>81</lpage>. <pub-id pub-id-type="doi">10.1146/annurev.genet.41.110306.130127</pub-id></citation></ref>
<ref id="B10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fachinetti</surname> <given-names>D.</given-names></name> <name><surname>Folco</surname> <given-names>H. D.</given-names></name> <name><surname>Nechemia-Arbely</surname> <given-names>Y.</given-names></name> <name><surname>Valente</surname> <given-names>L. P.</given-names></name> <name><surname>Nguyen</surname> <given-names>K.</given-names></name> <name><surname>Wong</surname> <given-names>A. J.</given-names></name><etal/></person-group> (<year>2013</year>). <article-title>A two-step mechanism for epigenetic specification of centromere identity and function.</article-title> <source><italic>Nat. Cell Biol.</italic></source> <volume>15</volume> <fpage>1056</fpage>&#x2013;<lpage>1066</lpage>. <pub-id pub-id-type="doi">10.1038/ncb2805</pub-id></citation></ref>
<ref id="B11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gent</surname> <given-names>J. I.</given-names></name> <name><surname>Schneider</surname> <given-names>K. L.</given-names></name> <name><surname>Topp</surname> <given-names>C. N.</given-names></name> <name><surname>Rodriguez</surname> <given-names>C.</given-names></name> <name><surname>Presting</surname> <given-names>G. G.</given-names></name> <name><surname>Dawe</surname> <given-names>R. K.</given-names></name></person-group> (<year>2011</year>). <article-title>Distinct influences of tandem repeats and retrotransposons on CENH3 nucleosome positioning.</article-title> <source><italic>Epigenetics Chromatin</italic></source> <volume>4</volume> <issue>3</issue>. <pub-id pub-id-type="doi">10.1186/1756-8935-4-3</pub-id></citation></ref>
<ref id="B12"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hayden</surname> <given-names>K. E.</given-names></name> <name><surname>Willard</surname> <given-names>H. F.</given-names></name></person-group> (<year>2012</year>). <article-title>Composition and organization of active centromere sequences in complex genomes.</article-title> <source><italic>BMC Genomics</italic></source> <volume>13</volume>:<issue>324</issue>. <pub-id pub-id-type="doi">10.1186/1471-2164-13-324</pub-id></citation></ref>
<ref id="B13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hirsch</surname> <given-names>C. D.</given-names></name> <name><surname>Wu</surname> <given-names>Y.</given-names></name> <name><surname>Yan</surname> <given-names>H.</given-names></name> <name><surname>Jiang</surname> <given-names>J.</given-names></name></person-group> (<year>2009</year>). <article-title>Lineage-specific adaptive evolution of the centromeric protein CENH3 in diploid and allotetraploid <italic>Oryza</italic> species.</article-title> <source><italic>Mol. Biol. Evol.</italic></source> <volume>26</volume> <fpage>2877</fpage>&#x2013;<lpage>2885</lpage>. <pub-id pub-id-type="doi">10.1093/molbev/msp208</pub-id></citation></ref>
<ref id="B14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hont</surname> <given-names>D.</given-names></name> <name><surname>Denoeud</surname> <given-names>F.</given-names></name> <name><surname>Aury</surname> <given-names>J.</given-names></name> <name><surname>Baurens</surname> <given-names>F.</given-names></name> <name><surname>Carreel</surname> <given-names>F.</given-names></name> <name><surname>Garsmeur</surname> <given-names>O.</given-names></name><etal/></person-group> (<year>2012</year>). <article-title>The banana (<italic>Musa acuminata</italic>) genome and the evolution of monocotyledonous plants.</article-title> <source><italic>Nature</italic></source> <volume>488</volume> <fpage>213</fpage>&#x2013;<lpage>217</lpage>. <pub-id pub-id-type="doi">10.1038/nature11241</pub-id></citation></ref>
<ref id="B15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hui</surname> <given-names>L.</given-names></name> <name><surname>Lu</surname> <given-names>L.</given-names></name> <name><surname>Heng</surname> <given-names>Y.</given-names></name> <name><surname>Qin</surname> <given-names>R.</given-names></name> <name><surname>Xing</surname> <given-names>Y.</given-names></name> <name><surname>Jin</surname> <given-names>W.</given-names></name></person-group> (<year>2010</year>). <article-title>Expression of CENH3 alleles in synthesized allopolyploid <italic>Oryza</italic> species.</article-title> <source><italic>J. Genet. Genomics</italic></source> <volume>37</volume> <fpage>703</fpage>&#x2013;<lpage>711</lpage>. <pub-id pub-id-type="doi">10.1016/S1673-8527(09)60088-6</pub-id></citation></ref>
<ref id="B16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kapustin</surname> <given-names>Y.</given-names></name> <name><surname>Souvorov</surname> <given-names>A.</given-names></name> <name><surname>Tatusova</surname> <given-names>T.</given-names></name> <name><surname>Lipman</surname> <given-names>D.</given-names></name></person-group> (<year>2008</year>). <article-title>Splign: algorithms for computing spliced alignments with identification of paralogs.</article-title> <source><italic>Biol. Direct</italic></source> <volume>3</volume>:<issue>20</issue>. <pub-id pub-id-type="doi">10.1186/1745-6150-3-20</pub-id></citation></ref>
<ref id="B17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kawabe</surname> <given-names>A.</given-names></name> <name><surname>Nasuda</surname> <given-names>S.</given-names></name> <name><surname>Charlesworth</surname> <given-names>D.</given-names></name></person-group> (<year>2006</year>). <article-title>Duplication of centromeric histone H3 (HTR12) gene in <italic>Arabidopsis halleri</italic> and <italic>A. lyrata</italic>, plant species with multiple centromeric satellite sequences.</article-title> <source><italic>Genetics</italic></source> <volume>174</volume> <fpage>2021</fpage>&#x2013;<lpage>2032</lpage>. <pub-id pub-id-type="doi">10.1534/genetics.106.063628</pub-id></citation></ref>
<ref id="B18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kearse</surname> <given-names>M.</given-names></name> <name><surname>Moir</surname> <given-names>R.</given-names></name> <name><surname>Wilson</surname> <given-names>A.</given-names></name> <name><surname>Stones-Havas</surname> <given-names>S.</given-names></name> <name><surname>Cheung</surname> <given-names>M.</given-names></name> <name><surname>Sturrock</surname> <given-names>S.</given-names></name><etal/></person-group> (<year>2012</year>). <article-title>Geneious Basic: an integrated and extendable desktop software platform for the organization and analysis of sequence data.</article-title> <source><italic>Bioinformatics</italic></source> <volume>28</volume> <fpage>1647</fpage>&#x2013;<lpage>1649</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/bts199</pub-id></citation></ref>
<ref id="B19"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kelliher</surname> <given-names>T.</given-names></name> <name><surname>Starr</surname> <given-names>D.</given-names></name> <name><surname>Wang</surname> <given-names>W.</given-names></name> <name><surname>McCuiston</surname> <given-names>J.</given-names></name> <name><surname>Zhong</surname> <given-names>H.</given-names></name> <name><surname>Nuccio</surname> <given-names>M. L.</given-names></name><etal/></person-group> (<year>2016</year>). <article-title>Maternal haploids are preferentially induced by CENH3-tailswap transgenic complementation in maize.</article-title> <source><italic>Front. Plant Sci.</italic></source> <volume>7</volume>:<issue>414</issue>. <pub-id pub-id-type="doi">10.3389/fpls.2016.00414</pub-id></citation></ref>
<ref id="B20"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kuppu</surname> <given-names>S.</given-names></name> <name><surname>Tan</surname> <given-names>E. H.</given-names></name> <name><surname>Nguyen</surname> <given-names>H.</given-names></name> <name><surname>Rodgers</surname> <given-names>A.</given-names></name> <name><surname>Comai</surname> <given-names>L.</given-names></name> <name><surname>Chan</surname> <given-names>S. W. L.</given-names></name><etal/></person-group> (<year>2015</year>). <article-title>Point mutations in centromeric histone induce post-zygotic incompatibility and uniparental inheritance.</article-title> <source><italic>PLoS Genet.</italic></source> <volume>11</volume>:<issue>e1005494</issue>. <pub-id pub-id-type="doi">10.1371/journal.pgen.1005494</pub-id></citation></ref>
<ref id="B21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Larkin</surname> <given-names>M. A.</given-names></name> <name><surname>Blackshields</surname> <given-names>G.</given-names></name> <name><surname>Brown</surname> <given-names>N. P.</given-names></name> <name><surname>Chenna</surname> <given-names>R.</given-names></name> <name><surname>Mcgettigan</surname> <given-names>P. A.</given-names></name> <name><surname>McWilliam</surname> <given-names>H.</given-names></name><etal/></person-group> (<year>2007</year>). <article-title>Clustal W and Clustal X version 2.0.</article-title> <source><italic>Bioinformatics</italic></source> <volume>23</volume> <fpage>2947</fpage>&#x2013;<lpage>2948</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btm404</pub-id></citation></ref>
<ref id="B22"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lermontova</surname> <given-names>I.</given-names></name> <name><surname>Sandmann</surname> <given-names>M.</given-names></name> <name><surname>Demidov</surname> <given-names>D.</given-names></name></person-group> (<year>2014</year>). <article-title>Centromeres and kinetochores of Brassicaceae.</article-title> <source><italic>Chromosome Res.</italic></source> <volume>22</volume> <fpage>135</fpage>&#x2013;<lpage>152</lpage>. <pub-id pub-id-type="doi">10.1007/s10577-014-9422-z</pub-id></citation></ref>
<ref id="B23"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Maheshwari</surname> <given-names>S.</given-names></name> <name><surname>Tan</surname> <given-names>E. H.</given-names></name> <name><surname>West</surname> <given-names>A.</given-names></name> <name><surname>Franklin</surname> <given-names>F. C. H.</given-names></name> <name><surname>Comai</surname> <given-names>L.</given-names></name> <name><surname>Chan</surname> <given-names>S. W. L.</given-names></name></person-group> (<year>2015</year>). <article-title>Naturally occurring differences in CENH3 affect chromosome segregation in zygotic mitosis of hybrids.</article-title> <source><italic>PLOS Genet.</italic></source> <volume>11</volume>:<issue>e1004970</issue>. <pub-id pub-id-type="doi">10.1371/journal.pgen.1004970</pub-id></citation></ref>
<ref id="B24"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Malik</surname> <given-names>H. S.</given-names></name> <name><surname>Henikoff</surname> <given-names>S.</given-names></name></person-group> (<year>2001</year>). <article-title>Adaptive evolution of Cid, a centromere-specific histone in Drosophila.</article-title> <source><italic>Genetics</italic></source> <volume>157</volume> <fpage>1293</fpage>&#x2013;<lpage>1298</lpage>.</citation></ref>
<ref id="B25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Malik</surname> <given-names>H. S.</given-names></name> <name><surname>Henikoff</surname> <given-names>S.</given-names></name></person-group> (<year>2009</year>). <article-title>Major evolutionary transitions in centromere complexity.</article-title> <source><italic>Cell</italic></source> <volume>138</volume> <fpage>1067</fpage>&#x2013;<lpage>1082</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2009.08.036</pub-id></citation></ref>
<ref id="B26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Masonbrink</surname> <given-names>R. E.</given-names></name> <name><surname>Gallagher</surname> <given-names>J. P.</given-names></name> <name><surname>Jareczek</surname> <given-names>J. J.</given-names></name> <name><surname>Renny-Byfield</surname> <given-names>S.</given-names></name> <name><surname>Grover</surname> <given-names>C. E.</given-names></name> <name><surname>Gong</surname> <given-names>L.</given-names></name><etal/></person-group> (<year>2014</year>). <article-title>CenH3 evolution in diploids and polyploids of three angiosperm genera.</article-title> <source><italic>BMC Plant Biol.</italic></source> <volume>14</volume>:<issue>383</issue>. <pub-id pub-id-type="doi">10.1186/s12870-014-0383-3</pub-id></citation></ref>
<ref id="B27"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Neumann</surname> <given-names>P.</given-names></name> <name><surname>Navr&#x00E1;tilov&#x00E1;</surname> <given-names>A.</given-names></name> <name><surname>Schroeder-Reiter</surname> <given-names>E.</given-names></name> <name><surname>Kobl&#x00ED;&#x017E;kov&#x00E1;</surname> <given-names>A.</given-names></name> <name><surname>Steinbauerov&#x00E1;</surname> <given-names>V.</given-names></name> <name><surname>Chocholov&#x00E1;</surname> <given-names>E.</given-names></name><etal/></person-group> (<year>2012</year>). <article-title>Stretching the rules: monocentric chromosomes with multiple centromere domains.</article-title> <source><italic>PLoS Genet.</italic></source> <volume>8</volume>:<issue>e1002777</issue>. <pub-id pub-id-type="doi">10.1371/journal.pgen.1002777</pub-id></citation></ref>
<ref id="B28"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pignatta</surname> <given-names>D.</given-names></name> <name><surname>Comai</surname> <given-names>L.</given-names></name></person-group> (<year>2009</year>). <article-title>Parental squabbles and genome expression: lessons from the polyploids.</article-title> <source><italic>J. Biol.</italic></source> <volume>8</volume> <issue>43</issue>. <pub-id pub-id-type="doi">10.1186/jbiol140</pub-id></citation></ref>
<ref id="B29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pillay</surname> <given-names>M.</given-names></name> <name><surname>Ssebuliba</surname> <given-names>R.</given-names></name> <name><surname>Hartman</surname> <given-names>J.</given-names></name> <name><surname>Vuylsteke</surname> <given-names>D.</given-names></name> <name><surname>Talengera</surname> <given-names>D.</given-names></name> <name><surname>Tushemereirwe</surname> <given-names>W.</given-names></name></person-group> (<year>2004</year>). <article-title>Conventional breeding strategies to enhance the sustainability of <italic>Musa</italic> biodiversity conservation for endemic cultivars.</article-title> <source><italic>Afr. Crop Sci. J.</italic></source> <volume>12</volume> <fpage>59</fpage>&#x2013;<lpage>65</lpage>. <pub-id pub-id-type="doi">10.4314/acsj.v12i1.27663</pub-id></citation></ref>
<ref id="B30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rapp</surname> <given-names>R. A.</given-names></name> <name><surname>Haigler</surname> <given-names>C. H.</given-names></name> <name><surname>Flagel</surname> <given-names>L.</given-names></name> <name><surname>Hovav</surname> <given-names>R. H.</given-names></name> <name><surname>Udall</surname> <given-names>J. A.</given-names></name> <name><surname>Wendel</surname> <given-names>J. F.</given-names></name></person-group> (<year>2010</year>). <article-title>Gene expression in developing fibres of Upland cotton (<italic>Gossypium hirsutum</italic> L.) was massively altered by domestication.</article-title> <source><italic>BMC Biol.</italic></source> <volume>8</volume>:<issue>139</issue>. <pub-id pub-id-type="doi">10.1186/1741-7007-8-139</pub-id></citation></ref>
<ref id="B31"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ravi</surname> <given-names>M.</given-names></name> <name><surname>Chan</surname> <given-names>S. W. L.</given-names></name></person-group> (<year>2010</year>). <article-title>Haploid plants produced by centromere-mediated genome elimination.</article-title> <source><italic>Nature</italic></source> <volume>464</volume> <fpage>615</fpage>&#x2013;<lpage>618</lpage>. <pub-id pub-id-type="doi">10.1038/nature08842</pub-id></citation></ref>
<ref id="B32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ravi</surname> <given-names>M.</given-names></name> <name><surname>Kwong</surname> <given-names>P. N.</given-names></name> <name><surname>Menorca</surname> <given-names>R. M. G.</given-names></name> <name><surname>Valencia</surname> <given-names>J. T.</given-names></name> <name><surname>Ramahi</surname> <given-names>J. S.</given-names></name> <name><surname>Stewart</surname> <given-names>J. L.</given-names></name><etal/></person-group> (<year>2010</year>). <article-title>The rapidly evolving centromere-specific histone has stringent functional requirements in Arabidopsis thaliana.</article-title> <source><italic>Genetics</italic></source> <volume>186</volume> <fpage>461</fpage>&#x2013;<lpage>471</lpage>. <pub-id pub-id-type="doi">10.1534/genetics.110.120337</pub-id></citation></ref>
<ref id="B33"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rice</surname> <given-names>P.</given-names></name> <name><surname>Longden</surname> <given-names>I.</given-names></name> <name><surname>Bleasby</surname> <given-names>A.</given-names></name></person-group> (<year>2000</year>). <article-title>EMBOSS: the european molecular biology open software suite.</article-title> <source><italic>Trends Genet.</italic></source> <volume>16</volume> <fpage>276</fpage>&#x2013;<lpage>277</lpage>. <pub-id pub-id-type="doi">10.1016/j.cocis.2008.07.002</pub-id></citation></ref>
<ref id="B34"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sanei</surname> <given-names>M.</given-names></name> <name><surname>Pickering</surname> <given-names>R.</given-names></name> <name><surname>Kumke</surname> <given-names>K.</given-names></name> <name><surname>Nasuda</surname> <given-names>S.</given-names></name> <name><surname>Houben</surname> <given-names>A.</given-names></name></person-group> (<year>2011</year>). <article-title>Loss of centromeric histone H3 (CENH3) from centromeres precedes uniparental chromosome elimination in interspecific barley hybrids.</article-title> <source><italic>Proc. Natl. Acad. Sci. U.S.A.</italic></source> <volume>108</volume> <fpage>E498</fpage>&#x2013;<lpage>E505</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1103190108</pub-id></citation></ref>
<ref id="B35"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Seymour</surname> <given-names>D. K.</given-names></name> <name><surname>Filiault</surname> <given-names>D. L.</given-names></name> <name><surname>Henry</surname> <given-names>I. M.</given-names></name> <name><surname>Monson-Miller</surname> <given-names>J.</given-names></name> <name><surname>Ravi</surname> <given-names>M.</given-names></name> <name><surname>Pang</surname> <given-names>A.</given-names></name><etal/></person-group> (<year>2012</year>). <article-title>Rapid creation of Arabidopsis doubled haploid lines for quantitative trait locus mapping.</article-title> <source><italic>Proc. Natl. Acad. Sci. U.S.A.</italic></source> <volume>109</volume> <fpage>4227</fpage>&#x2013;<lpage>4232</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1117277109</pub-id></citation></ref>
<ref id="B36"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tachiwana</surname> <given-names>H.</given-names></name> <name><surname>Kagawa</surname> <given-names>W.</given-names></name> <name><surname>Shiga</surname> <given-names>T.</given-names></name> <name><surname>Osakabe</surname> <given-names>A.</given-names></name> <name><surname>Miya</surname> <given-names>Y.</given-names></name> <name><surname>Saito</surname> <given-names>K.</given-names></name><etal/></person-group> (<year>2011</year>). <article-title>Crystal structure of the human centromeric nucleosome containing CENP-A.</article-title> <source><italic>Nature</italic></source> <volume>476</volume> <fpage>232</fpage>&#x2013;<lpage>235</lpage>. <pub-id pub-id-type="doi">10.1038/nature10258</pub-id></citation></ref>
<ref id="B37"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Talbert</surname> <given-names>P. B.</given-names></name> <name><surname>Bryson</surname> <given-names>T. D.</given-names></name> <name><surname>Henikoff</surname> <given-names>S.</given-names></name></person-group> (<year>2004</year>). <article-title>Adaptive evolution of centromere proteins in plants and animals.</article-title> <source><italic>J. Biol.</italic></source> <volume>3</volume> <issue>18</issue>. <pub-id pub-id-type="doi">10.1186/jbiol11</pub-id></citation></ref>
<ref id="B38"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tamura</surname> <given-names>K.</given-names></name> <name><surname>Stecher</surname> <given-names>G.</given-names></name> <name><surname>Peterson</surname> <given-names>D.</given-names></name> <name><surname>Filipski</surname> <given-names>A.</given-names></name> <name><surname>Kumar</surname> <given-names>S.</given-names></name></person-group> (<year>2013</year>). <article-title>MEGA6: Molecular Evolutionary Genetics Analysis version 6.0.</article-title> <source><italic>Mol. Biol. Evol.</italic></source> <volume>30</volume> <fpage>2725</fpage>&#x2013;<lpage>2729</lpage>. <pub-id pub-id-type="doi">10.1093/molbev/mst197</pub-id></citation></ref>
<ref id="B39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tan</surname> <given-names>E. H.</given-names></name> <name><surname>Henry</surname> <given-names>I. M.</given-names></name> <name><surname>Ravi</surname> <given-names>M.</given-names></name> <name><surname>Bradnam</surname> <given-names>K. R.</given-names></name> <name><surname>Mandakova</surname> <given-names>T.</given-names></name> <name><surname>Marimuthu</surname> <given-names>M. P. A.</given-names></name><etal/></person-group> (<year>2015</year>). <article-title>Catastrophic chromosomal restructuring during genome elimination in plants.</article-title> <source><italic>Elife</italic></source> <volume>4</volume>:<issue>e06516</issue>. <pub-id pub-id-type="doi">10.7554/eLife.06516</pub-id></citation></ref>
<ref id="B40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Verdaasdonk</surname> <given-names>J.</given-names></name> <name><surname>Bloom</surname> <given-names>K.</given-names></name></person-group> (<year>2011</year>). <article-title>Centromeres: unique chromatin structures that drive chromosome segregation.</article-title> <source><italic>Nat. Rev. Mol. Cell Biol.</italic></source> <volume>12</volume> <fpage>320</fpage>&#x2013;<lpage>332</lpage>. <pub-id pub-id-type="doi">10.1038/nrm3107</pub-id></citation></ref>
<ref id="B41"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>G.</given-names></name> <name><surname>He</surname> <given-names>Q.</given-names></name> <name><surname>Liu</surname> <given-names>F.</given-names></name> <name><surname>Cheng</surname> <given-names>Z.</given-names></name> <name><surname>Talbert</surname> <given-names>P.</given-names></name> <name><surname>Jin</surname> <given-names>W.</given-names></name></person-group> (<year>2011</year>). <article-title>Characterization of CENH3 proteins and centromere-associated DNA sequences in diploid and allotetraploid <italic>Brassica</italic> species.</article-title> <source><italic>Chromosoma</italic></source> <volume>120</volume> <fpage>353</fpage>&#x2013;<lpage>365</lpage>. <pub-id pub-id-type="doi">10.1007/s00412-011-0315-z</pub-id></citation></ref>
<ref id="B42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yoo</surname> <given-names>M.-J.</given-names></name> <name><surname>Szadkowski</surname> <given-names>E.</given-names></name> <name><surname>Wendel</surname> <given-names>J. F.</given-names></name></person-group> (<year>2013</year>). <article-title>Homoeolog expression bias and expression level dominance in allopolyploid cotton.</article-title> <source><italic>Heredity (Edinb).</italic></source> <volume>110</volume> <fpage>171</fpage>&#x2013;<lpage>180</lpage>. <pub-id pub-id-type="doi">10.1038/hdy.2012.94</pub-id></citation></ref>
<ref id="B43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yuan</surname> <given-names>J.</given-names></name> <name><surname>Guo</surname> <given-names>X.</given-names></name> <name><surname>Hu</surname> <given-names>J.</given-names></name> <name><surname>Lv</surname> <given-names>Z.</given-names></name> <name><surname>Han</surname> <given-names>F.</given-names></name></person-group> (<year>2015</year>). <article-title>Characterization of two CENH3 genes and their roles in wheat evolution.</article-title> <source><italic>New Phytol.</italic></source> <volume>206</volume> <fpage>839</fpage>&#x2013;<lpage>851</lpage>. <pub-id pub-id-type="doi">10.1111/nph.13235</pub-id></citation></ref>
<ref id="B44"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhong</surname> <given-names>C.</given-names></name> <name><surname>Marshall</surname> <given-names>J.</given-names></name> <name><surname>Topp</surname> <given-names>C.</given-names></name></person-group> (<year>2002</year>). <article-title>Centromeric retroelements and satellites interact with maize kinetochore protein CENH3.</article-title> <source><italic>Plant Cell</italic></source> <volume>14</volume> <fpage>2825</fpage>&#x2013;<lpage>2836</lpage>. <pub-id pub-id-type="doi">10.1105/tpc.006106</pub-id></citation></ref>
</ref-list>
</back>
</article>