Data availability
The genome assemblies and corresponding gene annotation results, raw sequencing datasets of US56-14-4, 10-9201, 10-9208 and the other six F1 hybrids have been deposited at the Genome Sequence Archive (GSA) database71 in the National Genomics Data Center, Beijing Institute of Genomics (BIG), Chinese Academy of Sciences and China National Center for Bioinformation (CNCB), under accession numbers PRJCA032574 (CRA020620 and CRA022922). The LA Purple genome is available from CNCB Genome Warehouse under accession GWHHOJF00000000.1. The raw sequencing fastq data of LA Purple are available from CNCB BioProject accession PRJCA039904.
Code availability
The code of an algorithm KLASSIFY that classify chimeric reads and identify breakpoints is available on GitHub https://github.com/tanghaibao/klassify (including full simulation and evaluation scripts with exact parameters, random seeds and example commands) and the code is archived on Zenodo at https://doi.org/10.5281/zenodo.20838810 (v0.1.6)72.
References
Bremer, G. A cytological investigation of some species and species-hybrids of the genus Saccharum. Genetica 5, 273–326 (1923).
Article Google Scholar
Bremer, G. Problems in breeding and cytology of sugar cane. IV. The origin of the increase of chromosome number in species hybrids of Saccharum. Euphytica 10, 325–342 (1961b).
Article Google Scholar
OECD/FAO. OECD-FAO Agricultural Outlook 2025–2034 (OECD Publishing, 2025).
Burr, G. O. et al. The sugarcane plant. Annu. Rev. Plant Physiol. 8, 275–308 (1957).
Article CAS Google Scholar
D’Hont, A. et al. Characterisation of the double genome structure of modern sugarcane cultivars (Saccharum spp.) by molecular cytogenetics. Mol. Gen. Genet. 250, 405–413 (1996).
Article PubMed Google Scholar
Piperidis, G., Piperidis, N. & D’Hont, A. Molecular cytogenetic investigation of chromosome composition and transmission in sugarcane. Mol. Genet. Genomics 284, 65–73 (2010).
Article PubMed CAS Google Scholar
D’Hont, A. Unraveling the genome structure of polyploids using FISH and GISH; examples of sugarcane and banana. Cytogenet. Genome Res. 109, 27–33 (2005).
Article PubMed Google Scholar
Price, S. Cytological studies in Saccharum and allied genera VII. Maternal chromosome transmission by S. officinarum in intra- and interspecific crosses. Bot. Gaz. 122, 298–305 (1961).
Article Google Scholar
Bielig, L. M., Mariani, A. & Berding, N. Cytological studies of 2n male gamete formation in sugarcane, Saccharum L. Euphytica 133, 117–124 (2003).
Article Google Scholar
Narayanaswami, S. Megasporogenesis and origin of triploids in Saccharum. Indian J. Agric. Sci. 10, 534–551 (1940).
Google Scholar
Zhang, J. et al. Allele-defined genome of the autopolyploid sugarcane Saccharum spontaneum L. Nat. Genet. 50, 1565–1573 (2018).
Article PubMed CAS Google Scholar
Zhang, Q. et al. Genomic insights into the recent chromosome reduction of autopolyploid sugarcane Saccharum spontaneum. Nat. Genet. 54, 885–896 (2022).
Article PubMed CAS Google Scholar
Healey, A. L. et al. The complex polyploid genome architecture of sugarcane. Nature 628, 804–810 (2024).
Article ADS PubMed PubMed Central CAS Google Scholar
Bao, Y. et al. A chromosomal-scale genome assembly of modern cultivated hybrid sugarcane provides insights into origination and evolution. Nat. Commun. 15, 3041 (2024).
Article ADS PubMed PubMed Central CAS Google Scholar
Zhang, J. et al. The highly allo-autopolyploid modern sugarcane genome and very recent allopolyploidization in Saccharum. Nat. Genet. 57, 242–253 (2025).
Article PubMed CAS Google Scholar
Wang, J. et al. Genetic architecture of sugarcane traits in a polyploid genomics framework. Nature 654, 994–1003 (2026).
Article PubMed PubMed Central Google Scholar
Rhie, A., Walenz, B., Koren, S. & Phillippy, A. Merqury: reference-free quality, completeness, and phasing assessment for genome assemblies. Genome Biol. 21, 245 (2020).
Article PubMed PubMed Central CAS Google Scholar
Bretagnolle, F. & Thompson, J. D. Gametes with the somatic chromosome number: mechanisms of their formation and role in the evolution of autopolyploid plants. New Phytol. 129, 1–22 (1995).
Article PubMed CAS Google Scholar
Bremer, G. Problems in breeding and cytology of sugar cane. II, the sugar cane breeding from a cytological view-point. Euphytica 10, 121–258 (1961).
Article Google Scholar
Bremer, G. Increase of chromosome number in species-hybrids of Saccharum in relation to the embryosac development. Bibliogr. Genet. 18, 99 (1959).
Google Scholar
Ramanna, M. S. & Jacobsen, E. Relevance of sexual polyploidization for crop improvement. Euphytica 133, 3–8 (2003).
Article Google Scholar
Wijnker, E. et al. The genomic landscape of meiotic crossovers and gene conversions in Arabidopsis thaliana. eLife 2, e01426 (2013).
Article PubMed PubMed Central Google Scholar
Choi, K. et al. Arabidopsis meiotic crossover hotspots overlap with H2A.Z nucleosomes at gene promoters. Nat. Genet. 45, 1327–1336 (2013).
Article PubMed PubMed Central CAS Google Scholar
Kreiner, J. M., Kron, P. & Husband, B. C. Frequency and maintenance of unreduced gametes in natural plant populations: associations with reproductive mode, life history and genome size. New Phytol. 214, 879–889 (2017).
Article PubMed CAS Google Scholar
Vieira, M. L. C. et al. Revisiting meiosis in sugarcane: chromosomal irregularities and the prevalence of bivalent configurations. Front. Genet. https://doi.org/10.3389/fgene.2018.00213 (2018).
Article PubMed PubMed Central Google Scholar
Oliveira, B. et al. New trends in sugarcane fertilization: Implications for NH3 volatilization, N2O emissions and crop yields. J. Environ. Manage. 342, 118233 (2023).
Article PubMed CAS Google Scholar
Price, S. Chromosome numbers in interspecific hybrids. Bot. Gaz. 118, 146–159 (1957).
Article Google Scholar
Chai, J. et al. All nonhomologous chromosomes and rearrangements in Saccharum officinarum × Saccharum spontaneum allopolyploids identified by oligo-based painting. Front. Plant Sci. https://doi.org/10.3389/fpls.2023.1176914 (2023).
Article PubMed PubMed Central Google Scholar
Higgins, J. et al. Unravelling mechanisms that govern meiotic crossover formation in wheat. Biochem. Soc. Trans. 50, 1179–1186 (2022).
Article PubMed PubMed Central CAS Google Scholar
Wenger, A. M. et al. Accurate circular consensus long-read sequencing improves variant detection and assembly of a human genome. Nat. Biotechnol. 37, 1155–1162 (2019).
Article ADS PubMed PubMed Central CAS Google Scholar
Burton, J. N. et al. Chromosome-scale scaffolding of de novo genome assemblies based on chromatin interactions. Nat. Biotechnol. 31, 1119–1125 (2013).
Article PubMed PubMed Central CAS Google Scholar
Chen, S., Zhou, Y., Chen, Y. & Gu, J. fastp: an ultra-fast all-in-one FASTQ preprocessor. Bioinformatics 34, i884–i890 (2018).
Article PubMed PubMed Central Google Scholar
Cheng, H. et al. Efficient near-telomere-to-telomere assembly of nanopore simplex reads. Nature 655, 166–173 (2026).
Article PubMed CAS Google Scholar
Cheng, H., Concepcion, G. T., Feng, X., Zhang, H. & Li, H. Haplotype-resolved de novo assembly using phased assembly graphs with hifiasm. Nat. Methods 18, 170–175 (2021).
Article PubMed PubMed Central CAS Google Scholar
Zeng, X. et al. Chromosome-level scaffolding of haplotype-resolved assemblies using Hi-C data without reference genomes. Nat. Plants 10, 1184–1200 (2024).
Article PubMed CAS Google Scholar
Durand, N. et al. Juicebox provides a visualization system for Hi-C contact maps with unlimited zoom. Cell Syst. 3, 99–101 (2016).
Article PubMed PubMed Central CAS Google Scholar
Li, H. Minimap2: pairwise alignment for nucleotide sequences. Bioinformatics 34, 3094–3100 (2018).
Article PubMed PubMed Central CAS Google Scholar
Chen, Y., Zhang, Y., Wang, A., Gao, M. & Chong, Z. Accurate long-read de novo assembly evaluation with Inspector. Genome Biol. 22, 312 (2021).
Article PubMed PubMed Central Google Scholar
Simão, F. A., Waterhouse, R. M., Ioannidis, P., Kriventseva, E. V. & Zdobnov, E. M. BUSCO: assessing genome assembly and annotation completeness with single-copy orthologs. Bioinformatics 31, 3210–3212 (2015).
Article PubMed Google Scholar
Li, H. & Durbin, R. Fast and accurate short read alignment with Burrows–Wheeler transform. Bioinformatics 25, 1754–1760 (2009).
Article PubMed PubMed Central CAS Google Scholar
Ou, S., Chen, J. & Jiang, N. Assessing genome assembly quality using the LTR Assembly Index (LAI). Nucleic Acids Res. 46, e126 (2018).
PubMed PubMed Central Google Scholar
Pedersen, B. S. & Quinlan, A. R. Mosdepth: quick coverage calculation for genomes and exomes. Bioinformatics 34, 867–868 (2018).
Article PubMed PubMed Central CAS Google Scholar
Krzywinski, M. et al. CIRCOS: an information aesthetic for comparative genomics. Genome Res. 19, 1639–1645 (2009).
Article PubMed PubMed Central CAS Google Scholar
Tang, H. et al. JCVI: A versatile toolkit for comparative genomics analysis. iMeta 3, e211 (2024).
Article PubMed PubMed Central CAS Google Scholar
Marcais, G. et al. MUMmer4: a fast and versatile genome alignment system. PLoS Comput. Biol. 14, e1005944 (2018).
Article PubMed PubMed Central Google Scholar
Flynn, J. M. et al. RepeatModeler2 for automated genomic discovery of transposable element families. Proc. Natl Acad. Sci. USA 117, 9451–9457 (2020).
Article ADS PubMed PubMed Central CAS Google Scholar
Li, Y., Jiang, N. & Sun, Y. AnnoSINE: a short interspersed nuclear elements annotation tool for plant genomes. Plant Physiol. 188, 955–970 (2022).
Article PubMed PubMed Central CAS Google Scholar
Rho, M. & Tang, H. MGEScan-non-LTR: computational identification and classification of autonomous non-LTR retrotransposons in eukaryotic genomes. Nucleic Acids Res. 37, e143–e143 (2009).
Article PubMed PubMed Central Google Scholar
Ou, S. et al. Benchmarking transposable element annotation methods for creation of a streamlined, comprehensive pipeline. Genome Biol. 20, 275 (2019).
Article PubMed PubMed Central CAS Google Scholar
Zhang, R.-G. et al. TEsorter: an accurate and fast method to classify LTR-retrotransposons in plant genomes. Hortic. Res. 9, uhac017 (2022).
Article PubMed PubMed Central Google Scholar
Yan, H., Bombarely, A. & Li, S. DeepTE: a computational method for de novo classification of transposons with convolutional neural network. Bioinformatics 36, 4269–4275 (2020).
Article PubMed CAS Google Scholar
Edgar, R. C. Search and clustering orders of magnitude faster than BLAST. Bioinformatics 26, 2460–2461 (2010).
Article PubMed CAS Google Scholar
Tarailo-Graovac, M. & Chen, N. Using RepeatMasker to identify repetitive elements in genomic sequences. Curr. Protoc. Bioinformatics 4, 4.10.11–14.10.14 (2009).
Google Scholar
Benson, G. Tandem repeats finder: a program to analyze DNA sequences. Nucleic Acids Res. 27, 573–580 (1999).
Article PubMed PubMed Central CAS Google Scholar
Lin, Y. et al. quarTeT: a telomere-to-telomere toolkit for gap-free genome assembly and centromeric repeat identification. Hortic. Res. 10, uhad127 (2023).
Article PubMed PubMed Central Google Scholar
Haas, B. J. et al. De novo transcript sequence reconstruction from RNA-seq using the Trinity platform for reference generation and analysis. Nat. Protoc. 8, 1494–1512 (2013).
Article PubMed PubMed Central CAS Google Scholar
Kim, D., Paggi, J. M., Park, C., Bennett, C. & Salzberg, S. L. Graph-based genome alignment and genotyping with HISAT2 and HISAT-genotype. Nat. Biotechnol. 37, 907–915 (2019).
Article PubMed PubMed Central CAS Google Scholar
Gabriel, L. et al. BRAKER3: Fully automated genome annotation using RNA-seq and protein evidence with GeneMark-ETP, AUGUSTUS, and TSEBRA. Genome Res. 34, 769–777 (2024).
Article PubMed PubMed Central CAS Google Scholar
Gabriel, L., Hoff, K., Bruna, T., Borodovsky, M. & Stanke, M. TSEBRA: transcript selector for BRAKER. BMC Bioinformatics 22, 566 (2021).
Article PubMed PubMed Central CAS Google Scholar
Holst, F. et al. ab initio prediction of primary eukaryotic gene models combining deep learning and a hidden Markov model. Nat. Methods 23, 732–739 (2026).
Article PubMed CAS Google Scholar
Haas, B. et al. Automated eukaryotic gene structure annotation using EVidenceModeler and the program to assemble spliced alignments. Genome Biol. 9, R7 (2008).
Article PubMed PubMed Central Google Scholar
Blum, M. et al. The InterPro protein families and domains database: 20 years on. Nucleic Acids Res. 49, D344–D354 (2021).
Article PubMed PubMed Central CAS Google Scholar
Cantalapiedra, C. P., Hernández-Plaza, A., Letunic, I., Bork, P. & Huerta-Cepas, J. eggNOG-mapper v2: functional annotation, orthology assignments, and domain prediction at the metagenomic scale. Mol. Biol. Evol. 38, 5825–5829 (2021).
Article PubMed PubMed Central CAS Google Scholar
Wang, Y. et al. MCScanX: a toolkit for detection and evolutionary analysis of gene synteny and collinearity. Nucleic Acids Res. 40, e49 (2012).
Article ADS PubMed PubMed Central CAS Google Scholar
Wu, T. D. & Watanabe, C. K. GMAP: a genomic mapping and alignment program for mRNA and EST sequences. Bioinformatics 21, 1859–1875 (2005).
Article PubMed CAS Google Scholar
Edgar, R. MUSCLE: a multiple sequence alignment method with reduced time and space complexity. BMC Bioinformatics 5, 113 (2004).
Article PubMed PubMed Central Google Scholar
Bailey, T. L. STREME: accurate and versatile sequence motif discovery. Bioinformatics 37, 2834–2840 (2021).
Article PubMed PubMed Central CAS Google Scholar
Jiao, W.-B. & Schneeberger, K. Chromosome-level assemblies of multiple Arabidopsis genomes reveal hotspots of rearrangements with altered evolutionary dynamics. Nat. Commun. 11, 989 (2020).
Article ADS PubMed PubMed Central CAS Google Scholar
Ono, Y., Hamada, M. & Asai, K. PBSIM3: a simulator for all types of PacBio and ONT long reads. NAR Genom. Bioinform. 4, lqac092 (2022).
Article PubMed PubMed Central Google Scholar
Baid, G. et al. DeepConsensus improves the accuracy of sequences with a gap-aware sequence transformer. Nat. Biotechnol. 41, 232–238 (2023).
Article PubMed CAS Google Scholar
CNCB–NGDC Members and Partners. Database resources of the National Genomics Data Center, China National Center for Bioinformation in 2026. Nucleic Acids Res. 54, D28–D47 (2026).
Article Google Scholar
Tang, H. KLASSIFY: classify chimeric reads based on unique k-mer contents (Version 0.1.6) [Computer software]. Zenodo https://doi.org/10.5281/zenodo.20838810 (2026).
Download references
Acknowledgements
We thank A. Paterson for restructuring and revising the manuscript.
Funding
This project was supported by the National Key Research and Development Program of China (2023YFD1200700, 2023YFD1200701) to R.M., (2024YFF1000800) to H. Tang and R.M., and the National Natural Science Foundation of China (W2531027) to R.M.
Ethics declarations
Competing interests
The authors declare no competing interests.
Peer review
Peer review information
Nature thanks André Marques who co-reviewed with Meng Zhang; and the other, anonymous reviewers for their contribution to the peer review of this work. Peer reviewer reports are available.
Additional information
Publisher’s note Springer Nature remains neutral with regard to jurisdictional claims in published maps and institutional affiliations.
Extended data figures and tables
Extended Data Fig. 1 SDR-mediated recombinant chromosome formation.
Assuming the second division restitution occurred in LA Purple, for a pair of homologous chromosomes, Chr1A and Chr1B, after normal first division and abnormal second division, the products would include a new recombinant chromosome and a non-recombinant chromosome identical to LA Purple. In the sequencing data of offspring, reads supporting both chromosomal structures around the crossover point are excepted. For example, two types of reads span the Chr1A breakpoint: one is recombinant reads, and other one is non-recombinant reads. Since no intact Chr1B exists in the offspring, only one type of reads can be aligned back to the Chr1B breakpoint. We define the breakpoints of Chr1A and Chr1B as “Type II” and “Type I” points, respectively. The Type I and Type II will always appear in pairs, and the number of pairs reflects the crossover events.
Extended Data Fig. 2 An actual IGV screenshot display of breakpoint in 9208 So-subgenome.
The breakpoint in So (S. officinarum) gametes, where it contains a Type I breakpoint paired with a Type II breakpoint between SoChr01B and SoChr01F.
Extended Data Fig. 3 The distribution of Type I and Type II breakpoints on Chr01 revealed by re-sequencing data of 8 F1s with female parent LA Purple as control.
The panel from top to bottom shows the depth profiles of LA Purple (So) Chr01A-H chromosomes using HiFi datasets from LA Purple and 8 F1 individuals. The blue and red vertical bars mark the location of Type I and Type II breakpoints identified by KLASSIFY, respectively. Type II breakpoints are exclusive to the So genome, while Type I can be observed in both So and Ss genomes. The gray hatched regions indicate the predicted centromeres.
Extended Data Fig. 4 An actual IGV screenshot display of breakpoint in 9208 Ss-subgenome.
The example breakpoint in Ss (S. spontaneum) gametes, where it contains a Type I breakpoint paired with another Type I breakpoint between SsChr04D and SsChr04G.
Extended Data Fig. 5 The distribution of Type I breakpoints on Chr01 revealed by re-sequencing data of 8 F1s with male parent US56-14-4 as control.
The top panel shows Chr01A-I chromosomes of US56-14-4 (Ss). Clear 0: 1: 2 and 0: 0.5: 1 depth ratio pattern are observed in the two sets of parental chromosomes, indicating distinct chromosomal composition in the gametes produced by female and male parents. The blue vertical bars mark the location of identified type I breakpoints, respectively. The type II breakpoints only appear in So chromosome. The gray regions indicate the predicted centromeres. Chr01E in US56-14-4 is a collapsed chromosome with a 2x of average depth. The regions of 0.5x depth in Ss-subgenome of F1s can be attributed to the collapsed chromosomal fragments as described in Table S2 and Fig. S6.
Full size table
Full size table
Supplementary information
Supplementary Information (download DOCX )
This file contains Supplementary Notes, Supplementary Figs. 1-15 and Supplementary Fig. 25, and Supplementary References.
Reporting Summary (download PDF )
Supplementary Fig. 16 (download PDF )
Distribution of genomic features across the 80 chromosomes of US56-14-4. Centromeric and telomeric regions were predicted by quartet. The density of DNA transposons, LINEs, SINEs, and satellites elements were calculated using a window size of 500 kb, and visualized using Rectchr v1.34 (https://github.com/hewm2008/RectChr).
Supplementary Fig. 17 (download PDF )
Distribution of genomic features across the 79 S. officinarum-derived chromosomes of 10-9208. The density of DNA transposons, LINEs, SINEs, and Gypsy, Copia elements are calculated using window size of 500-kb.
Supplementary Fig. 18 (download PDF )
Distribution of genomic features across the 39 S. spontaneum-derived chromosomes of 10-9208. The density of DNA transposons, LINEs, SINEs, and Gypsy, Copia elements are calculated using window size of 500-kb.
Supplementary Fig. 19 (download PDF )
Distribution of genomic features across the 77 S. officinarum-derived chromosomes of 10-9201. The density of DNA transposons, LINEs, SINEs, and Gypsy, Copia elements are calculated using window size of 500-kb.
Supplementary Fig. 20 (download PDF )
Distribution of genomic features across the 40 S. spontaneum-derived chromosomes of 10-9201. The density of DNA transposons, LINEs, SINEs, and Gypsy, Copia elements are calculated using window size of 500-kb.
Supplementary Fig. 21 (download PDF )
The schematic diagrams of all the S. officinarum-derived chromosomes of 10-9208 reveal the recombination relationships of maternal chromosomes, conforming to the role of SDR in gametogenesis in LA Purple. In the subplots (a), (c), (e), (g), (i), (k), (m), (o), and (q), the upper-left panel shows the 8 homologous chromosomes of LA Purple, whereas the bottom-left panel depicts the resultant chromosomes resulting from parental chromosomes recombination, along with the chromosomes not involved in fragments exchange. The vertical axis (from bottom to top) denotes the start and end coordinates of linear genomic sequences. The offspring chromosome IDs, labelled at the top of each ideogram, are named based on the composition and order of parental fragments, which was inferred from the DNA alignment results. Two parallel black lines next to the chromosome ideogram denote twice the average sequencing depth, while a single black line indicates the average depth. The subplots (b), (d), (f), (h), (j), (l), (n), (p) and (r) described the sequencing coverage of LA Purple genome using the HiFi reads of 10-9208.
Supplementary Fig. 22 (download PDF )
The schematic diagrams of all the S. officinarum-derived chromosomes of 10-9201 reveal the recombination relationships of maternal chromosomes. In every subplot, the top panel illustrates the 8 parental homologous chromosomes, whereas the bottom panel depicts the resultant chromosomes resulting from parental chromosomes recombination, along with the chromosomes not involved in fragments exchange. The vertical axis (from bottom to top) denotes the start and end coordinates of linear genomic sequences. The offspring chromosome IDs, labelled at the top of each ideogram, are named based on the composition and order of parental fragments, which was inferred from the DNA alignment results. Two parallel black lines next to the chromosome ideogram denote twice the average sequencing depth, while a single black line indicates the average depth. The asterisk indicates that the breakpoints can be supported by both the KLASSIFY algorithm and collinearity-based method.
Supplementary Fig. 23 (download PDF )
The schematic diagrams of all the S. spontaneum-derived chromosomes of 10-9208 reveal the recombination relationships of paternal chromosomes. In every subplot, the top panel represents 10 parental homologous chromosomes, while the bottom panel depicts the crossover region resulting from two parental chromosomes recombination and the chromosomes not involved in fragments exchange. The direction from bottom to top is the starting and ending points of linear sequences in the genome. The chromosome ID of offspring were named at the top of each chromosome ideogram based on the composition and order of parental fragments. The right panel shows the sequencing depth profile of US56-14-4 using binned S. spontaneum-derived HiFi reads of 10-9208. The dash lines mark the collapsed regions in the US56-14-4 assembly. The reads were randomly assigned to one of the two identical regions by aligner tools (minimap2 used in this study), thereby resulting in a sequencing depth half of the whole-genome average. But only one of the two copies was truly passed down to F1s, just like the other regions marked by solid lines.
Supplementary Fig. 24 (download PDF )
The schematic diagrams of all the S. spontaneum-derived chromosomes of 10-9201 reveal the recombination relationships of paternal chromosomes. In every subplot, the top panel represents 10 parental homologous chromosomes, while the bottom panel depicts the crossover region resulting from two parental chromosomes recombination and the chromosomes not involved in fragments exchange. The direction from bottom to top is the starting and ending points of linear sequences in the genome. The chromosome ID of offspring were named at the top of each chromosome ideogram based on the composition and order of parental fragments. The right panel shows the sequencing depth profile of US56-14-4 using binned S. spontaneum-derived HiFi reads of 10-9201. The dash lines mark the collapsed regions in the US56-14-4 assembly. The reads were randomly assigned to one of the two identical regions by aligner tools (minimap2 used in this study), thereby resulting in a sequencing depth half of the whole-genome average. But only one of the two copies was truly passed down to F1, just like the other regions marked by solid lines.
Supplementary Fig. 26 (download PDF )
Sequencing depth and breakpoints distribution plot of eight F1s on the two parental genomes (LA Purple and US56-14-4). The HiFi reads of 8 F1s were mapped to the merged parental genome using minimap25. Subplots (a) - (i) show 9 groups of S. officinarum chromosomes, while the rest are S. spontaneum chromosomes. We can see clear 0: 1: 2 and 0: 0.5: 1 depth patterns in two sets of parental data, which implies the gametes produced by the female and male parents are different. The blue and red vertical bars mark the location of identified type I and type II breakpoints, respectively. The type II breakpoints only appear in S. officinarum chromosomes. The gray hatched regions indicate the predicted centromeres.
Supplementary Fig. 27 (download PDF )
A complete demonstration of a Type II breakpoint. Fine-scale ‘zoom-in’ breakpoint sequences of this example region. The sequences with yellow and blue-coloured ID are intercepted from SoChr01B and SoChr01F flanking the breakpoints; the purple and green ones are recombinant and non-recombinant reads respectively; the brown one is intercepted from recombinant chromosome SoChr01FB in offspring 10-9208. The asterisks below each alignment denote sequence conservation. The area marked in grey between two polymorphic sites delineates the crossover boundary. Since no sequence variation exists between two parental chromosomes in this interval, DNA breakage and crossover event can occur at any coordinate. Upstream of the left boundary, all the reads match SoChr01B; downstream of the right boundary, the recombinant reads switch to match SoChr01F, while the non-recombinant reads remain consistent with SoChr01B.
Supplementary Tables (download XLSX )
This file contains Supplementary Tables 1-20.
Peer Review File (download PDF )
Rights and permissions
Springer Nature or its licensor (e.g. a society or other partner) holds exclusive rights to this article under a publishing agreement with the author(s) or other rightsholder(s); author self-archiving of the accepted manuscript version of this article is solely governed by the terms of such publishing agreement and applicable law.
Reprints and permissions
About this article
Cite this article
Zhu, S., Tang, H., Jones, T. et al. Uncovering the mechanism of female restitution in sugarcane hybrids. Nature (2026). https://doi.org/10.1038/s41586-026-10863-3
Download citation
Received:
Accepted:
Published:
Version of record:
DOI: https://doi.org/10.1038/s41586-026-10863-3