README FILE This file was generated on January 17th, 2026 by Julie Anne V.S. de Oliveira. A. GENERAL INFORMATION Title of the dataset: Tacca chantrieri genome sequence and annotation Brief description of the research project and its aims: Here we present the genome sequence and annotation of Tacca chantrieri (commonly known as black bat flower), famous for its dark, bat-shaped flowers with long filaments, making it an exotic and unusual ornamental plant. Using ONT long-read sequencing data, a highly continuous genome sequence was generated. This data publication contains the genome sequence with corresponding annotations. Author Information A. Investigator Contact Information Name: Julie Anne Vieira Salgado de Oliveira Institution: Plant Biotechnology and Bioinformatics, Institute for Cellular and Molecular Botany - IZMB, University of Bonn. Address: Kirschalle 1, 53115 Bonn, Germany. Email: jvieiras@uni-bonn.de B. Project Supervisor (Principal Investigator) Name: Prof. Dr. Boas Pucker Institution: Plant Biotechnology and Bioinformatics, Institute for Cellular and Molecular Botany - IZMB, University of Bonn. Address: Kirschalle 1, 53115 Bonn, Germany. Email: pucker@uni-bonn.de C. In case of questions related to this dataset, please contact: Name: Prof. Dr. Boas Pucker Institution: Plant Biotechnology and Bioinformatics, Institute for Cellular and Molecular Botany - IZMB, University of Bonn. Address:Kirschalle 1, 53115 Bonn, Germany. Email: pucker@uni-bonn.de Date of data collection: 2026-01-16 Acknowledgments: This work was supported by the de.NBI Cloud within the German Network for Bioinformatics Infrastructure (de.NBI) and ELIXIR-DE (Forschungszentrum Jülich and W-de.NBI-001, W-de.NBI-004, W-de.NBI-008, W-de.NBI-010, W-de.NBI-013, W-de.NBI-014, W-de.NBI-016, W-de.NBI-022). We are grateful for the excellent support provided by the team of the University of Bonn Botanic Gardens. Language of the dataset: English B. DATA & FILE OVERVIEW Tacca_chantrieri.genomic.fasta.gz Genome sequence of Tacca chantrieri. Assembly was performed with Hifiasm. Tacca_chantrieri.anno.gff.gz Structural annotation of the Tacca chantrieri genome sequence generated by GeMoMa v1.9. Tacca_chantrieri.cds.fasta.gz Protein encoding sequences of Tacca chantrieri derived from the structural annotation of the genome sequence. Tacca_chantrieri.pep.fasta.gz Polypeptide sequences of Tacca chantrieri inferred from the coding sequences. Tacca_chantrieri.anno.txt.gz Functional annotation predicted for the polypeptide sequences of Tacca chantrieri based on information available about Arabidopsis thaliana sequences. C. SHARING/ACCESS INFORMATION Data derived from another source: RNA-seq reads for the generation of hints for the gene prediction have been retrieved from the Sequence Read Archive. Licenses: CC BY 4.0 Links to related publications: n/a Other publicly accessible locations: n/a D. METHODOLOGICAL INFORMATION People involved in data collection, processing, analysis and/or submission: all authors Collection of data: The genome of a Tacca chantrieri plant (BONN-13679) was sequenced with nanopore long reads (R10.4.1 flow cells, MinION) based on high molecular weight DNA extracted with a previously established protocol (https://dx.doi.org/10.17504/protocols.io.bcvyiw7w). Processing of data: The genome sequence was assembled with Hifiasm-0.25.0-r726. The gene models were predicted by GeMoMa v1.9. The functional annotation was predicted based on sequence similarity to well characterized Arabidopsis thaliana sequences. Data structure: Genome, coding, and polypeptide sequences are provided in FASTA file format. The structural annotation is provided in the GFF3 format. The functional annotation is provided as a TAB-separated text file. All files are accessible with basic text editors. File formats: TXT, FASTA, GFF