Project description:Two PacBio Hifi sequencing runs from the kidney of a single male NMR sample used to make the mHetGla4.1.primary genome assembly. Specifically, we assembled a second NMR genome from an unrelated male of a separate captive colony in Toronto, Canada, using PacBio HiFi (155.8 Gb, read N50 = 11.45 Kb) and ONT-LSK (299.6 Gb, read N50 = 10.1 Kb) reads (contig N50 = 75.7 Mb, Compleasm S = 98%). This accession stores the Pacbio Hifi data for this independent assembly.
Project description:One PacBio Hifi sequencing run from the kidney of a single male CDMR (Bathyergus suillus) sample used to make the mBatSui1.1.primary genome assembly as an evolutionary comparator to our telomere-to-telomere naked mole-rat genome assembly. Specifically, we assembled a CDMR from a wild-derived sample in South African cape and sequenced in Toronto, Canada, using PacBio HiFi (89 Gb, read N50 = 18 Kb) and ONT-ULK (55 Gb, read N50 = 43 Kb) reads (contig N50 = 33 Mb, Compleasm S = 99%, QV = 71.0). This accession stores the Pacbio Hifi data for this assembly.
Project description:Two PacBio Hifi sequencing runs from the kidney of a single male NMR sample used to make the mHetGla4.1.primary genome assembly. Specifically, we assembled a second NMR genome from an unrelated male of a separate captive colony in Toronto, Canada, using PacBio HiFi (155.8 Gb, read N50 = 11.45 Kb) and ONT-LSK (299.6 Gb, read N50 = 10.1 Kb) reads (contig N50 = 75.7 Mb, Compleasm S = 98%). This accession stores the ONT-LSK data for this independent assembly.
Project description:The Yeonsan Ogye (Ogye) is the rare black chicken breed domesticated in Korean peninsula, which has been noted for entire black color upon its appearances including feather, skin, comb, eyes, shank, claws and internal organs. In this study, whole genome, transcriptome and epigenome sequencings of Ogye were performed using high-throughput NGS sequencing platforms. We have produced Illumina short-reads (Paired-End, Mate-Pair and FOSMID) and PacBio long-reads for whole genome sequencing (WGS), 1.4 billion reads for RNA-seq, and 123 million reads for RRBS (reduced representation bisulfite sequencing) data. Using WGS data, Ogye genome has been assembled, and coding/non-coding transcriptome maps were constructed on Ogye genome given largescale sequencing data. We have predicted 17,472 (3,550 newly annotated and 13,922 known) protein-coding transcripts, and 9,443 (6,689 novel and 2,754 known) long non-coding RNAs (lncRNAs).
Project description:The Yeonsan Ogye (Ogye) is the rare black chicken breed domesticated in Korean peninsula, which has been noted for entire black color upon its appearances including feather, skin, comb, eyes, shank, claws and internal organs. In this study, whole genome, transcriptome and epigenome sequencings of Ogye were performed using high-throughput NGS sequencing platforms. We have produced Illumina short-reads (Paired-End, Mate-Pair and FOSMID) and PacBio long-reads for whole genome sequencing (WGS), 1.4 billion reads for RNA-seq, and 123 million reads for RRBS (reduced representation bisulfite sequencing) data. Using WGS data, Ogye genome has been assembled, and coding/non-coding transcriptome maps were constructed on Ogye genome given largescale sequencing data. We have predicted 17,472 (3,550 newly annotated and 13,922 known) protein-coding transcripts, and 9,443 (6,689 novel and 2,754 known) long non-coding RNAs (lncRNAs).
Project description:Chromatin immunoprecipitation analysis of CENH3 in the Arabidopsis thaliana accessions Col-0, Ler-0, Cvi-0 and Tanz-1 was performed in order to align reads to PacBio HiFi genome assemblies which contain complete centromere repeat arrays.
Project description:One ONT-ULK sequencing run from the kidney of a single male CDMR (Bathyergus suillus) sample used to make the mBatSui1.1.primary genome assembly as an evolutionary comparator to our telomere-to-telomere naked mole-rat genome assembly. Specifically, we assembled a CDMR from a wild-derived sample in South African cape and sequenced in Toronto, Canada, using PacBio HiFi (89 Gb, read N50 = 18 Kb) and ONT-ULK (55 Gb, read N50 = 43 Kb) reads (contig N50 = 33 Mb, Compleasm S = 99%, QV = 71.0). This accession stores the ONT-ULK data for this assembly.
Project description:<p class='ql-align-justify'>Megasphaera hexanoica KCCM 43214T, isolated from cow rumen, is capable of producing medium-chain carboxylic acids such as n-caproate and n-caprylate. In this study, we present a high-quality genome assembly, along with intracellular metabolomic profiling and pangenomic analysis. Illumina sequencing generated 2.3 Mbp from 15,293,634 reads with a GC content of 49.5%, while PacBio HiFi sequencing produced 331.5 Mbp across 45,266 reads, with an average read length of 7,323 bp and a HiFi read N50 of 8,214 bp. Hybrid assembly of short and long reads resulted in a single 2.88 Mbp contig, containing 2,075-2,083 unique genes. A genome-scale metabolic model was constructed, to evaluate its metabolic capabilities under specific growth conditions. Intracellular metabolomic analysis of cells grown in fructose medium and lactate medium revealed key metabolic activities associated with chain elongation. Pangenomic analysis across nine annotated genomes identified 6,721 orthologous gene using OrthoMCL, emphasizing the genetic and functional diversity within the Megasphaera genus. This dataset offers valuable insights into the metabolism and biotechnological potential of M. hexanoica KCCM 43214T.</p>
Project description:We used PacBio data to identify more reliable transcripts from hESC, based on which we can estimate gene/transcript abundance better from Illumina data. PacBio long reads and Illumina short reads were generated from the same hESC cell line H1. PacBio reads were error-corrected by Illumina reads to identify transcripts. rSeq is used to estimate gene/transcript abundance of the identified transcriptome.