About Us
Research Watch
स्क्रिनमा देखिने चुरोट: सुर्तीजन्य हानि न्यूनीकरण नीतिमा दक्षिण एसियाले अझै के छुटाइरहेको छनेपालमा पिसाब नलीको संक्रमण र एन्टिबायोटिक प्रतिरोधको बढ्दो संकटFrontline Perspectives on Nursing Leadership in NepalProtecting the Smallest Lungs from the Hidden Grip of RSV in KathmanduThe Heavy Burden of Bullying on Student Wellbeing in NepalThe Emerging Landscape of Thyroid Health in Central NepalHow a Recent Western Nepal Study is Redefining Anemia DiagnosisHow H. Pylori is Impacting the Health of Karnali’s High-Altitude CommunitiesSweet Poison, Bitter Reality: The Unseen Diabetes Epidemic Among Nepal’s YouthHow Missing Checklists and Protocols are Costing Lives in Nepal’s ERsस्क्रिनमा देखिने चुरोट: सुर्तीजन्य हानि न्यूनीकरण नीतिमा दक्षिण एसियाले अझै के छुटाइरहेको छनेपालमा पिसाब नलीको संक्रमण र एन्टिबायोटिक प्रतिरोधको बढ्दो संकटFrontline Perspectives on Nursing Leadership in NepalProtecting the Smallest Lungs from the Hidden Grip of RSV in KathmanduThe Heavy Burden of Bullying on Student Wellbeing in NepalThe Emerging Landscape of Thyroid Health in Central NepalHow a Recent Western Nepal Study is Redefining Anemia DiagnosisHow H. Pylori is Impacting the Health of Karnali’s High-Altitude CommunitiesSweet Poison, Bitter Reality: The Unseen Diabetes Epidemic Among Nepal’s YouthHow Missing Checklists and Protocols are Costing Lives in Nepal’s ERs

Direct microhaplotype genotyping for GT-seq (Genotyping-in-Thousands by Sequencing) using a diploid abundance model.

Researchers

Nathan R Campbell, Amanda R Campbell, Shannon K Blair, Amanda J Finger

Abstract

GT-seq (Genotyping-in-Thousands by Sequencing) is widely used for high-throughput amplicon genotyping, but most analytical pipelines focus on single SNPs or rely on alignment-based variant calling. Here we present a direct microhaplotype genotyping framework that leverages the high read depth and low error rates typical of paired-end Illumina and Element sequencing. The pipeline first identifies primer-bounded reads and resolves paired-end sequences into quality-aware consensus amplicon sequences. Within each sample and locus, unique sequences are ranked by read abundance and the top one or two sequences are retained as directly observed haplotypes. These alleles are aggregated across samples to construct a catalog of observed haplotypes for each locus. In a second pass, reads are assigned to catalog haplotypes by exact sequence matching to produce diploid genotypes. Finally, catalog haplotype sequences are compared to identify phased SNP and collapsed indel variation. Optionally, catalog haplotypes may be aligned to a reference genome to project observed variants onto genomic coordinates and generate standards-compliant VCF output. This framework enables robust, microhaplotype genotyping directly from high-depth amplicon sequencing data. Comparison with an independent BWA/BCFtools alignment-based workflow demonstrated 99.67% genotype concordance across 102,520 genotype comparisons spanning 1,085 SNPs in 96 individuals. Genotype concordance remained above 99.4% even at the minimum supported sequencing depth of 10 reads per locus, demonstrating robust performance across a broad range of sequencing depths.
Source: PubMed (PMID: 42758791)View Original on PubMed