Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 89
Comparative genomics of dairy-associated Staphylococcus aureus from selected sub-Saharan African regions reveals milk as reservoir for human-and animal-derived strains and identifies a putative animal-related clade with presumptive novel siderophore.
PMID 36046020 · PMC9421002 · Frontiers in microbiology · 2022 · 7 claims · 8 setups
Milk serves as a reservoir for both human- and animal-derived S. aureus strains in sub-Saharan Africa
-
Has reproduction · 84
Genome of the endangered Guatemalan Beaded Lizard, Heloderma charlesbogerti, reveals evolutionary relationships of squamates and declines in effective population sizes.
PMID 36226801 · PMC9713440 · G3 (Bethesda, Md.) · 2022 · 8 claims · 7 setups
The assembled draft genome of H. charlesbogerti totals 2.31 Gb, similar in size to related species
-
Full-text index only
A statistical approach designed for finding mathematically defined repeats in shotgun data and determining the length distribution of clone-inserts.
PMID 15626332 · PMC5172250 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 6 setups
Repeats of different copy number have distinct probabilities of appearance in shotgun data, which can be modeled statistically to define recognition thresholds (MDRs) at different shotgun coverages.
-
Full-text index only
Identifying related L1 retrotransposons by analyzing 3' transduced sequences.
PMID 12734010 · PMC156586 · Genome biology · 2003 · 8 claims · 6 setups
L1 elements with transduction-derived 3' sequence (L1-TDs) can be computationally identified using RepeatMasker/TSDfinder and grouped into families sharing a common progenitor via BLAST comparison of downstream sequences.
-
Full-text index only
An analysis of the feasibility of short read sequencing.
PMID 16275781 · PMC1278949 · Nucleic acids research · 2005 · 8 claims · 8 setups
Re-sequencing and de novo sequencing of the majority of a bacterial genome is possible with read lengths of 20-30 nt.
-
Full-text index only
Investigating hookworm genomes by comparative analysis of two Ancylostoma species.
PMID 15854223 · PMC1112591 · BMC genomics · 2005 · 8 claims · 8 setups
Nearly 20,000 ESTs from 7 cDNA libraries define nearly 7,000 hookworm genes across A. caninum and A. ceylanicum
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
Genome-wide identification of specific oligonucleotides using artificial neural network and computational genomic analysis.
PMID 17518996 · PMC1892811 · BMC bioinformatics · 2007 · 7 claims · 4 setups
The IAB algorithm (integration of ANN and BLAST) identifies genome-wide specific oligos much faster than pure BLAST search while maintaining comparable success rate and cross homology
-
Full-text index only
The HuRef Browser: a web resource for individual human genomics.
PMID 19036787 · PMC2686481 · Nucleic acids research · 2009 · 7 claims · 6 setups
The HuRef Browser is a unified web application integrating assembly, annotation, and assembly-to-assembly comparison (ATAC) views for the diploid HuRef individual human genome.
-
Full-text index only
High-throughput sequencing provides insights into genome variation and evolution in Salmonella Typhi.
PMID 18660809 · PMC2652037 · Nature genetics · 2008 · 7 claims · 8 setups
Evolution in the Typhi population is characterized by ongoing loss of gene function (pseudogene accumulation) rather than gain of function or diversifying selection.
-
Full-text index only
High throughput sequencing and proteomics to identify immunogenic proteins of a new pathogen: the dirty genome approach.
PMID 20037647 · PMC2793016 · PloS one · 2009 · 7 claims · 7 setups
A dirty genome approach using unfinished, unclosed genome sequences combined with proteomics can rapidly identify immunogenic proteins useful for diagnostic tool development
-
Full-text index only
A new procedure for determining the genetic basis of a physiological process in a non-model species, illustrated by cold induced angiogenesis in the carp.
PMID 19852815 · PMC2771047 · BMC genomics · 2009 · 8 claims · 5 setups
The Conditional Stepped Reciprocal Best Hit (CSRBH) approach, combining direct RBH and zebrafish-stepped RBH (SRBH), outperformed other ortholog assignment methods and attained 8,726 carp-human functional homolog relationships for 16,650 carp contigs
-
Full-text index only
Gene and transcript abundances of bacterial type III secretion systems from the rumen microbiome are correlated with methane yield in sheep.
PMID 28789673 · PMC5549432 · BMC research notes · 2017 · 7 claims · 7 setups
Gene and transcript abundances of bacterial type III secretion system (T3SS) genes are positively correlated with methane yield in sheep
-
Has reproduction · 100
nf-core/mag: a best-practice pipeline for metagenome hybrid assembly and binning.
PMID 35118380 · PMC8808542 · NAR genomics and bioinformatics · 2022 · 8 claims · 7 setups
nf-core/mag is a Nextflow/nf-core pipeline for hybrid metagenome assembly, binning and taxonomic classification of MAGs.
-
Has reproduction · 71
polishCLR: A Nextflow Workflow for Polishing PacBio CLR Genome Assemblies.
PMID 36792366 · PMC9985148 · Genome biology and evolution · 2023 · 8 claims · 8 setups
polishCLR is a reproducible, containerized Nextflow workflow that implements best practices for polishing PacBio CLR genome assemblies.
-
Full-text index only
A rigorous method for multigenic families' functional annotation: the peptidyl arginine deiminase (PADs) proteins family example.
PMID 16271148 · PMC1310624 · BMC genomics · 2005 · 8 claims · 5 setups
Integrating EST-based expression data with phylogenetic analysis is a valid new method for functionally annotating multigenic protein families
-
Has reproduction · 81
Chromosome-scale Elaeis guineensis and E. oleifera assemblies: comparative genomics of oil palm and other Arecaceae.
PMID 38918881 · PMC11373658 · G3 (Bethesda, Md.) · 2024 · 8 claims · 8 setups
Improved E. guineensis genome assembly achieved with substantially increased continuity and completeness compared to prior assemblies
-
Full-text index only
Twin peaks: the draft human genome sequence.
PMID 11276423 · PMC138909 · Genome biology · 2001 · 8 claims · 8 setups
The predicted number of human genes (~26,000-40,000) is far lower than the widely assumed ~100,000, though downstream RNA/protein complexity can still generate substantial biological complexity.
-
Has reproduction · 83
VirPipe: an easy-to-use and customizable pipeline for detecting viral genomes from Nanopore sequencing.
PMID 37129547 · PMC10191607 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 1 setups
VirPipe is a new bioinformatics pipeline for detecting viral genomes from Nanopore or Illumina sequencing input with streamlined installation and customization.
-
Has reproduction · 91
De Novo Assembly and Annotation of the Larval Transcriptome of Two Spadefoot Toads Widely Divergent in Developmental Rate.
PMID 31217263 · PMC6686947 · G3 (Bethesda, Md.) · 2019 · 8 claims · 8 setups
De novo transcriptome assemblies were generated for larval P. cultripes and S. couchii, providing new genomic resources for spadefoot toads