Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The diploid genome sequence of an Asian individual.
PMID 18987735 · PMC2716080 · Nature · 2008 · 8 claims · 8 setups
First diploid genome sequence of an Asian (Han Chinese) individual generated using massively parallel Illumina sequencing
-
Has reproduction · 78
Metavisitor, a Suite of Galaxy Tools for Simple and Rapid Detection and Discovery of Viruses in Deep Sequence Data.
PMID 28045932 · PMC5207757 · PloS one · 2017 · 7 claims · 5 setups
Metavisitor is an open-source suite of modular Galaxy tools and preset workflows enabling non-experts to detect and assemble viral genomes from deep sequence data.
-
Full-text index only
Clustering of phosphorylation site recognition motifs can be exploited to predict the targets of cyclin-dependent kinase.
PMID 17316440 · PMC1852407 · Genome biology · 2007 · 8 claims · 6 setups
CDK consensus motifs are frequently clustered (closely spaced) in known CDK substrate proteins rather than uniformly distributed
-
Full-text index only
TFBScluster web server for the identification of mammalian composite regulatory elements.
PMID 16845063 · PMC1538905 · Nucleic acids research · 2006 · 7 claims · 5 setups
TFBScluster is a web server that identifies genome-wide clusters of TFBSs conserved in multiple mammalian species using human or mouse as the reference genome.
-
Full-text index only
Accurate whole human genome sequencing using reversible terminator chemistry.
PMID 18987734 · PMC2581791 · Nature · 2008 · 8 claims · 7 setups
A novel sequencing platform using fluorescent reversible terminator nucleotides on clonally amplified single-molecule DNA clusters generates several billion bases of accurate sequence per experiment at low cost.
-
Has reproduction · 80
TP53 engagement with the genome occurs in distinct local chromatin environments via pioneer factor activity.
PMID 25391375 · PMC4315292 · Genome research · 2015 · 8 claims · 8 setups
TP53 binding events fall into three distinct categories defined by the local chromatin environment: TSS (H3K4me3+), enhancer (H3K4me1+/H3K4me3-), and distal (H3K4me1-/H3K4me3-) peaks.
-
Full-text index only
Ensembl's 10th year.
PMID 19906699 · PMC2808936 · Nucleic acids research · 2010 · 8 claims · 8 setups
Ensembl provides comprehensive gene annotation and integrated genomic resources (variation, regulation, comparative genomics) across a growing set of chordate genomes
-
Has reproduction · 93
Population genomics of the Wolbachia endosymbiont in Drosophila melanogaster.
PMID 23284297 · PMC3527207 · PLoS genetics · 2012 · 8 claims · 8 setups
Wolbachia infection status can be accurately predicted in silico from whole-genome shotgun sequence of individual host strains, showing 99% concordance with diagnostic PCR.
-
Full-text index only
Simple models of genomic variation in human SNP density.
PMID 17553150 · PMC1919371 · BMC genomics · 2007 · 6 claims · 4 setups
Hierarchical Poisson model B, which allows both the mutation-rate proxy (Beta-distributed Λ) and the ARG-size proxy (Gamma-distributed T) to vary, fits the observed SNP density distribution significantly better than models with only one or neither varying.
-
Full-text index only
Mitochondrial diversity within modern human populations.
PMID 17439969 · PMC1888801 · Nucleic acids research · 2007 · 8 claims · 5 setups
Modern humans show extremely low divergence from the mitochondrial consensus sequence, differing on average by only 21.6 nucleotide sites
-
Full-text index only
PigGIS: Pig Genomic Informatics System.
PMID 17090590 · PMC1669765 · Nucleic acids research · 2007 · 7 claims · 7 setups
PigGIS identified 15,700 pig consensus sequences covering 18.5 Mb of homologous human exons
-
Full-text index only
Identification of the REST regulon reveals extensive transposable element-mediated binding site duplication.
PMID 16899447 · PMC1557810 · Nucleic acids research · 2006 · 8 claims · 8 setups
The RE1 PSSM identifies functional RE1 binding sites with greater sensitivity and selectivity than the previously used RE1 consensus sequence
-
Full-text index only
Identification of novel regulatory factor X (RFX) target genes by comparative genomics in Drosophila species.
PMID 17875208 · PMC2375033 · Genome biology · 2007 · 8 claims · 4 setups
A subset of C. elegans DAF-19 target genes have Drosophila homologs that are also regulated by dRFX, showing conservation of the RFX regulatory cascade between the two species.
-
Full-text index only
Genome-wide identification of specific oligonucleotides using artificial neural network and computational genomic analysis.
PMID 17518996 · PMC1892811 · BMC bioinformatics · 2007 · 7 claims · 4 setups
The IAB algorithm (integration of ANN and BLAST) identifies genome-wide specific oligos much faster than pure BLAST search while maintaining comparable success rate and cross homology
-
Has reproduction · 97
Determination of complete chromosomal haplotypes by bulk DNA sequencing.
PMID 33957932 · PMC8101039 · Genome biology · 2021 · 8 claims · 8 setups
A hierarchical computational strategy that first builds high-confidence local haplotype blocks from long-range/linked-read linkage and then concatenates them into whole-chromosome haplotypes using Hi-C contacts
-
Full-text index only
Molecular archeology of L1 insertions in the human genome.
PMID 12372140 · PMC134481 · Genome biology · 2002 · 8 claims · 4 setups
TSDfinder, a new algorithm, refines RepeatMasker-identified L1 boundaries by locating poly(A) tails, TSDs, and inversion breakpoints
-
Full-text index only
Using several pair-wise informant sequences for de novo prediction of alternatively spliced transcripts.
PMID 16925842 · PMC1810557 · Genome biology · 2006 · 8 claims · 4 setups
MARS, an extension of the Twinscan algorithm, uses multiple pairwise informant genomes to predict human alternatively spliced transcripts de novo without expressed sequence information.
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.
-
Full-text index only
Genome-wide identification of in vivo protein-DNA binding sites from ChIP-Seq data.
PMID 18684996 · PMC2532738 · Nucleic acids research · 2008 · 8 claims · 7 setups
SISSRs identifies binding sites from ChIP-Seq short reads with much higher resolution than the standard region-clustering approach
-
Has reproduction · 71
polishCLR: A Nextflow Workflow for Polishing PacBio CLR Genome Assemblies.
PMID 36792366 · PMC9985148 · Genome biology and evolution · 2023 · 8 claims · 8 setups
polishCLR is a reproducible, containerized Nextflow workflow that implements best practices for polishing PacBio CLR genome assemblies.