Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A simple and robust method for connecting small-molecule drugs using gene-expression signatures.
PMID 18518950 · PMC2464610 · BMC bioinformatics · 2008 · 8 claims · 4 setups
A new method for building reference gene-expression profiles and scoring/testing connections improves on the original Connectivity Map by enabling statistical significance testing of connections.
-
Has reproduction · 58
A comparative study of techniques for differential expression analysis on RNA-Seq data.
PMID 25119138 · PMC4132098 · PloS one · 2014 · 8 claims · 8 setups
edgeR performs slightly better than DESeq and Cuffdiff2 in terms of the ability to uncover true positives.
-
Has reproduction · 50
Polymorphism identification and improved genome annotation of Brassica rapa through Deep RNA sequencing.
PMID 25122667 · PMC4232532 · G3 (Bethesda, Md.) · 2014 · 8 claims · 8 setups
330,995 SNPs were identified in transcribed regions between B. rapa genotypes R500 and IMB211, at an average frequency of one SNP per 200 bases.
-
Full-text index only
Inverse symmetry in complete genomes and whole-genome inverse duplication.
PMID 19898631 · PMC2771390 · PloS one · 2009 · 8 claims · 5 setups
Reverse and complement symmetries are essentially absent in genomic sequences at all scales.
-
Full-text index only
Genetic variation at hair length candidate genes in elephants and the extinct woolly mammoth.
PMID 19747392 · PMC2754481 · BMC evolutionary biology · 2009 · 8 claims · 5 setups
The coding sequence of FGF5 is not the critical determinant of hair length differences among elephantids, including the woolly mammoth.
-
Full-text index only
Sequence variation in G-protein-coupled receptors: analysis of single nucleotide polymorphisms.
PMID 15784611 · PMC1069129 · Nucleic acids research · 2005 · 7 claims · 8 setups
Position-specific phylogenetic features describing evolutionary conservation at a site (e.g. SIFT score, normalized site entropy, residue frequency change) are the best individual discriminators of disease-causing versus neutral GPCR mutations.
-
Full-text index only
The stem cell population of the human colon crypt: analysis via methylation patterns.
PMID 17335343 · PMC1808490 · PLoS computational biology · 2007 · 8 claims · 3 setups
A coalescent-based, full probabilistic model with MCMC Bayesian inference provides a more powerful alternative to prior forward-simulation approaches for analyzing methylation pattern data from crypts.
-
Full-text index only
ADZE: a rarefaction approach for counting alleles private to combinations of populations.
PMID 18779233 · PMC2732282 · Bioinformatics (Oxford, England) · 2008 · 6 claims · 2 setups
A generalized rarefaction-based statistic can estimate the sample size-corrected number of distinct alleles private to any combination of populations, generalizing Kalinowski's (2004) private allelic richness to groups of populations.
-
Full-text index only
Validation of analytic methods for biomarkers used in drug development.
PMID 18829475 · PMC2744124 · Clinical cancer research : an official journal of the American Association for Cancer Research · 2008 · 8 claims · 8 setups
Analytical method validation (assessing assay performance/reproducibility) is a distinct process from clinical qualification (linking a biomarker to biological processes and clinical endpoints), though the two are intertwined.
-
Has reproduction
D2H2: diabetes data and hypothesis hub.
PMID 38107655 · PMC10723036 · Bioinformatics advances · 2023 · 6 claims · 6 setups
D2H2 is a web-based portal integrating hundreds of curated diabetes-relevant transcriptomics datasets with bioinformatics tools for gene/gene set queries
-
Has reproduction · 42
The electrostatic profile of consecutive Cβ atoms applied to protein structure quality assessment.
PMID 25506420 · PMC4257144 · F1000Research · 2013 · 8 claims · 8 setups
The EPD between Cβ atoms of consecutive residues provides unique signatures of amino acid pair types and can discriminate native from decoy protein structures.
-
Full-text index only
Duplication count distributions in DNA sequences.
PMID 19256873 · PMC3121164 · Physical review. E, Statistical, nonlinear, and soft matter physics · 2008 · 8 claims · 8 setups
Duplication count distributions N(c) for complex 40-mers show power-law-like decay for c roughly 3 to 50 (or higher) across human, C. elegans, A. thaliana, and D. melanogaster genomes.