Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Full-text index only
Human PAML browser: a database of positive selection on human genes using phylogenetic methods.
PMID 17962310 · PMC2238824 · Nucleic acids research · 2008 · 8 claims · 5 setups
The Human PAML Browser is a web-accessible database of codeml-based positive selection test results for 13,721 human genes with orthologs in UCSC multispecies alignments.
-
Full-text index only
Next-generation high-density self-assembling functional protein arrays.
PMID 18469824 · PMC3070491 · Nature methods · 2008 · 8 claims · 7 setups
A next-generation NAPPA method produces high-density protein microarrays displaying over 1500 unique proteins with >90% expression success
-
Full-text index only
Assessing individual differences in genome-wide gene expression in human whole blood: reliability over four hours and stability over 10 months.
PMID 19653838 · PMC3819565 · Twin research and human genetics : the official journal of the International Society for Twin Studies · 2009 · 8 claims · 5 setups
A subset of probesets (3,414) shows 4-hour test-retest reliability exceeding r=0.70 for detecting individual differences in gene expression.
-
Full-text index only
Novel mutations in the BRCA1 and BRCA2 genes in Iranian women with early-onset breast cancer.
PMID 12100744 · PMC116720 · Breast cancer research : BCR · 2002 · 8 claims · 4 setups
Germline BRCA1 and BRCA2 mutations are present in Iranian women with early-onset breast cancer, including four novel frameshift mutations
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
SeqDoC: rapid SNP and mutation detection by direct comparison of DNA sequence chromatograms.
PMID 15927052 · PMC1156871 · BMC bioinformatics · 2005 · 8 claims · 6 setups
SeqDoC generates a subtracted difference trace between a reference and test chromatogram that highlights single base changes
-
Full-text index only
Heterogeneous genomic molecular clocks in primates.
PMID 17029560 · PMC1592237 · PLoS genetics · 2006 · 7 claims · 7 setups
Non-CpG site substitutions show clear generation-time dependency, consistent with a replication-error origin
-
Full-text index only
The brain-derived neurotrophic factor rs6265 (Val66Met) polymorphism and depression in Mexican-Americans.
PMID 17632285 · PMC2686836 · Neuroreport · 2007 · 7 claims · 4 setups
BDNF SNP rs6265 (Val66Met) is significantly associated with diagnosis of major depression in Mexican-Americans
-
Full-text index only
Detecting purely epistatic multi-locus interactions by an omnibus permutation test on ensembles of two-locus analyses.
PMID 19761607 · PMC2759961 · BMC bioinformatics · 2009 · 8 claims · 5 setups
2LOmb performs an omnibus permutation test on ensembles of two-locus analyses via a four-step algorithm (two-locus analysis, permutation test, global p-value determination, progressive ensemble search)
-
Has reproduction · 71
polishCLR: A Nextflow Workflow for Polishing PacBio CLR Genome Assemblies.
PMID 36792366 · PMC9985148 · Genome biology and evolution · 2023 · 8 claims · 8 setups
polishCLR is a reproducible, containerized Nextflow workflow that implements best practices for polishing PacBio CLR genome assemblies.
-
Full-text index only
Nuclear beta-catenin expression is closely related to ulcerative growth of colorectal carcinoma.
PMID 11953860 · PMC2364167 · British journal of cancer · 2002 · 7 claims · 5 setups
Nuclear β-catenin expression is significantly associated with ulcerative growth of colorectal cancer
-
Full-text index only
The excess of 5' introns in eukaryotic genomes.
PMID 16314314 · PMC1292992 · Nucleic acids research · 2005 · 7 claims · 4 setups
All 21 eukaryotic genomes studied show a statistically significant 5′-biased distribution of introns in protein-coding genes
-
Full-text index only
Shooting darts: co-evolution and counter-adaptation in hermaphroditic snails.
PMID 15799778 · PMC1080126 · BMC evolutionary biology · 2005 · 8 claims · 6 setups
Dart shooting introduces an allohormone that inhibits digestion of donated sperm, increasing the amount reaching the spermathecae and fertilizing eggs, thereby manipulating the mating partner's sperm storage.
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
A genome-wide approach to identify genetic loci with a signature of natural selection in the Irish population.
PMID 16904005 · PMC1779589 · Genome biology · 2006 · 8 claims · 7 setups
Eight SNPs with extreme European-branch locus-specific branch length (LSBL) were selected from a genome-wide FST dataset as candidates for selection in Europe.
-
Full-text index only
Application of machine learning in SNP discovery.
PMID 16398931 · PMC1955739 · BMC bioinformatics · 2006 · 8 claims · 6 setups
PolyBayes produces high false-positive SNP predictions even with stringent parameters
-
Full-text index only
Detection of venous thromboembolism by proteomic serum biomarkers.
PMID 17579716 · PMC1891085 · PloS one · 2007 · 5 claims · 8 setups
A neural network-based classifier built from direct MALDI-TOF MS serum protein expression profiles can diagnose VTE with sensitivity/specificity that exceeds D-dimer assays
-
Full-text index only
PRESTO: rapid calculation of order statistic distributions and multiple-testing adjusted P-values via permutation for one and two-stage genetic association studies.
PMID 18620604 · PMC2483288 · BMC bioinformatics · 2008 · 8 claims · 4 setups
PRESTO is an order of magnitude faster than other existing permutation testing software for genetic association studies.
-
Full-text index only
GoMiner: a resource for biological interpretation of genomic and proteomic data.
PMID 12702209 · PMC154579 · Genome biology · 2003 · 8 claims · 4 setups
GoMiner organizes 'interesting' gene lists (e.g., differentially expressed genes) into the Gene Ontology hierarchy for biological interpretation, displaying results as both a tree and a directed acyclic graph (DAG).