Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
CROPPER: a metagene creator resource for cross-platform and cross-species compendium studies.
PMID 16995941 · PMC1592126 · BMC bioinformatics · 2006 · 7 claims · 5 setups
CROPPER is a web-based software resource that combines genomic data from heterogeneous sources using identifier and orthologous gene information from the Ensembl database.
-
Has reproduction · 69
Bioinformatics and system biology approaches to identify pathophysiological impact of COVID-19 to the progression and severity of neurological diseases.
PMID 34601390 · PMC8483812 · Computers in biology and medicine · 2021 · 8 claims · 8 setups
COVID-19 and neurological diseases (NDs) share pathophysiological connections that influence disease progression and severity
-
Full-text index only
Effective quantitative real-time polymerase chain reaction analysis of the parkin gene (PARK2) exon 1-12 dosage.
PMID 17324265 · PMC1810516 · BMC medical genetics · 2007 · 8 claims · 3 setups
Developed a real-time TaqMan PCR method that quantifies PARK2 exon 1-12 copy number by comparing amplification signal to the β-globin internal control gene
-
Full-text index only
A method for detecting epistasis in genome-wide studies using case-control multi-locus association analysis.
PMID 18667089 · PMC2533022 · BMC genomics · 2008 · 7 claims · 2 setups
HFCC is a method/software for genome-wide epistasis detection using case-control multi-locus association analysis, combining a fast computing algorithm with flexibility to test a variety of epistatic models.
-
Has reproduction · 78
A single-cell compendium of human cerebrospinal fluid identifies disease-associated immune cell populations.
PMID 39744938 · PMC11684814 · The Journal of clinical investigation · 2025 · 8 claims · 4 setups
Integration of public and newly generated scRNA-seq datasets yields a compendium of 139 subjects (193 samples, 403,973 immune cells) spanning CSF and blood across healthy controls and multiple neurologic diseases.
-
Full-text index only
MtSNPscore: a combined evidence approach for assessing cumulative impact of mitochondrial variations in disease.
PMID 19758471 · PMC2745589 · BMC bioinformatics · 2009 · 8 claims · 5 setups
MtSNPscore, a weighted scoring pipeline combining literature evidence, in silico predictions, and case/control frequency, can prioritize likely pathogenic mtDNA variations
-
Full-text index only
BayesRare: Bayesian mixture model for population-level rare cell type detection in multi-subject single-cell RNA sequencing data.
PMID 41632592 · PMC12867491 · Briefings in bioinformatics · 2026 · 8 claims · 4 setups
BayesRare is a hierarchical Bayesian mixture model framework for population-level rare cell type detection in multi-subject scRNA-seq data
-
Has reproduction · 61
miEAA 2.0: integrating multi-species microRNA enrichment analysis and workflow management systems.
PMID 32374865 · PMC7319446 · Nucleic acids research · 2020 · 8 claims · 5 setups
miEAA 2.0 extends miRNA enrichment analysis to ten species (previously only Homo sapiens), accepting both precursor and mature miRNA input.
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
Non-linear mapping for exploratory data analysis in functional genomics.
PMID 15661072 · PMC548129 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A relaxation method for non-linear mapping adapts one pair of points per step rather than all points at once, and was originally shown by Chang and Lee to outperform Sammon's mapping in cluster detection effectiveness and computational efficiency.
-
Full-text index only
Genome-wide association studies in neurological disorders.
PMID 18940696 · PMC2824165 · The Lancet. Neurology · 2008 · 8 claims · 6 setups
GWAS can identify common genetic variability associated with a trait across the whole genome, avoiding the bias and low throughput of candidate-gene studies
-
Full-text index only
Highly cost-efficient genome-wide association studies using DNA pools and dense SNP arrays.
PMID 18276640 · PMC2346606 · Nucleic acids research · 2008 · 8 claims · 5 setups
Illumina HumanHap300 arrays are substantially more efficient than Affymetrix Genechip HindIII arrays for DNA-pooling based GWAS
-
Full-text index only
Integrating proteomic, transcriptional, and interactome data reveals hidden components of signaling and regulatory networks.
PMID 19638617 · PMC2889494 · Science signaling · 2009 · 8 claims · 6 setups
Pathway reconstruction can be modeled as a prize-collecting Steiner tree problem, balancing penalties for excluding terminal nodes against costs for including edges, controlled by a parameter β.