Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 92
A network-guided protocol to discover susceptibility genes in genome-wide association studies using stability selection.
PMID 36609152 · PMC9850185 · STAR protocols · 2023 · 5 claims · 5 setups
The protocol identifies genes that are both statistically associated with a phenotype and functionally interconnected in a biological network
-
Full-text index only
Coxiella burnetii genotyping.
PMID 16102309 · PMC3320512 · Emerging infectious diseases · 2005 · 8 claims · 5 setups
Multispacer sequence typing (MST) is the first reliable method for typing Coxiella burnetii isolates
-
Full-text index only
BABELOMICS: a systems biology perspective in the functional annotation of genome-scale experiments.
PMID 16845052 · PMC1538844 · Nucleic acids research · 2006 · 8 claims · 8 setups
Babelomics is presented as an updated, complete suite of web tools for functional analysis of genome-scale experiments with new and improved modules
-
Has reproduction · 100
Smart spatial omics (S2-omics) optimizes region of interest selection to capture molecular heterogeneity in diverse tissues.
PMID 41298871 · PMC12662399 · Nature cell biology · 2025 · 7 claims · 6 setups
S2-omics is an end-to-end workflow that automatically selects ROIs from H&E histology images to maximize molecular information content for spatial omics profiling.
-
Has reproduction · 80
Differential analysis of RNA structure probing experiments at nucleotide resolution: uncovering regulatory functions of RNA structure.
PMID 35869080 · PMC9307511 · Nature communications · 2022 · 7 claims · 4 setups
DiffScan is a computational framework combining a Normalization module and a Scan module to identify SVRs at nucleotide resolution from SP data.
-
Full-text index only
A surrogate-based approach for post-genomic partner identification.
PMID 11602024 · PMC57814 · BMC biotechnology · 2001 · 8 claims · 5 setups
Peptide surrogates derived from random phage display libraries contain amino acid sequence information that identifies the natural biological partner of the panned target via database searching.
-
Full-text index only
Microarray analysis: genome-scale hypothesis scanning.
PMID 14551912 · PMC212694 · PLoS biology · 2003 · 8 claims · 5 setups
Microarrays can be used to both test and generate hypotheses, not merely to fish for candidate genes.
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Full-text index only
Evaluation of genome-wide chromatin library of Stat5 binding sites in human breast cancer.
PMID 15686596 · PMC549029 · Molecular cancer · 2005 · 8 claims · 5 setups
A chromatin library coupled with experimental validation can productively identify novel in vivo Stat5 chromatin binding sites in cancer, including abnormal regulatory sites in tumor-specific neochromatin.
-
Full-text index only
An evaluation of the performance of tag SNPs derived from HapMap in a Caucasian population.
PMID 16532062 · PMC1391920 · PLoS genetics · 2006 · 8 claims · 5 setups
CEU HapMap-derived tSNPs capture most of the genetic variation observed in the Estonian (EGP) population sample
-
Full-text index only
Motif discovery in promoters of genes co-localized and co-expressed during myeloid cells differentiation.
PMID 19059999 · PMC2632922 · Nucleic acids research · 2009 · 6 claims · 8 setups
A novel multi-step computational method (built on approximate pattern enumeration, binomial over-representation scoring with FDR correction, and k-medoids clustering) can identify over-represented motifs in a selected set of promoters relative to a background promoter set.
-
Full-text index only
The role of positive selection in determining the molecular cause of species differences in disease.
PMID 18837980 · PMC2576240 · BMC evolutionary biology · 2008 · 8 claims · 6 setups
Genes predicted to be under positive selection during human evolution are implicated in diseases (epithelial cancers, schizophrenia, autoimmune diseases, Alzheimer's disease) that differ in prevalence and symptomatology between humans and other mammals
-
Full-text index only
Identification of deleterious non-synonymous single nucleotide polymorphisms using sequence-derived information.
PMID 18588693 · PMC2446391 · BMC bioinformatics · 2008 · 8 claims · 5 setups
A decision tree built on 10 selected sequence-derived features classifies SAPs as Disease or Polymorphism with 82.6% accuracy and 0.607 MCC in cross-validation.
-
Full-text index only
Mapping proteins to disease terminologies: from UniProt to MeSH.
PMID 18460185 · PMC2367626 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Developed a three-step procedure (disease name extraction, exact matching, partial/similarity-based matching) to map UniProtKB/Swiss-Prot disease names to MeSH terms
-
Full-text index only
SNP selection for genes of iron metabolism in a study of genetic modifiers of hemochromatosis.
PMID 18366708 · PMC2289803 · BMC medical genetics · 2008 · 7 claims · 6 setups
Illumina validation/design scores above 0.6 are not strongly correlated with actual SNP genotyping performance (Gentrain score)
-
Full-text index only
Swarm intelligence based wavelet coefficient feature selection for mass spectral classification: an application to proteomics data.
PMID 19733729 · PMC2748225 · Analytica chimica acta · 2009 · 8 claims · 4 setups
ACA-based wavelet coefficient feature selection can achieve up to 100% classification accuracy on training, validating, and independent testing sets using only 5 selected features.
-
Full-text index only
Ontological Discovery Environment: a system for integrating gene-phenotype associations.
PMID 19733230 · PMC2783409 · Genomics · 2009 · 8 claims · 8 setups
ODE is a web-based system for storing, sharing, retrieving and analyzing phenotype-centered genomic data sets across species and experimental systems
-
Has reproduction · 99
The systematic assessment of completeness of public metadata accompanying omics studies in the Gene Expression Omnibus data repository.
PMID 40926267 · PMC12421755 · Genome biology · 2025 · 8 claims · 3 setups
Over 25% of critical metadata are omitted, with only 74.8% of relevant phenotypes available in publications or public repositories.
-
Full-text index only
BLASTO: a tool for searching orthologous groups.
PMID 17483516 · PMC1933156 · Nucleic acids research · 2007 · 7 claims · 2 setups
BLASTO treats each orthologous group as a unit and outputs a ranked list of orthologous groups instead of single sequences
-
Full-text index only
GenBank.
PMID 17202161 · PMC1781245 · Nucleic acids research · 2007 · 8 claims · 1 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI