Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
CorGen--measuring and generating long-range correlations for DNA sequence analysis.
PMID 16845099 · PMC1538783 · Nucleic acids research · 2006 · 8 claims · 3 setups
CorGen is a web server that measures long-range correlations in DNA sequences and generates random sequences with the same (or user-specified) correlation and composition parameters
-
Full-text index only
DBD--taxonomically broad transcription factor predictions: new content and functionality.
PMID 18073188 · PMC2238844 · Nucleic acids research · 2008 · 8 claims · 3 setups
DBD is a database of predicted sequence-specific DNA-binding transcription factors covering over 700 publicly available proteomes, up from 150 in the initial version.
-
Full-text index only
Rare germline mutations in the BRCA2 gene are associated with early-onset prostate cancer.
PMID 17700570 · PMC2360390 · British journal of cancer · 2007 · 8 claims · 3 setups
Germline protein-truncating BRCA2 mutations confer an elevated relative risk (~7.8-fold) of early-onset prostate cancer in Caucasian men
-
Full-text index only
Flanking p10 contribution and sequence bias in matrix based epitope prediction: revisiting the assumption of independent binding pockets.
PMID 18925947 · PMC2600787 · BMC structural biology · 2008 · 8 claims · 3 setups
The extended matrix PP10 (built from a proline-containing peptide library) shows significant improvement in binding prediction over the original nine-residue matrix P9
-
Full-text index only
iRefIndex: a consolidated protein interaction database with provenance.
PMID 18823568 · PMC2573892 · BMC bioinformatics · 2008 · 6 claims · 3 setups
A reproducible key (ROG) for each protein interactor and a corresponding key (RIG) for each interaction record can be generated by anyone using only primary sequence, taxonomy identifier, and the SHA-1 algorithm (SEGUID).
-
Full-text index only
A quantitative proteomic approach for identification of potential biomarkers in hepatocellular carcinoma.
PMID 18715028 · PMC3769105 · Journal of proteome research · 2008 · 8 claims · 3 setups
iTRAQ-based quantitative LC-MS/MS proteomics can identify and quantitate differentially expressed proteins between HCC tumor and adjacent noncancerous liver tissue
-
Full-text index only
Prediction of candidate primary immunodeficiency disease genes using a support vector machine learning approach.
PMID 19801557 · PMC2780952 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2009 · 6 claims · 3 setups
An SVM trained on 69 binary features of known PID genes can accurately classify PID vs non-PID genes and predict novel candidate PID genes
-
Has reproduction
Empowering integrative and collaborative exploration of single-cell and spatial multimodal data with SGS genome browser.
PMID 40233745 · PMC12143324 · Cell genomics · 2025 · 8 claims · 6 setups
SGS is a user-friendly, collaborative, versatile browser for integrative visualization of single-cell and spatial multimodal (scMulti-omics) data
-
Full-text index only
COMUS: Clinician-Oriented locus-specific MUtation detection and deposition System.
PMID 19958500 · PMC2788389 · BMC genomics · 2009 · 8 claims · 6 setups
COMUS is a bioinformatics system for detecting and depositing new mutations from patient DNA with a clinician-friendly interface
-
Full-text index only
GermVarX: A Robust Workflow for Joint Germline Variant Exploration in whole-exome sequencing cohorts.
PMID 41926483 · PMC13046259 · PloS one · 2026 · 8 claims · 8 setups
GermVarX is a fully automated, modular Nextflow DSL2 workflow for joint germline variant discovery and exploration in WES cohort studies
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Has reproduction · 89
MirDIP 5.2: tissue context annotation and novel microRNA curation.
PMID 36453996 · PMC9825511 · Nucleic acids research · 2023 · 7 claims · 6 setups
mirDIP 5.2 removed eight outdated resources, added miRNATIP, and ran five prediction algorithms against miRBase and mirGeneDB miRNAs to expand and improve interaction coverage
-
Full-text index only
Having a BLAST with bioinformatics (and avoiding BLASTphemy).
PMID 11597340 · PMC138974 · Genome biology · 2001 · 8 claims · 4 setups
BLAST is the most widely used tool for searching biological sequences for regions of local similarity
-
Full-text index only
The meso-genomic era.
PMID 11516332 · PMC139414 · Genome biology · 2001 · 8 claims · 8 setups
Linkage disequilibrium (LD) between SNPs extends much further in Northern European populations (~120 kb) than in a Nigerian population (<10 kb), reflecting differing population histories (bottlenecks vs. constant expansion).
-
Full-text index only
Variation resources at UC Santa Cruz.
PMID 17151077 · PMC1781230 · Nucleic acids research · 2007 · 8 claims · 8 setups
The UCSC Genome Browser variation resources integrate polymorphism data from public collections (dbSNP, HapMap, Affymetrix, Perlegen, SeattleSNPs) into a common format with additional annotations and genomic context.
-
Full-text index only
SUPERFAMILY--sophisticated comparative genomics, data mining, visualization and phylogeny.
PMID 19036790 · PMC2686452 · Nucleic acids research · 2009 · 7 claims · 6 setups
SUPERFAMILY provides structural, functional and evolutionary annotation for proteins from all completely sequenced genomes using SCOP-based hidden Markov models
-
Full-text index only
MBGD update 2010: toward a comprehensive resource for exploring microbial genome diversity.
PMID 19906735 · PMC2808943 · Nucleic acids research · 2010 · 8 claims · 6 setups
MBGD allows users to create ortholog groups using a specified subgroup of organisms, distinguishing it from other comparative genomics resources
-
Full-text index only
The chromosomal genome sequence of the common sea fan, Gorgonia ventalina (Linnaeus, 1758) (Malacalcyonacea: Gorgoniidae).
PMID 41625986 · PMC12856254 · Wellcome open research · 2026 · 7 claims · 6 setups
A chromosome-level genome assembly was produced for Gorgonia ventalina with a total length of 339.18 Mb
-
Full-text index only
The genome sequence of the Gold Spangle, Autographa bractea (Denis & Schiffermüller), 1775 (Lepidoptera: Noctuidae).
PMID 41710118 · PMC12910202 · Wellcome open research · 2026 · 8 claims · 8 setups
A chromosome-level, haplotype-resolved genome assembly was produced for Autographa bractea (Gold Spangle)
-
Full-text index only
The genome sequence of the Provence Hairstreak, Tomares ballus (Fabricius, 1787) (Lepidoptera: Lycaenidae).
PMID 41710119 · PMC12910201 · Wellcome open research · 2026 · 7 claims · 8 setups
A genome assembly was produced for a male Tomares ballus specimen consisting of two haplotypes (839.94 Mb and 831.10 Mb)