Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Machine-learning approaches for classifying haplogroup from Y chromosome STR data.
PMID 18551166 · PMC2396484 · PLoS computational biology · 2008 · 8 claims · 5 setups
Y-STR allelic variability is partitioned more by differences among haplogroups than by differences among populations, suggesting Y-STRs carry haplogroup information
-
Full-text index only
The evolutionary dynamics of a rapidly mutating virus within and between hosts: the case of hepatitis C virus.
PMID 19911046 · PMC2768904 · PLoS computational biology · 2009 · 8 claims · 3 setups
The replication rate of the strain that initiates an infection has a strong effect on the fitness of the infection at the between-host level, even though the virus evolves rapidly within the host.
-
Has reproduction · 74
Transcriptome profiling of Giardia intestinalis using strand-specific RNA-seq.
PMID 23555231 · PMC3610916 · PLoS computational biology · 2013 · 8 claims · 8 setups
Most of the G. intestinalis genome is transcribed in in vitro-grown trophozoites, but at vastly different expression levels.
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
ADaCGH: A parallelized web-based application and R package for the analysis of aCGH data.
PMID 17710137 · PMC1940324 · PloS one · 2007 · 8 claims · 4 setups
ADaCGH implements eight CNA detection methods, including the best-performing ones from recent reviews (CBS, GLAD, CGHseg, HMM)
-
Full-text index only
Identifying drug effects via pathway alterations using an integer linear programming optimization formulation on phosphoproteomic data.
PMID 19997482 · PMC2776985 · PLoS computational biology · 2009 · 7 claims · 4 setups
An ILP formulation of the Boolean pathway optimization problem fits phosphoproteomic data faster and more efficiently than the previously used genetic algorithm (GA) approach.
-
Has reproduction · 89
Spatial information matters: are traditional imputation methods effective for spatial transcriptomics data?
PMID 41627342 · PMC12862982 · Briefings in bioinformatics · 2026 · 7 claims · 3 setups
No single existing SOTA imputation method consistently performs well across newer SRT platforms/datasets
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
Cancer-specific high-throughput annotation of somatic mutations: computational prediction of driver missense mutations.
PMID 19654296 · PMC2763410 · Cancer research · 2009 · 7 claims · 7 setups
CHASM, a Random Forest-based computational method, was developed to identify and prioritize missense mutations likely to be functional drivers of tumor cell proliferation.
-
Full-text index only
Functional annotation and identification of candidate disease genes by computational analysis of normal tissue gene expression data.
PMID 18560577 · PMC2409962 · PloS one · 2008 · 7 claims · 5 setups
Ranked Coexpression Groups (RCG) built from k=6 nearest coexpressed genes, combined with a majority-rule functional characterization, integrate multiple datasets/coexpression measures to generate high-confidence functional annotation predictions
-
Full-text index only
Mutation analysis of the ATR gene in breast and ovarian cancer families.
PMID 15987455 · PMC1175065 · Breast cancer research : BCR · 2005 · 8 claims · 5 setups
ATR mediates the DNA damage response by phosphorylating tumor suppressors such as p53, BRCA1 and CHK1, making it a plausible candidate breast/ovarian cancer susceptibility gene
-
Full-text index only
The Princeton Protein Orthology Database (P-POD): a comparative genomics analysis tool for biologists.
PMID 17712414 · PMC1942082 · PloS one · 2007 · 8 claims · 5 setups
P-POD is the first comparative genomics database to combine results from multiple computational ortholog/homolog prediction methods with manually curated literature-derived experimental evidence of functional conservation.
-
Full-text index only
Detection of venous thromboembolism by proteomic serum biomarkers.
PMID 17579716 · PMC1891085 · PloS one · 2007 · 5 claims · 8 setups
A neural network-based classifier built from direct MALDI-TOF MS serum protein expression profiles can diagnose VTE with sensitivity/specificity that exceeds D-dimer assays
-
Full-text index only
Error-pooling-based statistical methods for identifying novel temporal replication profiles of human chromosomes observed by DNA tiling arrays.
PMID 17430969 · PMC1888820 · Nucleic acids research · 2007 · 8 claims · 4 setups
Developed an LPE-based error-pooling and weighted ANOVA modeling approach for statistical analysis of high-density tiling array data
-
Full-text index only
Computer identification of snoRNA genes using a Mammalian Orthologous Intron Database.
PMID 16093549 · PMC1184218 · Nucleic acids research · 2005 · 8 claims · 5 setups
Created the Mammalian Orthologous Intron Database (MOID) containing orthologous introns of human, mouse and rat identified via conserved reading-frame position
-
Full-text index only
Functional features of gene expression profiles differentiating gastrointestinal stromal tumours according to KIT mutations and expression.
PMID 19943934 · PMC2794290 · BMC cancer · 2009 · 8 claims · 5 setups
Hundreds of genes differentiate GISTs according to KIT versus PDGFRA mutation and expression status, despite no discriminative profile for clinical/pathological parameters.
-
Full-text index only
An SVM-based system for predicting protein subnuclear localizations.
PMID 16336650 · PMC1325059 · BMC bioinformatics · 2005 · 7 claims · 3 setups
New kernels defined on k-peptide vectors mapped by BLOSUM62-based high-scored pair matrices (D1, D2, D3) improve SVM discrimination of protein subnuclear localization compared to conventional k-peptide encodings.
-
Has reproduction · 93
Experimental identification and in silico prediction of bacterivory in green algae.
PMID 33649548 · PMC8245530 · The ISME journal · 2021 · 7 claims · 6 setups
Five prasinophyte strains (Pterosperma cristatum NIES626, Pyramimonas parkeae CCMP726, Pyramimonas parkeae NIES254, Nephroselmis pyriformis RCC618, Dolichomastix tenuilepis CCMP3274) ingest live fluorescently labeled bacteria, detected by microscopy and/or flow cytometry
-
Has reproduction · 91
Cotranscriptional RNA strand exchange underlies the gene regulation mechanism in a purine-sensing transcriptional riboswitch.
PMID 35348734 · PMC9756952 · Nucleic acids research · 2022 · 7 claims · 6 setups
A nascent intermediate central helix forms in the yxjA riboswitch that is mutually exclusive with both the aptamer's P1 helix and the expression platform's intrinsic terminator hairpin.
-
Full-text index only
GoMiner: a resource for biological interpretation of genomic and proteomic data.
PMID 12702209 · PMC154579 · Genome biology · 2003 · 8 claims · 4 setups
GoMiner organizes 'interesting' gene lists (e.g., differentially expressed genes) into the Gene Ontology hierarchy for biological interpretation, displaying results as both a tree and a directed acyclic graph (DAG).