Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Discovering cancer genes by integrating network and functional properties.
PMID 19765316 · PMC2758898 · BMC medical genomics · 2009 · 8 claims · 6 setups
Cancer genes have distinct PPI network topology (higher connectivity, higher clustering coefficient, shorter path length to known cancer genes) compared to non-cancer genes
-
Full-text index only
Identifying cis-regulatory sequences by word profile similarity.
PMID 19730735 · PMC2731932 · PloS one · 2009 · 8 claims · 8 setups
WPH-finder identifies putative co-regulated CRMs by scanning the genome for sequences with word profiles similar to a known CRM, without explicitly defining binding sites
-
Full-text index only
Apolipoprotein e, alcohol consumption, and risk of ischemic stroke: the Framingham Heart Study revisited.
PMID 19717024 · PMC2743951 · Journal of stroke and cerebrovascular diseases : the official journal of National Stroke Association · 2009 · 8 claims · 6 setups
ApoE E4 allele does not significantly modify the association between alcohol consumption and risk of ischemic stroke in subjects <65 years
-
Full-text index only
Identification of Molecular Subtypes of Clear-Cell Renal Cell Carcinoma in Patient-Derived Xenografts Using Multi-Omics.
PMID 40282537 · PMC12026142 · Cancers · 2025 · 8 claims · 7 setups
Each ccRCC PDX resembles one of the human molecular subtypes closely at both the transcript and protein levels.
-
Has reproduction · 85
Optimizing open data to support one health: best practices to ensure interoperability of genomic data from bacterial pathogens.
PMID 33103064 · PMC7568946 · One health outlook · 2020 · 8 claims · 3 setups
An open-access pathogen surveillance database (NCBI Pathogen Detection) plus contributor Best Practices enables FAIR, interoperable genomic data across human, animal, food, and environmental sources for One Health surveillance.
-
Full-text index only
The stem cell population of the human colon crypt: analysis via methylation patterns.
PMID 17335343 · PMC1808490 · PLoS computational biology · 2007 · 8 claims · 3 setups
A coalescent-based, full probabilistic model with MCMC Bayesian inference provides a more powerful alternative to prior forward-simulation approaches for analyzing methylation pattern data from crypts.
-
Full-text index only
Analyses and comparison of accuracy of different genotype imputation methods.
PMID 18958166 · PMC2569208 · PloS one · 2008 · 8 claims · 3 setups
Stronger LD produces higher imputation accuracy rates for all five methods
-
Full-text index only
An online database for brain disease research.
PMID 16594998 · PMC1489945 · BMC genomics · 2006 · 7 claims · 5 setups
SMRIDB is a comprehensive web-based database integrating gene expression data and clinical metadata to aid understanding of the genetic effects of brain disease (bipolar disorder, schizophrenia, depression)
-
Has reproduction · 52
Thermophilic and mesophilic sulfate reduction by rare biosphere bacteria in acidic metal-bearing mine wastes from the temperate climate zone.
PMID 41345487 · PMC12739157 · Scientific reports · 2025 · 7 claims · 8 setups
Sulfate reduction rates (SRR) at in-situ ambient temperature in acidic Bom-Gorkhon tailings sediments are unexpectedly high, up to 9.86 µmol SO4 cm-3 day-1.
-
Has reproduction · 45
Single-Cell Analysis Reveals Characterization of Infiltrating T Cells in Moderately Differentiated Colorectal Cancer.
PMID 33584715 · PMC7873865 · Frontiers in immunology · 2020 · 8 claims · 7 setups
Eight distinct T cell populations are identifiable in CRC tumor tissue and seven in peripheral blood by unsupervised clustering of scRNA-seq data.
-
Has reproduction · 67
binny: an automated binning algorithm to recover high-quality genomes from complex metagenomic datasets.
PMID 36239393 · PMC9677464 · Briefings in bioinformatics · 2022 · 8 claims · 8 setups
binny outperforms or is highly competitive with commonly used and state-of-the-art binning methods (MetaBAT2, MaxBin2, CONCOCT, VAMB, SemiBin, MetaDecoder)
-
Has reproduction · 81
Enabling Single-Cell Drug Response Annotations from Bulk RNA-Seq Using SCAD.
PMID 36762572 · PMC10104628 · Advanced science (Weinheim, Baden-Wurttemberg, Germany) · 2023 · 7 claims · 7 setups
SCAD, a transfer learning framework integrating adversarial discriminative domain adaptation (ADDA), can infer single-cell drug sensitivities by transferring knowledge from bulk RNA-seq pharmacogenomic data (GDSC) to scRNA-seq target domains
-
Has reproduction · 85
PowerBacGWAS: a computational pipeline to perform power calculations for bacterial genome-wide association studies.
PMID 35338232 · PMC8956664 · Communications biology · 2022 · 8 claims · 8 setups
Two computational approaches (sub-sampling and phenotype-simulation) can be implemented to perform power calculations for bacterial GWAS using existing genome collections, packaged as the PowerBacGWAS pipeline
-
Has reproduction · 44
Detecting DNA modifications from SMRT sequencing data by modeling sequence context dependence of polymerase kinetic.
PMID 23516341 · PMC3597545 · PLoS computational biology · 2013 · 8 claims · 7 setups
Local sequence context strongly determines position-specific polymerase kinetic rate: roughly 80% of IPD variation is explained by a 10 bp context (7 bases upstream, 2 bases downstream of the incorporation site), saturating at 7 bases upstream.
-
Full-text index only
Competitive enzymatic reaction to control allele-specific extensions.
PMID 15767273 · PMC1065263 · Nucleic acids research · 2005 · 6 claims · 7 setups
Protease-mediated allele-specific extension (PrASE) uses competition between polymerase activity and Proteinase K-mediated polymerase degradation to allow extension of perfectly matched primers while eliminating slower mismatched primer extension.
-
Full-text index only
Towards alignment independent quantitative assessment of homology detection.
PMID 17205117 · PMC1762415 · PloS one · 2006 · 8 claims · 6 setups
The Fhom Estimator uses the prevalence of a conserved protein feature (X) in two protein sets to estimate the fraction of true homologs among paired proteins, independent of alignment quality.
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Full-text index only
Organismal complexity, cell differentiation and gene expression: human over mouse.
PMID 17881362 · PMC2095826 · Nucleic acids research · 2007 · 8 claims · 7 setups
Human shows a greater fraction of tissue-specific genes and a greater ratio of total expression of tissue-specific to housekeeping genes than mouse across 32 homologous tissues