Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 82
Reusable building blocks in biological systems.
PMID 30958230 · PMC6303794 · Journal of the Royal Society, Interface · 2018 · 8 claims · 5 setups
Biological systems can be decomposed into phenotypic building blocks (PBBs) whose reusability ranges from single-use (condition-specific) to constitutive
-
Full-text index only
Identification of the REST regulon reveals extensive transposable element-mediated binding site duplication.
PMID 16899447 · PMC1557810 · Nucleic acids research · 2006 · 8 claims · 8 setups
The RE1 PSSM identifies functional RE1 binding sites with greater sensitivity and selectivity than the previously used RE1 consensus sequence
-
Has reproduction · 60
Integrating herbarium specimen observations into global phenology data systems.
PMID 30937223 · PMC6426164 · Applications in plant sciences · 2019 · 7 claims · 5 setups
A new PPO release adds terms and properties to relate observations of parts of plants to whole plants, enabling integration of herbarium phenology data with field observation data.
-
Has reproduction · 65
SPEAQeasy: a scalable pipeline for expression analysis and quantification for R/bioconductor-powered RNA-seq analyses.
PMID 33932985 · PMC8088074 · BMC bioinformatics · 2021 · 8 claims · 5 setups
SPEAQeasy is a portable, easy-to-install, Nextflow-powered RNA-seq processing pipeline that lowers the computational entry barrier for biologists/clinicians
-
Full-text index only
All systems GO for understanding mouse gene function.
PMID 15610553 · PMC549721 · Journal of biology · 2004 · 7 claims · 4 setups
Quantitative, multivariate cross-tissue expression measurements are powerfully predictive of gene function
-
Full-text index only
Predictive screening for regulators of conserved functional gene modules (gene batteries) in mammals.
PMID 15882449 · PMC1134656 · BMC genomics · 2005 · 8 claims · 4 setups
A predictive computational screen covering ~40% of annotated protein-coding genes identified 21 co-expressed gene clusters with statistically supported sharing of cis-regulatory motifs.
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
CanPredict: a computational tool for predicting cancer-associated missense mutations.
PMID 17537827 · PMC1933186 · Nucleic acids research · 2007 · 8 claims · 7 setups
CanPredict is a web application providing public access to a random forest classifier that combines SIFT, LogR.E-value, and GOSS scores to predict whether a missense mutation is cancer-associated
-
Has reproduction · 75
PHA4GE quality control contextual data tags: standardized annotations for sharing public health sequence datasets with known quality issues to facilitate testing and training.
PMID 38860884 · PMC11261899 · Microbial genomics · 2024 · 7 claims · 3 setups
PHA4GE developed a set of standardized contextual data tags (five fields plus controlled-vocabulary terms) for annotating pathogen sequence datasets with known quality issues.
-
Has reproduction · 66
Identification of the stress granule transcriptome via RNA-editing in single cells and in vivo.
PMID 35784648 · PMC9243631 · Cell reports methods · 2022 · 8 claims · 7 setups
A purification-free hyperTRIBE adaptation using FMR1-ADARcd-V5 identifies stress granule RNAs via condition-specific A-to-G editing read out by VASA-seq.
-
Has reproduction · 44
Dynamic Gene Attention Focus (DyGAF): Enhancing Biomarker Identification Through Dual-Model Attention Networks.
PMID 40160891 · PMC11951896 · Bioinformatics and biology insights · 2025 · 6 claims · 5 setups
DyGAF, a dual-model attention neural network (independent Model A + dependent Model B), identifies and ranks genes by significance for COVID-19 biomarker discovery more effectively than differential expression analysis (DEA) and random forest (RF) feature selection
-
Has reproduction · 74
Genetic architecture of natural variation of cardiac performance from flies to humans.
PMID 36383075 · PMC9668334 · eLife · 2022 · 8 claims · 7 setups
Natural genetic variation significantly influences cardiac performance traits (rhythmicity and contractility) across 167 DGRP lines
-
Full-text index only
Inferring combinatorial regulation of transcription in silico.
PMID 15647509 · PMC546154 · Nucleic acids research · 2005 · 8 claims · 5 setups
Combining Cluster-Buster (TFBS cluster prediction) with GOSSIP (rigorous GO enrichment statistics with multiple-testing/FDR correction) predicts biological functions controlled by combinatorial transcription factor action, without prior knowledge of factor targets
-
Full-text index only
Error-pooling-based statistical methods for identifying novel temporal replication profiles of human chromosomes observed by DNA tiling arrays.
PMID 17430969 · PMC1888820 · Nucleic acids research · 2007 · 8 claims · 4 setups
Developed an LPE-based error-pooling and weighted ANOVA modeling approach for statistical analysis of high-density tiling array data
-
Full-text index only
Mutation of ERBB2 provides a novel alternative mechanism for the ubiquitous activation of RAS-MAPK in ovarian serous low malignant potential tumors.
PMID 19010816 · PMC6953412 · Molecular cancer research : MCR · 2008 · 8 claims · 8 setups
Activating RAS-MAPK pathway mutations are present in >70% of serous LMP tumors versus ~12.5% of serous ovarian carcinomas
-
Full-text index only
A statistical framework for consolidating "sibling" probe sets for Affymetrix GeneChip data.
PMID 18435860 · PMC2397416 · BMC genomics · 2008 · 7 claims · 4 setups
A two-way ANOVA model with a treatment x probe-set interaction term can automatically determine whether sibling probe sets for a gene behave similarly (non-significant interaction, consolidate) or differently (significant interaction, treat as independent)
-
Has reproduction · 65
Cancer-predicting transcriptomic and epigenetic signatures revealed for ulcerative colitis in patient-derived epithelial organoids.
PMID 29983891 · PMC6033374 · Oncotarget · 2018 · 7 claims · 6 setups
UC patient-derived organoids histologically phenocopy primary UC tissue, showing disorganized stratified epithelium, reduced mucin/goblet cells, and non-uniform proliferation compared to non-IBD organoids
-
Has reproduction · 85
Digital sorting of complex tissues for cell type-specific gene expression profiles.
PMID 23497278 · PMC3626856 · BMC bioinformatics · 2013 · 8 claims · 8 setups
The Digital Sorting Algorithm (DSA) deconvolves mixed tissue expression into cell type-specific profiles using only marker genes, without requiring prior knowledge of cell type frequencies or in vitro pure-cell profiles.
-
Has reproduction · 30
Minimal metabolic pathway structure is consistent with associated biomolecular interactions.
PMID 24987116 · PMC4299494 · Molecular systems biology · 2014 · 8 claims · 8 setups
MinSpan, a mixed-integer linear optimization algorithm, computes the shortest, linearly independent pathways (sparsest basis of the null space of the stoichiometric matrix S) for genome-scale metabolic networks, which convex approaches (extreme pathways, elementary flux modes) cannot do at genome scale.
-
Has reproduction · 58
Revised annotations, sex-biased expression, and lineage-specific genes in the Drosophila melanogaster group.
PMID 25273863 · PMC4267930 · G3 (Bethesda, Md.) · 2014 · 8 claims · 6 setups
Revised RNA-seq-based gene models for D. ananassae, D. yakuba, and D. simulans include UTRs, empirically verified intron-exon boundaries, and previously unannotated novel exons, improving on r1.3 comparative-genomics annotations that lack UTRs.