Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 77
Representing and querying disease networks using graph databases.
PMID 27462371 · PMC4960687 · BioData mining · 2016 · 7 claims · 8 setups
Graph databases are well suited for representing biological information because it is typically highly connected, semi-structured and unpredictable, unlike relational databases which require rigid schemas.
-
Has reproduction · 58
A comparative study of techniques for differential expression analysis on RNA-Seq data.
PMID 25119138 · PMC4132098 · PloS one · 2014 · 8 claims · 8 setups
edgeR performs slightly better than DESeq and Cuffdiff2 in terms of the ability to uncover true positives.
-
Has reproduction · 88
AuPairWise: A Method to Estimate RNA-Seq Replicability through Co-expression.
PMID 27082953 · PMC4833304 · PLoS computational biology · 2016 · 7 claims · 6 setups
Sample-sample correlation of transcript abundances is a misleading measure of replicability for assessing differential expression, because it is dominated by gene-specific dynamic ranges rather than condition-dependent variation.
-
Has reproduction · 95
nf-rnaSeqCount: A Nextflow pipeline for obtaining raw read counts from RNA-seq data.
PMID 35574063 · PMC9097006 · South African computer journal = Suid-Afrikaanse rekenaartydskrif · 2021 · 7 claims · 5 setups
nf-rnaSeqCount is a portable, reproducible Nextflow pipeline that maps RNA-seq reads to a reference genome and quantifies gene abundance for differential expression analysis
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
TEDD: a comprehensive database for translation efficiency dynamics.
PMID 41217970 · PMC12807600 · Nucleic acids research · 2026 · 8 claims · 4 setups
TEDD integrates 1518 RNA-seq, Ribo-seq, and RNC-seq samples from 143 human projects (279 datasets) spanning 24 tissues/cell types, 74 cell lines, and 52 conditions.
-
Full-text index only
High-throughput sequencing provides insights into genome variation and evolution in Salmonella Typhi.
PMID 18660809 · PMC2652037 · Nature genetics · 2008 · 7 claims · 8 setups
Evolution in the Typhi population is characterized by ongoing loss of gene function (pseudogene accumulation) rather than gain of function or diversifying selection.
-
Full-text index only
DNA methylation biomarkers-based pan-cancer classifier: predictive modeling for cancer classification.
PMID 42152108 · PMC13185202 · Genome medicine · 2026 · 8 claims · 5 setups
Relatively simple ML models (logistic regression) outperform complex algorithms such as deep neural networks for methylation-based cancer classification
-
Full-text index only
A latent activated olfactory stem cell state revealed by single-cell transcriptomic and epigenomic profiling.
PMID 41512864 · PMC12903091 · Stem cell reports · 2026 · 7 claims · 8 setups
HBC-derived regeneration proceeds via three distinct lineages (rHBC, Sus, mOSN) marked by sequential, lineage-specific transcription factor (TF) expression cascades.