Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 49
Integrative transcriptomics and single-cell transcriptomics analyses reveal potential biomarkers and mechanisms of action in papillary thyroid carcinoma.
PMID 40520228 · PMC12162626 · Frontiers in genetics · 2025 · 8 claims · 8 setups
ENTPD1, SERPINA1, and TACSTD2 are potential transcriptomic biomarkers for PTC
-
Full-text index only
Mining expressed sequence tags identifies cancer markers of clinical interest.
PMID 17078886 · PMC1635568 · BMC bioinformatics · 2006 · 8 claims · 6 setups
An EST-mining approach (Fisher Exact Test on tumor vs. non-tumor library hit counts) identifies differentially expressed transcripts with an estimated false discovery rate below 22% when human and mouse screens are combined.
-
Full-text index only
Expoldb: expression linked polymorphism database with inbuilt tools for analysis of expression and simple repeats.
PMID 17038195 · PMC1618849 · BMC genomics · 2006 · 8 claims · 6 setups
EXPOLDB is a novel database integrating human gene expression variability data (including monozygotic twin comparisons) with (TG/CA)n repeat polymorphism information
-
Has reproduction · 92
Systematic review of human post-mortem immunohistochemical studies and bioinformatics analyses unveil the complexity of astrocyte reaction in Alzheimer's disease.
PMID 34297416 · PMC8766893 · Neuropathology and applied neurobiology · 2022 · 8 claims · 5 setups
Systematic review of 306 eligible articles identified 196 distinct proteins constituting the ADRA (AD reactive astrocyte) protein set
-
Has reproduction · 58
The Li2 mutation results in reduced subgenome expression bias in elongating fibers of allotetraploid cotton (Gossypium hirsutum L.).
PMID 24598808 · PMC3944810 · PloS one · 2014 · 8 claims · 7 setups
The Li2 mutation significantly reduces subgenome (homeolog) expression bias in the elongating fiber transcriptome.
-
Has reproduction · 96
Scalable Prediction of Acute Myeloid Leukemia Using High-Dimensional Machine Learning and Blood Transcriptomics.
PMID 31918046 · PMC6992905 · iScience · 2020 · 8 claims · 8 setups
Data-driven, high-dimensional ML approaches that learn multivariate signatures directly from genome-wide transcriptomic data (no prior gene selection) yield accurate and robust AML classifiers.
-
Full-text index only
Assessing individual differences in genome-wide gene expression in human whole blood: reliability over four hours and stability over 10 months.
PMID 19653838 · PMC3819565 · Twin research and human genetics : the official journal of the International Society for Twin Studies · 2009 · 8 claims · 5 setups
A subset of probesets (3,414) shows 4-hour test-retest reliability exceeding r=0.70 for detecting individual differences in gene expression.
-
Has reproduction · 65
SPEAQeasy: a scalable pipeline for expression analysis and quantification for R/bioconductor-powered RNA-seq analyses.
PMID 33932985 · PMC8088074 · BMC bioinformatics · 2021 · 8 claims · 5 setups
SPEAQeasy is a portable, easy-to-install, Nextflow-powered RNA-seq processing pipeline that lowers the computational entry barrier for biologists/clinicians
-
Has reproduction · 48
Comparative analysis of molecular signatures reveals a hybrid approach in breast cancer: Combining the Nottingham Prognostic Index with gene expressions into a hybrid signature.
PMID 35143511 · PMC8830616 · PloS one · 2022 · 8 claims · 6 setups
A hybrid signature combining the Nottingham Prognostic Index with SIS-selected gene expressions can be built in a data-driven fashion (NPI treated as a gene expression during feature selection).
-
Full-text index only
Expression profiling of drug response--from genes to pathways.
PMID 17117610 · PMC3181826 · Dialogues in clinical neuroscience · 2006 · 8 claims · 8 setups
Understanding individual response to a drug (efficacy/tolerability) is the major bottleneck in current drug development and clinical trials.
-
Full-text index only
WebGestalt: an integrated system for exploring gene sets in various biological contexts.
PMID 15980575 · PMC1160236 · Nucleic acids research · 2005 · 8 claims · 6 setups
WebGestalt is an integrated web-based system composed of four modules: gene set management, information retrieval, organization/visualization, and statistics.
-
Full-text index only
CRSD: a comprehensive web server for composite regulatory signature discovery.
PMID 16845073 · PMC1538777 · Nucleic acids research · 2006 · 7 claims · 5 setups
CRSD is a comprehensive web server integrating six large-scale databases (UniGene, mature microRNAs, putative promoter, TRANSFAC, pathway, GO) plus two newly constructed genome-wide databases (MRS and TRS) for composite regulatory signature discovery
-
Has reproduction · 83
Analyzing biomarker discovery: Estimating the reproducibility of biomarker sets.
PMID 35901020 · PMC9333302 · PloS one · 2022 · 7 claims · 3 setups
A Reproducibility Score, RS(D,BD), defined as the average Jaccard overlap between biomarker sets found by the same discovery process on comparable datasets from the same distribution, quantifies biomarker reproducibility on a 0-1 scale
-
Has reproduction · 81
SEMdag: Fast learning of Directed Acyclic Graphs via node or layer ordering.
PMID 39775401 · PMC11709272 · PloS one · 2025 · 8 claims · 5 setups
SEMdag() is a two-step order-based algorithm for fast learning of high-dimensional linear SEMs, using knowledge-based (KB) or data-driven bottom-up (BU) node/layer ordering followed by penalized (L1) DAG estimation
-
Has reproduction · 83
Gene-expression patterns in peripheral blood classify familial breast cancer susceptibility.
PMID 26538066 · PMC4634735 · BMC medical genomics · 2015 · 8 claims · 5 setups
A multigene peripheral-blood gene-expression biomarker accurately classifies which women from high-risk families develop familial breast cancer.
-
Has reproduction · 80
Colorectal Cancer Prediction Based on Weighted Gene Co-Expression Network Analysis and Variational Auto-Encoder.
PMID 32825264 · PMC7563725 · Biomolecules · 2020 · 6 claims · 7 setups
Combining WGCNA-derived hub genes with a VAE-derived 10-dimensional representation as features for an SVM classifier achieves high accuracy (0.9692) and AUC (0.9981) for colorectal cancer prediction.
-
Has reproduction · 85
ScLRTC: imputation for single-cell RNA-seq data via low-rank tensor completion.
PMID 34844559 · PMC8628418 · BMC genomics · 2021 · 8 claims · 8 setups
scLRTC imputes dropout entries closest to the original expression values on simulated datasets, outperforming other state-of-the-art methods by SSE and PCC.
-
Full-text index only
Nuclear beta-catenin expression is closely related to ulcerative growth of colorectal carcinoma.
PMID 11953860 · PMC2364167 · British journal of cancer · 2002 · 7 claims · 5 setups
Nuclear β-catenin expression is significantly associated with ulcerative growth of colorectal cancer
-
Full-text index only
Next-generation high-density self-assembling functional protein arrays.
PMID 18469824 · PMC3070491 · Nature methods · 2008 · 8 claims · 7 setups
A next-generation NAPPA method produces high-density protein microarrays displaying over 1500 unique proteins with >90% expression success
-
Full-text index only
Genetical genomic determinants of alcohol consumption in rats and humans.
PMID 19874574 · PMC2777866 · BMC biology · 2009 · 8 claims · 6 setups
A genetical genomics approach combining brain gene expression, bQTL, and eQTL analysis in HXB/BXH RI rats identifies candidate genes for alcohol consumption