Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Statistical challenges in preprocessing in microarray experiments in cancer.
PMID 18829474 · PMC3529914 · Clinical cancer research : an official journal of the American Association for Cancer Research · 2008 · 8 claims · 7 setups
Choice of pre-processing method materially changes which features are found significantly associated with survival in the Beer et al. lung cancer microarray dataset
-
Full-text index only
Integrated multi-level quality control for proteomic profiling studies using mass spectrometry.
PMID 19055809 · PMC2657802 · BMC bioinformatics · 2008 · 7 claims · 5 setups
QC processes for identifying and removing low-quality spectra are often overlooked in proteomic profiling studies
-
Full-text index only
Independent component analysis reveals new and biologically significant structures in micro array data.
PMID 16762055 · PMC1557674 · BMC bioinformatics · 2006 · 7 claims · 8 setups
ICA applied to three microarray datasets reveals many biologically significant components, including low-ranking ones not obvious by rank alone
-
Has reproduction · 50
Integrative analysis of transcriptomic data reveals a predictive gene signature for chemoradiotherapy response in rectal cancer.
PMID 41550766 · PMC12803930 · iScience · 2026 · 8 claims · 5 setups
A 186-gene signature derived from six GEO transcriptomic datasets predicts nCRT response in LARC with AUC 0.80 in cross-validation
-
Has reproduction · 96
Deep learning based protocol to construct an immune-related gene network of host-pathogen interactions in plants.
PMID 36525344 · PMC9791427 · STAR protocols · 2023 · 6 claims · 6 setups
A deep-learning protocol (DLNet) ranks genes by their contribution to classifying treatment versus control expression data, identifying genes involved in host defense against pathogens.
-
Has reproduction · 72
Neoadjuvant sintilimab plus chemotherapy in EGFR-mutant NSCLC: Phase 2 trial interim results (NEOTIDE/CTONG2104).
PMID 38897205 · PMC11293361 · Cell reports. Medicine · 2024 · 8 claims · 8 setups
Neoadjuvant sintilimab plus carboplatin/nab-paclitaxel is clinically feasible and tolerable in resectable EGFR-mutant NSCLC, with all 18 patients completing treatment and undergoing radical surgery.
-
Full-text index only
Differential analysis for high density tiling microarray data.
PMID 17892592 · PMC2231405 · BMC bioinformatics · 2007 · 8 claims · 6 setups
gSAM, a generalized extension of Significance Analysis of Microarrays (SAM), uses a piece-wise function to segment genome-wide differential response by protein-coding vs non-coding regions and by 5' vs 3' vs intra-genic bias within genes, rather than treating a gene as an atomic unit.
-
Full-text index only
Predictive screening for regulators of conserved functional gene modules (gene batteries) in mammals.
PMID 15882449 · PMC1134656 · BMC genomics · 2005 · 8 claims · 4 setups
A predictive computational screen covering ~40% of annotated protein-coding genes identified 21 co-expressed gene clusters with statistically supported sharing of cis-regulatory motifs.
-
Has reproduction · 81
Enabling Single-Cell Drug Response Annotations from Bulk RNA-Seq Using SCAD.
PMID 36762572 · PMC10104628 · Advanced science (Weinheim, Baden-Wurttemberg, Germany) · 2023 · 7 claims · 7 setups
SCAD, a transfer learning framework integrating adversarial discriminative domain adaptation (ADDA), can infer single-cell drug sensitivities by transferring knowledge from bulk RNA-seq pharmacogenomic data (GDSC) to scRNA-seq target domains
-
Full-text index only
Quadratic regression analysis for gene discovery and pattern recognition for non-cyclic short time-course microarray experiments.
PMID 15850479 · PMC1127068 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A step-down quadratic regression method (fitting quadratic, then linear, then null models per gene) identifies differentially expressed genes and classifies them into 9 temporal expression patterns using continuous time information.
-
Full-text index only
Evolution of organelle-associated protein profiling.
PMID 19110081 · PMC2680700 · Journal of proteomics · 2009 · 8 claims · 8 setups
Traditional biochemical organelle isolation followed by MS cataloguing suffers high false-positive rates because organelles cannot be purified to homogeneity and are structurally heterogeneous
-
Full-text index only
Cardiovascular genetic medicine: genomic assessment of prognosis and diagnosis in patients with cardiomyopathy and heart failure.
PMID 20559924 · PMC4745893 · Journal of cardiovascular translational research · 2008 · 8 claims · 6 setups
Molecular signature analysis (MSA) uses machine-learning/classification methods (e.g., PAM/nearest shrunken centroids) on gene expression patterns to classify samples by phenotype for diagnosis, prognosis, or therapy response.
-
Has reproduction · 86
The selection of software and database for metagenomics sequence analysis impacts the outcome of microbial profiling and pathogen detection.
PMID 37027361 · PMC10081788 · PloS one · 2023 · 7 claims · 7 setups
Obtaining an accurate species-level microbial profile using current direct-read metagenomics profiling software is still a challenging task.
-
Has reproduction · 90
A Decentralized Kidney Transplant Biopsy Classifier for Transplant Rejection Developed Using Genes of the Banff-Human Organ Transplant Panel.
PMID 35619722 · PMC9128066 · Frontiers in immunology · 2022 · 6 claims · 6 setups
A random forest model trained solely on B-HOT panel genes (B-HOT Model) accurately classifies kidney transplant biopsies as NR, ABMR, or TCMR.
-
Full-text index only
New perspectives on an old disease: proteomics in cancer research.
PMID 17472735 · PMC1895992 · Genome biology · 2007 · 8 claims · 8 setups
The HUPO Plasma Proteome Project has catalogued over 3,020 non-redundant gene products (>7,000 proteins/isoforms) in human plasma, many originating from tissues/organs rather than being plasma-intrinsic.
-
Full-text index only
The continuing search for cancer-causing somatic mutations.
PMID 17319975 · PMC1851399 · Breast cancer research : BCR · 2007 · 8 claims · 3 setups
A screen of 91 breast cancers uncovered 87 somatic variants spread across 16 of the 21 genes examined
-
Has reproduction · 68
Regulatory and evolutionary adaptation of yeast to acute lethal ethanol stress.
PMID 33170850 · PMC7654773 · PloS one · 2020 · 8 claims · 4 setups
Yeast cells activate a rapid transcriptional reprogramming process following acute lethal ethanol stress that is likely adaptive for post-stress survival.
-
Has reproduction · 61
Multi-omics analyses identify mannose phosphate isomerase-centered hypoxia-induced angiogenesis signature in colorectal cancer.
PMID 41204349 · PMC12595641 · Journal of translational medicine · 2025 · 8 claims · 8 setups
Twelve HIA-related genes were identified that are transcriptionally activated by HIFs and functionally implicated in angiogenesis in CRC.
-
Full-text index only
The promise of biomarkers in colorectal cancer detection.
PMID 15322316 · PMC3839323 · Disease markers · 2004 · 8 claims · 8 setups
A panel/combination of biomarkers is likely to provide better predictive value for early CRC detection than any single biomarker.
-
Full-text index only
Re-evaluating early breast neoplasia.
PMID 18279539 · PMC2374963 · Breast cancer research : BCR · 2008 · 8 claims · 7 setups
The classic single linear model of breast cancer progression requires revision based on high-throughput molecular genetic and gene expression data.