Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 57
Identification and Mechanisms of Osteocyte Subsets Involved in the Pathological Progression of Osteoporosis.
PMID 41250977 · PMC12850396 · Advanced science (Weinheim, Baden-Wurttemberg, Germany) · 2026 · 8 claims · 8 setups
Six distinct osteocyte subsets (C1-C6) exist in mouse bone, identified by single-cell sequencing.
-
Has reproduction · 67
Heterogeneity and Differentiation Trajectories of Infiltrating CD8+ T Cells in Lung Adenocarcinoma.
PMID 36358600 · PMC9658355 · Cancers · 2022 · 7 claims · 8 setups
Infiltrating CD8+ T cells in LUAD can be divided into ten transcriptionally distinct subsets: eight cytotoxic (CTL) subsets, one naive-like (NTL) subset, and one exhausted (ETL) subset.
-
Full-text index only
Towards precise classification of cancers based on robust gene functional expression profiles.
PMID 15774002 · PMC1274255 · BMC bioinformatics · 2005 · 6 claims · 7 setups
Functional expression profiles (FEPs) achieve comparable or better classification performance than conventional gene expression profiles (GEPs) across four public microarray datasets
-
Has reproduction · 50
Integrative analysis of transcriptomic data reveals a predictive gene signature for chemoradiotherapy response in rectal cancer.
PMID 41550766 · PMC12803930 · iScience · 2026 · 8 claims · 5 setups
A 186-gene signature derived from six GEO transcriptomic datasets predicts nCRT response in LARC with AUC 0.80 in cross-validation
-
Full-text index only
Unravelling the hidden heterogeneities of diffuse large B-cell lymphoma based on coupled two-way clustering.
PMID 17888167 · PMC2082044 · BMC genomics · 2007 · 8 claims · 6 setups
A proposed coupled two-way clustering (CTWC/SPC) method combined with a GO-based functional concept consistency score can identify compact, robust gene subsets that define clinically meaningful DLBCL subtypes
-
Full-text index only
Computer-aided identification of polymorphism sets diagnostic for groups of bacterial and viral genetic variants.
PMID 17672919 · PMC1973086 · BMC bioinformatics · 2007 · 6 claims · 8 setups
The Not-N algorithm, incorporated into the Minimum SNPs program, identifies small marker sets diagnostic for user-defined subgroups of genetic variants with 0% false negatives
-
Has reproduction · 59
Refining breast cancer biomarker discovery and drug targeting through an advanced data-driven approach.
PMID 38253993 · PMC10810249 · BMC bioinformatics · 2024 · 8 claims · 8 setups
The BGWO_SA_Ens algorithm (hybrid BGWO + simulated annealing with an ensemble classifier objective function) selects predictive breast cancer biomarker genes with high classification performance
-
Has reproduction · 71
Gene Set Enrichment Analysis Reveals Individual Variability in Host Responses in Tuberculosis Patients.
PMID 34421903 · PMC8375662 · Frontiers in immunology · 2021 · 8 claims · 8 setups
TB patients show substantial individual variability in the intensity of hallmark IFN responses, as well as in complement system, metabolic, and other pathway responses.
-
Has reproduction · 71
Parsimonious Gene Correlation Network Analysis (PGCNA): a tool to define modular gene co-expression for refined molecular stratification in cancer.
PMID 30993001 · PMC6459838 · NPJ systems biology and applications · 2019 · 8 claims · 7 setups
Retaining only the top ~3 most correlated edges per gene (EPG3) combined with FastUnfold clustering (termed PGCNA) produces gene co-expression modules with significantly better separation and enrichment of known biology than using all edges or other clustering methods.
-
Full-text index only
Analysis of the glutathione S-transferase (GST) gene family.
PMID 15607001 · PMC3500200 · Human genomics · 2004 · 8 claims · 4 setups
The complete human GST gene family comprises 16 genes in six subfamilies: alpha (GSTA), mu (GSTM), omega (GSTO), pi (GSTP), theta (GSTT) and zeta (GSTZ).
-
Full-text index only
AUGUSTUS at EGASP: using EST, protein and genomic alignments for improved gene prediction in the human genome.
PMID 16925833 · PMC1810548 · Genome biology · 2006 · 8 claims · 5 setups
AUGUSTUS predicted significantly more genes correctly than any other ab initio program in EGASP
-
Full-text index only
Multiplex amplification of all coding sequences within 10 cancer genes by Gene-Collector.
PMID 17317684 · PMC1874629 · Nucleic acids research · 2007 · 7 claims · 7 setups
Gene-Collector is a method for multiplex nucleic acid amplification that specifically circularizes only correctly paired (cognate) PCR primer products on a Collector probe, degrading non-cognate artifacts by exonuclease treatment.
-
Full-text index only
SNP haplotype tagging from DNA pools of two individuals.
PMID 12709267 · PMC156884 · BMC bioinformatics · 2003 · 8 claims · 3 setups
An algorithm can reconstruct haplotypes from pools of two individuals' DNA under very general conditions, without requiring Hardy-Weinberg equilibrium.
-
Has reproduction · 96
Scalable Prediction of Acute Myeloid Leukemia Using High-Dimensional Machine Learning and Blood Transcriptomics.
PMID 31918046 · PMC6992905 · iScience · 2020 · 8 claims · 8 setups
Data-driven, high-dimensional ML approaches that learn multivariate signatures directly from genome-wide transcriptomic data (no prior gene selection) yield accurate and robust AML classifiers.
-
Has reproduction · 100
Identification of a PRDM1-regulated T cell network to regulate atherosclerotic plaque inflammation.
PMID 41039608 · PMC12490039 · Genome medicine · 2025 · 6 claims · 7 setups
A distinct gene co-expression module with a prominent T cell signature is enriched in unstable plaques and distinguishes high-risk from low-risk lesions.
-
Full-text index only
Association between telomere length and V(H) gene mutation status in chronic lymphocytic leukaemia: clinical and biological implications.
PMID 12592375 · PMC2377180 · British journal of cancer · 2003 · 7 claims · 5 setups
Unmutated VH gene CLL cases have significantly shorter telomeres than mutated VH gene CLL cases
-
Full-text index only
Prevalence of von Hippel-Lindau gene mutations in sporadic renal cell carcinoma: results from The Netherlands cohort study.
PMID 15932632 · PMC1177929 · BMC cancer · 2005 · 7 claims · 4 setups
VHL mutations were detected in 61% (114/187) of sporadic clear-cell RCC patients
-
Full-text index only
Relative impact of nucleotide and copy number variation on gene expression phenotypes.
PMID 17289997 · PMC2665772 · Science (New York, N.Y.) · 2007 · 8 claims · 5 setups
SNPs and CNVs capture largely non-overlapping signals of genetic variation affecting gene expression
-
Full-text index only
Genome-wide prediction of functional gene-gene interactions inferred from patterns of genetic differentiation in mice and men.
PMID 18270580 · PMC2217631 · PloS one · 2008 · 8 claims · 6 setups
Pairs of unlinked SNPs showing excess genetic differentiation (LD in mouse RILs, Fst in human populations) beyond what simulations/coalescent models predict by chance represent candidate functionally interacting (epistatic) gene pairs.
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance