Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 51
A platelet-related signature for predicting the prognosis and immunotherapy benefit in bladder cancer based on machine learning combinations.
PMID 39280688 · PMC11399026 · Translational andrology and urology · 2024 · 8 claims · 8 setups
An Enet (alpha=0.4) machine-learning algorithm built from 10 platelet-related genes yields the optimal platelet-related signature (PRS) for bladder cancer prognosis, with average C-index 0.73
-
Has reproduction · 83
Multiomic machine learning on lactylation for molecular typing and prognosis of lung adenocarcinoma.
PMID 39856156 · PMC11760357 · Scientific reports · 2025 · 8 claims · 8 setups
Ten multiomics clustering algorithms identify two distinct lactylation cancer subtypes (CS1 and CS2) in LUAD
-
Has reproduction · 81
Enabling Single-Cell Drug Response Annotations from Bulk RNA-Seq Using SCAD.
PMID 36762572 · PMC10104628 · Advanced science (Weinheim, Baden-Wurttemberg, Germany) · 2023 · 7 claims · 7 setups
SCAD, a transfer learning framework integrating adversarial discriminative domain adaptation (ADDA), can infer single-cell drug sensitivities by transferring knowledge from bulk RNA-seq pharmacogenomic data (GDSC) to scRNA-seq target domains
-
Has reproduction · 84
Foster thy young: enhanced prediction of orphan genes in assembled genomes.
PMID 34928390 · PMC9023268 · Nucleic acids research · 2022 · 7 claims · 8 setups
Each of the five gene-prediction pipelines under-predicts orphan genes, with detection as low as 11% under one prediction scenario.
-
Has reproduction · 35
Molecular and Clinical Relevance of ZBTB38 Expression Levels in Prostate Cancer.
PMID 32365491 · PMC7281456 · Cancers · 2020 · 8 claims · 7 setups
ZBTB38 expression is decreased in localised prostate cancer and further decreased in metastatic prostate cancer compared to non-cancerous prostate tissue
-
Full-text index only
Integrating complex genomic datasets and tumour cell sensitivity profiles to address a 'simple' question: which patients should get this drug?
PMID 20003409 · PMC2799438 · BMC medicine · 2009 · 8 claims · 5 setups
A panel of 48 genomically characterized breast cancer cell lines can model patient tumour heterogeneity to identify biomarkers predicting response to PG-11047
-
Has reproduction · 71
Assessment tool based on fatty acid metabolic signatures for predicting the prognosis and treatment response in bladder cancer.
PMID 38076064 · PMC10703629 · Heliyon · 2023 · 8 claims · 8 setups
Consensus clustering of prognosis-related fatty acid metabolism genes (FAMGs) identifies three molecular subtypes of BLCA (FAMC1, FAMC2, FAMC3) with distinct prognoses and tumor microenvironments
-
Full-text index only
EGASP: the human ENCODE Genome Annotation Assessment Project.
PMID 16925836 · PMC1810551 · Genome biology · 2006 · 8 claims · 6 setups
Best-performing computational gene prediction methods correctly predict at least one transcript for close to 70% of annotated genes in the ENCODE regions.
-
Full-text index only
Exogean: a framework for annotating protein-coding genes in eukaryotic genomic DNA.
PMID 16925841 · PMC1810556 · Genome biology · 2006 · 8 claims · 5 setups
Exogean is a framework using directed acyclic coloured multigraphs (DACMs) to represent biological objects (mRNA, ESTs, protein alignments, exons) and iteratively combine them into complex protein-coding transcript models.
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Full-text index only
Ab initio identification of human microRNAs based on structure motifs.
PMID 18088431 · PMC2238772 · BMC bioinformatics · 2007 · 8 claims · 7 setups
MiRPred predicts miRNA precursors ab initio using only predicted secondary structure motifs, ignoring nucleotide sequence
-
Full-text index only
Non-EST-based prediction of novel alternatively spliced cassette exons with cell signaling function in Caenorhabditis elegans and human.
PMID 17452356 · PMC1904267 · Nucleic acids research · 2007 · 8 claims · 7 setups
PASE (Prediction of Alternative Signaling Exons) is a computational algorithm combining Markov splice-site models, a Bayesian classifier, species conservation, and Scansite motif scoring to identify novel alternative cassette exons involved in cell signaling.
-
Has reproduction · 84
Improving recombinant protein production by yeast through genome-scale modeling using proteome constraints.
PMID 35624178 · PMC9142503 · Nature communications · 2022 · 7 claims · 5 setups
pcSecYeast, a proteome-constrained genome-scale model integrating metabolism, translation, and detailed secretory pathway processing (translocation, PTMs, folding, misfolding, degradation), was constructed for S. cerevisiae
-
Full-text index only
Prediction of candidate primary immunodeficiency disease genes using a support vector machine learning approach.
PMID 19801557 · PMC2780952 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2009 · 6 claims · 3 setups
An SVM trained on 69 binary features of known PID genes can accurately classify PID vs non-PID genes and predict novel candidate PID genes
-
Full-text index only
Twin peaks: the draft human genome sequence.
PMID 11276423 · PMC138909 · Genome biology · 2001 · 8 claims · 8 setups
The predicted number of human genes (~26,000-40,000) is far lower than the widely assumed ~100,000, though downstream RNA/protein complexity can still generate substantial biological complexity.
-
Full-text index only
GeneAlign: a coding exon prediction tool based on phylogenetical comparisons.
PMID 16845010 · PMC1538901 · Nucleic acids research · 2006 · 8 claims · 5 setups
GeneAlign predicts coding exons by using signal detection (GeneSplicer/WMM) combined with CORAL, a heuristic linear-time alignment tool, to align candidate signal-flanked regions against annotated exons of a homologous organism's genes
-
Full-text index only
ProMiR II: a web server for the probabilistic prediction of clustered, nonclustered, conserved and nonconserved microRNAs.
PMID 16845048 · PMC1538778 · Nucleic acids research · 2006 · 6 claims · 4 setups
ProMiR II improves on the original ProMiR by integrating free energy, G/C ratio, conservation score and entropy for more controllable miRNA prediction
-
Full-text index only
Using multiple alignments to improve seeded local alignment algorithms.
PMID 16100379 · PMC1185574 · Nucleic acids research · 2005 · 8 claims · 2 setups
Using information implicit in a multiple alignment to dynamically build a spaced-seed index weighted toward promising regions increases sensitivity of local alignment search compared to indexing a sequence alone
-
Has reproduction · 78
Enhancing chemotherapy response prediction via matched colorectal tumor-organoid gene expression analysis and network-based biomarker selection.
PMID 39754813 · PMC11754497 · Translational oncology · 2025 · 6 claims · 8 setups
A consensus WGCNA approach combining matched tumor-organoid and independent organoid drug-response expression data identifies gene modules and hub genes predictive of 5-FU chemotherapy response
-
Full-text index only
DiRE: identifying distant regulatory elements of co-expressed genes.
PMID 18487623 · PMC2447744 · Nucleic acids research · 2008 · 8 claims · 4 setups
DiRE predicts distant regulatory elements by combining gene co-expression data, comparative genomics and TFBS profiles to determine TFBS-association signatures