Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Exhaustive prediction of disease susceptibility to coding base changes in the human genome.
PMID 18793467 · PMC2537574 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Inter-species conservation is the strongest single predictor of disease-associated coding mutations among the factors tested.
-
Has reproduction · 50
Wx: a neural network-based feature selection algorithm for transcriptomic data.
PMID 31324856 · PMC6642261 · Scientific reports · 2019 · 8 claims · 8 setups
The Wx algorithm ranks genes using a DI score reflecting each gene's classification power for distinguishing two groups.
-
Full-text index only
DupyliCate: mining, classifying, and characterizing gene duplications.
PMID 42209743 · PMC13219399 · Scientific reports · 2026 · 8 claims · 8 setups
DupyliCate is a Python tool for identifying and classifying gene duplication arrays, using BUSCO-based species-specific thresholds and offering integrated expression divergence and Ka/Ks analysis.
-
Has reproduction · 57
Genome-wide kinetic properties of transcriptional bursting in mouse embryonic stem cells.
PMID 32596448 · PMC7299619 · Science advances · 2020 · 8 claims · 8 setups
Transcriptional bursting kinetics is regulated by a combination of promoter- and gene body-binding proteins, including the polycomb repressive complex 2 (PRC2) and transcription elongation factors
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Full-text index only
Statistical learning of peptide retention behavior in chromatographic separations: a new kernel-based approach for computational proteomics.
PMID 18053132 · PMC2254445 · BMC bioinformatics · 2007 · 6 claims · 5 setups
The paired oligo-border kernel (POBK) combined with SVMs predicts peptide adsorption/elution in SAX-SPE and retention time in IP-RP-HPLC more accurately than existing methods.
-
Full-text index only
Uncovering information on expression of natural antisense transcripts in Affymetrix MOE430 datasets.
PMID 17598913 · PMC1929078 · BMC genomics · 2007 · 8 claims · 4 setups
Standard Affymetrix expression GeneChips (MOE430, HG-U133) contain probe sets that detect natural antisense transcripts (NATs)
-
Full-text index only
In silico analysis of missense substitutions using sequence-alignment based methods.
PMID 18951440 · PMC3431198 · Human mutation · 2008 · 8 claims · 7 setups
Carefully validated PMSA-based computational algorithms can achieve predictive values of ~75-95% for classifying missense substitutions as pathogenic or neutral.
-
Full-text index only
Swarm intelligence based wavelet coefficient feature selection for mass spectral classification: an application to proteomics data.
PMID 19733729 · PMC2748225 · Analytica chimica acta · 2009 · 8 claims · 4 setups
ACA-based wavelet coefficient feature selection can achieve up to 100% classification accuracy on training, validating, and independent testing sets using only 5 selected features.
-
Has reproduction
DAGFormer: A graph-based domain adaptation approach for single-cell cancer drug response prediction.
PMID 41417875 · PMC12795466 · PLoS computational biology · 2025 · 7 claims · 5 setups
DAGFormer constructs cellular neighbor graphs using diverse topological strategies to represent intercellular interactions for drug response prediction
-
Has reproduction · 59
Refining breast cancer biomarker discovery and drug targeting through an advanced data-driven approach.
PMID 38253993 · PMC10810249 · BMC bioinformatics · 2024 · 8 claims · 8 setups
The BGWO_SA_Ens algorithm (hybrid BGWO + simulated annealing with an ensemble classifier objective function) selects predictive breast cancer biomarker genes with high classification performance
-
Has reproduction · 59
Application of Machine Learning in Predicting Hepatic Metastasis or Primary Site in Gastroenteropancreatic Neuroendocrine Tumors.
PMID 37887568 · PMC10605255 · Current oncology (Toronto, Ont.) · 2023 · 8 claims · 7 setups
Multi-gene random forest models classify primary tumor vs. liver metastasis samples with 100% accuracy in training/test cohorts and >90% accuracy in an independent validation cohort
-
Has reproduction
Developing a thyroid cancer differentiation state classification system using deep residual networks and metabolic signature profiling.
PMID 40993300 · PMC12460824 · NPJ digital medicine · 2025 · 8 claims · 8 setups
Metabolic status is closely tied to tumor differentiation state, with distinct metabolic reprogramming occurring during thyroid cancer dedifferentiation
-
Full-text index only
Characterization of the Schistosoma transcriptome opens up the world of helminth genomics.
PMID 14709167 · PMC395727 · Genome biology · 2003 · 8 claims · 5 setups
Near-complete transcriptome complements have now been described for S. japonicum and S. mansoni
-
Full-text index only
Serum diagnosis of diffuse large B-cell lymphomas and further identification of response to therapy using SELDI-TOF-MS and tree analysis patterning.
PMID 18163913 · PMC2242801 · BMC cancer · 2007 · 8 claims · 8 setups
SELDI-TOF-MS serum proteomic patterns analyzed by decision tree classification (Biomarker Pattern Software) can discriminate DLBCL patients from healthy controls with high sensitivity and specificity.
-
Full-text index only
Pan-Cancer Single-Cell RNA Sequencing Analysis Refines Multi-Origin Monocyte and Macrophage Lineages.
PMID 41231218 · PMC12865363 · Cancer immunology research · 2026 · 6 claims · 8 setups
TAMs arise from two distinct origins: C1QC+ TAMs likely derive from resident tissue macrophages, while SPP1+ TAMs and ISG15+ TAMs likely originate from circulating monocytes.
-
Has reproduction · 66
Integrative bioinformatics and artificial intelligence analyses of transcriptomics data identified genes associated with major depressive disorders including NRG1.
PMID 37583471 · PMC10423927 · Neurobiology of stress · 2023 · 7 claims · 5 setups
Differentially expressed genes in MDD patients are enriched in immune response, inflammatory response, neurodegeneration, and cerebellar atrophy pathways.
-
Full-text index only
Interaction preferences across protein-protein interfaces of obligatory and non-obligatory components are different.
PMID 16105176 · PMC1201154 · BMC structural biology · 2005 · 8 claims · 5 setups
Interaction patterns across obligatory and non-obligatory interfaces are different, with obligatory contacts predominantly non-polar
-
Full-text index only
Impact of short-read sequencing on the misassembly of a plant genome.
PMID 33530937 · PMC7852129 · BMC genomics · 2021 · 7 claims · 6 setups
Short-read tomato assembly has substantial high-coverage (0.6%, 5.1 Mb) and low-coverage (9.7%, 79.6 Mb) regions relative to background coverage
-
Full-text index only
Diverse Cyanopeptides follow distinct temporal succession patterns in freshwater harmful algal blooms.
PMID 41711085 · PMC13196583 · The ISME journal · 2026 · 8 claims · 7 setups
Shifts from Microcystis to Dolichospermum dominance occur later in the bloom season, coinciding with lower temperatures