Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Reliable Inference of Phylogenomic Relationship via Assembly-Based Strategy Accommodating Raw Reads and Proteins.
PMID 41800729 · PMC12969758 · Molecular ecology resources · 2026 · 7 claims · 8 setups
VEHoP infers protein-coding regions from diverse input types (raw reads, draft genomes, transcriptomes, annotated genomes) and automates generation of orthologous alignments, concatenated supermatrices, and phylogenetic trees in a single pipeline run.
-
Full-text index only
sCellST predicts single-cell gene expression from H& E images.
PMID 41513659 · PMC12858858 · Nature communications · 2026 · 7 claims · 6 setups
sCellST is a weakly supervised (Multiple Instance Learning) deep learning framework that predicts single-cell gene expression from H&E images alone, trained using paired spatial transcriptomics (Visium) and H&E slides
-
Has reproduction · 84
Foster thy young: enhanced prediction of orphan genes in assembled genomes.
PMID 34928390 · PMC9023268 · Nucleic acids research · 2022 · 8 claims · 6 setups
Each of the five tested gene prediction pipelines under-predicts orphan genes, as few as 11% detected under one scenario
-
Full-text index only
A global view of protein expression in human cells, tissues, and organs.
PMID 20029370 · PMC2824494 · Molecular systems biology · 2009 · 7 claims · 6 setups
A high fraction (>65%) of proteins is expressed in most human cells and tissues, while very few proteins (<2%) are detected in any single cell type.
-
Has reproduction · 90
Inferring a spatial code of cell-cell interactions across a whole animal body.
PMID 36395331 · PMC9714814 · PLoS computational biology · 2022 · 8 claims · 6 setups
cell2cell computes cell-cell interaction (CCI) potential using a novel modified Bray-Curtis score based on complementary coexpression of ligand-receptor pairs between cells
-
Full-text index only
Structural evolution of the protein kinase-like superfamily.
PMID 16244704 · PMC1261164 · PLoS computational biology · 2005 · 8 claims · 5 setups
All kinases in the superfamily share a 'universal core' domain consisting only of the regions required for ATP binding and the phosphotransfer reaction.
-
Has reproduction · 84
Single-cell protein activity analysis reveals aberrant myogenesis and IGF2-PI3K pathway dependencies in MYOD1-mutant rhabdomyosarcoma.
PMID 41758938 · PMC12947870 · Science advances · 2026 · 8 claims · 8 setups
MYOD1 L122R-mutant SRMS tumors contain three coexisting, conserved cell states (progenitor, transition, differentiated) reflecting aberrant myogenic differentiation
-
Full-text index only
Calibrating tissue level PDE models of ligand dynamics using single cell and spatial transcriptomics data.
PMID 41714655 · PMC13039149 · NPJ systems biology and applications · 2026 · 8 claims · 8 setups
scRNA-seq and spatial transcriptomics data provide a rich, underused source of information for calibrating tissue-scale PDE models of ligand dynamics.
-
Full-text index only
FLYNC: a machine-learning-driven framework for discovering long noncoding RNAs in Drosophila melanogaster.
PMID 41551930 · PMC12805895 · NAR genomics and bioinformatics · 2026 · 7 claims · 8 setups
FLYNC, an explainable boosting machine (EBM) model, accurately predicts the probability that a newly identified RNA transcript in D. melanogaster is a lncRNA
-
Full-text index only
Genome comparison without alignment using shortest unique substrings.
PMID 15910684 · PMC1166540 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A number of sequence comparison tasks, including detection of unique genomic regions, can be accomplished efficiently without an alignment step using shortest unique substrings.
-
Full-text index only
CanPredict: a computational tool for predicting cancer-associated missense mutations.
PMID 17537827 · PMC1933186 · Nucleic acids research · 2007 · 8 claims · 7 setups
CanPredict is a web application providing public access to a random forest classifier that combines SIFT, LogR.E-value, and GOSS scores to predict whether a missense mutation is cancer-associated
-
Full-text index only
A computational screen for type I polyketide synthases in metagenomics shotgun data.
PMID 18953415 · PMC2568958 · PloS one · 2008 · 8 claims · 6 setups
Combining HMM domain searches with maximum-likelihood phylogenetic trees can discriminate true PKS I sequences from evolutionarily related but functionally different enzymes (e.g., FAS I) in metagenomic data.
-
Full-text index only
In silico analysis of missense substitutions using sequence-alignment based methods.
PMID 18951440 · PMC3431198 · Human mutation · 2008 · 8 claims · 7 setups
Carefully validated PMSA-based computational algorithms can achieve predictive values of ~75-95% for classifying missense substitutions as pathogenic or neutral.
-
Full-text index only
Protocol for quantifying interaction patterns among genomic alterations in cancer.
PMID 41686643 · PMC12915222 · STAR protocols · 2026 · 6 claims · 5 setups
Background-aware permutation strategies that constrain permutation per gene and per sample enable robust, scalable inference of condition-specific (context-aware) genetic interactions across cancer cohorts
-
Has reproduction · 53
PulmonDB: a curated lung disease gene expression database.
PMID 31949184 · PMC6965635 · Scientific reports · 2020 · 7 claims · 6 setups
PulmonDB is a curated, web-based relational database (plus R package) integrating microarray and RNA-seq gene expression data with controlled-vocabulary annotation for COPD and IPF.
-
Full-text index only
Human synthetic lethal inference as potential anti-cancer target gene detection.
PMID 20015360 · PMC2804737 · BMC systems biology · 2009 · 7 claims · 8 setups
Targeting the synthetic lethal partner of a gene mutated in cancer selectively damages tumor cells while sparing healthy cells, offering a rationale for anti-cancer drug design