Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Systems biology of SNPs.
PMID 16820779 · PMC1681509 · Molecular systems biology · 2006 · 8 claims · 2 setups
Co-sets are groups of enzymatic reactions that are perfectly correlated (correlation coefficient of 1) in a reconstructed metabolic network and represent functional modules.
-
Full-text index only
Identification and analysis of co-occurrence networks with NetCutter.
PMID 18781200 · PMC2526157 · PloS one · 2008 · 8 claims · 4 setups
Random sampling from a complete permutation set of the bipartite graph permits co-occurrence analysis with optimal stringency, and the edge-swapping (ES) model closely approximates this and is the preferred null-model among six tested.
-
Has reproduction · 79
Interpretable prediction models for widespread m6A RNA modification across cell lines and tissues.
PMID 37995291 · PMC10697738 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 6 setups
CLSM6A, a CNN-based model set, predicts single-nucleotide-resolution m6A RNA modification sites across eight cell lines and three tissues in H. sapiens
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
Wiggle-predicting functionally flexible regions from primary sequence.
PMID 16839194 · PMC1500818 · PLoS computational biology · 2006 · 7 claims · 6 setups
A GNM-derived, correlation-weighted 'FF score' can objectively define functionally flexible regions (FFRs) that match experimentally confirmed flexible/functional regions (hinges, recognition loops, catalytic loops).
-
Has reproduction · 100
Computational modeling demonstrates that glioblastoma cells can survive spatial environmental challenges through exploratory adaptation.
PMID 31836713 · PMC6911112 · Nature communications · 2019 · 8 claims · 6 setups
Stochastic exploration of the gene-regulatory network structure confers enhanced adaptive capacity, enabling GBM cells to converge to new target phenotypes in novel environments.
-
Full-text index only
ADaCGH: A parallelized web-based application and R package for the analysis of aCGH data.
PMID 17710137 · PMC1940324 · PloS one · 2007 · 8 claims · 4 setups
ADaCGH implements eight CNA detection methods, including the best-performing ones from recent reviews (CBS, GLAD, CGHseg, HMM)
-
Full-text index only
Importance sampling for the infinite sites model.
PMID 18976228 · PMC2832804 · Statistical applications in genetics and molecular biology · 2008 · 7 claims · 2 setups
A new importance sampling proposal distribution for the ISM, derived from a new result on exact sampling from a single segregating site, generally shows greater efficiency than the GT and SD proposals.
-
Has reproduction · 100
DeepRNA-Reg: a deep-learning based approach for comparative analysis of CLIP experiments.
PMID 41055236 · PMC12505516 · RNA biology · 2025 · 7 claims · 5 setups
DeepRNA-Reg, a recurrent neural network-based algorithm, predicts differentially enriched sites in paired HITS-CLIP datasets and outperforms dCLIP 1.7.
-
Full-text index only
Inferring combinatorial regulation of transcription in silico.
PMID 15647509 · PMC546154 · Nucleic acids research · 2005 · 8 claims · 5 setups
Combining Cluster-Buster (TFBS cluster prediction) with GOSSIP (rigorous GO enrichment statistics with multiple-testing/FDR correction) predicts biological functions controlled by combinatorial transcription factor action, without prior knowledge of factor targets
-
Full-text index only
ARED 3.0: the large and diverse AU-rich transcriptome.
PMID 16381826 · PMC1347415 · Nucleic acids research · 2006 · 7 claims · 6 setups
ARED 3.0 computationally mapped more than 4000 ARE-mRNAs to the human genome, representing 5-8% of human genes.
-
Full-text index only
Endonuclease-independent insertion provides an alternative pathway for L1 retrotransposition in the human genome.
PMID 17517773 · PMC1920257 · Nucleic acids research · 2007 · 8 claims · 5 setups
An endonuclease-independent pathway (NCLI) for L1 insertion has been active in recent human genome evolution
-
Full-text index only
Motif discovery in promoters of genes co-localized and co-expressed during myeloid cells differentiation.
PMID 19059999 · PMC2632922 · Nucleic acids research · 2009 · 6 claims · 8 setups
A novel multi-step computational method (built on approximate pattern enumeration, binomial over-representation scoring with FDR correction, and k-medoids clustering) can identify over-represented motifs in a selected set of promoters relative to a background promoter set.
-
Has reproduction · 60
Core transcriptional signatures of phase change in the migratory locust.
PMID 31292921 · PMC6881432 · Protein & cell · 2019 · 8 claims · 7 setups
PhaseCore genes defined by AC-PCA contribution to phase differentiation predict phase status with >87.5% accuracy
-
Has reproduction · 95
Increased prevalence of hybrid epithelial/mesenchymal state and enhanced phenotypic heterogeneity in basal breast cancer.
PMID 38974967 · PMC11225361 · iScience · 2024 · 7 claims · 7 setups
Luminal breast cancer gene expression signature is closely/positively associated with an epithelial signature
-
Full-text index only
Detection of venous thromboembolism by proteomic serum biomarkers.
PMID 17579716 · PMC1891085 · PloS one · 2007 · 5 claims · 8 setups
A neural network-based classifier built from direct MALDI-TOF MS serum protein expression profiles can diagnose VTE with sensitivity/specificity that exceeds D-dimer assays
-
Full-text index only
Identifying alternative hyper-splicing signatures in MG-thymoma by exon arrays.
PMID 18545673 · PMC2409220 · PloS one · 2008 · 8 claims · 6 setups
An integrative ad-hoc functional GO analysis combining threshold-based (Fisher exact/hypergeometric) and threshold-free (Kolmogorov-Smirnov) statistics, plus term-to-parent comparisons, detects disease-relevant splicing events from exon array data.
-
Full-text index only
A comprehensive modular map of molecular interactions in RB/E2F pathway.
PMID 18319725 · PMC2290939 · Molecular systems biology · 2008 · 8 claims · 4 setups
A comprehensive, curated map of RB/E2F pathway molecular interactions was built using SBGN notation in CellDesigner and converted to BioPAX 2.0 format
-
Full-text index only
Cancer-specific high-throughput annotation of somatic mutations: computational prediction of driver missense mutations.
PMID 19654296 · PMC2763410 · Cancer research · 2009 · 7 claims · 7 setups
CHASM, a Random Forest-based computational method, was developed to identify and prioritize missense mutations likely to be functional drivers of tumor cell proliferation.
-
Full-text index only
Reconstruction of pathways associated with amino acid metabolism in human mitochondria.
PMID 18267298 · PMC5054205 · Genomics, proteomics & bioinformatics · 2007 · 8 claims · 5 setups
Out of 20 amino acids, the metabolic pathways of 17 utilize mitochondrial enzymes, and dysfunction of these enzymes causes over 40 known human mitochondrial diseases/disorders