Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 80
Colorectal Cancer Prediction Based on Weighted Gene Co-Expression Network Analysis and Variational Auto-Encoder.
PMID 32825264 · PMC7563725 · Biomolecules · 2020 · 6 claims · 7 setups
Combining WGCNA hub genes and VAE 10-dimensional representation as features for an SVM classifier achieves high accuracy in predicting CRC
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
Iterative class discovery and feature selection using Minimal Spanning Trees.
PMID 15355552 · PMC520744 · BMC bioinformatics · 2004 · 7 claims · 5 setups
Iterating between MST-based clustering and t-statistic feature selection removes noise genes step-wise while sharpening the sample clustering
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.
-
Full-text index only
A non-parametric meta-analysis approach for combining independent microarray datasets: application using two microarray datasets pertaining to chronic allograft nephropathy.
PMID 18302764 · PMC2276496 · BMC genomics · 2008 · 8 claims · 6 setups
A novel non-parametric meta-analysis approach for combining independent microarray datasets is presented, requiring no distributional assumptions and being logically intuitive.
-
Full-text index only
Local combinational variables: an approach used in DNA-binding helix-turn-helix motif prediction with sequence information.
PMID 19651875 · PMC2761287 · Nucleic acids research · 2009 · 8 claims · 7 setups
The LCV approach predicts HTH motifs with 93.29% accuracy, 93.93% sensitivity and 92.66% specificity using only primary sequence information
-
Full-text index only
3D-super-enhancers are condensate-associated cis-regulatory communities.
PMID 41797539 · PMC12968393 · Nucleic acids research · 2026 · 8 claims · 8 setups
BOUQUET integrates genome topology, chromatin occupancy, and graph theory (label propagation) to assign CREs and transcription protein machinery to target genes and identify protein-rich 'communities' associated with condensates.
-
Has reproduction · 67
Adaptive learning embedding features to improve the predictive performance of SARS-CoV-2 phosphorylation sites.
PMID 37847658 · PMC10628388 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 6 setups
PSPred-ALE outperforms state-of-the-art SARS-CoV-2 phosphorylation site predictors (e.g. DeepIPs) and handcrafted feature-based methods in benchmarking comparisons
-
Has reproduction · 87
Forseti: a mechanistic and predictive model of the splicing status of scRNA-seq reads.
PMID 38940130 · PMC11256924 · Bioinformatics (Oxford, England) · 2024 · 7 claims · 5 setups
Forseti is the first probabilistic model for resolving the splicing status of exonic scRNA-seq reads by scoring putative fragments linking read alignments to proximate priming sites
-
Full-text index only
PPC: an algorithm for accurate estimation of SNP allele frequencies in small equimolar pools of DNA using data from high density microarrays.
PMID 16199750 · PMC1240117 · Nucleic acids research · 2005 · 7 claims · 6 setups
The PPC algorithm, which applies a probe-pair-specific second-degree polynomial correction, increases the accuracy of allele frequency estimates from pooled DNA compared with previously described algorithms
-
Full-text index only
Complex germline and somatic mutation processes at a haploid human minisatellite shown by single-molecule analysis.
PMID 18929582 · PMC2599865 · Mutation research · 2008 · 8 claims · 5 setups
Overall MSY1 mutation frequencies in sperm (2.68%) and blood (1.88%) are not significantly different
-
Full-text index only
Whole genome distribution and ethnic differentiation of copy number variation in Caucasian and Asian populations.
PMID 19956714 · PMC2776354 · PloS one · 2009 · 8 claims · 5 setups
3,019 CNVs (2,381 autosomal, 638 X chromosome) were identified across 985 Caucasian and 692 Asian individuals using the Affymetrix 500K array
-
Full-text index only
An SVM-based system for predicting protein subnuclear localizations.
PMID 16336650 · PMC1325059 · BMC bioinformatics · 2005 · 7 claims · 3 setups
New kernels defined on k-peptide vectors mapped by BLOSUM62-based high-scored pair matrices (D1, D2, D3) improve SVM discrimination of protein subnuclear localization compared to conventional k-peptide encodings.
-
Has reproduction · 89
Graph Random Forest: A Graph Embedded Algorithm for Identifying Highly Connected Important Features.
PMID 37509188 · PMC10377046 · Biomolecules · 2023 · 8 claims · 6 setups
GRF identifies effective features that form highly connected sub-graphs on the underlying biological network
-
Full-text index only
Characterization of the linkage disequilibrium structure and identification of tagging-SNPs in five DNA repair genes.
PMID 16091150 · PMC1208870 · BMC cancer · 2005 · 7 claims · 5 setups
Three of the five DNA repair genes (MRE11A, RAD50, XRCC4) do not conform to a contiguous haplotype block structure; instead SNPs in high LD can be non-contiguous, fitting a more flexible LD group paradigm
-
Full-text index only
HapMap-based study of the 17q21 ERBB2 amplicon in susceptibility to breast cancer.
PMID 17117180 · PMC2360759 · British journal of cancer · 2006 · 6 claims · 5 setups
Common genetic variation (tSNPs and haplotypes) across the 400-kb 17q21 ERBB2 amplicon is not associated with breast cancer risk in British women.
-
Full-text index only
Data driven network inference and longitudinal transcriptomics unveil dynamic regulation in Chronic Lymphocytic Leukaemia models.
PMID 41540084 · PMC12894745 · NPJ systems biology and applications · 2026 · 8 claims · 8 setups
The presence of immune cells in the environment significantly alters CLL cell activation
-
Has reproduction · 96
GC-biased gene conversion conceals the prediction of the nearly neutral theory in avian genomes.
PMID 30616647 · PMC6322265 · Genome biology · 2019 · 8 claims · 6 setups
gBGC conceals the correlation between life-history traits and dN/dS in birds; accounting for it reveals correlations consistent with nearly neutral theory
-
Full-text index only
Rare-type mutations of MMAC1 tumor suppressor gene in human glioma cell lines and their tumors of origin.
PMID 10551321 · PMC5926156 · Japanese journal of cancer research : Gann · 1999 · 8 claims · 6 setups
6 of 10 glioma cell lines examined showed MMAC1 mutations with presumed loss of heterozygosity (LOH)
-
Full-text index only
Pancreatic tumours: molecular pathways implicated in ductal cancer are involved in ampullary but not in exocrine nonductal or endocrine tumorigenesis.
PMID 11161385 · PMC2363700 · British journal of cancer · 2001 · 8 claims · 6 setups
PDC shows frequent alterations of K-ras, p53, p16 and DPC4, confirming these as the core molecular fingerprint of ductal cancer