Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Prediction of catalytic residues using Support Vector Machine with selected protein sequence and structural properties.
PMID 16790052 · PMC1534064 · BMC bioinformatics · 2006 · 8 claims · 7 setups
The Sequential Minimal Optimization (SMO) SVM algorithm was the best-performing classifier among 26 WEKA classifiers for predicting catalytic residues
-
Full-text index only
Machine-learning approaches for classifying haplogroup from Y chromosome STR data.
PMID 18551166 · PMC2396484 · PLoS computational biology · 2008 · 8 claims · 5 setups
Y-STR allelic variability is partitioned more by differences among haplogroups than by differences among populations, suggesting Y-STRs carry haplogroup information
-
Has reproduction · 100
miRbiom: Machine-learning on Bayesian causal nets of RBP-miRNA interactions successfully predicts miRNA profiles.
PMID 34637468 · PMC8509996 · PloS one · 2021 · 7 claims · 6 setups
RBPs beyond Drosha/DGCR8/Dicer are involved in regulating miRNA biogenesis and explain its spatio-temporal nature
-
Full-text index only
Application of machine learning in SNP discovery.
PMID 16398931 · PMC1955739 · BMC bioinformatics · 2006 · 8 claims · 6 setups
PolyBayes produces high false-positive SNP predictions even with stringent parameters
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Has reproduction · 44
Dynamic Gene Attention Focus (DyGAF): Enhancing Biomarker Identification Through Dual-Model Attention Networks.
PMID 40160891 · PMC11951896 · Bioinformatics and biology insights · 2025 · 6 claims · 5 setups
DyGAF, a dual-model attention neural network (independent Model A + dependent Model B), identifies and ranks genes by significance for COVID-19 biomarker discovery more effectively than differential expression analysis (DEA) and random forest (RF) feature selection
-
Has reproduction · 67
Adaptive learning embedding features to improve the predictive performance of SARS-CoV-2 phosphorylation sites.
PMID 37847658 · PMC10628388 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 3 setups
PSPred-ALE, a deep learning predictor using a self-adaptive learning embedding algorithm, automatically extracts contextual sequence features and identifies SARS-CoV-2 phosphorylation sites without feature engineering.
-
Has reproduction · 83
A temporal classifier predicts histopathology state and parses acute-chronic phasing in inflammatory bowel disease patients.
PMID 36694043 · PMC9873918 · Communications biology · 2023 · 8 claims · 7 setups
The DSS phenotype-by-time interaction defines parsimonious temporal (dynamic) expression and splicing signatures of acute and chronic colitis distinct from time-specific differential expression.
-
Full-text index only
Integrated analysis of genetic and proteomic data identifies biomarkers associated with adverse events following smallpox vaccination.
PMID 18923431 · PMC2692715 · Genes and immunity · 2009 · 7 claims · 6 setups
A two-stage strategy (Random Forest filtering followed by decision tree modeling) can integrate categorical genetic and continuous proteomic data to identify biomarkers of AE risk
-
Full-text index only
The impact of peptide abundance and dynamic range on stable-isotope-based quantitative proteomic analyses.
PMID 18798661 · PMC2746028 · Journal of proteome research · 2008 · 8 claims · 7 setups
Over half of confidently identified peptides in complex mixtures have S/N ratios below 10 on both FT-ICR and Orbitrap instruments
-
Has reproduction · 84
COXPRESdb v8: an animal gene coexpression database navigating from a global view to detailed investigations.
PMID 36350658 · PMC9825429 · Nucleic acids research · 2023 · 8 claims · 6 setups
COXPRESdb version 8 adds CoexMap (UMAP-based genome-scale coexpression visualization), KEGG pathway enrichment summaries, and CoexPub (literature-linking tool) as new analysis features.