Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A non-parametric meta-analysis approach for combining independent microarray datasets: application using two microarray datasets pertaining to chronic allograft nephropathy.
PMID 18302764 · PMC2276496 · BMC genomics · 2008 · 8 claims · 6 setups
A novel non-parametric meta-analysis approach for combining independent microarray datasets is presented, requiring no distributional assumptions and being logically intuitive.
-
Full-text index only
SePaCS--a web-based application for classification of seroreactivity profiles.
PMID 17478503 · PMC1933220 · Nucleic acids research · 2007 · 8 claims · 4 setups
SePaCS is a freely available web-based tool that trains and applies multiple classification methods (4 Naive Bayes variants, SVM with RBF kernel, LDA, DLDA) to seroreactivity profiles and outputs results as a summary table plus a detailed PDF report
-
Has reproduction · 90
A Decentralized Kidney Transplant Biopsy Classifier for Transplant Rejection Developed Using Genes of the Banff-Human Organ Transplant Panel.
PMID 35619722 · PMC9128066 · Frontiers in immunology · 2022 · 6 claims · 6 setups
A random forest model trained solely on B-HOT panel genes (B-HOT Model) accurately classifies kidney transplant biopsies as NR, ABMR, or TCMR.
-
Has reproduction · 74
Evaluation of classification and forecasting methods on time series gene expression data.
PMID 33156855 · PMC7647064 · PloS one · 2020 · 6 claims · 3 setups
Deep learning based methods generally outperform traditional approaches for time series gene expression classification
-
Full-text index only
PA-GOSUB: a searchable database of model organism protein sequences with their predicted Gene Ontology molecular function and subcellular localization.
PMID 15608166 · PMC540074 · Nucleic acids research · 2005 · 7 claims · 4 setups
PA-GOSUB significantly extends the coverage of GO molecular function and subcellular localization annotations for 10 model organism proteomes compared with existing databases (GOA, Swiss-Prot).
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Full-text index only
An SVM-based system for predicting protein subnuclear localizations.
PMID 16336650 · PMC1325059 · BMC bioinformatics · 2005 · 7 claims · 3 setups
New kernels defined on k-peptide vectors mapped by BLOSUM62-based high-scored pair matrices (D1, D2, D3) improve SVM discrimination of protein subnuclear localization compared to conventional k-peptide encodings.
-
Has reproduction · 85
Predicting the pathogenicity of missense variants using features derived from AlphaFold2.
PMID 37084271 · PMC10203375 · Bioinformatics (Oxford, England) · 2023 · 6 claims · 8 setups
AlphaFold2-derived structural features (solvent accessibility, amino acid network features, physicochemical environment, pLDDT) can be used to train a random forest classifier (AlphScore) that distinguishes proxy-benign from proxy-pathogenic missense variants.
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
Prodepth: predict residue depth by support vector regression approach from protein sequences only.
PMID 19759917 · PMC2742725 · PloS one · 2009 · 8 claims · 8 setups
Residue depth can be reliably predicted solely from protein primary sequence using support vector regression on sequence-derived features.
-
Full-text index only
In Silico screening for functional candidates amongst hypothetical proteins.
PMID 19754976 · PMC2758874 · BMC bioinformatics · 2009 · 7 claims · 6 setups
An in silico selection strategy combining subcellular targeting-signal prediction with protein domain identification can enrich for true functional proteins among hypothetical proteins
-
Full-text index only
miRGator: an integrated system for functional annotation of microRNAs.
PMID 17942429 · PMC2238850 · Nucleic acids research · 2008 · 8 claims · 8 setups
miRGator integrates target prediction, functional enrichment analysis (GO/pathway/disease), and expression data (miRNA/mRNA/protein) into one system for functional annotation of miRNAs
-
Full-text index only
Genomic variation in myeloma: design, content, and initial application of the Bank On A Cure SNP Panel to detect associations with progression-free survival.
PMID 18778477 · PMC2553089 · BMC medicine · 2008 · 7 claims · 7 setups
A custom BOAC SNP panel of 3404 SNPs in 983 genes was developed using the Affymetrix GeneChip Targeted Genotyping Platform, focused on non-synonymous coding SNPs and regulatory-region SNPs in candidate genes.
-
Full-text index only
A genome-wide survey of Major Histocompatibility Complex (MHC) genes and their paralogues in zebrafish.
PMID 16271140 · PMC1309616 · BMC genomics · 2005 · 8 claims · 4 setups
149 putative MHC gene loci and their paralogues were identified in the zebrafish genome using sequence similarity searches against the Zv4 draft assembly.
-
Full-text index only
Intricate targeting of immunoglobulin somatic hypermutation maximizes the efficiency of affinity maturation.
PMID 15867095 · PMC2213188 · The Journal of experimental medicine · 2005 · 7 claims · 6 setups
IgVH genes have evolved precise placement of coding-strand Cs so that AID-induced C-to-T mutations are predominantly silent, especially in the CDRs.
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
Somatic mutations in mitochondria: the chicken or the egg?
PMID 16207343 · PMC1257449 · Arthritis research & therapy · 2005 · 6 claims · 6 setups
Patients with RA have a higher incidence of somatic mtDNA mutations (in MT-ND1 transcripts) in synoviocytes and synovial tissue compared with OA patients
-
Full-text index only
GeneMark: web software for gene finding in prokaryotes, eukaryotes and viruses.
PMID 15980510 · PMC1160247 · Nucleic acids research · 2005 · 8 claims · 2 setups
The GeneMark website provides web interfaces to the GeneMark family of ab initio gene-finding programs for prokaryotic, eukaryotic and viral genomic sequences
-
Has reproduction · 49
oPOSSUM-3: advanced analysis of regulatory motif over-representation across genes or ChIP-Seq datasets.
PMID 22973536 · PMC3429929 · G3 (Bethesda, Md.) · 2012 · 8 claims · 6 setups
oPOSSUM-3 is a web-accessible system that identifies over-represented TFBS and TFBS families in DNA sequences of co-expressed genes or in sequences from high-throughput methods such as ChIP-Seq.
-
Full-text index only
LOCATE: a mammalian protein subcellular localization database.
PMID 17986452 · PMC2238969 · Nucleic acids research · 2008 · 8 claims · 6 setups
LOCATE is a curated, web-accessible database housing membrane organization and subcellular localization data for mouse and human proteins.