Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
MALDI profiling of human lung cancer subtypes.
PMID 19890392 · PMC2767501 · PloS one · 2009 · 8 claims · 8 setups
PIMAC/MALDI-TOF peptide profiles combined with classification models can distinguish normal lung from tumor and differentiate NSCLC histological subtypes
-
Full-text index only
Predicting phenotype and emerging strains among Chlamydia trachomatis infections.
PMID 19788805 · PMC2819883 · Emerging infectious diseases · 2009 · 8 claims · 7 setups
A 7-locus MLST scheme selected from conserved housekeeping genes shared across 4 Chlamydiaceae species (7 genomes) can genotype diverse C. trachomatis reference and clinical isolates.
-
Full-text index only
Proteome analysis enables separate clustering of normal breast, benign breast and breast cancer tissues.
PMID 12865921 · PMC2394238 · British journal of cancer · 2003 · 6 claims · 3 setups
Hierarchical cluster analysis of 2-DE proteome data can distinguish normal breast, benign breast, and breast cancer tissues based on protein expression profiles
-
Has reproduction · 81
Whole-Genome Sequence of Cervid atadenovirus A from the Initial Cases of an Adenovirus Hemorrhagic Disease Epizootic of Black-Tailed Deer in Canada.
PMID 36129291 · PMC9584332 · Microbiology resource announcements · 2022 · 7 claims · 4 setups
A complete 30,616-nucleotide genome of Cervid atadenovirus A was determined from lung tissue of black-tailed deer that died of AHD in British Columbia in 2020
-
Full-text index only
Mining novel biomarkers for prognosis of gastric cancer with serum proteomics.
PMID 19740432 · PMC2753349 · Journal of experimental & clinical cancer research : CR · 2009 · 7 claims · 4 setups
A 5-peak prognosis pattern (4474, 4542, 6443/6643, 4988, 6685 Da) predicts poor vs good prognosis in GC with higher sensitivity/specificity than CEA and TNM stage
-
Has reproduction · 83
Accurate prediction of metagenome-assembled genome completeness by MAGISTA, a random forest model built on alignment-free intra-bin statistics.
PMID 35248155 · PMC8898458 · Environmental microbiome · 2022 · 7 claims · 7 setups
MAGISTA, a random forest model built on alignment-free intra-bin distance-distribution statistics, can estimate MAG completeness and purity without relying on reference marker genes.
-
Has reproduction · 48
Bioinformatics Analysis of the Characteristics and Correlation of m6A Methylation in Breast Cancer Progression.
PMID 35655723 · PMC9148239 · Contrast media & molecular imaging · 2022 · 8 claims · 8 setups
Breast cancer samples can be divided into 4 m6A subtypes (quiescent, m6A methylation, protein-binding, mixed) based on expression of m6A-related genes, with consistent proportions across datasets
-
Full-text index only
ProMiR II: a web server for the probabilistic prediction of clustered, nonclustered, conserved and nonconserved microRNAs.
PMID 16845048 · PMC1538778 · Nucleic acids research · 2006 · 6 claims · 4 setups
ProMiR II improves on the original ProMiR by integrating free energy, G/C ratio, conservation score and entropy for more controllable miRNA prediction
-
Full-text index only
KEGG for linking genomes to life and the environment.
PMID 18077471 · PMC2238879 · Nucleic acids research · 2008 · 8 claims · 4 setups
KEGG provides a reference knowledge base for linking genomes to life via PATHWAY mapping and to the environment via BRITE mapping.
-
Full-text index only
Logical Analysis of Data (LAD) model for the early diagnosis of acute ischemic stroke.
PMID 18616825 · PMC2492849 · BMC medical informatics and decision making · 2008 · 7 claims · 5 setups
An LAD classification model built from a support-set of 3 peptide peaks can distinguish stroke patients from controls with 75% accuracy on an independent validation set
-
Full-text index only
Predicting deleterious nsSNPs: an analysis of sequence and structural attributes.
PMID 16630345 · PMC1489951 · BMC bioinformatics · 2006 · 8 claims · 7 setups
Sequence conservation (PSIC score difference) at the nsSNP position is the single most useful attribute for predicting deleterious vs neutral status.
-
Has reproduction · 60
TRAPID 2.0: a web application for taxonomic and functional analysis of de novo transcriptomes.
PMID 34197621 · PMC8464036 · Nucleic acids research · 2021 · 8 claims · 8 setups
TRAPID 2.0 is a web application performing global characterization of de novo transcriptomes via structural, functional, and taxonomic annotation in an initial processing phase, followed by an exploratory phase of downstream analyses.
-
Full-text index only
In silico analysis of missense substitutions using sequence-alignment based methods.
PMID 18951440 · PMC3431198 · Human mutation · 2008 · 8 claims · 7 setups
Carefully validated PMSA-based computational algorithms can achieve predictive values of ~75-95% for classifying missense substitutions as pathogenic or neutral.
-
Full-text index only
Comprehensive splice-site analysis using comparative genomics.
PMID 16914448 · PMC1557818 · Nucleic acids research · 2006 · 8 claims · 6 setups
Over half a million splice sites were collected from five species (H. sapiens, M. musculus, D. melanogaster, C. elegans, A. thaliana) and classified into four main subtypes: U2-type GT-AG and GC-AG, and U12-type GT-AG and AT-AC.
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Full-text index only
Boosting accuracy of automated classification of fluorescence microscope images for location proteomics.
PMID 15207009 · PMC449699 · BMC bioinformatics · 2004 · 8 claims · 8 setups
New classifiers (SVMs, ensembles) and new wavelet-derived (Gabor, Daubechies) features improve recognition of protein subcellular location patterns over the previous neural network approach
-
Full-text index only
CanPredict: a computational tool for predicting cancer-associated missense mutations.
PMID 17537827 · PMC1933186 · Nucleic acids research · 2007 · 8 claims · 7 setups
CanPredict is a web application providing public access to a random forest classifier that combines SIFT, LogR.E-value, and GOSS scores to predict whether a missense mutation is cancer-associated
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Has reproduction · 66
Integrative bioinformatics and artificial intelligence analyses of transcriptomics data identified genes associated with major depressive disorders including NRG1.
PMID 37583471 · PMC10423927 · Neurobiology of stress · 2023 · 7 claims · 5 setups
Differentially expressed genes in MDD patients are enriched in immune response, inflammatory response, neurodegeneration, and cerebellar atrophy pathways.
-
Full-text index only
Analysis of the glutathione S-transferase (GST) gene family.
PMID 15607001 · PMC3500200 · Human genomics · 2004 · 8 claims · 4 setups
The complete human GST gene family comprises 16 genes in six subfamilies: alpha (GSTA), mu (GSTM), omega (GSTO), pi (GSTP), theta (GSTT) and zeta (GSTZ).