Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
High fidelity of whole-genome amplified DNA on high-density single nucleotide polymorphism arrays.
PMID 18786630 · PMC2659594 · Genomics · 2008 · 8 claims · 7 setups
WGA product performs well on the Affymetrix 250K SNP array compared to genomic DNA, especially with the BRLMM calling algorithm.
-
Full-text index only
Optimality driven nearest centroid classification from genomic data.
PMID 17912341 · PMC1991588 · PloS one · 2007 · 7 claims · 5 setups
A theoretical result determines the subset of features of a given size that minimizes the misclassification rate for a nearest-centroid (LDA) classifier, based on equation (4).
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Full-text index only
Next-generation high-density self-assembling functional protein arrays.
PMID 18469824 · PMC3070491 · Nature methods · 2008 · 8 claims · 7 setups
A next-generation NAPPA method produces high-density protein microarrays displaying over 1500 unique proteins with >90% expression success
-
Full-text index only
An SVM-based system for predicting protein subnuclear localizations.
PMID 16336650 · PMC1325059 · BMC bioinformatics · 2005 · 7 claims · 3 setups
New kernels defined on k-peptide vectors mapped by BLOSUM62-based high-scored pair matrices (D1, D2, D3) improve SVM discrimination of protein subnuclear localization compared to conventional k-peptide encodings.
-
Full-text index only
Bayesian model accounting for within-class biological variability in Serial Analysis of Gene Expression (SAGE).
PMID 15339345 · PMC517707 · BMC bioinformatics · 2004 · 7 claims · 5 setups
A Bayesian mixture model is proposed to account for within-class biological variability in SAGE/Digital-Northern/MPSS tag counting data.
-
Full-text index only
Identification of HLA-DRPhebeta47 as the susceptibility marker of hypersensitivity to beryllium in individuals lacking the berylliosis-associated supratypic marker HLA-DPGlubeta69.
PMID 16098233 · PMC1198259 · Respiratory research · 2005 · 7 claims · 4 setups
HLA-DPGlu69 is the primary marker of Be-hypersensitivity, present at higher frequency in berylliosis patients than in Be-sensitized subjects and Be-exposed controls
-
Full-text index only
The use of edge-betweenness clustering to investigate biological function in protein interaction networks.
PMID 15740614 · PMC555937 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Edge-Betweenness clustering separates protein interaction graphs into subgraphs whose GO term distributions show significant correlations, revealing biologically meaningful functional modules.
-
Full-text index only
Human PAML browser: a database of positive selection on human genes using phylogenetic methods.
PMID 17962310 · PMC2238824 · Nucleic acids research · 2008 · 8 claims · 5 setups
The Human PAML Browser is a web-accessible database of codeml-based positive selection test results for 13,721 human genes with orthologs in UCSC multispecies alignments.
-
Full-text index only
Identification and characterization of HLA-A*0301 epitopes in HIV-1 gag proteins using a novel approach.
PMID 19903485 · PMC2836169 · Journal of immunological methods · 2010 · 7 claims · 7 setups
PS mutations V7I and I34L (p17) and K403R (p7) in HIV-1 gag significantly correlate with HLA-A*0301
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
ARED Organism: expansion of ARED reveals AU-rich element cluster variations between human and mouse.
PMID 17984078 · PMC2238997 · Nucleic acids research · 2008 · 6 claims · 4 setups
ARED Organism and ARED-Integrated are new/updated public databases cataloguing ARE-containing mRNAs/genes in human, mouse and rat
-
Full-text index only
Genomic transcriptional profiling identifies a candidate blood biomarker signature for the diagnosis of septicemic melioidosis.
PMID 19903332 · PMC3091321 · Genome biology · 2009 · 6 claims · 5 setups
A candidate 37-transcript diagnostic signature distinguishes septicemic melioidosis from sepsis caused by other organisms with 100% accuracy in the training set and 78%/80% accuracy in two independent validation sets
-
Has reproduction · 60
Integrating herbarium specimen observations into global phenology data systems.
PMID 30937223 · PMC6426164 · Applications in plant sciences · 2019 · 7 claims · 5 setups
A new PPO release adds terms and properties to relate observations of parts of plants to whole plants, enabling integration of herbarium phenology data with field observation data.
-
Full-text index only
Satellog: a database for the identification and prioritization of satellite repeats in disease association studies.
PMID 15949044 · PMC1181805 · BMC bioinformatics · 2005 · 7 claims · 6 setups
Satellog is a database cataloging all pure 1-16 unit satellite repeats in the human genome with supplementary polymorphism, gene-location, and expression data for prioritizing repeats in disease-association studies.
-
Full-text index only
Genomic divergences among cattle, dog and human estimated from large-scale alignments of genomic sequences.
PMID 16759380 · PMC1525190 · BMC genomics · 2006 · 8 claims · 6 setups
Overall pairwise genomic divergences among cattle, dog and human are relatively constant (0.32–0.37 change/site)
-
Full-text index only
Bias of selection on human copy-number variants.
PMID 16482228 · PMC1366494 · PLoS genetics · 2006 · 8 claims · 8 setups
Human CNVs are significantly overrepresented near telomeres and centromeres and enriched in simple tandem repeats relative to the genome as a whole
-
Full-text index only
A non-parametric meta-analysis approach for combining independent microarray datasets: application using two microarray datasets pertaining to chronic allograft nephropathy.
PMID 18302764 · PMC2276496 · BMC genomics · 2008 · 8 claims · 6 setups
A novel non-parametric meta-analysis approach for combining independent microarray datasets is presented, requiring no distributional assumptions and being logically intuitive.
-
Has reproduction · 94
Deep learning from phylogenies to uncover the epidemiological dynamics of outbreaks.
PMID 35794110 · PMC9258765 · Nature communications · 2022 · 8 claims · 5 setups
Deep learning (FFNN-SS and CNN-CBLV) enables accurate and fast likelihood-free estimation of epidemiological parameters and model selection from phylogenies
-
Full-text index only
Protein length in eukaryotic and prokaryotic proteomes.
PMID 15951512 · PMC1150220 · Nucleic acids research · 2005 · 7 claims · 5 setups
Eukaryotic proteins are significantly longer than prokaryotic proteins across virtually all functional categories and the majority of protein families