Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Boosting accuracy of automated classification of fluorescence microscope images for location proteomics.
PMID 15207009 · PMC449699 · BMC bioinformatics · 2004 · 8 claims · 8 setups
New classifiers (SVMs, ensembles) and new wavelet-derived (Gabor, Daubechies) features improve recognition of protein subcellular location patterns over the previous neural network approach
-
Full-text index only
Selecting additional tag SNPs for tolerating missing data in genotyping.
PMID 16259642 · PMC1316880 · BMC bioinformatics · 2005 · 7 claims · 6 setups
There exists a subset of SNPs (robust tag SNPs) that can distinguish all distinct haplotypes even when up to m SNPs are missing
-
Full-text index only
Computational tradeoffs in multiplex PCR assay design for SNP genotyping.
PMID 16042802 · PMC1190169 · BMC genomics · 2005 · 7 claims · 6 setups
Achieving high-multiplexing/high-coverage multiplex PCR designs is subject to a computational phase transition as the SNP-pair compatibility probability crosses a critical threshold
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
Computational verification of protein-protein interactions by orthologous co-expression.
PMID 15740634 · PMC555590 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Co-expression of orthologous protein pairs across multiple species can verify/predict S. cerevisiae PPIs with better performance than S. cerevisiae co-expression alone.
-
Full-text index only
The use of edge-betweenness clustering to investigate biological function in protein interaction networks.
PMID 15740614 · PMC555937 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Edge-Betweenness clustering separates protein interaction graphs into subgraphs whose GO term distributions show significant correlations, revealing biologically meaningful functional modules.
-
Full-text index only
Towards alignment independent quantitative assessment of homology detection.
PMID 17205117 · PMC1762415 · PloS one · 2006 · 8 claims · 6 setups
The Fhom Estimator uses the prevalence of a conserved protein feature (X) in two protein sets to estimate the fraction of true homologs among paired proteins, independent of alignment quality.
-
Full-text index only
Selection of target sites for mobile DNA integration in the human genome.
PMID 17166054 · PMC1664696 · PLoS computational biology · 2006 · 8 claims · 8 setups
A comprehensive bioinformatic method was developed to annotate every base pair in the human genome for its likelihood of hosting integration by each of seven mobile DNA elements, using >200 genomic feature variables.
-
Full-text index only
Biomarkers that discriminate multiple myeloma patients with or without skeletal involvement detected using SELDI-TOF mass spectrometry and statistical and machine learning tools.
PMID 17124346 · PMC3862287 · Disease markers · 2006 · 8 claims · 5 setups
SELDI-TOF MS serum profiling can discriminate MM patients with vs without skeletal (bone lesion) involvement using peak biomarkers
-
Full-text index only
Antibody binding loop insertions as diversity elements.
PMID 17023486 · PMC1635297 · Nucleic acids research · 2006 · 7 claims · 8 setups
A lysozyme-binding VHH CDR3 loop can be grafted into two surface-exposed loops of superfolder GFP, conferring lysozyme-binding activity while the protein remains fluorescent.
-
Full-text index only
GENCODE: producing a reference annotation for ENCODE.
PMID 16925838 · PMC1810553 · Genome biology · 2006 · 8 claims · 8 setups
GENCODE annotation combines initial manual annotation by HAVANA, experimental validation, and refinement based on results to identify protein-coding genes in ENCODE regions
-
Full-text index only
Willing to do the math: an interview with David Botstein. Interview by Jane Gitschier.
PMID 16733551 · PMC1464829 · PLoS genetics · 2006 · 8 claims · 5 setups
Highly polymorphic, multiallelic DNA markers spaced across the genome could be used to build a complete human genetic linkage map
-
Full-text index only
Genetic requirement for pneumococcal ear infection.
PMID 18670623 · PMC2593789 · PloS one · 2007 · 7 claims · 8 setups
STM screening of 5,280 S. pneumoniae ST556 mutants in a chinchilla middle ear infection model identified 169 genes required for ear infection
-
Full-text index only
Imputation of missing genotypes: an empirical evaluation of IMPUTE.
PMID 19077279 · PMC2636842 · BMC genetics · 2008 · 8 claims · 7 setups
IMPUTE achieves 97% median genotype imputation accuracy in Caucasian (NNC) subjects when <10% of SNPs are untyped
-
Full-text index only
Accuracy of predicting the genetic risk of disease using a genome-wide approach.
PMID 18852893 · PMC2561058 · PloS one · 2008 · 8 claims · 4 setups
Deterministic equations can predict the accuracy (r_gĝ) of genome-wide genetic risk/value prediction for continuous, dichotomous, and case-control study designs.
-
Full-text index only
The impact of peptide abundance and dynamic range on stable-isotope-based quantitative proteomic analyses.
PMID 18798661 · PMC2746028 · Journal of proteome research · 2008 · 8 claims · 7 setups
Over half of confidently identified peptides in complex mixtures have S/N ratios below 10 on both FT-ICR and Orbitrap instruments
-
Full-text index only
A simple and robust method for connecting small-molecule drugs using gene-expression signatures.
PMID 18518950 · PMC2464610 · BMC bioinformatics · 2008 · 8 claims · 4 setups
A new method for building reference gene-expression profiles and scoring/testing connections improves on the original Connectivity Map by enabling statistical significance testing of connections.
-
Full-text index only
Evaluation of two methods for computational HLA haplotypes inference using a real dataset.
PMID 18230173 · PMC2268655 · BMC bioinformatics · 2008 · 8 claims · 5 setups
PHASE v2.1.1 had the best overall performance in both haplotype construction and frequency calculation compared to Arlequin V3.0
-
Full-text index only
Widespread dysregulation of MiRNAs by MYCN amplification and chromosomal imbalances in neuroblastoma: association of miRNA expression with survival.
PMID 19924232 · PMC2773120 · PloS one · 2009 · 8 claims · 5 setups
37 miRNAs are significantly differentially expressed between MYCN-amplified (MNA) and non-MNA neuroblastoma tumors, suggesting direct or indirect regulation by MYCN