Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Constructing support vector machine ensembles for cancer classification based on proteomic profiling.
PMID 16689692 · PMC5173238 · Genomics, proteomics & bioinformatics · 2005 · 7 claims · 4 setups
CSVME, built by selecting a subset of base SVMs via SVM-RFE ranking and fusing them with a trained upper-layer SVM, achieves better classification performance than an ensemble of all base SVMs.
-
Full-text index only
BRCA1 and BRCA2 mutation predictions using the BOADICEA and BRCAPRO models and penetrance estimation in high-risk French-Canadian families.
PMID 16417652 · PMC1413985 · Breast cancer research : BCR · 2006 · 8 claims · 7 setups
BOADICEA predicts accurately the number of BRCA1 and BRCA2 mutations across family groups and discriminates well between carriers and noncarriers
-
Full-text index only
A common missense variant in BRCA2 predisposes to early onset breast cancer.
PMID 16280055 · PMC1410744 · Breast cancer research : BCR · 2005 · 7 claims · 4 setups
BRCA2 C5972T homozygosity (TT genotype) is rare but confers a roughly five-fold increased risk of breast cancer.
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Full-text index only
Phylogenomic approaches to common problems encountered in the analysis of low copy repeats: the sulfotransferase 1A gene family example.
PMID 15752422 · PMC555591 · BMC evolutionary biology · 2005 · 8 claims · 8 setups
A previously unidentified fourth human SULT1A gene (SULT1A4) exists on chromosome 16 and is transcriptionally active
-
Full-text index only
Computational verification of protein-protein interactions by orthologous co-expression.
PMID 15740634 · PMC555590 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Co-expression of orthologous protein pairs across multiple species can verify/predict S. cerevisiae PPIs with better performance than S. cerevisiae co-expression alone.
-
Full-text index only
The use of edge-betweenness clustering to investigate biological function in protein interaction networks.
PMID 15740614 · PMC555937 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Edge-Betweenness clustering separates protein interaction graphs into subgraphs whose GO term distributions show significant correlations, revealing biologically meaningful functional modules.
-
Full-text index only
Evaluation of genome-wide chromatin library of Stat5 binding sites in human breast cancer.
PMID 15686596 · PMC549029 · Molecular cancer · 2005 · 8 claims · 5 setups
A chromatin library coupled with experimental validation can productively identify novel in vivo Stat5 chromatin binding sites in cancer, including abnormal regulatory sites in tumor-specific neochromatin.
-
Full-text index only
A high throughput method for genome-wide analysis of retroviral integration.
PMID 17028098 · PMC1636494 · Nucleic acids research · 2006 · 8 claims · 8 setups
VITA uses MmeI to cleave DNA at a fixed distance from its recognition site, generating 21-22 bp genomic tags that serve as signatures of lentiviral integration sites.
-
Full-text index only
Evolution of variants of yeast site-specific recombinase Flp that utilize native genomic sequences as recombination target sites.
PMID 17003057 · PMC1635253 · Nucleic acids research · 2006 · 8 claims · 8 setups
Stepwise directed evolution using chimeric FLRT (FRT/genomic hybrid) intermediate sites can generate Flp variants capable of recombining native genomic FRT-like sequences from the human IL10 gene (FL-IL10A, FL-IL10B).
-
Full-text index only
In silico and in vivo splicing analysis of MLH1 and MSH2 missense mutations shows exon- and tissue-specific effects.
PMID 16995940 · PMC1590028 · BMC genomics · 2006 · 8 claims · 6 setups
In silico ESE-prediction algorithms (ESEfinder, RescueESE, PESX) do not reliably predict actual in vivo splicing behavior of missense mutations
-
Full-text index only
Automatic annotation of eukaryotic genes, pseudogenes and promoters.
PMID 16925832 · PMC1810547 · Genome biology · 2006 · 8 claims · 6 setups
Fgenesh++ gene prediction pipeline identifies 91% of coding nucleotides with 90% specificity
-
Full-text index only
MACSIMS: multiple alignment of complete sequences information management system.
PMID 16792820 · PMC1539025 · BMC bioinformatics · 2006 · 8 claims · 5 setups
MACSIMS is a multiple alignment-based information management system combining knowledge-based database mining with ab initio sequence predictions
-
Full-text index only
Large-scale and high-confidence proteomic analysis of human seminal plasma.
PMID 16709260 · PMC1779515 · Genome biology · 2006 · 8 claims · 6 setups
923 proteins were identified with high confidence in seminal plasma from a single individual, combining results from three ejaculate samples
-
Full-text index only
A clustering property of highly-degenerate transcription factor binding sites in the mammalian genome.
PMID 16670430 · PMC1456330 · Nucleic acids research · 2006 · 8 claims · 7 setups
Highly-degenerate RE1 sites are significantly enriched in promoters of validated and putative REST target genes compared to control promoters
-
Full-text index only
CARAT: a novel method for allelic detection of DNA copy number changes using high density oligonucleotide arrays.
PMID 16504045 · PMC1402331 · BMC bioinformatics · 2006 · 8 claims · 5 setups
CARAT is a novel algorithm that uses SNP probe intensity and genotype-based allelic dosage response in a regression framework to estimate allele-specific copy number genome-wide.
-
Full-text index only
Comprehensive genome analysis of 203 genomes provides structural genomics with new insights into protein family space.
PMID 16481312 · PMC1373602 · Nucleic acids research · 2006 · 8 claims · 7 setups
The number of protein families continues to expand steadily as more genomes are sequenced, showing no sign of saturation.
-
Full-text index only
A novel approach for rapid screening of mitochondrial D310 polymorphism.
PMID 16433919 · PMC1388229 · BMC cancer · 2006 · 7 claims · 4 setups
A single-step BsaXI RFLP assay can rapidly determine 7-C carriers at the D310 first polyC stretch of mtDNA without sequencing
-
Full-text index only
Ab initio identification of human microRNAs based on structure motifs.
PMID 18088431 · PMC2238772 · BMC bioinformatics · 2007 · 8 claims · 7 setups
MiRPred predicts miRNA precursors ab initio using only predicted secondary structure motifs, ignoring nucleotide sequence