Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
COMUS: Clinician-Oriented locus-specific MUtation detection and deposition System.
PMID 19958500 · PMC2788389 · BMC genomics · 2009 · 8 claims · 6 setups
COMUS is a bioinformatics system for detecting and depositing new mutations from patient DNA with a clinician-friendly interface
-
Full-text index only
A compatible exon-exon junction database for the identification of exon skipping events using tandem mass spectrum data.
PMID 19087293 · PMC2636810 · BMC bioinformatics · 2008 · 6 claims · 6 setups
A theoretical exon-exon junction protein database accounting for all in-phase (frame-preserving) exon combinations can be built from the Ensembl Core Database using Perl/Bioperl/MySQL/Ensembl API.
-
Full-text index only
X:Map: annotation and visualization of genome structure for Affymetrix exon array analysis.
PMID 17932061 · PMC2238884 · Nucleic acids research · 2008 · 7 claims · 4 setups
X:Map is a genome annotation database that maps every Affymetrix exon array probeset to Ensembl genome features (genes, ESTs, GenScan predictions) and supports both high-throughput and gene-centric analysis.
-
Full-text index only
The HIV positive selection mutation database.
PMID 17108357 · PMC1669717 · Nucleic acids research · 2007 · 8 claims · 5 setups
The database provides codon-level Ka/Ks selection pressure maps for HIV protease and the first 381 codons of RT, built from a novel ~50,000-sample clinical dataset.
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Full-text index only
Human SNPs resulting in premature stop codons and protein truncation.
PMID 16595072 · PMC3500177 · Human genomics · 2006 · 8 claims · 6 setups
Genome-wide screening of dbSNP identified 28 validated X-SNPs from 28 genes with known minor allele frequencies.
-
Full-text index only
HaploSNPer: a web-based allele and SNP detection tool.
PMID 18307806 · PMC2288614 · BMC genetics · 2008 · 6 claims · 2 setups
HaploSNPer is a web-based tool integrating BLASTN, CAP3/PHRAP, and QualitySNP into a single pipeline for allele and SNP detection from diploid and polyploid species
-
Full-text index only
MILANO--custom annotation of microarray results using automatic literature searches.
PMID 15661078 · PMC547913 · BMC bioinformatics · 2005 · 7 claims · 4 setups
MILANO annotates microarray gene lists by counting literature co-occurrences of each gene with user-defined secondary terms
-
Full-text index only
Prediction of missed cleavage sites in tryptic peptides aids protein identification in proteomics.
PMID 17203985 · PMC2664920 · Journal of proteome research · 2007 · 8 claims · 4 setups
An information-theoretic log-likelihood scoring method can predict experimentally observed missed cleavage sites from amino acid sequence alone with up to 90% accuracy.
-
Full-text index only
QTL MatchMaker: a multi-species quantitative trait loci (QTL) database and query system for annotation of genes and QTL.
PMID 16381937 · PMC1347390 · Nucleic acids research · 2006 · 8 claims · 5 setups
QTL MatchMaker integrates QTL information with physical, genetic and cytogenetic maps across human, mouse and rat genomes
-
Full-text index only
The relationship of potential G-quadruplex sequences in cis-upstream regions of the human genome to SP1-binding elements.
PMID 18353860 · PMC2377421 · Nucleic acids research · 2008 · 7 claims · 1 setups
A large number of upstream PQSSs incorporate the SP1-binding element, establishing a clear link between PQSS occurrence and SP1 elements
-
Full-text index only
Bases and spaces: resources on the web for accessing the draft human genome.
PMID 11178254 · PMC138875 · Genome biology · 2000 · 8 claims · 8 setups
By combining currently available genomic databases and mapping resources (GenBank/Entrez, UniGene, RH maps, BAC fingerprint maps, Ensembl, NIX), it is possible to devise strategies that fully exploit the fragmentary draft human genome sequence.
-
Full-text index only
Satellog: a database for the identification and prioritization of satellite repeats in disease association studies.
PMID 15949044 · PMC1181805 · BMC bioinformatics · 2005 · 7 claims · 6 setups
Satellog is a database cataloging all pure 1-16 unit satellite repeats in the human genome with supplementary polymorphism, gene-location, and expression data for prioritizing repeats in disease-association studies.
-
Full-text index only
How to find soluble proteins: a comprehensive analysis of alpha/beta hydrolases for recombinant expression in E. coli.
PMID 15804363 · PMC1079826 · BMC genomics · 2005 · 7 claims · 7 setups
Predicted solubility in E. coli (via CV-CV') depends on hydrolase size, phylogenetic origin, homologous family, and superfamily
-
Full-text index only
SECIS elements in the coding regions of selenoprotein transcripts are functional in higher eukaryotes.
PMID 17169995 · PMC1802603 · Nucleic acids research · 2007 · 8 claims · 5 setups
SECIS elements located within coding regions of selenoprotein mRNAs support functional Sec insertion in mammalian cells
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
Comparative sequence analysis of leucine-rich repeats (LRRs) within vertebrate toll-like receptors.
PMID 17517123 · PMC1899181 · BMC genomics · 2007 · 8 claims · 4 setups
A new method combining known LRR structures, multiple sequence alignment, and secondary structure prediction identifies and aligns LRRs in TLRs more accurately than PFAM/InterPro/SMART
-
Full-text index only
Large-scale discovery of insertion hotspots and preferential integration sites of human transposed elements.
PMID 20008508 · PMC2836564 · Nucleic acids research · 2010 · 8 claims · 6 setups
Most TEs insert within specific 'hotspots' along the targeted TE rather than uniformly.
-
Full-text index only
MEROPS: the peptidase database.
PMID 19892822 · PMC2808883 · Nucleic acids research · 2010 · 8 claims · 5 setups
MEROPS is a manually curated hierarchical classification of peptidases and protein inhibitors organized into protein species, families, and clans based on sequence and structural homology.