Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Proteomic-based identification of haptoglobin-1 precursor as a novel circulating biomarker of ovarian cancer.
PMID 15199385 · PMC2364749 · British journal of cancer · 2004 · 7 claims · 7 setups
Six serum protein spots (~40 kDa, pI 5.9–6.6) significantly overexpressed in grade 1, 2 and 3 ovarian cancer patients were identified by MALDI-TOFMS and n-ESIQ(q)TOFMS as isoforms of haptoglobin-1 precursor (HAP1)
-
Full-text index only
Skittle: a 2-dimensional genome visualization tool.
PMID 20042093 · PMC2817707 · BMC bioinformatics · 2009 · 7 claims · 6 setups
Skittle is a 2D genome visualization tool combining a color-coded Nucleotide Display, a Repeat Map, a Repeat Overview, and an Alignment Cylinder to reveal genomic patterns at multiple scales
-
Full-text index only
TassDB: a database of alternative tandem splice sites.
PMID 17142241 · PMC1669710 · Nucleic acids research · 2007 · 7 claims · 3 setups
TassDB is a relational database storing GYNGYN donor and NAGNAG acceptor tandem splice sites across eight species
-
Full-text index only
MitoVariome: a variome database of human mitochondrial DNA.
PMID 19958475 · PMC2788364 · BMC genomics · 2009 · 8 claims · 5 setups
MitoVariome is a web-based, integrated variome database for human mitochondrial DNA that unifies sequence variation, haplogroup, and disease annotation information not jointly available in prior databases (MITOMAP, mtDB, Mitome, MitoRes).
-
Full-text index only
Prediction of missed cleavage sites in tryptic peptides aids protein identification in proteomics.
PMID 17203985 · PMC2664920 · Journal of proteome research · 2007 · 8 claims · 4 setups
An information-theoretic log-likelihood scoring method can predict experimentally observed missed cleavage sites from amino acid sequence alone with up to 90% accuracy.
-
Full-text index only
Satellog: a database for the identification and prioritization of satellite repeats in disease association studies.
PMID 15949044 · PMC1181805 · BMC bioinformatics · 2005 · 7 claims · 6 setups
Satellog is a database cataloging all pure 1-16 unit satellite repeats in the human genome with supplementary polymorphism, gene-location, and expression data for prioritizing repeats in disease-association studies.
-
Full-text index only
Short tandem repeat sequences in the Mycoplasma genitalium genome and their use in a multilocus genotyping system.
PMID 18664269 · PMC2515158 · BMC microbiology · 2008 · 8 claims · 7 setups
18 STR loci (1-5 base repeat units, copy number 4-26) were identified in the M. genitalium G37 genome via bioinformatics analysis.
-
Full-text index only
Implementation of a data repository-driven approach for targeted proteomics experiments by multiple reaction monitoring.
PMID 19121650 · PMC2706936 · Journal of proteomics · 2009 · 7 claims · 5 setups
A new MRM worksheet was implemented in The Global Proteome Machine database (GPMDB) that provides all information needed to design MRM transitions based solely on archived observations from previous experiments by other researchers.
-
Full-text index only
Sequence context affects the rate of short insertions and deletions in flies and primates.
PMID 18291026 · PMC2374710 · Genome biology · 2008 · 8 claims · 6 setups
The rate of insertion or deletion of specific lengths can vary by more than 100-fold depending on the surrounding sequence context
-
Full-text index only
AgBase: a functional genomics resource for agriculture.
PMID 16961921 · PMC1618847 · BMC genomics · 2006 · 8 claims · 7 setups
AgBase is a curated, web-accessible database providing structural and functional (GO) annotation for agricultural genomes and their pathogens.
-
Full-text index only
Comparative genomic profiling of Dutch clinical Bordetella pertussis isolates using DNA microarrays: identification of genes absent from epidemic strains.
PMID 18590534 · PMC2481270 · BMC genomics · 2008 · 8 claims · 6 setups
B. pertussis strains carrying the ptxP3 allele gradually replaced ptxP1 strains and became predominant from 1998, coinciding with the Dutch pertussis resurgence
-
Full-text index only
The most frequent short sequences in non-coding DNA.
PMID 19966278 · PMC2831315 · Nucleic acids research · 2010 · 8 claims · 2 setups
Short frequent sequences (9-14 bases) in non-coding DNA may play a role in maintaining chromosome structure and function
-
Full-text index only
Polymorphic segmental duplications at 8p23.1 challenge the determination of individual defensin gene repertoires and the assembly of a contiguous human reference sequence.
PMID 15588320 · PMC544879 · BMC genomics · 2004 · 8 claims · 8 setups
The hg16 automatic assembly of the 8p23.1 DEF locus contains misassemblies caused by segmental duplications and interindividual/intraindividual genetic variation
-
Full-text index only
Genomic sequence and analysis of a vaccinia virus isolate from a patient with a smallpox vaccine-related complication.
PMID 17062162 · PMC1635044 · Virology journal · 2006 · 8 claims · 6 setups
VACV-DUKE is most similar to VACV-ACAM2000 and VACV-CLONE3, confirming it as a Dryvax-derived isolate
-
Full-text index only
Novel gene and gene model detection using a whole genome open reading frame analysis in proteomics.
PMID 16646984 · PMC1557991 · Genome biology · 2006 · 8 claims · 4 setups
A six-frame genomic ORF translation used as an MS search database can detect novel peptides absent from standard protein databases, revealing incomplete genome annotation.
-
Full-text index only
Anopheles gambiae genome reannotation through synthesis of ab initio and comparative gene prediction algorithms.
PMID 16569258 · PMC1557760 · Genome biology · 2006 · 8 claims · 7 setups
An exon-gene-union (EGU) algorithm followed by an open-reading-frame-selection algorithm can synthesize ab initio (GENSCAN, GeneMark, SNAP) and comparative (Ensembl/Genewise) predictions into a single, more complete CDS set
-
Full-text index only
G2Cdb: the Genes to Cognition database.
PMID 18984621 · PMC2686544 · Nucleic acids research · 2009 · 7 claims · 7 setups
G2Cdb integrates experimentally validated synapse proteome datasets with mouse/human genomic annotation, phenotype, and human disease data in a gene-centric database.
-
Full-text index only
RAId_DbS: mass-spectrometry based peptide identification web server with knowledge integration.
PMID 18954448 · PMC2605478 · BMC genomics · 2008 · 7 claims · 4 setups
Constructed enhanced protein databases integrating annotated SAPs, PTMs, and disease associations for 17 organisms.
-
Full-text index only
Decision tree-driven tandem mass spectrometry for shotgun proteomics.
PMID 18931669 · PMC2597439 · Nature methods · 2008 · 8 claims · 5 setups
A decision tree (DT) algorithm that selects CAD or ETD per precursor based on z and m/z yields more peptide identifications than either CAD or ETD alone
-
Full-text index only
High accuracy mass spectrometry analysis as a tool to verify and improve gene annotation using Mycobacterium tuberculosis as an example.
PMID 18597682 · PMC2483986 · BMC genomics · 2008 · 8 claims · 5 setups
High-accuracy MS proteomics (LTQ-Orbitrap) can be used to verify and improve gene annotation by identifying peptides specific to one of two competing annotation datasets.