Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins.
PMID 15608248 · PMC539979 · Nucleic acids research · 2005 · 7 claims · 5 setups
RefSeq provides a curated, non-redundant, explicitly linked collection of genomic, transcript and protein sequences spanning prokaryotes, eukaryotes and viruses.
-
Full-text index only
A survey of integral alpha-helical membrane proteins.
PMID 19760129 · PMC2780624 · Journal of structural and functional genomics · 2009 · 8 claims · 8 setups
An automated annotation pipeline defines the integral membrane genome and family associations for 21,379 proteins from 34 genomes, most belonging to 598 Pfam-derived membrane protein families.
-
Full-text index only
Genome annotation errors in pathway databases due to semantic ambiguity in partial EC numbers.
PMID 16034025 · PMC1179732 · Nucleic acids research · 2005 · 7 claims · 4 setups
Partial EC numbers are semantically ambiguous, and databases that assign a gene to all reactions sharing the same partial EC number make a faulty inference, causing systematic misannotation.
-
Full-text index only
Investigation of the human tear film proteome using multiple proteomic approaches.
PMID 18334958 · PMC2268847 · Molecular vision · 2008 · 8 claims · 6 setups
Tear collection method (capillary vs. Schirmer strip) significantly impacts which proteins are detected in the tear film proteome.
-
Full-text index only
MEROPS: the peptidase database.
PMID 19892822 · PMC2808883 · Nucleic acids research · 2010 · 8 claims · 5 setups
MEROPS is a manually curated hierarchical classification of peptidases and protein inhibitors organized into protein species, families, and clans based on sequence and structural homology.
-
Full-text index only
Comprehensive splice-site analysis using comparative genomics.
PMID 16914448 · PMC1557818 · Nucleic acids research · 2006 · 8 claims · 6 setups
Over half a million splice sites were collected from five species (H. sapiens, M. musculus, D. melanogaster, C. elegans, A. thaliana) and classified into four main subtypes: U2-type GT-AG and GC-AG, and U12-type GT-AG and AT-AC.
-
Full-text index only
Putting proteins in one place.
PMID 12790144 · PMC1316915 · Environmental health perspectives · 2003 · 8 claims · 7 setups
Rapamycin inhibits TOR, causing the silencing protein Sir3 to detach from chromatin at stress-response genes, triggering a coordinated multigene stress response that halts cancer cell proliferation
-
Full-text index only
Polymorphism discovery and association analyses of the interferon genes in type 1 diabetes.
PMID 16504056 · PMC1402321 · BMC genetics · 2006 · 7 claims · 8 setups
No statistical evidence of a major association between T1D and any of the interferon or interferon-related genes tested (IFNA cluster, IFNB1, IFNW1, IFNG, ICSBP1)
-
Full-text index only
Genetic diversity of clinical isolates of Bacillus cereus using multilocus sequence typing.
PMID 18990211 · PMC2585095 · BMC microbiology · 2008 · 8 claims · 7 setups
The 55 clinical B. cereus isolates were phylogenetically diverse, comprising 38 sequence types (STs) distributed across two of three previously described clades.
-
Full-text index only
Extraction of human kinase mutations from literature, databases and genotyping studies.
PMID 19758464 · PMC2745582 · BMC bioinformatics · 2009 · 7 claims · 6 setups
A literature mining pipeline combining MutationFinder, false-positive filtering, and SVM-based classification can extract and disambiguate single-point mutation mentions from abstracts and full text
-
Full-text index only
Genomics and the prevention and control of common chronic diseases: emerging priorities for public health action.
PMID 15888216 · PMC1327699 · Preventing chronic disease · 2005 · 8 claims · 6 setups
Family history is the most consistent risk factor for almost all human diseases across the lifespan.
-
Full-text index only
Complete genome sequence and comparative analysis of the wild-type commensal Escherichia coli strain SE11 isolated from a healthy adult.
PMID 18931093 · PMC2608844 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2008 · 8 claims · 6 setups
The SE11 genome comprises a 4.8 Mb chromosome encoding 4679 protein-coding genes and six plasmids encoding 323 protein-coding genes
-
Full-text index only
SilkDB v2.0: a platform for silkworm (Bombyx mori ) genome biology.
PMID 19793867 · PMC2808975 · Nucleic acids research · 2010 · 8 claims · 8 setups
A new 8.5x-coverage silkworm genome assembly with N50 scaffold size of ~3.7 Mb over a 432 Mb genome represents a significant quality improvement over the prior draft.
-
Full-text index only
Automated recognition of retroviral sequences in genomic data--RetroTector.
PMID 17636050 · PMC1976444 · Nucleic acids research · 2007 · 8 claims · 8 setups
RetroTector uses 'fragment threading' (detection of chains of conserved retroviral motifs satisfying distance constraints) combined with LTR detection and protein reconstruction to identify ERVs in genomic sequences
-
Full-text index only
Identification and characterization of insect-specific proteins by genome data analysis.
PMID 17407609 · PMC1852559 · BMC genomics · 2007 · 8 claims · 7 setups
Comparative genome analysis across five holometabolous insects and three non-insect eukaryotes (opisthokonts) identifies 154 insect-specific orthologous groups (refined to 51 proteins) and 466 eukaryote/opisthokont-core orthologous groups
-
Has reproduction · 100
Viral Diagnostics in Plants Using Next Generation Sequencing: Computational Analysis in Practice.
PMID 29123534 · PMC5662881 · Frontiers in plant science · 2017 · 8 claims · 8 setups
NGS/RNA-seq enables unbiased, hypothesis-free detection of multiple known and emergent plant viruses, unlike RT-PCR which only detects one or a few known viruses per test.