Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Better smelling through genetics: mammalian odor perception.
PMID 18938244 · PMC2590501 · Current opinion in neurobiology · 2008 · 8 claims · 8 setups
Odorant receptor (OR) gene repertoire size and pseudogene fraction vary dramatically across mammalian species
-
Full-text index only
Distribution and effects of nonsense polymorphisms in human genes.
PMID 18852891 · PMC2561068 · PloS one · 2008 · 8 claims · 8 setups
Nonsense SNPs occur at a lower density than nonsynonymous SNPs, indicating stronger purifying selection against premature stop codons than amino acid changes.
-
Full-text index only
TB database: an integrated platform for tuberculosis research.
PMID 18835847 · PMC2686437 · Nucleic acids research · 2009 · 8 claims · 8 setups
TBDB is an integrated database providing access to TB genomic data and resources relevant to discovery/development of TB drugs, vaccines and biomarkers.
-
Full-text index only
A general definition and nomenclature for alternative splicing events.
PMID 18688268 · PMC2467475 · PLoS computational biology · 2008 · 6 claims · 4 setups
Existing AS nomenclatures (Malko et al.'s 5-letter strings, Nagasaki et al.'s bit matrices, and the ASD/ATD/AEdb system) are redundant, ambiguous, or incapable of representing complex or large splicing variations.
-
Full-text index only
Genomic and bioinformatics analysis of human adenovirus type 37: new insights into corneal tropism.
PMID 18471294 · PMC2397415 · BMC genomics · 2008 · 7 claims · 7 setups
The complete genome of HAdV-37 was sequenced and annotated (35,213 bp, 56.6% GC content, 35 predicted coding sequences plus 8 hypothetical ORFs)
-
Full-text index only
Prediction-based approaches to characterize bidirectional promoters in the mammalian genome.
PMID 18366609 · PMC2386062 · BMC genomics · 2008 · 8 claims · 7 setups
The mapping algorithm identified 5,647 candidate bidirectional promoter regions in the mouse genome, similar in number to those previously found in human.
-
Full-text index only
Variation analysis and gene annotation of eight MHC haplotypes: the MHC Haplotype Project.
PMID 18193213 · PMC2206249 · Immunogenetics · 2008 · 8 claims · 6 setups
Comparison of eight HLA-homozygous MHC haplotype sequences identified >44,000 variations (substitutions and indels), submitted to dbSNP
-
Full-text index only
The mammalian phenotype ontology: enabling robust annotation and comparative analysis.
PMID 20052305 · PMC2801442 · Wiley interdisciplinary reviews. Systems biology and medicine · 2009 · 8 claims · 6 setups
The Mammalian Phenotype (MP) Ontology enables classification and organization of phenotypic data for mouse and other mammalian species in a computationally useful, standardized manner.
-
Full-text index only
The Bifidobacterium dentium Bd1 genome sequence reflects its genetic adaptation to the human oral cavity.
PMID 20041198 · PMC2788695 · PLoS genetics · 2009 · 8 claims · 8 setups
The B. dentium Bd1 genome was sequenced to completion, revealing a single circular 2,636,368 bp chromosome with 2,143 predicted ORFs
-
Full-text index only
Evolutionary trace annotation of protein function in the structural proteome.
PMID 20036248 · PMC2831211 · Journal of molecular biology · 2010 · 8 claims · 7 setups
ET-ranked residue clusters can be used to build 3D templates that predict GO function in enzymes and non-enzymes alike, without prior knowledge of functional mechanism.
-
Full-text index only
Genome reannotation of Escherichia coli CFT073 with new insights into virulence.
PMID 19930606 · PMC2785843 · BMC genomics · 2009 · 8 claims · 7 setups
Reannotation excluded 608 CDSs from the original RefSeq annotation, mostly unfunctional 'hypothetical'/'putative' genes
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
BioGPS: an extensible and customizable portal for querying and organizing gene annotation resources.
PMID 19919682 · PMC3091323 · Genome biology · 2009 · 8 claims · 4 setups
BioGPS aggregates distributed, third-party gene annotation resources into a single customizable portal for human, mouse, and rat genes.
-
Full-text index only
The complete genome sequence of Vibrio cholerae: a tale of two chromosomes and of two lifestyles.
PMID 11178241 · PMC138858 · Genome biology · 2000 · 8 claims · 4 setups
The V. cholerae O1 (El Tor) genome consists of two chromosomes with asymmetrically distributed gene functions
-
Has reproduction · 38
Genomic capacities for Reactive Oxygen Species metabolism across marine phytoplankton.
PMID 37098087 · PMC10128935 · PloS one · 2023 · 8 claims · 3 setups
Genes encoding superoxide (O2•−) scavenging are ubiquitous across phytoplankton, but their fractional gene allocation decreases with increasing cell radius, consistent with a nearly fixed core gene set.
-
Full-text index only
Target SNP selection in complex disease association studies.
PMID 15248903 · PMC487897 · BMC bioinformatics · 2004 · 7 claims · 3 setups
A computational pipeline can retrieve gene sequence, collect SNP variation data, and annotate SNPs falling in functional motifs (promoter, exon-intron structure, AU-rich elements, TF binding sites, splice sites) with expression in target tissue
-
Full-text index only
Inherited disorder phenotypes: controlled annotation and statistical analysis for knowledge mining from gene lists.
PMID 16351744 · PMC1866390 · BMC bioinformatics · 2005 · 5 claims · 3 setups
OMIM Clinical Synopsis free-text phenotype and location names can be normalized and hierarchically structured into a controlled vocabulary suitable for computational analysis
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Full-text index only
pTARGET: a web server for predicting protein subcellular localization.
PMID 16844995 · PMC1538910 · Nucleic acids research · 2006 · 7 claims · 3 setups
pTARGET web server predicts nine distinct subcellular localizations in eukaryotic non-plant proteins using an algorithm based on location-specific Pfam domain occurrence patterns and amino acid composition (AAC)
-
Full-text index only
DBD--taxonomically broad transcription factor predictions: new content and functionality.
PMID 18073188 · PMC2238844 · Nucleic acids research · 2008 · 8 claims · 3 setups
DBD is a database of predicted sequence-specific DNA-binding transcription factors covering over 700 publicly available proteomes, up from 150 in the initial version.