Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome annotation errors in pathway databases due to semantic ambiguity in partial EC numbers.
PMID 16034025 · PMC1179732 · Nucleic acids research · 2005 · 7 claims · 4 setups
Partial EC numbers are semantically ambiguous, and databases that assign a gene to all reactions sharing the same partial EC number make a faulty inference, causing systematic misannotation.
-
Full-text index only
A genetic variation map for chicken with 2.8 million single-nucleotide polymorphisms.
PMID 15592405 · PMC2263125 · Nature · 2004 · 8 claims · 8 setups
A genetic variation map of 2.8 million SNPs was constructed for chicken by comparing 3 domestic breeds to Red Jungle Fowl
-
Full-text index only
AceView: a comprehensive cDNA-supported gene and transcripts annotation.
PMID 16925834 · PMC1810549 · Genome biology · 2006 · 8 claims · 4 setups
At the mRNA level, AceView transcripts are the closest match to Gencode transcripts among all evaluated methods, including alternative splice variants
-
Full-text index only
A macaque's-eye view of human insertions and deletions: differences in mechanisms.
PMID 17941704 · PMC1976337 · PLoS computational biology · 2007 · 7 claims · 4 setups
Insertion and deletion rates are differentially associated with replication- versus recombination-related genomic features, indicating the two mutation types are driven in part by distinct mechanisms
-
Full-text index only
'Unknown' proteins and 'orphan' enzymes: the missing half of the engineering parts list--and how to find it.
PMID 20001958 · PMC3022307 · The Biochemical journal · 2009 · 8 claims · 8 setups
Comparative genomics is the single most effective strategy for predicting functions of unknown proteins and finding genes for orphan enzymes
-
Full-text index only
Sequence context affects the rate of short insertions and deletions in flies and primates.
PMID 18291026 · PMC2374710 · Genome biology · 2008 · 8 claims · 6 setups
The rate of insertion or deletion of specific lengths can vary by more than 100-fold depending on the surrounding sequence context
-
Full-text index only
Comparative genomics and understanding of microbial biology.
PMID 10998382 · PMC2627966 · Emerging infectious diseases · 2000 · 8 claims · 7 setups
GC content varies widely among prokaryotic genomes (29% in B. burgdorferi to 68% in M. tuberculosis) and shapes codon usage and amino acid composition.
-
Full-text index only
Eurasian and African mitochondrial DNA influences in the Saudi Arabian population.
PMID 17331239 · PMC1810519 · BMC evolutionary biology · 2007 · 8 claims · 4 setups
The majority (85%) of Saudi Arab mtDNA lineages have a western Asia (Eurasian) provenance
-
Full-text index only
Characterizing natural variation using next-generation sequencing technologies.
PMID 19801172 · PMC3994700 · Trends in genetics : TIG · 2009 · 8 claims · 8 setups
Next-generation sequencing enables complete, genome-wide surveys of genetic variation at unprecedented resolution, overcoming limitations of genotyping panels and microarrays.
-
Full-text index only
Molecular analysis of Plasmodium ovale variants.
PMID 15324543 · PMC3323326 · Emerging infectious diseases · 2004 · 8 claims · 5 setups
P. ovale isolates separate into two genetically distinct types, classic (Nigerian I/CDC) and variant (LS), consistent across four independent gene loci.
-
Full-text index only
G, N, and P gene-based analysis of Chandipura viruses, India.
PMID 15705335 · PMC3294343 · Emerging infectious diseases · 2005 · 8 claims · 4 setups
The 2003 epidemic CHPV isolates are closely related to, and not very divergent from, the 1965 isolate, indicating the outbreak was not associated with extensive mutations in the G, N, and P genes.
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Full-text index only
Molecular phylogeny of the kelch-repeat superfamily reveals an expansion of BTB/kelch proteins in animals.
PMID 13678422 · PMC222960 · BMC bioinformatics · 2003 · 8 claims · 8 setups
The human genome encodes at least 71 kelch-repeat proteins
-
Full-text index only
Evolutionary trace annotation of protein function in the structural proteome.
PMID 20036248 · PMC2831211 · Journal of molecular biology · 2010 · 8 claims · 7 setups
ET-ranked residue clusters can be used to build 3D templates that predict GO function in enzymes and non-enzymes alike, without prior knowledge of functional mechanism.
-
Full-text index only
Molecular analysis of a leprosy immunotherapeutic bacillus provides insights into Mycobacterium evolution.
PMID 17912347 · PMC1989137 · PloS one · 2007 · 8 claims · 8 setups
MIP is the evolutionary predecessor/ancestor of the pathogenic Mycobacterium avium intracellulare complex (MAIC), having retained a free-living lifestyle rather than undergoing parasitic reductive genome evolution