Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Heterogeneous genomic molecular clocks in primates.
PMID 17029560 · PMC1592237 · PLoS genetics · 2006 · 7 claims · 7 setups
Non-CpG site substitutions show clear generation-time dependency, consistent with a replication-error origin
-
Full-text index only
The UCSC Genome Browser Database: update 2006.
PMID 16381938 · PMC1347506 · Nucleic acids research · 2006 · 8 claims · 8 setups
The UCSC Genome Browser Database (GBD) provides integrated sequence and annotation data, with web tools (Genome Browser, Table Browser, Proteome Browser, Gene Sorter, BLAT, In Silico PCR) for visualizing and querying genomes of about a dozen vertebrate species and several model organisms.
-
Full-text index only
Pairagon+N-SCAN_EST: a model-based gene annotation pipeline.
PMID 16925839 · PMC1810554 · Genome biology · 2006 · 7 claims · 5 setups
Pairagon+N-SCAN_EST, using only native alignments, was as accurate as ENSEMBL and ExoGean in the EGASP mRNA/EST evidence assessment
-
Full-text index only
Analyses of deep mammalian sequence alignments and constraint predictions for 1% of the human genome.
PMID 17567995 · PMC1891336 · Genome research · 2007 · 7 claims · 3 setups
Four different alignment methods show large-scale consistency but substantial differences in small-scale rearrangements, sensitivity, and specificity.
-
Full-text index only
The ENCODE Project at UC Santa Cruz.
PMID 17166863 · PMC1781110 · Nucleic acids research · 2007 · 8 claims · 4 setups
The UCSC ENCODE portal serves as the primary repository and access point for sequence-based ENCODE pilot phase data
-
Full-text index only
Vertebrate gene finding from multiple-species alignments using a two-level strategy.
PMID 16925840 · PMC1810555 · Genome biology · 2006 · 8 claims · 5 setups
DOGFISH cleanly separates a multi-species alignment classifier (RVM cascade) from an HMM-based structure predictor, avoiding tight coupling of alignment complexity with HMM formalism
-
Full-text index only
Integrative functional genomics.
PMID 15239826 · PMC463286 · Genome biology · 2004 · 8 claims · 8 setups
Ultra-conserved noncoding elements exist across human, mouse and rat genomes at very high sequence identity, often far from genes
-
Full-text index only
Ensembl 2007.
PMID 17148474 · PMC1761443 · Nucleic acids research · 2007 · 8 claims · 7 setups
Ensembl added 18 new chordate genomes this year, increasing total genomes available from 15 to 33, the largest yearly increase to date.
-
Full-text index only
Using several pair-wise informant sequences for de novo prediction of alternatively spliced transcripts.
PMID 16925842 · PMC1810557 · Genome biology · 2006 · 8 claims · 4 setups
MARS, an extension of the Twinscan algorithm, uses multiple pairwise informant genomes to predict human alternatively spliced transcripts de novo without expressed sequence information.
-
Full-text index only
CONTRAST: a discriminative, phylogeny-free approach to multiple informant de novo gene prediction.
PMID 18096039 · PMC2246271 · Genome biology · 2007 · 8 claims · 5 setups
CONTRAST predicts exact coding region structures for 65% more human genes than the previous state-of-the-art de novo predictor (N-SCAN)
-
Full-text index only
Exogean: a framework for annotating protein-coding genes in eukaryotic genomic DNA.
PMID 16925841 · PMC1810556 · Genome biology · 2006 · 8 claims · 5 setups
Exogean is a framework using directed acyclic coloured multigraphs (DACMs) to represent biological objects (mRNA, ESTs, protein alignments, exons) and iteratively combine them into complex protein-coding transcript models.
-
Full-text index only
AUGUSTUS at EGASP: using EST, protein and genomic alignments for improved gene prediction in the human genome.
PMID 16925833 · PMC1810548 · Genome biology · 2006 · 8 claims · 5 setups
AUGUSTUS predicted significantly more genes correctly than any other ab initio program in EGASP
-
Full-text index only
Ensembl 2009.
PMID 19033362 · PMC2686571 · Nucleic acids research · 2009 · 8 claims · 6 setups
Ensembl provides comprehensive, consistently annotated genome information for chordate genomes with automatically generated genesets and comparative genomics data
-
Full-text index only
The vertebrate genome annotation (Vega) database.
PMID 18003653 · PMC2238886 · Nucleic acids research · 2008 · 8 claims · 8 setups
Vega is a database for viewing manual genome annotation of human, mouse and zebrafish genomic sequences produced at the Wellcome Trust Sanger Institute.
-
Full-text index only
Reference based annotation with GeneMapper.
PMID 16600017 · PMC1557983 · Genome biology · 2006 · 7 claims · 6 setups
GeneMapper transfers reference gene annotations to target genomes with higher accuracy than GeneWise and Projector