Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Identification and evolutionary analysis of novel exons and alternative splicing events using cross-species EST-to-genome comparisons in human, mouse and rat.
PMID 16536879 · PMC1479377 · BMC bioinformatics · 2006 · 8 claims · 6 setups
ENACE, a cross-species EST-to-genome comparison algorithm, can identify novel cassette-on exons and retained introns for EST-scanty species and distinguish conserved vs lineage-specific exons
-
Full-text index only
Analysis of expressed sequence tags from Actinidia: applications of a cross species EST database for gene discovery in the areas of flavor, health, color and ripening.
PMID 18655731 · PMC2515324 · BMC genomics · 2008 · 7 claims · 6 setups
A collection of 132,577 ESTs from four Actinidia species was generated and clustered into 41,858 non-redundant clusters (18,070 TCs and 23,788 singletons)
-
Full-text index only
Phylogenetic profiling of the Arabidopsis thaliana proteome: what proteins distinguish plants from other organisms?
PMID 15287975 · PMC507878 · Genome biology · 2004 · 8 claims · 6 setups
3,848 Arabidopsis proteins were identified as likely plant-specific based on phylogenetic profiling and EST confirmation in multiple plant species
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
Proceedings of the First International Conference on Phylogenomics. March 15-19, 2006. Quebec, Canada.
PMID 17288567 · PMC1796603 · BMC evolutionary biology · 2007 · 8 claims · 8 setups
Gene tree parsimony applied to EST data with widespread gene duplication can infer an organismal phylogeny in excellent agreement with the expected angiosperm phylogeny.
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences
-
Full-text index only
Trans-natural antisense transcripts including noncoding RNAs in 10 species: implications for expression regulation.
PMID 18653530 · PMC2528163 · Nucleic acids research · 2008 · 8 claims · 7 setups
A new computational pipeline identifies trans-SAs using ESTs (not just mRNAs) across 10 animal species, improving coverage over prior methods
-
Full-text index only
NEIBank: genomics and bioinformatics resources for vision research.
PMID 18648525 · PMC2480482 · Molecular vision · 2008 · 8 claims · 7 setups
NEIBank is an integrated genomics and bioinformatics resource for vision research, combining EST/cDNA clone data, SAGE expression data, and eye disease gene databases.
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
The ASAP II database: analysis and comparative genomics of alternative splicing in 15 animal species.
PMID 17108355 · PMC1669709 · Nucleic acids research · 2007 · 8 claims · 4 setups
ASAP II expands human alternative splicing data ~3-fold over the previous ASAP database, to ~89,078 distinct alternative splicing relationships in 11,717 genes
-
Full-text index only
TranspoGene and microTranspoGene: transposed elements influence on the transcriptome of seven vertebrates and invertebrates.
PMID 17986453 · PMC2238949 · Nucleic acids research · 2008 · 8 claims · 5 setups
TranspoGene catalogs TEs within protein-coding genes of seven species (human, mouse, chicken, zebrafish, fruit fly, nematode, sea squirt), classified as proximal promoter, exonized, exonic, or intronic TEs.
-
Full-text index only
PolyA_DB 2: mRNA polyadenylation sites in vertebrate genes.
PMID 17202160 · PMC1899096 · Nucleic acids research · 2007 · 7 claims · 5 setups
PolyA_DB 2 catalogs poly(A) sites for genes in human, mouse, rat, chicken and zebrafish, identified by aligning cDNA/ESTs with genome sequences
-
Full-text index only
Polymorphix: a sequence polymorphism database.
PMID 15608242 · PMC540030 · Nucleic acids research · 2005 · 8 claims · 5 setups
Polymorphix is an ACNUC-structured database that organizes EMBL/GenBank sequences into within-species homologous sequence families using similarity and bibliographic criteria, with alignments, outgroups and phylogenetic trees provided.
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Full-text index only
Genome-wide survey for biologically functional pseudogenes.
PMID 16680195 · PMC1456316 · PLoS computational biology · 2006 · 8 claims · 6 setups
A subset of ancient, cross-species-conserved pseudogenes (30 of 1,453 candidate quartets) show evidence consistent with retained biological function
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.
-
Full-text index only
The fate of the duplicated androgen receptor in fishes: a late neofunctionalization event?
PMID 19094205 · PMC2637867 · BMC evolutionary biology · 2008 · 8 claims · 4 setups
AR was duplicated into two paralogs, AR-A and AR-B, during a teleost-specific whole genome duplication (WGD), after the split of Acipenseriformes but before the divergence of Osteoglossiformes.
-
Full-text index only
Genomic structure and expression of Jmjd6 and evolutionary analysis in the context of related JmjC domain containing proteins.
PMID 18564434 · PMC2453528 · BMC genomics · 2008 · 8 claims · 6 setups
Jmjd6 has been misleadingly annotated as a transmembrane receptor for engulfment of apoptotic cells; recent evidence contradicts this transmembrane receptor function
-
Full-text index only
SelenoDB 1.0 : a database of selenoprotein genes, proteins and SECIS elements.
PMID 18174224 · PMC2238826 · Nucleic acids research · 2008 · 6 claims · 5 setups
Standard genome annotation pipelines misannotate selenoprotein genes because they rely on UGA as a universal stop codon, failing to recognize its dual role as the selenocysteine-recoding codon.
-
Full-text index only
The truth about mouse, human, worms and yeast.
PMID 15601543 · PMC3525071 · Human genomics · 2004 · 8 claims · 8 setups
Comparing genomes in pairs or larger sets (mouse-human, C. elegans-C. briggsae, multiple Saccharomyces, human-pufferfish, etc.) reveals unsuspected genes and helps eliminate false-positive gene predictions