Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.
-
Full-text index only
In silico whole-genome screening for cancer-related single-nucleotide polymorphisms located in human mRNA untranslated regions.
PMID 17201911 · PMC1774567 · BMC genomics · 2007 · 8 claims · 5 setups
A computational EST-based pipeline can identify UTR-SNPs that are statistically over-represented in cancerous versus normal tissue libraries
-
Full-text index only
NEIBank: genomics and bioinformatics resources for vision research.
PMID 18648525 · PMC2480482 · Molecular vision · 2008 · 8 claims · 7 setups
NEIBank is an integrated genomics and bioinformatics resource for vision research, combining EST/cDNA clone data, SAGE expression data, and eye disease gene databases.
-
Full-text index only
Comprehensive annotation of bidirectional promoters identifies co-regulation among breast and ovarian cancer genes.
PMID 17447839 · PMC1853124 · PLoS computational biology · 2007 · 8 claims · 8 setups
A new algorithm using spliced ESTs (cross-validated against Known Genes and GenBank mRNA) comprehensively maps bidirectional promoters in the human genome
-
Full-text index only
ARED 3.0: the large and diverse AU-rich transcriptome.
PMID 16381826 · PMC1347415 · Nucleic acids research · 2006 · 7 claims · 6 setups
ARED 3.0 computationally mapped more than 4000 ARE-mRNAs to the human genome, representing 5-8% of human genes.
-
Full-text index only
Visualizing the genome: techniques for presenting human genome data and annotations.
PMID 12149135 · PMC119855 · BMC bioinformatics · 2002 · 8 claims · 4 setups
Web-based client-server genome browsers (e.g., LocusLink evidence viewer, UCSC genome browser) are limited by lack of true interactivity, requiring server round-trips for navigation
-
Full-text index only
Analysis of the prostate cancer cell line LNCaP transcriptome using a sequencing-by-synthesis approach.
PMID 17010196 · PMC1592491 · BMC genomics · 2006 · 8 claims · 7 setups
High-throughput 454 sequencing-by-synthesis of LNCaP cDNA can profile transcript abundance across the transcriptome
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
GeneTide--Terra Incognita Discovery Endeavor: a new transcriptome focused member of the GeneCards/GeneNote suite of databases.
PMID 15608261 · PMC540076 · Nucleic acids research · 2005 · 8 claims · 7 setups
GeneTide integrates UniGene, DoTS, AceView, BLAT/GeneLoc genomic alignment, and GeneAnnot probe-set data into a unified Consensus/Uniqueness/Score scheme to associate ESTs with GeneCards genes
-
Full-text index only
TranspoGene and microTranspoGene: transposed elements influence on the transcriptome of seven vertebrates and invertebrates.
PMID 17986453 · PMC2238949 · Nucleic acids research · 2008 · 8 claims · 5 setups
TranspoGene catalogs TEs within protein-coding genes of seven species (human, mouse, chicken, zebrafish, fruit fly, nematode, sea squirt), classified as proximal promoter, exonized, exonic, or intronic TEs.
-
Full-text index only
Gene prediction in eukaryotes with a generalized hidden Markov model that uses hints from external sources.
PMID 16469098 · PMC1409804 · BMC bioinformatics · 2006 · 7 claims · 3 setups
AUGUSTUS+ extends the AUGUSTUS GHMM by combining intrinsic sequence information with extrinsic hints via an extended emission alphabet, so the GHMM jointly models the DNA sequence, gene structure, and hint collection.
-
Full-text index only
Widely variable endogenous retroviral methylation levels in human placenta.
PMID 17617638 · PMC1950553 · Nucleic acids research · 2007 · 6 claims · 6 setups
Three HERV-E LTRs that function as alternative gene promoters (LTR-PTN, LTR-EBR, LTR-MID1) are unmethylated in placenta but heavily methylated in blood cells, where they are not active promoters
-
Full-text index only
Mice and more.
PMID 14519193 · PMC328447 · Genome biology · 2003 · 8 claims · 7 setups
A multispecies weighted conservation score, which accounts for each species' divergence rate, can identify conserved non-coding sequences (multispecies conserved sequences, MCSs) likely to be biologically significant
-
Full-text index only
Automatic annotation of eukaryotic genes, pseudogenes and promoters.
PMID 16925832 · PMC1810547 · Genome biology · 2006 · 8 claims · 6 setups
Fgenesh++ gene prediction pipeline identifies 91% of coding nucleotides with 90% specificity
-
Full-text index only
Inference of transcriptional regulation using gene expression data from the bovine and human genomes.
PMID 17683551 · PMC1978505 · BMC genomics · 2007 · 7 claims · 8 setups
Using human reference promoter sequences is a useful approach for studying gene expression regulation in species with limited or non-existing genomic sequence, such as cattle.
-
Full-text index only
Widespread ultraconservation divergence in primates.
PMID 18492662 · PMC2464743 · Molecular biology and evolution · 2008 · 8 claims · 4 setups
The number of UCEs has decreased throughout primate evolution, from ~1,000 in ancestral primates to 635 in modern humans.
-
Full-text index only
CONTRAST: a discriminative, phylogeny-free approach to multiple informant de novo gene prediction.
PMID 18096039 · PMC2246271 · Genome biology · 2007 · 8 claims · 5 setups
CONTRAST predicts exact coding region structures for 65% more human genes than the previous state-of-the-art de novo predictor (N-SCAN)
-
Full-text index only
MODBASE: a database of annotated comparative protein structure models and associated resources.
PMID 16381869 · PMC1347422 · Nucleic acids research · 2006 · 8 claims · 7 setups
MODBASE is a database of automatically calculated comparative protein structure models covering all UniProt sequences matchable to a known structure
-
Full-text index only
Catalogues of mammalian long noncoding RNAs: modest conservation and incompleteness.
PMID 19895688 · PMC3091318 · Genome biology · 2009 · 8 claims · 6 setups
MacroRNA and lincRNA exons are subject to the same relatively low degree of sequence constraint, contrary to prior reports that lincRNAs are far more conserved
-
Full-text index only
The dystrobrevin binding protein 1 (DTNBP1) gene is associated with schizophrenia in the Irish Case Control Study of Schizophrenia (ICCSS) sample.
PMID 19800201 · PMC2783814 · Schizophrenia research · 2009 · 8 claims · 7 setups
Common alleles at DTNBP1 SNPs, particularly rs760761, are associated with schizophrenia in the ICCSS sample