Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A space-efficient and accurate method for mapping and aligning cDNA sequences onto genomic sequence.
PMID 18344523 · PMC2377433 · Nucleic acids research · 2008 · 7 claims · 6 setups
Spaln maps and aligns large cDNA sequence sets onto whole mammalian genomes using substantially less memory than comparable existing tools
-
Full-text index only
The ASAP II database: analysis and comparative genomics of alternative splicing in 15 animal species.
PMID 17108355 · PMC1669709 · Nucleic acids research · 2007 · 8 claims · 4 setups
ASAP II expands human alternative splicing data ~3-fold over the previous ASAP database, to ~89,078 distinct alternative splicing relationships in 11,717 genes
-
Full-text index only
Evidence for a preferential targeting of 3'-UTRs by cis-encoded natural antisense transcripts.
PMID 16204454 · PMC1243798 · Nucleic acids research · 2005 · 8 claims · 4 setups
Cis-encoded natural antisense RNAs show striking preferential complementarity to 3′-UTRs of their target genes in human and mouse genomes
-
Full-text index only
ECgene: genome annotation for alternative splicing.
PMID 15608289 · PMC540072 · Nucleic acids research · 2005 · 8 claims · 5 setups
ECgene combines genome-based EST clustering with a graph-theoretic transcript assembly procedure to predict gene models including alternative splicing events.
-
Full-text index only
CRSD: a comprehensive web server for composite regulatory signature discovery.
PMID 16845073 · PMC1538777 · Nucleic acids research · 2006 · 7 claims · 5 setups
CRSD is a comprehensive web server integrating six large-scale databases (UniGene, mature microRNAs, putative promoter, TRANSFAC, pathway, GO) plus two newly constructed genome-wide databases (MRS and TRS) for composite regulatory signature discovery
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.
-
Full-text index only
Satellog: a database for the identification and prioritization of satellite repeats in disease association studies.
PMID 15949044 · PMC1181805 · BMC bioinformatics · 2005 · 7 claims · 6 setups
Satellog is a database cataloging all pure 1-16 unit satellite repeats in the human genome with supplementary polymorphism, gene-location, and expression data for prioritizing repeats in disease-association studies.
-
Full-text index only
Phylogenomic approaches to common problems encountered in the analysis of low copy repeats: the sulfotransferase 1A gene family example.
PMID 15752422 · PMC555591 · BMC evolutionary biology · 2005 · 8 claims · 8 setups
A previously unidentified fourth human SULT1A gene (SULT1A4) exists on chromosome 16 and is transcriptionally active
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
Identification and characterization of insect-specific proteins by genome data analysis.
PMID 17407609 · PMC1852559 · BMC genomics · 2007 · 8 claims · 7 setups
Comparative genome analysis across five holometabolous insects and three non-insect eukaryotes (opisthokonts) identifies 154 insect-specific orthologous groups (refined to 51 proteins) and 466 eukaryote/opisthokont-core orthologous groups
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.
-
Full-text index only
EPD in its twentieth year: towards complete promoter coverage of selected model organisms.
PMID 16381980 · PMC1347508 · Nucleic acids research · 2006 · 7 claims · 4 setups
EPD is an annotated, non-redundant collection of experimentally defined eukaryotic POL II promoters accessed via genome position pointers.
-
Full-text index only
PolyA_DB 2: mRNA polyadenylation sites in vertebrate genes.
PMID 17202160 · PMC1899096 · Nucleic acids research · 2007 · 7 claims · 5 setups
PolyA_DB 2 catalogs poly(A) sites for genes in human, mouse, rat, chicken and zebrafish, identified by aligning cDNA/ESTs with genome sequences
-
Full-text index only
ECgene: an alternative splicing database update.
PMID 17132829 · PMC1716719 · Nucleic acids research · 2007 · 8 claims · 5 setups
ECgene provides functional annotation (domain, GO, expression pattern) for alternatively spliced genes
-
Full-text index only
Discovering multiple transcripts of human hepatocytes using massively parallel signature sequencing (MPSS).
PMID 17601345 · PMC1929076 · BMC genomics · 2007 · 8 claims · 8 setups
MPSS detected 10,279 UniGene clusters, representing 7,475 known genes, in human hepatocytes
-
Full-text index only
ARED 3.0: the large and diverse AU-rich transcriptome.
PMID 16381826 · PMC1347415 · Nucleic acids research · 2006 · 7 claims · 6 setups
ARED 3.0 computationally mapped more than 4000 ARE-mRNAs to the human genome, representing 5-8% of human genes.
-
Full-text index only
PLANdbAffy: probe-level annotation database for Affymetrix expression microarrays.
PMID 19906711 · PMC2808952 · Nucleic acids research · 2010 · 6 claims · 4 setups
PLANdbAffy is a database of Affymetrix probe alignments to the human genome for five widely used arrays (HG-U133A, HG-U133B, HG-U133 Plus 2.0, Human Exon 1.0, Human Gene 1.0)
-
Full-text index only
Transcriptome annotation using tandem SAGE tags.
PMID 17709346 · PMC2034470 · Nucleic acids research · 2007 · 8 claims · 7 setups
A novel algorithm pairs tandem SAGE tags anchored on two different restriction sites (CATG and GATC) to define tag-delimited genomic sequences (TDGS)
-
Full-text index only
The other side of comparative genomics: genes with no orthologs between the cow and other mammalian species.
PMID 20003425 · PMC2808326 · BMC genomics · 2009 · 7 claims · 4 setups
3,801 bovine genes have no orthologs in human, mouse and dog, and 1,010 human genes have no orthologs in cow despite having orthologs in mouse and dog