Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Rapid detection and curation of conserved DNA via enhanced-BLAT and EvoPrinterHD analysis.
PMID 18307801 · PMC2268679 · BMC genomics · 2008 · 8 claims · 8 setups
eBLAT detects up to 75% more conserved bases than original BLAT alignments, with the largest gains between evolutionarily distant orthologs
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
ChimerDB--a knowledgebase for fusion sequences.
PMID 16381848 · PMC1347382 · Nucleic acids research · 2006 · 8 claims · 6 setups
ChimerDB integrates bioinformatics analysis of mRNA/EST sequences, manually collected literature data, and OMIM translocation data into a single fusion sequence knowledgebase
-
Full-text index only
Computational comparison of two mouse draft genomes and the human golden path.
PMID 12537546 · PMC151282 · Genome biology · 2003 · 8 claims · 7 setups
The Celera and public mouse genome assemblies differ in about 10% of the mouse genome, with complementary strengths (Celera higher base-pair accuracy and overall coverage; public assembly higher quality in some finished BAC regions and freely accessible)
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences
-
Full-text index only
Trans-natural antisense transcripts including noncoding RNAs in 10 species: implications for expression regulation.
PMID 18653530 · PMC2528163 · Nucleic acids research · 2008 · 8 claims · 7 setups
A new computational pipeline identifies trans-SAs using ESTs (not just mRNAs) across 10 animal species, improving coverage over prior methods
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs
-
Full-text index only
WebScipio: an online tool for the determination of gene structures using protein sequences.
PMID 18801164 · PMC2644328 · BMC genomics · 2008 · 7 claims · 4 setups
WebScipio, a web interface to Scipio, determines gene structure from a query protein sequence against an assembled eukaryotic genome with quality approaching manual annotation.
-
Full-text index only
Genomic analysis of the TRIM family reveals two groups of genes with distinct evolutionary properties.
PMID 18673550 · PMC2533329 · BMC evolutionary biology · 2008 · 8 claims · 6 setups
The human TRIM family is split into two groups (group 1 and group 2) that differ in domain structure, genomic organization, and evolutionary properties.
-
Full-text index only
ChimerDB 2.0--a knowledgebase for fusion genes updated.
PMID 19906715 · PMC2808913 · Nucleic acids research · 2010 · 8 claims · 4 setups
ChimerDB 2.0 is an updated knowledgebase integrating fusion transcripts from GenBank transcriptome analysis with Sanger CGP, OMIM, PubMed, and Mitelman's database data.
-
Full-text index only
The UCSC Proteome Browser.
PMID 15608236 · PMC540054 · Nucleic acids research · 2005 · 8 claims · 5 setups
The UCSC Proteome Browser is tightly integrated with the UCSC Genome Browser, giving users simultaneous access to genome and proteome data.
-
Full-text index only
Gene losses during human origins.
PMID 16464126 · PMC1361800 · PLoS biology · 2006 · 7 claims · 7 setups
A comparative genomic screen identified 67 new human-specific nonprocessed pseudogenes, bringing the total (with 13 from prior literature) to 80 human-specific pseudogenes.
-
Full-text index only
PLANdbAffy: probe-level annotation database for Affymetrix expression microarrays.
PMID 19906711 · PMC2808952 · Nucleic acids research · 2010 · 6 claims · 4 setups
PLANdbAffy is a database of Affymetrix probe alignments to the human genome for five widely used arrays (HG-U133A, HG-U133B, HG-U133 Plus 2.0, Human Exon 1.0, Human Gene 1.0)
-
Full-text index only
MutScreener: primer design tool for PCR-direct sequencing.
PMID 16845093 · PMC1538803 · Nucleic acids research · 2006 · 8 claims · 4 setups
MutScreener is a web-based application that automates PCR-direct sequencing assay design by annotating gene structure and designing PCR and sequencing primers.
-
Full-text index only
Phylogenomic approaches to common problems encountered in the analysis of low copy repeats: the sulfotransferase 1A gene family example.
PMID 15752422 · PMC555591 · BMC evolutionary biology · 2005 · 8 claims · 8 setups
A previously unidentified fourth human SULT1A gene (SULT1A4) exists on chromosome 16 and is transcriptionally active
-
Full-text index only
How accurately is ncRNA aligned within whole-genome multiple alignments?
PMID 17963514 · PMC2206062 · BMC bioinformatics · 2007 · 7 claims · 4 setups
MULTIZ does a fairly accurate job of aligning ncRNA regions across 17 vertebrate genomes, but better alignments exist in some regions.
-
Full-text index only
PolyA_DB 2: mRNA polyadenylation sites in vertebrate genes.
PMID 17202160 · PMC1899096 · Nucleic acids research · 2007 · 7 claims · 5 setups
PolyA_DB 2 catalogs poly(A) sites for genes in human, mouse, rat, chicken and zebrafish, identified by aligning cDNA/ESTs with genome sequences
-
Full-text index only
A space-efficient and accurate method for mapping and aligning cDNA sequences onto genomic sequence.
PMID 18344523 · PMC2377433 · Nucleic acids research · 2008 · 7 claims · 6 setups
Spaln maps and aligns large cDNA sequence sets onto whole mammalian genomes using substantially less memory than comparable existing tools
-
Has reproduction · 67
Optimal scaling of digital transcriptomes.
PMID 24223126 · PMC3819321 · PloS one · 2013 · 8 claims · 8 setups
Fifteen existing and novel transcript-count normalization algorithms can be compared with two novel, mutually independent metrics: the number of "uniform" genes (sufficiently low coefficient of variation after normalization) and low average Spearman correlation between normalized expression profiles of gene pairs.
-
Full-text index only
Retropseudogenes derived from the human Ro/SS-A autoantigen-associated hY RNAs.
PMID 15817567 · PMC1074747 · Nucleic acids research · 2005 · 8 claims · 8 setups
966 pseudogenes derived from the four human Y (hY) RNAs were characterized in the human genome