Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
ARED Organism: expansion of ARED reveals AU-rich element cluster variations between human and mouse.
PMID 17984078 · PMC2238997 · Nucleic acids research · 2008 · 6 claims · 4 setups
ARED Organism and ARED-Integrated are new/updated public databases cataloguing ARE-containing mRNAs/genes in human, mouse and rat
-
Full-text index only
Identification, characterization and comparative genomics of chimpanzee endogenous retroviruses.
PMID 16805923 · PMC1779541 · Genome biology · 2006 · 8 claims · 6 setups
The chimpanzee genome contains at least 42 separate families of endogenous retroviruses, 9 newly identified
-
Full-text index only
SVC: structured visualization of evolutionary sequence conservation.
PMID 15991338 · PMC1160265 · Nucleic acids research · 2005 · 7 claims · 5 setups
SVC aligns protein-coding sequences of orthologous gene pairs and maps them back onto their encoding exons/introns to generate a scaffold of conserved gene structure.
-
Full-text index only
Systematic identification of pseudogenes through whole genome expression evidence profiling.
PMID 16945953 · PMC1636364 · Nucleic acids research · 2006 · 8 claims · 8 setups
Developed a novel bioinformatics method that identifies pseudogenes by profiling whole-genome transcript and protein expression evidence
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs
-
Full-text index only
The truth about mouse, human, worms and yeast.
PMID 15601543 · PMC3525071 · Human genomics · 2004 · 8 claims · 8 setups
Comparing genomes in pairs or larger sets (mouse-human, C. elegans-C. briggsae, multiple Saccharomyces, human-pufferfish, etc.) reveals unsuspected genes and helps eliminate false-positive gene predictions
-
Full-text index only
Uncovering information on expression of natural antisense transcripts in Affymetrix MOE430 datasets.
PMID 17598913 · PMC1929078 · BMC genomics · 2007 · 8 claims · 4 setups
Standard Affymetrix expression GeneChips (MOE430, HG-U133) contain probe sets that detect natural antisense transcripts (NATs)
-
Full-text index only
Prioritization of candidate cancer genes--an aid to oncogenomic studies.
PMID 18710882 · PMC2566894 · Nucleic acids research · 2008 · 8 claims · 8 setups
Computational classifiers using combinations of protein conservation, gene structure, protein domains, protein interactions, and regulatory data can distinguish known cancer genes (CD/CR) from unlabelled human genes
-
Full-text index only
Statistical Viewer: a tool to upload and integrate linkage and association data as plots displayed within the Ensembl genome browser.
PMID 15826305 · PMC1087836 · BMC bioinformatics · 2005 · 8 claims · 3 setups
Statistical Viewer is a plug-in package for Ensembl that displays disease study-specific linkage and/or association data as 2D plots within Ensembl's Contig View and Cyto View pages.
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
Benchmarking ortholog identification methods using functional genomics data.
PMID 16613613 · PMC1557999 · Genome biology · 2006 · 8 claims · 7 setups
InParanoid is the best overall ortholog identification method for identifying functionally equivalent proteins when sensitivity and selectivity are combined into an overall score.
-
Full-text index only
Computing Ka and Ks with a consideration of unequal transitional substitutions.
PMID 16740169 · PMC1552089 · BMC evolutionary biology · 2006 · 7 claims · 7 setups
MYN, a modified version of the Yang-Nielsen (YN) algorithm based on the Tamura-Nei Model, allows unequal transitional substitution rates between purines (κR) and pyrimidines (κY) plus codon frequency bias
-
Full-text index only
Recent segmental and gene duplications in the mouse genome.
PMID 12914656 · PMC193640 · Genome biology · 2003 · 8 claims · 8 setups
33.6 Mb (1.2%) of the February 2003 mouse genome assembly (2,695 Mb) is involved in recent segmental duplications
-
Full-text index only
Vertebrate gene finding from multiple-species alignments using a two-level strategy.
PMID 16925840 · PMC1810555 · Genome biology · 2006 · 8 claims · 5 setups
DOGFISH cleanly separates a multi-species alignment classifier (RVM cascade) from an HMM-based structure predictor, avoiding tight coupling of alignment complexity with HMM formalism
-
Full-text index only
miRBase: tools for microRNA genomics.
PMID 17991681 · PMC2238936 · Nucleic acids research · 2008 · 8 claims · 6 setups
miRBase release 10.0 contains 5071 miRNA hairpin loci from 58 species, expressing 5922 distinct mature miRNA sequences, a growth of over 2000 sequences in 2 years
-
Full-text index only
Expansion of the Bactericidal/Permeability Increasing-like (BPI-like) protein locus in cattle.
PMID 17362520 · PMC1839098 · BMC genomics · 2007 · 8 claims · 8 setups
The bovine BPI-like locus spans 470 kbp and contains 14 contiguous genes (13 intact + 1 pseudogene); 9 are orthologous to human/mouse BPI-like genes and 4 (named BSP30A, BSP30B, BSP30C, BSP30D) arose through cattle-specific duplication of the PSP gene
-
Has reproduction · 94
A Deluge of Complex Repeats: The Solanum Genome.
PMID 26241045 · PMC4524691 · PloS one · 2015 · 8 claims · 7 setups
~50–60% of the S. tuberosum and S. lycopersicum genomes are composed of repetitive elements
-
Full-text index only
Sushi gets serious: the draft genome sequence of the pufferfish Fugu rubripes.
PMID 12225591 · PMC139409 · Genome biology · 2002 · 8 claims · 7 setups
The Fugu rubripes draft genome sequence was generated by whole-genome shotgun sequencing assembled to ~5.6x coverage using the JAZZ pipeline.
-
Has reproduction · 67
Optimal scaling of digital transcriptomes.
PMID 24223126 · PMC3819321 · PloS one · 2013 · 8 claims · 8 setups
Fifteen existing and novel transcript-count normalization algorithms can be compared with two novel, mutually independent metrics: the number of "uniform" genes (sufficiently low coefficient of variation after normalization) and low average Spearman correlation between normalized expression profiles of gene pairs.
-
Full-text index only
Large-scale analysis of human alternative protein isoforms: pattern classification and correlation with subcellular localization signals.
PMID 15860772 · PMC1087780 · Nucleic acids research · 2005 · 8 claims · 8 setups
Constructed a large-scale dataset of 6876 human alternative protein isoforms from 2624 genes by combining H-Invitational full-length cDNA data and SwissProt VARSPLIC entries