Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The other side of comparative genomics: genes with no orthologs between the cow and other mammalian species.
PMID 20003425 · PMC2808326 · BMC genomics · 2009 · 7 claims · 4 setups
3,801 bovine genes have no orthologs in human, mouse and dog, and 1,010 human genes have no orthologs in cow despite having orthologs in mouse and dog
-
Full-text index only
The functional importance of disease-associated mutation.
PMID 12220483 · PMC128831 · BMC bioinformatics · 2002 · 6 claims · 1 setups
Disease-associated mutations occur in conserved regions of genes and can be used to identify likely disease-causing mutations
-
Full-text index only
Evola: Ortholog database of all human genes in H-InvDB with manual curation of phylogenetic trees.
PMID 17982176 · PMC2238928 · Nucleic acids research · 2008 · 6 claims · 7 setups
Evola combines genome synteny-based computational ortholog detection with manual curation of phylogenetic trees by experts to yield more reliable orthologs than automated pairwise methods
-
Full-text index only
Network of Cancer Genes: a web resource to analyze duplicability, orthology and network properties of cancer genes.
PMID 19906700 · PMC2808873 · Nucleic acids research · 2010 · 7 claims · 4 setups
NCG is a web database integrating duplicability, orthology, evolutionary appearance, and network topology data for 736 human cancer genes
-
Full-text index only
In silico discovery of gene-coding variants in murine quantitative trait loci using strain-specific genome sequence databases.
PMID 12537567 · PMC151180 · Genome biology · 2002 · 6 claims · 4 setups
Strain-specific mouse genome sequence databases can be used in a high-throughput in silico pipeline to discover gene-coding variants within murine QTLs, without de novo sequencing.
-
Full-text index only
The post-genomic era for a select few.
PMID 14759254 · PMC395745 · Genome biology · 2004 · 8 claims · 8 setups
The Exofish comparative-genomics tool identifies protein-coding DNA segments by comparing two genome sequences and was used to compare pufferfish (Takifugu, Tetraodon) genomes with mammalian genomes, improving annotation of the human and mouse genomes.
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences
-
Full-text index only
Linking disease-associated genes to regulatory networks via promoter organization.
PMID 15701758 · PMC549397 · Nucleic acids research · 2005 · 8 claims · 7 setups
Pairs of TFBSs conserved both vertically (orthologous genes) and horizontally (co-regulated genes) can serve as seeds to build promoter models representing potential co-regulation networks
-
Full-text index only
GenomeTrafac: a whole genome resource for the detection of transcription factor binding site clusters associated with conventional and microRNA encoding genes conserved between mouse and human gene orthologs.
PMID 17178752 · PMC1781107 · Nucleic acids research · 2007 · 8 claims · 5 setups
GenomeTrafac is a web-accessible database enabling genome-wide detection of conserved cis-element clusters in human-mouse gene orthologs, covering both conventional and microRNA genes
-
Full-text index only
Systematic identification of pseudogenes through whole genome expression evidence profiling.
PMID 16945953 · PMC1636364 · Nucleic acids research · 2006 · 8 claims · 8 setups
Developed a novel bioinformatics method that identifies pseudogenes by profiling whole-genome transcript and protein expression evidence
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
Multiple whole genome alignments and novel biomedical applications at the VISTA portal.
PMID 17488840 · PMC1933192 · Nucleic acids research · 2007 · 8 claims · 4 setups
A novel multiple whole-genome alignment algorithm treats all genomes symmetrically, avoiding dependence on a single base/reference genome
-
Has reproduction · 84
Foster thy young: enhanced prediction of orphan genes in assembled genomes.
PMID 34928390 · PMC9023268 · Nucleic acids research · 2022 · 7 claims · 8 setups
Each of the five gene-prediction pipelines under-predicts orphan genes, with detection as low as 11% under one prediction scenario.
-
Full-text index only
MutDB: update on development of tools for the biochemical analysis of genetic variation.
PMID 17827212 · PMC2238958 · Nucleic acids research · 2008 · 7 claims · 5 setups
MutDB integrates dbSNP and Swiss-Prot genetic variation data with protein structural information, functional disruption prediction scores, and clinical phenotype links (OMIM, dbGAP)
-
Full-text index only
Comparative genomic analysis of prion genes.
PMID 17199895 · PMC1781936 · BMC genomics · 2007 · 8 claims · 8 setups
SPRN and PRNP homologues are present in all vertebrates, whereas PRND is restricted to tetrapods and PRNT is restricted to primates
-
Full-text index only
Genome-wide analysis of human disease alleles reveals that their locations are correlated in paralogous proteins.
PMID 18989397 · PMC2565504 · PLoS computational biology · 2008 · 7 claims · 5 setups
The locations of sequence variants are correlated between paralogous human proteins more than expected by chance.
-
Full-text index only
In silico analysis of missense substitutions using sequence-alignment based methods.
PMID 18951440 · PMC3431198 · Human mutation · 2008 · 8 claims · 7 setups
Carefully validated PMSA-based computational algorithms can achieve predictive values of ~75-95% for classifying missense substitutions as pathogenic or neutral.
-
Full-text index only
Exonic remnants of whole-genome duplication reveal cis-regulatory function of coding exons.
PMID 19969543 · PMC2831330 · Nucleic acids research · 2010 · 8 claims · 8 setups
38 candidate cis-regulatory coding exons (RCEs) with predicted target genes were identified genome-wide
-
Full-text index only
Discovery of novel human transcript variants by analysis of intronic single-block EST with polyadenylation site.
PMID 19906316 · PMC2784480 · BMC genomics · 2009 · 8 claims · 7 setups
Intronic single-block ESTs with poly(A/T) tails reveal previously unidentified novel transcript variants missed by existing databases.
-
Full-text index only
EGenBio: a data management system for evolutionary genomics and biodiversity.
PMID 17118150 · PMC1683573 · BMC bioinformatics · 2006 · 7 claims · 7 setups
EGenBio is a web-based system for integrated management, filtering, curation, and visualization of large-scale genomic sequences, alignments, and phylogenetic trees for evolutionary genomics and biodiversity research.