Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The sequence and de novo assembly of the giant panda genome.
PMID 20010809 · PMC3951497 · Nature · 2010 · 8 claims · 8 setups
A draft giant panda genome was successfully generated and assembled de novo using only Illumina Genome Analyser short-read sequencing
-
Full-text index only
Genomic divergences among cattle, dog and human estimated from large-scale alignments of genomic sequences.
PMID 16759380 · PMC1525190 · BMC genomics · 2006 · 8 claims · 6 setups
Overall pairwise genomic divergences among cattle, dog and human are relatively constant (0.32–0.37 change/site)
-
Full-text index only
Improving the specificity of exon prediction using comparative genomics.
PMID 18831778 · PMC2559877 · BMC genomics · 2008 · 8 claims · 6 setups
A log-odds ratio scoring method based on codon conservation across human-mouse/human-dog alignments and adjacent-codon dependency can classify putative exons as coding vs non-coding.
-
Full-text index only
Using multiple alignments to improve seeded local alignment algorithms.
PMID 16100379 · PMC1185574 · Nucleic acids research · 2005 · 8 claims · 2 setups
Using information implicit in a multiple alignment to dynamically build a spaced-seed index weighted toward promising regions increases sensitivity of local alignment search compared to indexing a sequence alone
-
Full-text index only
Design factors that influence PCR amplification success of cross-species primers among 1147 mammalian primer pairs.
PMID 17029642 · PMC1635982 · BMC genomics · 2006 · 8 claims · 7 setups
The number of index-species (IS) mismatches in a primer pair significantly reduces amplification success, with an estimated 6-8% decrease in success rate per additional mismatch.
-
Full-text index only
Vertebrate gene finding from multiple-species alignments using a two-level strategy.
PMID 16925840 · PMC1810555 · Genome biology · 2006 · 8 claims · 5 setups
DOGFISH cleanly separates a multi-species alignment classifier (RVM cascade) from an HMM-based structure predictor, avoiding tight coupling of alignment complexity with HMM formalism
-
Full-text index only
Evola: Ortholog database of all human genes in H-InvDB with manual curation of phylogenetic trees.
PMID 17982176 · PMC2238928 · Nucleic acids research · 2008 · 6 claims · 7 setups
Evola combines genome synteny-based computational ortholog detection with manual curation of phylogenetic trees by experts to yield more reliable orthologs than automated pairwise methods
-
Full-text index only
A methodological framework for the reconstruction of contiguous regions of ancestral genomes and its application to mammalian genomes.
PMID 19043541 · PMC2580819 · PLoS computational biology · 2008 · 8 claims · 5 setups
A general model-free methodological framework is proposed for reconstructing Contiguous Ancestral Regions (CARs) from conserved syntenies, generalizing prior computational and cytogenetic approaches
-
Full-text index only
The other side of comparative genomics: genes with no orthologs between the cow and other mammalian species.
PMID 20003425 · PMC2808326 · BMC genomics · 2009 · 7 claims · 4 setups
3,801 bovine genes have no orthologs in human, mouse and dog, and 1,010 human genes have no orthologs in cow despite having orthologs in mouse and dog
-
Full-text index only
Developments in CORG: a gene-centric comparative genomics resource.
PMID 17135197 · PMC1751536 · Nucleic acids research · 2007 · 7 claims · 4 setups
CORG provides pairwise and multiple sequence alignments of upstream promoter regions and whole gene loci across 10 vertebrate species.
-
Full-text index only
The UCSC genome browser database: update 2007.
PMID 17142222 · PMC1669757 · Nucleic acids research · 2007 · 8 claims · 8 setups
The UCSC Genome Browser Database provides sequence and annotation data for 13 vertebrate and 19 invertebrate species as of September 2006.
-
Full-text index only
The UCSC Genome Browser Database: update 2006.
PMID 16381938 · PMC1347506 · Nucleic acids research · 2006 · 8 claims · 8 setups
The UCSC Genome Browser Database (GBD) provides integrated sequence and annotation data, with web tools (Genome Browser, Table Browser, Proteome Browser, Gene Sorter, BLAT, In Silico PCR) for visualizing and querying genomes of about a dozen vertebrate species and several model organisms.
-
Full-text index only
Dcode.org anthology of comparative genomic tools.
PMID 15980535 · PMC1160116 · Nucleic acids research · 2005 · 8 claims · 7 setups
The dcode.org suite (zPicture, Mulan, eShadow, rVista 2.0, multiTF, Creme 2.0, ECR Browser) provides integrated tools for comparative genomic analysis and non-coding regulatory element discovery.
-
Full-text index only
Phylogenomic approaches to common problems encountered in the analysis of low copy repeats: the sulfotransferase 1A gene family example.
PMID 15752422 · PMC555591 · BMC evolutionary biology · 2005 · 8 claims · 8 setups
A previously unidentified fourth human SULT1A gene (SULT1A4) exists on chromosome 16 and is transcriptionally active
-
Full-text index only
Computational approaches for predicting the biological effect of p53 missense mutations: a comparison of three sequence analysis based methods.
PMID 16522644 · PMC1390679 · Nucleic acids research · 2006 · 7 claims · 6 setups
Align-GVGD predicts loss of transactivation activity with high specificity (~88%) but lower sensitivity (67.9-71.2%) for neutral mutants
-
Has reproduction · 61
lncEvo: automated identification and conservation study of long noncoding RNAs.
PMID 33563213 · PMC7871587 · BMC bioinformatics · 2021 · 8 claims · 5 setups
lncEvo is an integrated Nextflow/Docker pipeline combining transcriptome assembly, lncRNA identification, and cross-species conservation analysis into a single workflow.
-
Full-text index only
Recurring genomic breaks in independent lineages support genomic fragility.
PMID 17090315 · PMC1636669 · BMC evolutionary biology · 2006 · 6 claims · 6 setups
The propensity of a chromosomal region to break is significantly correlated among independent lineages, even after accounting for covariates like region length and functional class.
-
Full-text index only
Genome-wide identification of human functional DNA using a neutral indel model.
PMID 16410828 · PMC1326222 · PLoS computational biology · 2006 · 8 claims · 8 setups
A neutral indel model predicting a geometric distribution of intergap segment (IGS) lengths fits human-mouse ancestral repeat (AR) alignment data excellently
-
Full-text index only
How accurately is ncRNA aligned within whole-genome multiple alignments?
PMID 17963514 · PMC2206062 · BMC bioinformatics · 2007 · 7 claims · 4 setups
MULTIZ does a fairly accurate job of aligning ncRNA regions across 17 vertebrate genomes, but better alignments exist in some regions.
-
Full-text index only
Multiple whole genome alignments and novel biomedical applications at the VISTA portal.
PMID 17488840 · PMC1933192 · Nucleic acids research · 2007 · 8 claims · 4 setups
A novel multiple whole-genome alignment algorithm treats all genomes symmetrically, avoiding dependence on a single base/reference genome