Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Sushi gets serious: the draft genome sequence of the pufferfish Fugu rubripes.
PMID 12225591 · PMC139409 · Genome biology · 2002 · 8 claims · 7 setups
The Fugu rubripes draft genome sequence was generated by whole-genome shotgun sequencing assembled to ~5.6x coverage using the JAZZ pipeline.
-
Full-text index only
Genomic organization of zebrafish microRNAs.
PMID 18510755 · PMC2427041 · BMC genomics · 2008 · 8 claims · 6 setups
Using sequence conservation and prediction algorithms, 35 new zebrafish miRNAs were identified, bringing the total to 415.
-
Full-text index only
DNA sequence and analysis of human chromosome 9.
PMID 15164053 · PMC2734081 · Nature · 2004 · 8 claims · 8 setups
The finished euchromatic sequence of chromosome 9 comprises 109,044,351 base pairs, representing >99.6% of the region.
-
Full-text index only
The DNA sequence and analysis of human chromosome 13.
PMID 15057823 · PMC2665288 · Nature · 2004 · 8 claims · 8 setups
95.5 Mb of finished sequence from chromosome 13 was completed, containing 633 genes and 296 pseudogenes.
-
Full-text index only
Differences in the evolutionary history of disease genes affected by dominant or recessive mutations.
PMID 16817963 · PMC1534034 · BMC genomics · 2006 · 8 claims · 8 setups
Dominant disease genes are more conserved at the protein level (mouse orthologues) than recessive disease genes.
-
Full-text index only
Target SNP selection in complex disease association studies.
PMID 15248903 · PMC487897 · BMC bioinformatics · 2004 · 7 claims · 3 setups
A computational pipeline can retrieve gene sequence, collect SNP variation data, and annotate SNPs falling in functional motifs (promoter, exon-intron structure, AU-rich elements, TF binding sites, splice sites) with expression in target tissue
-
Full-text index only
Dcode.org anthology of comparative genomic tools.
PMID 15980535 · PMC1160116 · Nucleic acids research · 2005 · 8 claims · 7 setups
The dcode.org suite (zPicture, Mulan, eShadow, rVista 2.0, multiTF, Creme 2.0, ECR Browser) provides integrated tools for comparative genomic analysis and non-coding regulatory element discovery.
-
Full-text index only
DG-CST (Disease Gene Conserved Sequence Tags), a database of human-mouse conserved elements associated to disease genes.
PMID 15608249 · PMC539965 · Nucleic acids research · 2005 · 5 claims · 8 setups
Comparative human-mouse genome analysis identifies conserved sequence tags (CSTs, >=70% identity over >=100bp) that frequently correspond to non-coding elements with putative regulatory or structural roles
-
Full-text index only
Molecular phylogeny of the antiangiogenic and neurotrophic serpin, pigment epithelium derived factor in vertebrates.
PMID 17020603 · PMC1609119 · BMC genomics · 2006 · 8 claims · 8 setups
A single PEDF gene is present in all examined vertebrate species but is absent from invertebrates (D. melanogaster, C. elegans, C. intestinalis)
-
Full-text index only
Vertebrate gene finding from multiple-species alignments using a two-level strategy.
PMID 16925840 · PMC1810555 · Genome biology · 2006 · 8 claims · 5 setups
DOGFISH cleanly separates a multi-species alignment classifier (RVM cascade) from an HMM-based structure predictor, avoiding tight coupling of alignment complexity with HMM formalism
-
Full-text index only
The whole alignment and nothing but the alignment: the problem of spurious alignment flanks.
PMID 18796526 · PMC2566872 · Nucleic acids research · 2008 · 8 claims · 4 setups
Some common scoring schemes tend to overextend alignments, generating spurious alignment flanks up to hundreds of bp/amino acids in length
-
Full-text index only
Molecular evolution of Cide family proteins: novel domain formation in early vertebrates and the subsequent divergence.
PMID 18500987 · PMC2426694 · BMC evolutionary biology · 2008 · 8 claims · 5 setups
Sequences homologous to the CIDE-N domain/NCD show a wide phylogenetic distribution, from hydra and sea anemone to mammals, while true Cide proteins are restricted to vertebrates.
-
Full-text index only
SVC: structured visualization of evolutionary sequence conservation.
PMID 15991338 · PMC1160265 · Nucleic acids research · 2005 · 7 claims · 5 setups
SVC aligns protein-coding sequences of orthologous gene pairs and maps them back onto their encoding exons/introns to generate a scaffold of conserved gene structure.
-
Full-text index only
The DNA sequence of the human X chromosome.
PMID 15772651 · PMC2665286 · Nature · 2005 · 8 claims · 8 setups
The euchromatic sequence of the human X chromosome was determined to 99.3% completeness (~155 Mb total)
-
Full-text index only
How accurately is ncRNA aligned within whole-genome multiple alignments?
PMID 17963514 · PMC2206062 · BMC bioinformatics · 2007 · 7 claims · 4 setups
MULTIZ does a fairly accurate job of aligning ncRNA regions across 17 vertebrate genomes, but better alignments exist in some regions.
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
Adapting to a changing world: RAG genomics and evolution.
PMID 16004728 · PMC3525258 · Human genomics · 2005 · 8 claims · 7 setups
RAG-1/RAG-2 origin is a foundational hallmark of adaptive immunity, enabling V(D)J recombination of antigen receptor genes.
-
Full-text index only
Comparative mapping of expressed sequence tags containing microsatellites in rainbow trout (Oncorhynchus mykiss).
PMID 15836796 · PMC1090573 · BMC genomics · 2005 · 8 claims · 7 setups
89 polymorphic microsatellite markers were developed from rainbow trout EST-derived cDNA clones
-
Full-text index only
Combining comparative genomics with de novo motif discovery to identify human transcription factor DNA-binding motifs.
PMID 17217514 · PMC1780116 · BMC bioinformatics · 2006 · 6 claims · 4 setups
A novel method combining 8-species comparative genomics with de novo motif discovery identifies human TF DNA-binding motifs overrepresented and conserved in upstream regions of co-regulated genes
-
Full-text index only
VISTA Enhancer Browser--a database of tissue-specific human enhancers.
PMID 17130149 · PMC1716724 · Nucleic acids research · 2007 · 8 claims · 2 setups
Comparative genome analysis can identify candidate human enhancer elements whose tissue-specific in vivo activity can then be experimentally validated in transgenic mice.