Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
SNPmasker: automatic masking of SNPs and repeats across eukaryotic genomes.
PMID 16845091 · PMC1538889 · Nucleic acids research · 2006 · 8 claims · 4 setups
SNPmasker is a web service combining SNP masking and repeat masking, supporting both coordinate-defined and homology-search-defined input regions, a combination not offered by prior tools
-
Full-text index only
A statistical approach designed for finding mathematically defined repeats in shotgun data and determining the length distribution of clone-inserts.
PMID 15626332 · PMC5172250 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 6 setups
Repeats of different copy number have distinct probabilities of appearance in shotgun data, which can be modeled statistically to define recognition thresholds (MDRs) at different shotgun coverages.
-
Full-text index only
Pigs in sequence space: a 0.66X coverage pig genome survey based on shotgun sequencing.
PMID 15885146 · PMC1142312 · BMC genomics · 2005 · 8 claims · 7 setups
Pig sequence is closer to human than mouse is, across exons, UTRs, introns, intergenic regions, ultra-conserved elements, and miRNAs
-
Full-text index only
Skittle: a 2-dimensional genome visualization tool.
PMID 20042093 · PMC2817707 · BMC bioinformatics · 2009 · 7 claims · 6 setups
Skittle is a 2D genome visualization tool combining a color-coded Nucleotide Display, a Repeat Map, a Repeat Overview, and an Alignment Cylinder to reveal genomic patterns at multiple scales
-
Full-text index only
Recent segmental and gene duplications in the mouse genome.
PMID 12914656 · PMC193640 · Genome biology · 2003 · 8 claims · 8 setups
33.6 Mb (1.2%) of the February 2003 mouse genome assembly (2,695 Mb) is involved in recent segmental duplications
-
Full-text index only
PigGIS: Pig Genomic Informatics System.
PMID 17090590 · PMC1669765 · Nucleic acids research · 2007 · 7 claims · 7 setups
PigGIS identified 15,700 pig consensus sequences covering 18.5 Mb of homologous human exons
-
Full-text index only
CEAS: cis-regulatory element annotation system.
PMID 16845068 · PMC1538818 · Nucleic acids research · 2006 · 7 claims · 5 setups
CEAS is the first web server to streamline genome-scale ChIP-chip downstream analyses for biologists without strong bioinformatics support
-
Full-text index only
Computational comparison of two mouse draft genomes and the human golden path.
PMID 12537546 · PMC151282 · Genome biology · 2003 · 8 claims · 7 setups
The Celera and public mouse genome assemblies differ in about 10% of the mouse genome, with complementary strengths (Celera higher base-pair accuracy and overall coverage; public assembly higher quality in some finished BAC regions and freely accessible)
-
Full-text index only
A genome-wide survey of segmental duplications that mediate common human genetic variation of chromosomal architecture.
PMID 15588494 · PMC3525102 · Human genomics · 2004 · 8 claims · 5 setups
PSD-mediated genomic architecture analogous to the 8p23/4p16 inversion regions is not unique to those loci but recurs genome-wide.
-
Has reproduction · 55
Genome-Wide Survey and Development of the First Microsatellite Markers Database (AnCorDB) in Anemone coronaria L.
PMID 35328546 · PMC8949970 · International journal of molecular sciences · 2022 · 8 claims · 8 setups
Generated the first draft genome assembly of A. coronaria by Illumina sequencing a haploid androgenetic plant
-
Full-text index only
Pegasys: software for executing and integrating analyses of biological sequences.
PMID 15096276 · PMC406494 · BMC bioinformatics · 2004 · 8 claims · 7 setups
Pegasys is a flexible, modular, customizable software system for executing and integrating heterogeneous biological sequence analysis tools
-
Full-text index only
The Vertebrate Genome Annotation (Vega) database.
PMID 15608237 · PMC540089 · Nucleic acids research · 2005 · 8 claims · 8 setups
Vega is a community database for browsing manual annotation of finished vertebrate genome sequences, based on an extended Ensembl-style schema.
-
Has reproduction · 92
Chromosome-scale genome sequencing, assembly and annotation of six genomes from subfamily Leishmaniinae.
PMID 34489462 · PMC8421402 · Scientific data · 2021 · 8 claims · 8 setups
Chromosome-scale genomes of six Leishmaniinae species (five L. (Mundinia) species and one Porcisia species) were sequenced, assembled and annotated, providing genome, proteome, transcriptome and GFF outputs for taxa previously lacking public reference genomes
-
Full-text index only
A clustering property of highly-degenerate transcription factor binding sites in the mammalian genome.
PMID 16670430 · PMC1456330 · Nucleic acids research · 2006 · 8 claims · 7 setups
Highly-degenerate RE1 sites are significantly enriched in promoters of validated and putative REST target genes compared to control promoters
-
Full-text index only
Comparative genomic study reveals a transition from TA richness in invertebrates to GC richness in vertebrates at CpG flanking sites: an indication for context-dependent mutagenicity of methylated CpG sites.
PMID 19329065 · PMC5054122 · Genomics, proteomics & bioinformatics · 2008 · 8 claims · 8 setups
Nucleotide preference at CpG flanking sites transitions from 5' T (invertebrates) to 5' A (vertebrates) at the invertebrate-vertebrate boundary
-
Full-text index only
Predicting failure rate of PCR in large genomes.
PMID 18492719 · PMC2441781 · Nucleic acids research · 2008 · 7 claims · 8 setups
The number of predicted primer-binding sites in genomic DNA is the most important factor determining PCR failure.
-
Has reproduction · 99
A haplotype-resolved genome assembly of the bocaccio rockfish, Sebastes paucispinis.
PMID 40323688 · PMC12584591 · The Journal of heredity · 2025 · 6 claims · 8 setups
This paper presents the first de novo, haplotype-resolved reference-quality genome assembly of Sebastes paucispinis (bocaccio rockfish).