Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome-wide detection of segmental duplications and potential assembly errors in the human genome sequence.
PMID 12702206 · PMC154576 · Genome biology · 2003 · 8 claims · 6 setups
Segmental duplications comprise 3.53% (107.4/3,043.1 Mb) of the June 2002 human genome assembly
-
Full-text index only
Recent segmental and gene duplications in the mouse genome.
PMID 12914656 · PMC193640 · Genome biology · 2003 · 8 claims · 8 setups
33.6 Mb (1.2%) of the February 2003 mouse genome assembly (2,695 Mb) is involved in recent segmental duplications
-
Full-text index only
Microarray-based DNA methylation profiling: technology and applications.
PMID 16428248 · PMC1345696 · Nucleic acids research · 2006 · 7 claims · 6 setups
A microarray-based method enriching unmethylated and methylated DNA fractions via methylation-sensitive restriction enzymes followed by hybridization enables high-throughput DNA methylation profiling of large genomic regions.
-
Full-text index only
WebGestalt: an integrated system for exploring gene sets in various biological contexts.
PMID 15980575 · PMC1160236 · Nucleic acids research · 2005 · 8 claims · 6 setups
WebGestalt is an integrated web-based system composed of four modules: gene set management, information retrieval, organization/visualization, and statistics.
-
Full-text index only
Genome comparison without alignment using shortest unique substrings.
PMID 15910684 · PMC1166540 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A number of sequence comparison tasks, including detection of unique genomic regions, can be accomplished efficiently without an alignment step using shortest unique substrings.
-
Full-text index only
Comparative population genomics reveals convergent and divergent selection in the apricot-peach-plum-mei complex.
PMID 38883333 · PMC11179850 · Horticulture research · 2024 · 7 claims · 7 setups
A haplotype-resolved telomere-to-telomere (T2T) genome of plum (P. salicina cv. 'Fengtangli') was assembled into two gap-free haplotypes of 251.25 and 251.29 Mb.
-
Full-text index only
The DNA sequence of the human X chromosome.
PMID 15772651 · PMC2665286 · Nature · 2005 · 8 claims · 8 setups
The euchromatic sequence of the human X chromosome was determined to 99.3% completeness (~155 Mb total)
-
Has reproduction · 61
TEMP: a computational method for analyzing transposable element polymorphism in populations.
PMID 24753423 · PMC4066757 · Nucleic acids research · 2014 · 8 claims · 8 setups
TEMP combines pair-end (discordant) read and split (soft-clipped) read information to identify both presence and absence of TE insertions in genomic DNA from heterogeneous/pooled samples.
-
Full-text index only
Genome-scale validation of deep-sequencing libraries.
PMID 19002256 · PMC2577887 · PloS one · 2008 · 6 claims · 4 setups
Mab-seq allows a small aliquot of a ChIP-seq sequencing library to be labeled and hybridized to commercial microarrays for quality control before deep sequencing, without compromising the library for subsequent sequencing.
-
Full-text index only
Genome-wide identification of human functional DNA using a neutral indel model.
PMID 16410828 · PMC1326222 · PLoS computational biology · 2006 · 8 claims · 8 setups
A neutral indel model predicting a geometric distribution of intergap segment (IGS) lengths fits human-mouse ancestral repeat (AR) alignment data excellently
-
Full-text index only
A human genome-wide library of local phylogeny predictions for whole-genome inference problems.
PMID 18710563 · PMC2556685 · BMC genomics · 2008 · 7 claims · 5 setups
A genome-wide library of nearly 16 million local maximum parsimony phylogenies was constructed from HapMap CEU and YRI SNP data across all human autosomes
-
Full-text index only
Sequencing the regulatory genome.
PMID 18598374 · PMC2481419 · Genome biology · 2008 · 8 claims · 8 setups
Nuclear-lamina-associated domains (LADs) define chromatin regions with distinct transcriptional characteristics (fewer, lower-expressed genes, low RNA Pol II occupancy, H3K27me3-enriched borders)
-
Full-text index only
A genome-wide screen for noncoding elements important in primate evolution.
PMID 18215302 · PMC2242780 · BMC evolutionary biology · 2008 · 8 claims · 4 setups
A new likelihood ratio test (LRT) method, using nearby ancestral repeats to control for local mutation rate, can identify noncoding elements with lineage-specific accelerated substitution rates.
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.
-
Full-text index only
Genome-wide prioritization of disease genes and identification of disease-disease associations from an integrated human functional linkage network.
PMID 19728866 · PMC2768980 · Genome biology · 2009 · 6 claims · 6 setups
Integrating 16 genomic features (32 sub-features) via a naïve Bayes classifier produces a genome-scale FLN of 21,657 human genes and 22,388,609 weighted links that outperforms any individual data source for inferring functional linkages.
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.
-
Has reproduction · 73
Vespucci: a system for building annotated databases of nascent transcripts.
PMID 24304890 · PMC3936758 · Nucleic acids research · 2014 · 8 claims · 7 setups
Existing ChIP-seq and RNA-seq analysis platforms (e.g. Cufflinks, peak callers) are unsuited to GRO-seq because they assume spliced/exonic reads, uniform density and paired-end data, and cannot identify transcriptional units de novo across the whole genome.
-
Full-text index only
Genome-wide tracking of unmethylated DNA Alu repeats in normal and cancer cells.
PMID 18084025 · PMC2241897 · Nucleic acids research · 2008 · 5 claims · 7 setups
QUMA (quantitative real-time PCR) and AUMA (fingerprinting PCR) methods can quantify and individually identify unmethylated Alu elements on a genomic scale using the methylation-sensitive SmaI site as a surrogate marker