Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Continued colonization of the human genome by mitochondrial DNA.
PMID 15361937 · PMC515365 · PLoS biology · 2004 · 7 claims · 6 setups
NUMT insertion into nuclear chromosomes is an ongoing process shaped by double-strand-break repair (as shown in yeast) and continuing in humans.
-
Full-text index only
Sequence occurrence and structural uniqueness of a G-quadruplex in the human c-kit promoter.
PMID 17720713 · PMC2034477 · Nucleic acids research · 2007 · 8 claims · 4 setups
The native 22-nt c-kit87 sequence occurs only once in the entire human genome.
-
Full-text index only
Integrative analysis of the human cis-antisense gene pairs, miRNAs and their transcription regulation patterns.
PMID 19906709 · PMC2811022 · Nucleic acids research · 2010 · 8 claims · 5 setups
A genome-wide catalog of up to ~9000 overlapping antisense loci (23,782 non-redundant SAT pairs, clustered into 8894) was compiled and stored in the USAGP database
-
Full-text index only
Cryptic loxP sites in mammalian genomes: genome-wide distribution and relevance for the efficiency of BAC/PAC recombineering techniques.
PMID 17284462 · PMC1865043 · Nucleic acids research · 2007 · 6 claims · 6 setups
Cryptic lox P sites occur frequently and are homogeneously distributed across the mouse genome (1.2 primary sites per megabase).
-
Full-text index only
Stable in a genome of instability: an interview with Evan Eichler. Interview by Jane Gitschier.
PMID 18654618 · PMC2442658 · PLoS genetics · 2008 · 8 claims · 5 setups
Loss of AGG interruptions in CGG repeat tracts predisposes FMR1 alleles to instability and faster progression toward premutation/disease state
-
Full-text index only
Protein structure and function by the sea.
PMID 11983051 · PMC139342 · Genome biology · 2002 · 8 claims · 8 setups
High-throughput structural genomics (X-ray crystallography and NMR) can rapidly expand the number of solved protein structures far beyond what is currently in the Protein Data Bank.
-
Full-text index only
In vitro and in silico analysis reveals an efficient algorithm to predict the splicing consequences of mutations at the 5' splice sites.
PMID 17726045 · PMC2094079 · Nucleic acids research · 2007 · 8 claims · 6 setups
Two exonic mutations, PINK1 E417G and PARK7 E64D, disrupt binding to U1 snRNA and cause skipping of the mutation-harboring exon
-
Full-text index only
Modeling ChIP sequencing in silico with applications.
PMID 18725927 · PMC2507756 · PLoS computational biology · 2008 · 8 claims · 4 setups
Observed ChIP-seq tag counts follow an initial power-law distribution followed by a long right tail.
-
Full-text index only
Genome-wide detection of segmental duplications and potential assembly errors in the human genome sequence.
PMID 12702206 · PMC154576 · Genome biology · 2003 · 8 claims · 6 setups
Segmental duplications comprise 3.53% (107.4/3,043.1 Mb) of the June 2002 human genome assembly
-
Full-text index only
Protein kinases of the human malaria parasite Plasmodium falciparum: the kinome of a divergent eukaryote.
PMID 15479470 · PMC526369 · BMC genomics · 2004 · 8 claims · 4 setups
65 ePK sequences were identified in the P. falciparum genome and classified via phylogenetic analysis relative to the seven established ePK groups
-
Full-text index only
The relationship of potential G-quadruplex sequences in cis-upstream regions of the human genome to SP1-binding elements.
PMID 18353860 · PMC2377421 · Nucleic acids research · 2008 · 7 claims · 1 setups
A large number of upstream PQSSs incorporate the SP1-binding element, establishing a clear link between PQSS occurrence and SP1 elements
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.
-
Full-text index only
The most frequent short sequences in non-coding DNA.
PMID 19966278 · PMC2831315 · Nucleic acids research · 2010 · 8 claims · 2 setups
Short frequent sequences (9-14 bases) in non-coding DNA may play a role in maintaining chromosome structure and function
-
Full-text index only
Design and analysis issues in genome-wide somatic mutation studies of cancer.
PMID 18692126 · PMC2820387 · Genomics · 2009 · 6 claims · 4 setups
Two-stage (discovery + validation) sequencing designs efficiently allocate resources and can produce highly informative candidate driver gene lists even with relatively small sample sizes.
-
Full-text index only
Genomic signatures of human versus avian influenza A viruses.
PMID 17073083 · PMC3294750 · Emerging infectious diseases · 2006 · 8 claims · 6 setups
52 validated 'species-associated' amino acid positions distinguish human from avian influenza A viruses
-
Full-text index only
CompMoby: comparative MobyDick for detection of cis-regulatory motifs.
PMID 18950538 · PMC2605473 · BMC bioinformatics · 2008 · 7 claims · 4 setups
CompMoby identifies cis-regulatory binding sites at both transcriptional and post-transcriptional levels in metazoans without prior knowledge of the trans-acting factor
-
Has reproduction · 79
Enriched domain detector: a program for detection of wide genomic enrichment domains robust against local variations.
PMID 24782521 · PMC4066758 · Nucleic acids research · 2014 · 8 claims · 5 setups
EDD is a new algorithm that detects broad (megabase-size) enrichment domains from ChIP-seq data of widely distributed chromatin proteins such as A- and B-type lamins.
-
Full-text index only
Identifying related L1 retrotransposons by analyzing 3' transduced sequences.
PMID 12734010 · PMC156586 · Genome biology · 2003 · 8 claims · 6 setups
L1 elements with transduction-derived 3' sequence (L1-TDs) can be computationally identified using RepeatMasker/TSDfinder and grouped into families sharing a common progenitor via BLAST comparison of downstream sequences.
-
Full-text index only
A genome-wide survey of segmental duplications that mediate common human genetic variation of chromosomal architecture.
PMID 15588494 · PMC3525102 · Human genomics · 2004 · 8 claims · 5 setups
PSD-mediated genomic architecture analogous to the 8p23/4p16 inversion regions is not unique to those loci but recurs genome-wide.
-
Has reproduction · 89
Evolution of Highly Repetitive Silk Genes in the Luna Moth, Actias luna.
PMID 41738778 · PMC12962854 · Genome biology and evolution · 2026 · 8 claims · 8 setups
Eight sericin genes were identified in the Actias luna genome, including two clusters of closely related paralogs.