Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The biological function of some human transcription factor binding motifs varies with position relative to the transcription start site.
PMID 18367472 · PMC2377430 · Nucleic acids research · 2008 · 8 claims · 5 setups
1226 eight-letter DNA words show statistically significant positional preferences relative to the TSS across 7914 human promoter regions
-
Full-text index only
SNP@Promoter: a database of human SNPs (single nucleotide polymorphisms) within the putative promoter regions.
PMID 18315851 · PMC2259403 · BMC bioinformatics · 2008 · 8 claims · 4 setups
SNP@Promoter is a database of human SNPs within putative promoter regions and predicted transcription factor binding sites
-
Full-text index only
Finding signals that regulate alternative splicing in the post-genomic era.
PMID 12429065 · PMC244920 · Genome biology · 2002 · 8 claims · 8 setups
Alternative splicing generates protein and regulatory diversity from a limited number of genes and modulates isoform levels in a cell-context-specific manner
-
Full-text index only
Predicting positive p53 cancer rescue regions using Most Informative Positive (MIP) active learning.
PMID 19756158 · PMC2742196 · PLoS computational biology · 2009 · 8 claims · 4 setups
MIP active learning is a novel active learning method that preferentially seeks informative Positive (functionally active) examples rather than only maximizing classifier accuracy.
-
Full-text index only
Improving the specificity of exon prediction using comparative genomics.
PMID 18831778 · PMC2559877 · BMC genomics · 2008 · 8 claims · 6 setups
A log-odds ratio scoring method based on codon conservation across human-mouse/human-dog alignments and adjacent-codon dependency can classify putative exons as coding vs non-coding.
-
Full-text index only
ARED 3.0: the large and diverse AU-rich transcriptome.
PMID 16381826 · PMC1347415 · Nucleic acids research · 2006 · 7 claims · 6 setups
ARED 3.0 computationally mapped more than 4000 ARE-mRNAs to the human genome, representing 5-8% of human genes.
-
Full-text index only
QuadBase: genome-wide database of G4 DNA--occurrence and conservation in human, chimpanzee, mouse and rat promoters and 146 microbes.
PMID 17962308 · PMC2238983 · Nucleic acids research · 2008 · 8 claims · 3 setups
QuadBase is a compendium of G4 DNA (quadruplex) motifs focused on their occurrence and conservation in promoters, composed of EuQuad and ProQuad
-
Full-text index only
Evidence for a preferential targeting of 3'-UTRs by cis-encoded natural antisense transcripts.
PMID 16204454 · PMC1243798 · Nucleic acids research · 2005 · 8 claims · 4 setups
Cis-encoded natural antisense RNAs show striking preferential complementarity to 3′-UTRs of their target genes in human and mouse genomes
-
Full-text index only
Paired-end mapping reveals extensive structural variation in the human genome.
PMID 17901297 · PMC2674581 · Science (New York, N.Y.) · 2007 · 8 claims · 8 setups
Paired-end mapping (PEM) combining 3-kb fragment paired-end capture, massive 454 sequencing, and computational mapping detects SVs ~3 kb or larger with an average breakpoint resolution of 644 bp
-
Full-text index only
Identification and analysis of co-occurrence networks with NetCutter.
PMID 18781200 · PMC2526157 · PloS one · 2008 · 8 claims · 4 setups
Random sampling from a complete permutation set of the bipartite graph permits co-occurrence analysis with optimal stringency, and the edge-swapping (ES) model closely approximates this and is the preferred null-model among six tested.
-
Full-text index only
The Princeton Protein Orthology Database (P-POD): a comparative genomics analysis tool for biologists.
PMID 17712414 · PMC1942082 · PloS one · 2007 · 8 claims · 5 setups
P-POD is the first comparative genomics database to combine results from multiple computational ortholog/homolog prediction methods with manually curated literature-derived experimental evidence of functional conservation.
-
Has reproduction · 64
Starvation-induced transgenerational inheritance of small RNAs in C. elegans.
PMID 25018105 · PMC4377509 · Cell · 2014 · 8 claims · 7 setups
L1 starvation induces changes in endogenous 22G small RNAs (STGs) that are inherited for at least three generations in fed descendants.
-
Full-text index only
Shotgun haplotyping: a novel method for surveying allelic sequence variation.
PMID 16221968 · PMC1253838 · Nucleic acids research · 2005 · 8 claims · 7 setups
A novel shotgun haplotyping method generates haplotypic sequences from long PCR products by shotgun sequencing both alleles concurrently and using read-pair information to separate alleles during assembly
-
Full-text index only
Screening for microsatellite instability identifies frequent 3'-untranslated region mutation of the RB1-inducible coiled-coil 1 gene in colon tumors.
PMID 19888451 · PMC2766054 · PloS one · 2009 · 7 claims · 4 setups
Somatic mutation frequency (%MSI) of 3'UTR microsatellites in MSI-H colorectal tumors correlates significantly with microsatellite length (r=0.86, p=7.2×10−13), following an exponential growth model.
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.