Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
EPD in its twentieth year: towards complete promoter coverage of selected model organisms.
PMID 16381980 · PMC1347508 · Nucleic acids research · 2006 · 7 claims · 4 setups
EPD is an annotated, non-redundant collection of experimentally defined eukaryotic POL II promoters accessed via genome position pointers.
-
Full-text index only
A clustering property of highly-degenerate transcription factor binding sites in the mammalian genome.
PMID 16670430 · PMC1456330 · Nucleic acids research · 2006 · 8 claims · 7 setups
Highly-degenerate RE1 sites are significantly enriched in promoters of validated and putative REST target genes compared to control promoters
-
Full-text index only
Promoting human promoters.
PMID 16760901 · PMC1681504 · Molecular systems biology · 2006 · 8 claims · 5 setups
A computational protocol using MARS on known sequence motifs can quantify transcription factor effects on human gene expression despite noise and data size challenges
-
Full-text index only
NetworKIN: a resource for exploring cellular phosphorylation networks.
PMID 17981841 · PMC2238868 · Nucleic acids research · 2008 · 8 claims · 4 setups
NetworKIN integrates consensus substrate motifs with probabilistic network context modelling to predict cellular kinase-substrate relations.
-
Full-text index only
Correlating novel variable and conserved motifs in the Hemagglutinin protein with significant biological functions.
PMID 18681973 · PMC2553082 · Virology journal · 2008 · 8 claims · 6 setups
14 MEME blocks were identified in the HA protein of H3N2 strains (1968-1999), with blocks 1, 2, 3, and 7 correlating with several biological functions
-
Full-text index only
The biological function of some human transcription factor binding motifs varies with position relative to the transcription start site.
PMID 18367472 · PMC2377430 · Nucleic acids research · 2008 · 8 claims · 5 setups
1226 eight-letter DNA words show statistically significant positional preferences relative to the TSS across 7914 human promoter regions
-
Has reproduction · 62
Application of alternative de novo motif recognition models for analysis of structural heterogeneity of transcription factor binding sites: a case study of FOXA2 binding sites.
PMID 34547062 · PMC8408018 · Vavilovskii zhurnal genetiki i selektsii · 2021 · 6 claims · 7 setups
Combining four de novo models (PWM, diPWM, BaMM, InMoDe) significantly increases the fraction of recognized peaks versus PWM alone (by 26.3%).
-
Full-text index only
CoMoDis: composite motif discovery in mammalian genomes.
PMID 17130158 · PMC1702496 · Nucleic acids research · 2007 · 7 claims · 4 setups
CoMoDis is a new bioinformatics tool that streamlines computational identification of novel regulatory modules starting from a single seed motif
-
Full-text index only
Divergence of exonic splicing elements after gene duplication and the impact on gene structures.
PMID 19883501 · PMC3091315 · Genome biology · 2009 · 8 claims · 7 setups
ESEs and ESSs diverge especially fast shortly after gene duplication, correlating with time since duplication (Ks)
-
Full-text index only
What makes species unique? The contribution of proteins with obscure features.
PMID 16859532 · PMC1779552 · Genome biology · 2006 · 7 claims · 8 setups
POFs constitute 18-38% (average 26%) of a typical eukaryotic proteome
-
Full-text index only
High-throughput chromatin information enables accurate tissue-specific prediction of transcription factor binding sites.
PMID 18988630 · PMC2662491 · Nucleic acids research · 2009 · 8 claims · 8 setups
Incorporating H3K4me3 chromatin modification estimates greatly improves the accuracy of in silico prediction of in vivo TF binding for a wide range of TFs in human and mouse
-
Full-text index only
G-quadruplexes: the beginning and end of UTRs.
PMID 18832370 · PMC2577360 · Nucleic acids research · 2008 · 8 claims · 5 setups
UTRs show significant strand asymmetry with C-PQS more common than G-PQS, consistent with general depletion of G-quadruplex-forming RNA
-
Full-text index only
Molecular evolution of Cide family proteins: novel domain formation in early vertebrates and the subsequent divergence.
PMID 18500987 · PMC2426694 · BMC evolutionary biology · 2008 · 8 claims · 5 setups
Sequences homologous to the CIDE-N domain/NCD show a wide phylogenetic distribution, from hydra and sea anemone to mammals, while true Cide proteins are restricted to vertebrates.
-
Full-text index only
The i-motif in the bcl-2 P1 promoter forms an unexpectedly stable structure with a unique 8:5:7 loop folding pattern.
PMID 19908860 · PMC2787777 · Journal of the American Chemical Society · 2009 · 8 claims · 6 setups
The full-length bcl-2 C-rich promoter sequence (Py39WT) forms one major intramolecular i-motif structure with a transitional pH of 6.6
-
Has reproduction · 61
TEMP: a computational method for analyzing transposable element polymorphism in populations.
PMID 24753423 · PMC4066757 · Nucleic acids research · 2014 · 8 claims · 8 setups
TEMP combines pair-end (discordant) read and split (soft-clipped) read information to identify both presence and absence of TE insertions in genomic DNA from heterogeneous/pooled samples.
-
Has reproduction · 88
Transcriptomic Data Meta-Analysis Sheds Light on High Light Response in Arabidopsis thaliana L.
PMID 35457273 · PMC9026532 · International journal of molecular sciences · 2022 · 7 claims · 6 setups
Meta-analysis of five transcriptomic experiments identified 1151 differentially expressed genes that compose a coordinated gene network responding to high light stress
-
Full-text index only
A distinct epigenetic signature at targets of a leukemia protein.
PMID 17266773 · PMC1796549 · BMC genomics · 2007 · 7 claims · 7 setups
Combining gene expression microarray analysis with bioinformatic search for AML1-consensus sequences identifies direct AML1 targets that expression analysis alone cannot resolve
-
Full-text index only
Conserved elements with potential to form polymorphic G-quadruplex structures in the first intron of human genes.
PMID 18187510 · PMC2275096 · Nucleic acids research · 2008 · 8 claims · 6 setups
G-richness downstream of the TSS is strand-biased, concentrated on the nontemplate strand, with a peak at +200 to +300 bp
-
Has reproduction · 80
PanglaoDB: a web server for exploration of mouse and human single-cell RNA sequencing data.
PMID 30951143 · PMC6450036 · Database : the journal of biological databases and curation · 2019 · 7 claims · 7 setups
PanglaoDB is a web server providing pre-processed and pre-computed analyses of >1054 single-cell experiments (>4 million cells) from mouse and human across many tissues and platforms.
-
Has reproduction · 84
Integrative Transcriptomic and Evolutionary Analysis of Drought and Heat Stress Responses in Solanum tuberosum and Solanum lycopersicum.
PMID 41470732 · PMC12736803 · Plants (Basel, Switzerland) · 2025 · 7 claims · 8 setups
Drought and heat stress induce coordinated transcriptional reprogramming in potato and tomato: induction of molecular chaperone activity, oxidative stress responses, and immune signaling, with repression of photosynthetic and primary metabolic pathways reflecting energy reallocation.