Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Defective splicing, disease and therapy: searching for master checkpoints in exon definition.
PMID 16855287 · PMC1524908 · Nucleic acids research · 2006 · 8 claims · 8 setups
Splicing-affecting genomic variations can account for up to 50% of mutations leading to gene dysfunction in some genes
-
Full-text index only
An oncogenomics-based in vivo RNAi screen identifies tumor suppressors in liver cancer.
PMID 19012953 · PMC2990916 · Cell · 2008 · 7 claims · 8 setups
shRNA pools targeting genes recurrently deleted in human HCC accelerate hepatocarcinogenesis in vivo, whereas randomly selected shRNA pools do not.
-
Full-text index only
Random amino acid mutations and protein misfolding lead to Shannon limit in sequence-structure communication.
PMID 18769673 · PMC2518838 · PloS one · 2008 · 8 claims · 6 setups
The protein sequence-structure map behaves as a noisy digital communication channel whose capacity C exceeds the transmission rate R for native structures, satisfying Shannon's noisy channel theorem
-
Full-text index only
Evolutionary distance estimation and fidelity of pair wise sequence alignment.
PMID 15840174 · PMC1087827 · BMC bioinformatics · 2005 · 8 claims · 8 setups
Evolutionary distance estimation is relatively unaffected by alignment error as long as 50% or more of homologous sites remain identical between sequences
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Full-text index only
The whole alignment and nothing but the alignment: the problem of spurious alignment flanks.
PMID 18796526 · PMC2566872 · Nucleic acids research · 2008 · 8 claims · 4 setups
Some common scoring schemes tend to overextend alignments, generating spurious alignment flanks up to hundreds of bp/amino acids in length
-
Full-text index only
A surrogate-based approach for post-genomic partner identification.
PMID 11602024 · PMC57814 · BMC biotechnology · 2001 · 8 claims · 5 setups
Peptide surrogates derived from random phage display libraries contain amino acid sequence information that identifies the natural biological partner of the panned target via database searching.
-
Full-text index only
Comparative genomics of Drosophila and human core promoters.
PMID 16827941 · PMC1779564 · Genome biology · 2006 · 8 claims · 6 setups
Drosophila core promoters contain 298 highly significant (p≤1e-16) non-randomly positioned 8-mers within 100 bp of the TSS, grouped into 15 distinct DNA motifs
-
Full-text index only
The fragile breakage versus random breakage models of chromosome evolution.
PMID 16501665 · PMC1378107 · PLoS computational biology · 2006 · 8 claims · 6 setups
Sankoff and Trinh's synteny block identification algorithm (ST-Synteny) is flawed, producing erroneous block identifications even in small toy examples.
-
Full-text index only
The biological function of some human transcription factor binding motifs varies with position relative to the transcription start site.
PMID 18367472 · PMC2377430 · Nucleic acids research · 2008 · 8 claims · 5 setups
1226 eight-letter DNA words show statistically significant positional preferences relative to the TSS across 7914 human promoter regions
-
Has reproduction · 68
Molecular subtype of recurrent implantation failure reveals distinct endometrial etiology of female infertility.
PMID 40660214 · PMC12257665 · Journal of translational medicine · 2025 · 8 claims · 8 setups
RIF endometrial samples segregate into two reproducible molecular subtypes: an immune-driven subtype (RIF-I) and a metabolic-driven subtype (RIF-M)
-
Full-text index only
Genomic divergences among cattle, dog and human estimated from large-scale alignments of genomic sequences.
PMID 16759380 · PMC1525190 · BMC genomics · 2006 · 8 claims · 6 setups
Overall pairwise genomic divergences among cattle, dog and human are relatively constant (0.32–0.37 change/site)
-
Has reproduction · 77
Electroacupuncture reshapes the microbial co-occurrence networks related to the behavioral and psychological symptoms of dementia in Alzheimer's disease.
PMID 41676443 · PMC12806058 · iMetaOmics · 2025 · 6 claims · 7 setups
Electroacupuncture reshapes microbial co-occurrence network topology and drives keystone species in AD-related BPSD, with R. gnavus emerging as a likely keystone species post-intervention.
-
Full-text index only
Predicting preferential DNA vector insertion sites: implications for functional genomics and gene therapy.
PMID 18047689 · PMC2106846 · Genome biology · 2007 · 8 claims · 6 setups
Vector insertion site preferences differ substantially between viral vectors and transposons, affecting both oncogenic risk in gene therapy and utility for functional genomics
-
Full-text index only
Widespread dysregulation of MiRNAs by MYCN amplification and chromosomal imbalances in neuroblastoma: association of miRNA expression with survival.
PMID 19924232 · PMC2773120 · PloS one · 2009 · 8 claims · 5 setups
37 miRNAs are significantly differentially expressed between MYCN-amplified (MNA) and non-MNA neuroblastoma tumors, suggesting direct or indirect regulation by MYCN
-
Has reproduction · 53
Blood RNA signature RISK4LEP predicts leprosy years before clinical onset.
PMID 34090257 · PMC8182229 · EBioMedicine · 2021 · 7 claims · 5 setups
A 4-gene blood RNA signature (RISK4LEP: MT-ND2, REX1BD, TPGS1, UBC) predicts leprosy development 4–61 months before clinical diagnosis
-
Full-text index only
Optimal step length EM algorithm (OSLEM) for the estimation of haplotype frequency and its application in lipoprotein lipase genotyping.
PMID 12529185 · PMC149347 · BMC bioinformatics · 2003 · 5 claims · 4 setups
OSLEM (Optimal Step Length EM), which approximates an optimal step length via a fixed-point search (D_N = D_{N-1} + λ(D_preN - D_{N-1})), runs about twice as fast as standard EM while producing the same haplotype frequency estimates.
-
Full-text index only
The distribution of SNPs in human gene regulatory regions.
PMID 16209714 · PMC1260019 · BMC genomics · 2005 · 8 claims · 6 setups
SNPs occur with higher density closer to the transcriptional start site within gene promoter regions than in further upstream regions
-
Full-text index only
Selection of target sites for mobile DNA integration in the human genome.
PMID 17166054 · PMC1664696 · PLoS computational biology · 2006 · 8 claims · 8 setups
A comprehensive bioinformatic method was developed to annotate every base pair in the human genome for its likelihood of hosting integration by each of seven mobile DNA elements, using >200 genomic feature variables.
-
Full-text index only
The impact of peptide abundance and dynamic range on stable-isotope-based quantitative proteomic analyses.
PMID 18798661 · PMC2746028 · Journal of proteome research · 2008 · 8 claims · 7 setups
Over half of confidently identified peptides in complex mixtures have S/N ratios below 10 on both FT-ICR and Orbitrap instruments