Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
CRSD: a comprehensive web server for composite regulatory signature discovery.
PMID 16845073 · PMC1538777 · Nucleic acids research · 2006 · 7 claims · 5 setups
CRSD is a comprehensive web server integrating six large-scale databases (UniGene, mature microRNAs, putative promoter, TRANSFAC, pathway, GO) plus two newly constructed genome-wide databases (MRS and TRS) for composite regulatory signature discovery
-
Full-text index only
Spontaneous symmetry breaking in genome evolution.
PMID 18367477 · PMC2377439 · Nucleic acids research · 2008 · 6 claims · 3 setups
Exon size distributions in sequenced genomes follow a lognormal pattern typical of a random Kolmogoroff fractioning process
-
Full-text index only
Risk of pancreatic cancer in families with Lynch syndrome.
PMID 19861671 · PMC4091624 · JAMA · 2009 · 7 claims · 3 setups
Families with germline MMR gene mutations (Lynch Syndrome) have an 8.6-fold increased risk of pancreatic cancer compared to the general U.S. population.
-
Full-text index only
Finding signals that regulate alternative splicing in the post-genomic era.
PMID 12429065 · PMC244920 · Genome biology · 2002 · 8 claims · 8 setups
Alternative splicing generates protein and regulatory diversity from a limited number of genes and modulates isoform levels in a cell-context-specific manner
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
Random amino acid mutations and protein misfolding lead to Shannon limit in sequence-structure communication.
PMID 18769673 · PMC2518838 · PloS one · 2008 · 8 claims · 6 setups
The protein sequence-structure map behaves as a noisy digital communication channel whose capacity C exceeds the transmission rate R for native structures, satisfying Shannon's noisy channel theorem
-
Full-text index only
Identification and characterization of HLA-A*0301 epitopes in HIV-1 gag proteins using a novel approach.
PMID 19903485 · PMC2836169 · Journal of immunological methods · 2010 · 7 claims · 7 setups
PS mutations V7I and I34L (p17) and K403R (p7) in HIV-1 gag significantly correlate with HLA-A*0301
-
Full-text index only
Association of poly-purine/poly-pyrimidine sequences with meiotic recombination hot spots.
PMID 16846522 · PMC1543642 · BMC genomics · 2006 · 7 claims · 6 setups
PPT frequency is significantly elevated in yeast meiotic recombination hot spots compared with cold spots
-
Full-text index only
Empirical Bayes analysis of quantitative proteomics experiments.
PMID 19829701 · PMC2759080 · PloS one · 2009 · 8 claims · 4 setups
Developed a new empirical Bayes framework that models log2 SILAC protein ratios and is robust to non-Gaussian tails and data sparsity, unlike Gaussian mixture models or Efron's original spline-based approach
-
Full-text index only
A simple and efficient algorithm for genome-wide homozygosity analysis in disease.
PMID 19756043 · PMC2758715 · Molecular systems biology · 2009 · 8 claims · 4 setups
A genome-wide AH analysis (GAHA) algorithm can identify disease-associated loci by comparing frequencies of homozygous segments between cases and controls using a z-statistic proportion test
-
Full-text index only
A note on generalized Genome Scan Meta-Analysis statistics.
PMID 15717930 · PMC551600 · BMC bioinformatics · 2005 · 7 claims · 3 setups
An Edgeworth series approximation to the null distribution of the weighted GSMA statistic provides a more accurate representation than the normal approximation, especially in the tails
-
Full-text index only
Analysis of nucleotide diversity of NAT2 coding region reveals homogeneity across Native American populations and high intra-population diversity.
PMID 16847467 · PMC3099416 · The pharmacogenomics journal · 2007 · 8 claims · 6 setups
NAT2 variants are homogeneously distributed across native populations of the American continent
-
Full-text index only
Localized-statistical quantification of human serum proteome associated with type 2 diabetes.
PMID 18795103 · PMC2529402 · PloS one · 2008 · 8 claims · 5 setups
Developed LSPAD (localized statistics of protein abundance distribution) to calculate statistical significance of protein-abundance bias between two serum cohorts using a local Fisher's exact test window
-
Full-text index only
Inherited disorder phenotypes: controlled annotation and statistical analysis for knowledge mining from gene lists.
PMID 16351744 · PMC1866390 · BMC bioinformatics · 2005 · 5 claims · 3 setups
OMIM Clinical Synopsis free-text phenotype and location names can be normalized and hierarchically structured into a controlled vocabulary suitable for computational analysis
-
Full-text index only
CYCLONET--an integrated database on cell cycle regulation and carcinogenesis.
PMID 17202170 · PMC1899094 · Nucleic acids research · 2007 · 7 claims · 4 setups
Cyclonet is a web-based integrated database combining 'omics' and chemoinformatics data on mammalian cell cycle regulation in normal and pathological (cancer) states, built on a systems biology approach.
-
Full-text index only
A statistical approach designed for finding mathematically defined repeats in shotgun data and determining the length distribution of clone-inserts.
PMID 15626332 · PMC5172250 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 6 setups
Repeats of different copy number have distinct probabilities of appearance in shotgun data, which can be modeled statistically to define recognition thresholds (MDRs) at different shotgun coverages.
-
Full-text index only
Error-pooling-based statistical methods for identifying novel temporal replication profiles of human chromosomes observed by DNA tiling arrays.
PMID 17430969 · PMC1888820 · Nucleic acids research · 2007 · 8 claims · 4 setups
Developed an LPE-based error-pooling and weighted ANOVA modeling approach for statistical analysis of high-density tiling array data
-
Full-text index only
Genome-wide copy number profiling on high-density bacterial artificial chromosomes, single-nucleotide polymorphisms, and oligonucleotide microarrays: a platform comparison based on statistical power analysis.
PMID 17363414 · PMC2779891 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2007 · 8 claims · 6 setups
High-density oligonucleotide/SNP platforms are superior to the BAC platform for genome-wide detection of copy-number variations smaller than 1 Mb
-
Full-text index only
Analysis of variants in DNA damage signalling genes in bladder cancer.
PMID 18638378 · PMC2488326 · BMC medical genetics · 2008 · 7 claims · 5 setups
SNPs in DSB signalling genes may modulate predisposition to bladder cancer and influence effects of environmental exposures
-
Full-text index only
Analysis of virulence factors of Helicobacter pylori isolated from a Vietnamese population.
PMID 19698173 · PMC2739534 · BMC microbiology · 2009 · 8 claims · 5 setups
Three distinct deletion patterns (39-bp, 18-bp, no deletion) exist upstream of the cagA EPIYA repeat region (pre-EPIYA region), providing a novel genotyping marker.