Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Information-theoretic identification of predictive SNPs and supervised visualization of genome-wide association studies.
PMID 16899448 · PMC1557808 · Nucleic acids research · 2006 · 7 claims · 4 setups
3D VizStruct (DFT-based radial mapping + KLD as z-axis) can identify SNPs/polymorphic markers that are predictive of underlying biological class distinctions across diverse datasets
-
Full-text index only
BTW: a web server for Boltzmann time warping of gene expression time series.
PMID 16845055 · PMC1538860 · Nucleic acids research · 2006 · 5 claims · 4 setups
Symmetric time warping distance is more flexible than Euclidean distance or correlation coefficient for identifying genes with similar temporal expression profiles, especially across sequences of different length.
-
Full-text index only
TreeFam: a curated database of phylogenetic trees of animal gene families.
PMID 16381935 · PMC1347480 · Nucleic acids research · 2006 · 7 claims · 6 setups
Tree-based inference of orthologs and paralogs is more robust than BLAST-based methods because evolutionary rates (and thus pairwise BLAST scores) vary across gene family members
-
Full-text index only
Broad network-based predictability of Saccharomyces cerevisiae gene loss-of-function phenotypes.
PMID 18053250 · PMC2246260 · Genome biology · 2007 · 8 claims · 4 setups
Loss-of-function phenotypes in yeast are predictable from a gene's connections in a functional gene network via guilt-by-association.
-
Full-text index only
Direct maximum parsimony phylogeny reconstruction from genotype data.
PMID 18053244 · PMC2222657 · BMC bioinformatics · 2007 · 6 claims · 4 setups
The paper presents the first practical method for computing maximum parsimony phylogenies directly from genotype data, using integer linear programming.
-
Full-text index only
Statistical learning of peptide retention behavior in chromatographic separations: a new kernel-based approach for computational proteomics.
PMID 18053132 · PMC2254445 · BMC bioinformatics · 2007 · 6 claims · 5 setups
The paired oligo-border kernel (POBK) combined with SVMs predicts peptide adsorption/elution in SAX-SPE and retention time in IP-RP-HPLC more accurately than existing methods.
-
Full-text index only
Inconsistencies in Neanderthal genomic DNA sequences.
PMID 17937503 · PMC2014787 · PLoS genetics · 2007 · 8 claims · 6 setups
The Noonan et al. and Green et al. Neanderthal nuclear DNA datasets yield mutually inconsistent estimates of population split time and Neanderthal admixture proportion when analyzed with the same method
-
Full-text index only
Unravelling the hidden heterogeneities of diffuse large B-cell lymphoma based on coupled two-way clustering.
PMID 17888167 · PMC2082044 · BMC genomics · 2007 · 8 claims · 6 setups
A proposed coupled two-way clustering (CTWC/SPC) method combined with a GO-based functional concept consistency score can identify compact, robust gene subsets that define clinically meaningful DLBCL subtypes
-
Full-text index only
Transcriptome annotation using tandem SAGE tags.
PMID 17709346 · PMC2034470 · Nucleic acids research · 2007 · 8 claims · 7 setups
A novel algorithm pairs tandem SAGE tags anchored on two different restriction sites (CATG and GATC) to define tag-delimited genomic sequences (TDGS)
-
Full-text index only
g:Profiler--a web-based toolset for functional profiling of gene lists from large-scale experiments.
PMID 17478515 · PMC1933153 · Nucleic acids research · 2007 · 8 claims · 5 setups
g:Profiler integrates four modules (g:Profiler core, g:Convert, g:Orth, g:Sorter) into a single cross-linked web tool for gene list analysis
-
Full-text index only
Exploiting noise in array CGH data to improve detection of DNA copy number change.
PMID 17272296 · PMC1994778 · Nucleic acids research · 2007 · 7 claims · 4 setups
When aberrations are present, noise in BAC, 19k oligo, and 385k oligo array-CGH data is highly non-Gaussian and shows long-range spatial correlations.
-
Full-text index only
The genomic analysis of lactic acidosis and acidosis response in human cancers.
PMID 19057672 · PMC2585811 · PLoS genetics · 2008 · 8 claims · 8 setups
Lactic acidosis and hypoxia induce largely distinct gene expression programs in HMECs, with lactic acidosis producing a much larger and more dramatic response
-
Full-text index only
Sys-BodyFluid: a systematical database for human body fluid proteome research.
PMID 18978022 · PMC2686600 · Nucleic acids research · 2009 · 6 claims · 4 setups
Sys-BodyFluid is a web-based database integrating proteomic data from 11 human body fluids (plasma/serum, urine, cerebrospinal fluid, saliva, bronchoalveolar lavage fluid, synovial fluid, nipple aspirate fluid, tear fluid, seminal fluid, milk, amniotic fluid), containing over 10,000 proteins
-
Full-text index only
The proteogenomic path towards biomarker discovery.
PMID 18764911 · PMC2574627 · Pediatric transplantation · 2008 · 8 claims · 8 setups
Serum creatinine is a widely used but non-ideal biomarker for renal transplant monitoring because it lacks specificity and sensitivity for graft injury
-
Full-text index only
"Reverse ecology" and the power of population genomics.
PMID 18752601 · PMC2626434 · Evolution; international journal of organic evolution · 2008 · 8 claims · 7 setups
Population genomic data can be used to rapidly identify genes targeted by adaptive natural selection, an approach termed 'reverse ecology'.
-
Full-text index only
Genomic and proteomic profiling of oxidative stress response in human diploid fibroblasts.
PMID 18654835 · PMC4973579 · Biogerontology · 2009 · 8 claims · 7 setups
Treatment of HCA3 human diploid fibroblasts with a mild dose of H2O2 (600 μM) induces premature senescence without killing the cells.
-
Full-text index only
Simultaneous analysis of all SNPs in genome-wide and re-sequencing association studies.
PMID 18654633 · PMC2464715 · PLoS genetics · 2008 · 8 claims · 5 setups
A Bayesian-inspired penalised maximum likelihood stochastic search method can simultaneously analyse all SNPs (up to 500K) from a GWA study in a few hours on a desktop workstation
-
Full-text index only
Bayesian survival analysis in genetic association studies.
PMID 18617538 · PMC2530885 · Bioinformatics (Oxford, England) · 2008 · 7 claims · 5 setups
A novel Bayesian method (BETA-Surv) extends prior case-control haplotype-clustering work to censored survival outcomes by clustering haplotypes via gene tree/perfect phylogeny topology and relative mutation age.
-
Full-text index only
High accuracy mass spectrometry analysis as a tool to verify and improve gene annotation using Mycobacterium tuberculosis as an example.
PMID 18597682 · PMC2483986 · BMC genomics · 2008 · 8 claims · 5 setups
High-accuracy MS proteomics (LTQ-Orbitrap) can be used to verify and improve gene annotation by identifying peptides specific to one of two competing annotation datasets.
-
Full-text index only
Identifying alternative hyper-splicing signatures in MG-thymoma by exon arrays.
PMID 18545673 · PMC2409220 · PloS one · 2008 · 8 claims · 6 setups
An integrative ad-hoc functional GO analysis combining threshold-based (Fisher exact/hypergeometric) and threshold-free (Kolmogorov-Smirnov) statistics, plus term-to-parent comparisons, detects disease-relevant splicing events from exon array data.