Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Full-text index only
Design and analysis issues in genome-wide somatic mutation studies of cancer.
PMID 18692126 · PMC2820387 · Genomics · 2009 · 6 claims · 4 setups
Two-stage (discovery + validation) sequencing designs efficiently allocate resources and can produce highly informative candidate driver gene lists even with relatively small sample sizes.
-
Has reproduction · 85
Digital sorting of complex tissues for cell type-specific gene expression profiles.
PMID 23497278 · PMC3626856 · BMC bioinformatics · 2013 · 8 claims · 8 setups
The Digital Sorting Algorithm (DSA) deconvolves mixed tissue expression into cell type-specific profiles using only marker genes, without requiring prior knowledge of cell type frequencies or in vitro pure-cell profiles.
-
Full-text index only
Non-EST based prediction of exon skipping and intron retention events using Pfam information.
PMID 16204458 · PMC1243800 · Nucleic acids research · 2005 · 7 claims · 5 setups
A novel ab initio method predicts exon skipping and intron retention events using only Pfam domain annotation, via a Viterbi-like dynamic programming algorithm applied to the Pfam alignment.
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.
-
Full-text index only
Clustering of phosphorylation site recognition motifs can be exploited to predict the targets of cyclin-dependent kinase.
PMID 17316440 · PMC1852407 · Genome biology · 2007 · 8 claims · 6 setups
CDK consensus motifs are frequently clustered (closely spaced) in known CDK substrate proteins rather than uniformly distributed
-
Full-text index only
The DAVID Gene Functional Classification Tool: a novel biological module-centric algorithm to functionally analyze large gene lists.
PMID 17784955 · PMC2375021 · Genome biology · 2007 · 8 claims · 6 setups
Gene-gene functional similarity can be measured using kappa statistics applied to a binary gene-annotation-term matrix built from 14 annotation categories.
-
Full-text index only
Effect of read-mapping biases on detecting allele-specific expression from RNA-sequencing data.
PMID 19808877 · PMC2788925 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 6 setups
Reads mapped to the reference genome show a significant bias toward the reference allele at heterozygous SNPs
-
Has reproduction · 63
Community assessment of methods to deconvolve cellular composition from bulk gene expression.
PMID 39191725 · PMC11350143 · Nature communications · 2024 · 8 claims · 4 setups
Most deconvolution methods accurately predict coarse-grained immune/stromal cell populations from bulk expression.
-
Full-text index only
A high-throughput method for quantifying alleles and haplotypes of the malaria vaccine candidate Plasmodium falciparum merozoite surface protein-1 19 kDa.
PMID 16626494 · PMC1459863 · Malaria journal · 2006 · 8 claims · 5 setups
Pyrosequencing, after adjustment to a standard curve, provides accurate and precise estimates of allele frequencies in mixed MSP-1_19 infections
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Full-text index only
Exploring the immunome: A brave new world for human vaccine development.
PMID 20009527 · PMC2919815 · Human vaccines · 2009 · 7 claims · 7 setups
Screening the Mtb proteome in silico for epitopes ('fishing for antigens using epitopes as bait') revealed a remarkable diversity of human immune responses to Mtb proteins without an ascribed function, suggesting human immune response to Mtb is omnivorous rather than focused on single immunodominant proteins.
-
Full-text index only
Pathway analysis of kidney cancer using proteomics and metabolic profiling.
PMID 17123452 · PMC1665458 · Molecular cancer · 2006 · 8 claims · 8 setups
31 proteins are differentially expressed with high statistical significance (p<0.05) in ccRCC tumor tissue compared to adjacent non-malignant kidney tissue
-
Full-text index only
Variations in the transcriptome of Alzheimer's disease reveal molecular networks involved in cardiovascular diseases.
PMID 18842138 · PMC2760875 · Genome biology · 2008 · 8 claims · 6 setups
AD-related genes (APOE, A2M, PON2, MAP4) and CVD-associated genes (COMT, CBS, WNK1) congregate in a single co-expression module, linking AD and CVD at the transcriptional level