Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Overview of microarray analysis of gene expression and its applications to cervical cancer investigation.
PMID 18182341 · PMC7129792 · Taiwanese journal of obstetrics & gynecology · 2007 · 6 claims · 5 setups
Oligonucleotide microarray and cDNA microarray are the two main microarray platforms used to study gene expression genome-wide.
-
Has reproduction · 85
Digital sorting of complex tissues for cell type-specific gene expression profiles.
PMID 23497278 · PMC3626856 · BMC bioinformatics · 2013 · 8 claims · 8 setups
The Digital Sorting Algorithm (DSA) deconvolves mixed tissue expression into cell type-specific profiles using only marker genes, without requiring prior knowledge of cell type frequencies or in vitro pure-cell profiles.
-
Has reproduction · 83
Gene Expression Analysis Platform (GEAP): A highly customizable, fast, versatile and ready-to-use microarray analysis platform.
PMID 34927664 · PMC8754388 · Genetics and molecular biology · 2021 · 8 claims · 2 setups
GEAP is a GUI-based microarray analysis platform combining a C# front-end with an R (RTerm) back-end via the rgeap package, enabling analysis independent of manufacturer/platform.
-
Has reproduction · 87
A robust data scaling algorithm to improve classification accuracies in biomedical data.
PMID 27612635 · PMC5016890 · BMC bioinformatics · 2016 · 8 claims · 2 setups
Models trained on data scaled by the GL algorithm outperform models trained on data scaled by the Min-max or Z-score algorithms across 16 binary classification tasks, measured by AUROC and percentage of correct classification
-
Has reproduction · 75
Identification and analysis of genes associated with epithelial ovarian cancer by integrated bioinformatics methods.
PMID 34143800 · PMC8213194 · PloS one · 2021 · 7 claims · 6 setups
306 overlapping DEGs (265 up-regulated, 41 down-regulated) were identified across three independent GEO datasets in EOC vs normal ovarian tissue.
-
Full-text index only
Cardiovascular genetic medicine: genomic assessment of prognosis and diagnosis in patients with cardiomyopathy and heart failure.
PMID 20559924 · PMC4745893 · Journal of cardiovascular translational research · 2008 · 8 claims · 6 setups
Molecular signature analysis (MSA) uses machine-learning/classification methods (e.g., PAM/nearest shrunken centroids) on gene expression patterns to classify samples by phenotype for diagnosis, prognosis, or therapy response.
-
Full-text index only
Have microarrays failed to deliver for developmental biology?
PMID 12225576 · PMC139405 · Genome biology · 2002 · 8 claims · 8 setups
Despite predictions that microarrays would transform biology, very few published developmental biology microarray studies have generated novel insights.
-
Full-text index only
The GermOnline cross-species systems browser provides comprehensive information on genes and gene products relevant for sexual reproduction.
PMID 17145711 · PMC1751528 · Nucleic acids research · 2007 · 7 claims · 5 setups
GermOnline is a cross-species systems browser providing comprehensive curated information on genes and gene products relevant for sexual reproduction across nine model organisms including human.
-
Full-text index only
CLEAN: CLustering Enrichment ANalysis.
PMID 19640299 · PMC2734555 · BMC bioinformatics · 2009 · 8 claims · 4 setups
The gene-specific CLEAN score improves reproducibility of cluster analysis conclusions across independent datasets compared to the traditional cluster-wide score (cwCLEAN).
-
Full-text index only
Genomic approaches to the genetics of alcoholism.
PMID 12875046 · PMC6683845 · Alcohol research & health : the journal of the National Institute on Alcohol Abuse and Alcoholism · 2002 · 8 claims · 4 setups
Alcoholism is a complex disease that develops from a combination of numerous genetic and environmental factors, unlike single-gene disorders such as cystic fibrosis or Huntington's disease.
-
Full-text index only
Genomics, molecular imaging, bioinformatics, and bio-nano-info integration are synergistic components of translational medicine and personalized healthcare research.
PMID 18831773 · PMC3226104 · BMC genomics · 2008 · 8 claims · 8 setups
Genomics, molecular imaging, bioinformatics, and bio-nano-info integration are synergistic components of translational medicine and personalized healthcare
-
Has reproduction · 71
Comprehensive comparison of gene expression diversity among a variety of human stem cells.
PMID 36458020 · PMC9706419 · NAR genomics and bioinformatics · 2022 · 6 claims · 7 setups
iPSC gene expression is more strongly influenced by tissue origin than other stem cell types, whereas ESCs and somatic stem cells (MSCs, HSCs) are more strongly impacted by culture condition.
-
Full-text index only
Multiplex amplification enabled by selective circularization of large sets of genomic DNA fragments.
PMID 15860768 · PMC1087789 · Nucleic acids research · 2005 · 8 claims · 4 setups
Selector oligonucleotides can circularize many distinct genomic restriction fragments in one reaction, enabling parallel amplification with a single universal primer pair instead of multiple target-specific primer pairs
-
Full-text index only
Meeting highlights: beyond the genome 2000: the 18th International Congress of Biochemistry and Molecular Biology.
PMID 11119309 · PMC2448388 · Yeast (Chichester, England) · 2000 · 8 claims · 8 setups
Celera sequenced a human genome to ~45-fold coverage from one donor and used high-quality sequence stretches to define ~6 million SNPs
-
Full-text index only
Feature context-dependency and complexity-reduction in probability landscapes for integrative genomics.
PMID 18783599 · PMC2559821 · Theoretical biology & medical modelling · 2008 · 8 claims · 2 setups
Probability landscapes permit systematic detection, analysis, and utilization of feature context-dependency in genomic data.
-
Full-text index only
Uncovering information on expression of natural antisense transcripts in Affymetrix MOE430 datasets.
PMID 17598913 · PMC1929078 · BMC genomics · 2007 · 8 claims · 4 setups
Standard Affymetrix expression GeneChips (MOE430, HG-U133) contain probe sets that detect natural antisense transcripts (NATs)
-
Full-text index only
A new procedure for determining the genetic basis of a physiological process in a non-model species, illustrated by cold induced angiogenesis in the carp.
PMID 19852815 · PMC2771047 · BMC genomics · 2009 · 8 claims · 5 setups
The Conditional Stepped Reciprocal Best Hit (CSRBH) approach, combining direct RBH and zebrafish-stepped RBH (SRBH), outperformed other ortholog assignment methods and attained 8,726 carp-human functional homolog relationships for 16,650 carp contigs
-
Has reproduction · 91
Genome-wide identification of conserved and novel microRNAs in one bud and two tender leaves of tea plant (Camellia sinensis) by small RNA sequencing, microarray-based hybridization and genome survey scaffold sequences.
PMID 29157210 · PMC5697157 · BMC plant biology · 2017 · 7 claims · 8 setups
175 conserved and 83 novel miRNAs were identified mainly in one bud and two tender leaves of tea plant via small RNA sequencing combined with genome survey data
-
Has reproduction · 100
Integrative transcriptome sequencing identifies trans-splicing events with important roles in human embryonic stem cell pluripotency.
PMID 24131564 · PMC3875859 · Genome research · 2014 · 8 claims · 8 setups
TSscan, a computational pipeline integrating long- and short-read transcriptome sequencing from multiple hESC lines, can detect trans-splicing while minimizing false positives from experimental artifacts and genetic rearrangements.
-
Full-text index only
Optimality driven nearest centroid classification from genomic data.
PMID 17912341 · PMC1991588 · PloS one · 2007 · 7 claims · 5 setups
A theoretical result determines the subset of features of a given size that minimizes the misclassification rate for a nearest-centroid (LDA) classifier, based on equation (4).