Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Application of proteomics in the study of tumor metastasis.
PMID 15862116 · PMC5172469 · Genomics, proteomics & bioinformatics · 2004 · 8 claims · 8 setups
Cell function is directly regulated through proteins, not genes or mRNA, so metastasis-related gene findings need protein-level validation via proteomics.
-
Full-text index only
Comparing whole genomes using DNA microarrays.
PMID 18347592 · PMC7097741 · Nature reviews. Genetics · 2008 · 8 claims · 6 setups
DNA microarrays offer a relatively inexpensive and efficient alternative to genome sequencing for comparing all known classes of genomic diversity between closely related genomes.
-
Has reproduction · 83
Analyzing biomarker discovery: Estimating the reproducibility of biomarker sets.
PMID 35901020 · PMC9333302 · PloS one · 2022 · 7 claims · 3 setups
A Reproducibility Score, RS(D,BD), defined as the average Jaccard overlap between biomarker sets found by the same discovery process on comparable datasets from the same distribution, quantifies biomarker reproducibility on a 0-1 scale
-
Has reproduction · 45
NUSAP1 Could be a Potential Target for Preventing NAFLD Progression to Liver Cancer.
PMID 35431924 · PMC9010788 · Frontiers in pharmacology · 2022 · 8 claims · 8 setups
NUSAP1 is a hub gene linking NAFLD fibrosis progression and HCC and may be a therapeutic target to prevent NAFLD progression to liver cancer
-
Full-text index only
High fidelity of whole-genome amplified DNA on high-density single nucleotide polymorphism arrays.
PMID 18786630 · PMC2659594 · Genomics · 2008 · 8 claims · 7 setups
WGA product performs well on the Affymetrix 250K SNP array compared to genomic DNA, especially with the BRLMM calling algorithm.
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Has reproduction · 96
Scalable Prediction of Acute Myeloid Leukemia Using High-Dimensional Machine Learning and Blood Transcriptomics.
PMID 31918046 · PMC6992905 · iScience · 2020 · 8 claims · 8 setups
Data-driven, high-dimensional ML approaches that learn multivariate signatures directly from genome-wide transcriptomic data (no prior gene selection) yield accurate and robust AML classifiers.
-
Full-text index only
PromoterPlot: a graphical display of promoter similarities by pattern recognition.
PMID 15980503 · PMC1160174 · Nucleic acids research · 2005 · 7 claims · 4 setups
PromoterPlot is a web-based tool that displays and processes TransFac transcription factor search results as an interactive SVG page
-
Full-text index only
A qualitative assessment of direct-labeled cDNA products prior to microarray analysis.
PMID 15762992 · PMC1079821 · BMC genomics · 2005 · 5 claims · 5 setups
The Agilent 2100 Bioanalyzer can be used in a novel assay to assess the quality/quantity of direct-labeled Cy-dye cDNA prior to microarray hybridization
-
Full-text index only
Meeting highlights: beyond the genome 2000: the 18th International Congress of Biochemistry and Molecular Biology.
PMID 11119309 · PMC2448388 · Yeast (Chichester, England) · 2000 · 8 claims · 8 setups
Celera sequenced a human genome to ~45-fold coverage from one donor and used high-quality sequence stretches to define ~6 million SNPs
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Full-text index only
Optimality driven nearest centroid classification from genomic data.
PMID 17912341 · PMC1991588 · PloS one · 2007 · 7 claims · 5 setups
A theoretical result determines the subset of features of a given size that minimizes the misclassification rate for a nearest-centroid (LDA) classifier, based on equation (4).
-
Full-text index only
A new procedure for determining the genetic basis of a physiological process in a non-model species, illustrated by cold induced angiogenesis in the carp.
PMID 19852815 · PMC2771047 · BMC genomics · 2009 · 8 claims · 5 setups
The Conditional Stepped Reciprocal Best Hit (CSRBH) approach, combining direct RBH and zebrafish-stepped RBH (SRBH), outperformed other ortholog assignment methods and attained 8,726 carp-human functional homolog relationships for 16,650 carp contigs
-
Full-text index only
Improvements to GALA and dbERGE II: databases featuring genomic sequence alignment, annotation and experimental results.
PMID 15608239 · PMC539999 · Nucleic acids research · 2005 · 8 claims · 8 setups
GALA is now a set of interlinked relational databases covering five vertebrate species: human, chimpanzee, mouse, rat and chicken.
-
Full-text index only
Stemming cancer: functional genomics of cancer stem cells in solid tumors.
PMID 18561035 · PMC2758383 · Stem cell reviews · 2008 · 8 claims · 8 setups
Cancer stem cells are a minority tumor subpopulation that alone can maintain indefinite tumor growth, as shown by serial transplantation experiments
-
Full-text index only
Large-scale analysis of Macaca fascicularis transcripts and inference of genetic divergence between M. fascicularis and M. mulatta.
PMID 18294402 · PMC2287170 · BMC genomics · 2008 · 8 claims · 6 setups
Constructed full-length-enriched cDNA libraries and determined 85,721 EST sequences and 9407 full-insert sequences from cynomolgus macaque brain (7 regions), testis, and liver
-
Full-text index only
Neuroscience in the era of functional genomics and systems biology.
PMID 19829370 · PMC3645852 · Nature · 2009 · 8 claims · 7 setups
Omics/discovery-based approaches do not eschew hypotheses but elevate hypothesis testing to high-throughput hypothesis generation and prioritization.
-
Full-text index only
An "omics" approach to uropathogenic Escherichia coli vaccinology.
PMID 19758805 · PMC2770165 · Trends in microbiology · 2009 · 8 claims · 8 setups
An 'omics'-based screening strategy integrating genomic, proteomic, and metabolomic data can identify PASivE UPEC proteins as vaccine candidates
-
Full-text index only
Identifying synonymous regulatory elements in vertebrate genomes.
PMID 15980499 · PMC1160227 · Nucleic acids research · 2005 · 7 claims · 4 setups
SynoR is a tool that performs de novo genome-wide identification of synonymous regulatory elements (SREs) using evolutionarily conserved TFBS modules as seeds
-
Full-text index only
An emerging cyberinfrastructure for biodefense pathogen and pathogen-host data.
PMID 17984082 · PMC2239001 · Nucleic acids research · 2008 · 8 claims · 7 setups
The Biodefense Proteomics Resource Center (RC) is a public cyberinfrastructure that stores, integrates, and disseminates experimental data from seven Proteomics Research Centers (PRCs) on biodefense pathogens and host interactions