Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Comparative genomics and understanding of microbial biology.
PMID 10998382 · PMC2627966 · Emerging infectious diseases · 2000 · 8 claims · 7 setups
GC content varies widely among prokaryotic genomes (29% in B. burgdorferi to 68% in M. tuberculosis) and shapes codon usage and amino acid composition.
-
Full-text index only
Application of proteomics methods for pathogen discovery.
PMID 19751767 · PMC7119679 · Journal of virological methods · 2010 · 8 claims · 6 setups
Proteomic techniques can detect and characterize unknown infectious agents in cell culture without prior knowledge of the pathogen
-
Has reproduction · 50
MEDUSA: A Pipeline for Sensitive Taxonomic Classification and Flexible Functional Annotation of Metagenomic Shotgun Sequences.
PMID 35330728 · PMC8940201 · Frontiers in genetics · 2022 · 6 claims · 6 setups
MEDUSA is an automated, Conda-installable and Snakemake-managed pipeline performing preprocessing, assembly, alignment, taxonomic classification, and functional annotation on shotgun data.
-
Full-text index only
SNPmasker: automatic masking of SNPs and repeats across eukaryotic genomes.
PMID 16845091 · PMC1538889 · Nucleic acids research · 2006 · 8 claims · 4 setups
SNPmasker is a web service combining SNP masking and repeat masking, supporting both coordinate-defined and homology-search-defined input regions, a combination not offered by prior tools
-
Full-text index only
A model-based approach to selection of tag SNPs.
PMID 16776821 · PMC1525207 · BMC bioinformatics · 2006 · 7 claims · 5 setups
The Li and Stephens hidden Markov model outperforms other tested models (simple Markov, two-state HMM, HMM-4D, greedy GR-1/GR-2) in description code-length, tag set information content, and prediction of tagged SNPs.
-
Has reproduction · 86
Molecular Classification Models for Triple Negative Breast Cancer Subtype Using Machine Learning.
PMID 34575658 · PMC8472680 · Journal of personalized medicine · 2021 · 6 claims · 4 setups
A training gene set of 719 unique upregulated DEGs (subtype-specific) can be used to build ML models that classify TNBC into BLIA, BLIS, MES, and LAR subtypes.
-
Full-text index only
Identifying drug effects via pathway alterations using an integer linear programming optimization formulation on phosphoproteomic data.
PMID 19997482 · PMC2776985 · PLoS computational biology · 2009 · 7 claims · 4 setups
An ILP formulation of the Boolean pathway optimization problem fits phosphoproteomic data faster and more efficiently than the previously used genetic algorithm (GA) approach.