Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 95
OptiType: precision HLA typing from next-generation sequencing data.
PMID 25143287 · PMC4441069 · Bioinformatics (Oxford, England) · 2014 · 8 claims · 8 setups
OptiType, an ILP-based HLA genotyping algorithm, produces accurate four-digit HLA-I predictions from NGS data not enriched for the HLA cluster.
-
Full-text index only
Repurposing public sarcoma multi-omics for neoantigen discovery.
PMID 42012689 · PMC13100081 · Cancer immunology, immunotherapy : CII · 2026 · 8 claims · 7 setups
Reanalysis of legacy CKS multi-omic data shows that standard genome-wide metrics frequently underestimate the true immunogenic potential of these tumors.
-
Full-text index only
SVNeoPP: A Workflow for Structural-Variant-Derived Neoantigen Prediction and Prioritization Using Multi-Omics Data.
PMID 41892252 · PMC13024079 · Biology · 2026 · 8 claims · 7 setups
SVNeoPP is an end-to-end Snakemake workflow that takes WGS and RNA-seq as input to call/annotate SVs, reconstruct altered transcripts and coding sequences in an isoform-aware, traceable manner, and generate candidate peptides.
-
Full-text index only
Cleanifier: contamination removal from microbial sequences using spaced seeds of a human pangenome index.
PMID 41252442 · PMC12758600 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 4 setups
Cleanifier is a fast, memory-frugal alignment-free tool for detecting and removing human contamination using gapped k-mers (spaced seeds) and a human pangenome index.
-
Full-text index only
Inference of SARS-CoV-2 exposure biomarkers using large-scale T-cell repertoire profiling.
PMID 41680899 · PMC12903587 · Genome medicine · 2026 · 7 claims · 6 setups
A novel batch-effect correction method (log-normal gene usage modeling, Z-score normalization, and roulette-wheel resampling) allows combining AIRR-seq data from different batches and protocols.