Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 84
randPedPCA: rapid approximation of principal components from large pedigrees.
PMID 40877802 · PMC12392600 · Genetics, selection, evolution : GSE · 2025 · 7 claims · 2 setups
Matrix-vector multiplication with the dense additive relationship matrix A can be performed implicitly and efficiently via forward/backward substitution using the sparse inverse relationship (Cholesky) factor L^-1, avoiding explicit construction of A.
-
Has reproduction · 76
Tracing human genetic histories and natural selection with precise local ancestry inference.
PMID 40379651 · PMC12084304 · Nature communications · 2025 · 7 claims · 7 setups
Orchestra, a two-stage LAI method combining a recombination-distance base layer with a deep learning (convolutional + attention) smoothing module, outperforms RFmix, FLARE and Gnomix in precision and recall across simulated admixture generations.
-
Full-text index only
Application of qualifying variants for genomic analysis.
PMID 41570118 · PMC12926777 · Bioinformatics (Oxford, England) · 2026 · 7 claims · 4 setups
QVs should be treated as dynamic, multifaceted elements permeating the entire analysis workflow, not as a single static filtering step
-
Full-text index only
EXPLANA: a user-friendly workflow for EXPLoratory ANAlysis and feature selection in cross-sectional and longitudinal microbiome studies.
PMID 41416890 · PMC12766912 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 3 setups
EXPLANA is a feature selection workflow for longitudinal microbiome studies (LMS) that supports numerical and categorical data and also accommodates cross-sectional studies.
-
Full-text index only
Tractor workflow: a scalable Nextflow framework for local ancestry-aware genome-wide association studies.
PMID 41838407 · PMC13197121 · Bioinformatics (Oxford, England) · 2026 · 7 claims · 6 setups
Developed a scalable Nextflow workflow that automates phasing, local ancestry inference (LAI), and Tractor GWAS into a reproducible end-to-end pipeline
-
Full-text index only
DoBSeqWF: a framework for sensitive detection of individual genetic variation in pooled sequencing data.
PMID 41704565 · PMC12907731 · NAR genomics and bioinformatics · 2026 · 7 claims · 5 setups
DoBSeqWF, a Nextflow-based pipeline, processes pooled DoBSeq sequencing data through alignment, variant calling, machine-learning-based filtering, and variant pinpointing/assignment to individuals.
-
Full-text index only
PPC: an algorithm for accurate estimation of SNP allele frequencies in small equimolar pools of DNA using data from high density microarrays.
PMID 16199750 · PMC1240117 · Nucleic acids research · 2005 · 7 claims · 6 setups
The PPC algorithm, which applies a probe-pair-specific second-degree polynomial correction, increases the accuracy of allele frequency estimates from pooled DNA compared with previously described algorithms
-
Has reproduction · 61
TEMP: a computational method for analyzing transposable element polymorphism in populations.
PMID 24753423 · PMC4066757 · Nucleic acids research · 2014 · 8 claims · 8 setups
TEMP combines pair-end (discordant) read and split (soft-clipped) read information to identify both presence and absence of TE insertions in genomic DNA from heterogeneous/pooled samples.
-
Has reproduction · 78
A case study for large-scale human microbiome analysis using JCVI's metagenomics reports (METAREP).
PMID 22719821 · PMC3374610 · PloS one · 2012 · 8 claims · 7 setups
METAREP version 1.3.1 is an open-source, scalable tool for querying, browsing and comparing extremely large volumes of metagenomic annotations, with an extended data model, dynamic weighting, distributed searches and advanced clustering.