Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 90
Prioritized mass spectrometry increases the depth, sensitivity and data completeness of single-cell proteomics.
PMID 37012480 · PMC10172113 · Nature methods · 2023 · 8 claims · 5 setups
pSCoPE (prioritized precursor selection via MaxQuant.Live) increases sensitivity, data completeness, and proteome coverage more than twofold over shotgun single-cell proteomics
-
Full-text index only
Optimal step length EM algorithm (OSLEM) for the estimation of haplotype frequency and its application in lipoprotein lipase genotyping.
PMID 12529185 · PMC149347 · BMC bioinformatics · 2003 · 5 claims · 4 setups
OSLEM (Optimal Step Length EM), which approximates an optimal step length via a fixed-point search (D_N = D_{N-1} + λ(D_preN - D_{N-1})), runs about twice as fast as standard EM while producing the same haplotype frequency estimates.
-
Full-text index only
FEDRANN: effective long-read overlap detection based on dimensionality reduction and approximate nearest neighbors.
PMID 42102720 · PMC13201080 · GigaScience · 2026 · 8 claims · 6 setups
A pipeline combining IDF transformation, sparse random projection (SRP), and NNDescent (the FEDRANN strategy) enables accurate overlap detection across diverse long-read datasets
-
Full-text index only
BFAST: an alignment tool for large scale genome resequencing.
PMID 19907642 · PMC2770639 · PloS one · 2009 · 7 claims · 4 setups
BFAST is a new algorithm and freely available software tool for aligning large-scale short-read sequencing data to large reference genomes with user-customizable speed and accuracy
-
Full-text index only
DANST enables cell-type deconvolution in spatial transcriptomics using deep domain adversarial neural networks.
PMID 41663685 · PMC12996496 · Communications biology · 2026 · 7 claims · 6 setups
DANST, a deconvolution framework using deep domain adversarial neural networks, achieves superior cell-type deconvolution accuracy compared with existing methods on human and mouse benchmark datasets
-
Full-text index only
scGACL: a generative adversarial network with multi-scale contrastive learning for accurate single-cell RNA sequencing imputation.
PMID 41632596 · PMC12866930 · Briefings in bioinformatics · 2026 · 8 claims · 6 setups
scGACL, a GAN integrated with multi-scale contrastive learning, is proposed to overcome the over-smoothing problem in scRNA-seq imputation
-
Has reproduction · 51
SGCP: a spectral self-learning method for clustering genes in co-expression networks.
PMID 38956463 · PMC11221046 · BMC bioinformatics · 2024 · 7 claims · 4 setups
SGCP, a spectral self-learning method, yields gene co-expression modules with higher GO enrichment than WGCNA, CoExpNets, and CEMiTool across 12 real gene expression datasets.
-
Full-text index only
nf-core/viralmetagenome: A novel pipeline for untargeted viral genome reconstruction.
PMID 42057295 · PMC13141149 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 5 setups
nf-core/viralmetagenome is a Nextflow pipeline that automates untargeted reconstruction and variant analysis of eukaryotic DNA and RNA viruses from short-read metagenomic or hybridisation-capture data.
-
Has reproduction · 86
LMAS: evaluating metagenomic short de novo assembly methods through defined communities.
PMID 36576131 · PMC9795473 · GigaScience · 2022 · 8 claims · 5 setups
LMAS (Last Metagenomic Assembler Standing) is a flexible, Nextflow-based, Docker-containerized automated workflow for benchmarking de novo metagenomic assemblers against defined mock communities, producing an interactive HTML report.
-
Has reproduction · 76
Tracing human genetic histories and natural selection with precise local ancestry inference.
PMID 40379651 · PMC12084304 · Nature communications · 2025 · 7 claims · 7 setups
Orchestra, a two-stage LAI method combining a recombination-distance base layer with a deep learning (convolutional + attention) smoothing module, outperforms RFmix, FLARE and Gnomix in precision and recall across simulated admixture generations.
-
Full-text index only
metaFun: An analysis pipeline for metagenomic big data with fast and unified functional searches.
PMID 41530917 · PMC12818822 · Gut microbes · 2026 · 8 claims · 8 setups
metaFun is an open-source, end-to-end Nextflow/Apptainer pipeline integrating quality control, taxonomic profiling, functional profiling, de novo assembly, binning, genome assessment, comparative genomics, network analysis, and strain-level microdiversity analysis into a unified framework
-
Full-text index only
Spider: a flexible and unified framework for simulating spatial transcriptomics data.
PMID 41237053 · PMC12790819 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 6 setups
Spider simulates ST data without requiring real ST data as a reference
-
Has reproduction · 100
poreCov-An Easy to Use, Fast, and Robust Workflow for SARS-CoV-2 Genome Reconstruction via Nanopore Sequencing.
PMID 34394197 · PMC8355734 · Frontiers in genetics · 2021 · 8 claims · 8 setups
poreCov is an easy-to-use, fast, and robust Nextflow-based workflow for reference-based SARS-CoV-2 genome reconstruction and lineage determination from nanopore sequencing data
-
Has reproduction · 29
MOSAIK: a hash-based algorithm for accurate next-generation sequencing short-read mapping.
PMID 24599324 · PMC3944147 · PloS one · 2014 · 8 claims · 8 setups
MOSAIK is the only aligner that consistently aligns reads from all major sequencing platforms (Illumina, AB SOLiD, Roche 454, Ion Torrent, Pacific Biosciences SMRT) using the same algorithmic approach.
-
Full-text index only
Computational tradeoffs in multiplex PCR assay design for SNP genotyping.
PMID 16042802 · PMC1190169 · BMC genomics · 2005 · 7 claims · 6 setups
Achieving high-multiplexing/high-coverage multiplex PCR designs is subject to a computational phase transition as the SNP-pair compatibility probability crosses a critical threshold
-
Full-text index only
ANOMALY: a Snakemake pipeline for identifying NuMTs from long-read sequencing data.
PMID 41647924 · PMC12869244 · NAR genomics and bioinformatics · 2026 · 8 claims · 8 setups
ANOMALY is a novel Snakemake pipeline for detecting NuMTs from long-read sequencing data
-
Has reproduction · 50
RNA-Seq alignment to individualized genomes improves transcript abundance estimates in multiparent populations.
PMID 25236449 · PMC4174954 · Genetics · 2014 · 8 claims · 7 setups
Genetic variants distinguishing an individual genome from the reference cause read misalignment and biased transcript abundance estimates, and fine-tuning of alignment algorithms does not correct this problem.
-
Full-text index only
Analyses and comparison of accuracy of different genotype imputation methods.
PMID 18958166 · PMC2569208 · PloS one · 2008 · 8 claims · 3 setups
Stronger LD produces higher imputation accuracy rates for all five methods
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
bakR: uncovering differential RNA synthesis and degradation kinetics transcriptome-wide with Bayesian hierarchical modeling.
PMID 37028916 · PMC10275263 · RNA (New York, N.Y.) · 2023 · 8 claims · 4 setups
bakR uses Bayesian hierarchical modeling to share information (specifically a replicate variability vs. read count trend) across transcripts, increasing statistical power for differential kinetic analysis