Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 80
SLDMS: A Tool for Calculating the Overlapping Regions of Sequences.
PMID 35046988 · PMC8761809 · Frontiers in plant science · 2021 · 8 claims · 5 setups
SLDMS is a novel method for computing overlapping regions of sequencing reads using suffix array (SA), longest common prefix (LCP) array, document array (DA), and a monotonic stack.
-
Has reproduction · 84
Fractional ridge regression: a fast, interpretable reparameterization of ridge regression.
PMID 33252656 · PMC7702219 · GigaScience · 2020 · 8 claims · 2 setups
Ridge regression can be reparameterized in terms of γ, the ratio between the L2-norms of the regularized and unregularized (OLS) coefficient solutions, defining 'fractional ridge regression' (FRR)
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Has reproduction · 30
Minimal metabolic pathway structure is consistent with associated biomolecular interactions.
PMID 24987116 · PMC4299494 · Molecular systems biology · 2014 · 8 claims · 8 setups
MinSpan, a mixed-integer linear optimization algorithm, computes the shortest, linearly independent pathways (sparsest basis of the null space of the stoichiometric matrix S) for genome-scale metabolic networks, which convex approaches (extreme pathways, elementary flux modes) cannot do at genome scale.
-
Full-text index only
htSNPer1.0: software for haplotype block partition and htSNPs selection.
PMID 15740612 · PMC1274247 · BMC bioinformatics · 2005 · 6 claims · 1 setups
The GBB algorithm finds the globally optimal minimal htSNP set with far less computing time than exhaustive/enumeration search.
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
Detecting unannotated splicing events in short-read RNA-seq with SAMI, a UMI-aware Nextflow pipeline.
PMID 42166739 · PMC13242923 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 5 setups
SAMI is a UMI-aware, Singularity-contained Nextflow pipeline that detects splicing events diverging from transcript annotations directly from raw FASTQ files.
-
Full-text index only
FLASH-MM: fast and scalable single-cell differential expression analysis using linear mixed-effects models.
PMID 41644528 · PMC12982622 · Nature communications · 2026 · 8 claims · 6 setups
FLASH-MM produces LMM parameter estimates identical to lmer (lme4) up to the sixth decimal place while being 50- to 140-fold faster as sample size increases from 20,000 to 120,000 cells
-
Has reproduction · 71
RNAmountAlign: Efficient software for local, global, semiglobal pairwise and multiple RNA sequence/structure alignment.
PMID 31978147 · PMC6980424 · PloS one · 2020 · 7 claims · 6 setups
RNAmountAlign performs pairwise local, global, and semiglobal (query search) alignment and progressive multiple alignment (global and local) using incremental ensemble mountain height, running in O(n^3) time and O(n^2) space for two sequences of length n
-
Full-text index only
DupyliCate: mining, classifying, and characterizing gene duplications.
PMID 42209743 · PMC13219399 · Scientific reports · 2026 · 8 claims · 8 setups
DupyliCate is a Python tool for identifying and classifying gene duplication arrays, using BUSCO-based species-specific thresholds and offering integrated expression divergence and Ka/Ks analysis.
-
Full-text index only
SGCRNA: spectral clustering-guided co-expression network analysis without scale-free constraints for multi-omic data.
PMID 41615289 · PMC12856952 · Briefings in bioinformatics · 2026 · 8 claims · 8 setups
WGCNA's reliance on a scale-free topology assumption is problematic because real co-expression networks do not consistently exhibit scale-free properties