Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Processing and population genetic analysis of multigenic datasets with ProSeq3 software.
PMID 19797407 · PMC2778335 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 7 setups
ProSeq3 is a program with a graphic user interface that simplifies preparation and basic population genetic analysis of multigenic DNA polymorphism datasets
-
Has reproduction · 60
Integrating herbarium specimen observations into global phenology data systems.
PMID 30937223 · PMC6426164 · Applications in plant sciences · 2019 · 7 claims · 5 setups
A new PPO release adds terms and properties to relate observations of parts of plants to whole plants, enabling integration of herbarium phenology data with field observation data.
-
Has reproduction · 100
Shiny-Calorie: a context-aware application for indirect calorimetry data analysis and visualization using R.
PMID 41640623 · PMC12867577 · Bioinformatics advances · 2026 · 8 claims · 8 setups
Shiny-Calorie is an open-source interactive application for transparent data and metadata integration, statistical analysis, and visualization of indirect calorimetry datasets.
-
Full-text index only
Grammar-based distance in progressive multiple sequence alignment.
PMID 18616828 · PMC2478692 · BMC bioinformatics · 2008 · 7 claims · 3 setups
A grammar-based (LZ complexity) distance metric can be used to determine the order in which sequences are progressively pairwise aligned
-
Full-text index only
Digital evolution.
PMID 14551915 · PMC212697 · PLoS biology · 2003 · 7 claims · 5 setups
Digital organisms (Avidians) evolved from simple self-replicators, through an unexpected transitional form, to complex performers of many logic functions, with the full genealogy traceable and no missing links.
-
Full-text index only
Bayesian survival analysis in genetic association studies.
PMID 18617538 · PMC2530885 · Bioinformatics (Oxford, England) · 2008 · 7 claims · 5 setups
A novel Bayesian method (BETA-Surv) extends prior case-control haplotype-clustering work to censored survival outcomes by clustering haplotypes via gene tree/perfect phylogeny topology and relative mutation age.
-
Has reproduction · 71
Sustainable data analysis with Snakemake.
PMID 34035898 · PMC8114187 · F1000Research · 2021 · 8 claims · 4 setups
Reproducibility alone is insufficient for sustainable data analysis; transparency and adaptability are equally important additional properties.
-
Full-text index only
PRESTO: rapid calculation of order statistic distributions and multiple-testing adjusted P-values via permutation for one and two-stage genetic association studies.
PMID 18620604 · PMC2483288 · BMC bioinformatics · 2008 · 8 claims · 4 setups
PRESTO is an order of magnitude faster than other existing permutation testing software for genetic association studies.
-
Has reproduction · 71
Hyb: a bioinformatics pipeline for the analysis of CLASH (crosslinking, ligation and sequencing of hybrids) data.
PMID 24211736 · PMC3969109 · Methods (San Diego, Calif.) · 2014 · 8 claims · 6 setups
The 'hyb' pipeline detects, calls, folds and annotates chimeric reads from CLASH high-throughput sequencing data.
-
Has reproduction · 59
Identification of herpesvirus transcripts from genomic regions around the replication origins.
PMID 37773348 · PMC10541914 · Scientific reports · 2023 · 8 claims · 8 setups
Herpesviruses display distinct patterns of transcriptional overlaps near or at the replication origins (Oris)
-
Full-text index only
GoMiner: a resource for biological interpretation of genomic and proteomic data.
PMID 12702209 · PMC154579 · Genome biology · 2003 · 8 claims · 4 setups
GoMiner organizes 'interesting' gene lists (e.g., differentially expressed genes) into the Gene Ontology hierarchy for biological interpretation, displaying results as both a tree and a directed acyclic graph (DAG).
-
Has reproduction · 43
Compression of structured high-throughput sequencing data.
PMID 24260313 · PMC3832420 · PloS one · 2013 · 8 claims · 7 setups
Leveraging an explicit data schema (separate field encoding, field modeling, template compression, domain modeling) enables stronger compression of HTS alignment data than general-purpose compression of serialized bytes.
-
Full-text index only
A novel wavelet-based thresholding method for the pre-processing of mass spectrometry data that accounts for heterogeneous noise.
PMID 18615428 · PMC2855839 · Proteomics · 2008 · 6 claims · 4 setups
Noise in SELDI-TOF/MALDI-TOF mass spectrometry data is heteroscedastic across the m/z range, with larger variance at lower m/z values, contrary to the homogeneous noise assumption of existing wavelet denoising methods.
-
Full-text index only
PrimerZ: streamlined primer design for promoters, exons and human SNPs.
PMID 17537812 · PMC1933185 · Nucleic acids research · 2007 · 6 claims · 3 setups
PrimerZ automates primer design for gene promoters, exons, and human SNPs by integrating Ensembl sequence retrieval with Primer3 design in a single web workflow
-
Has reproduction · 78
A case study for large-scale human microbiome analysis using JCVI's metagenomics reports (METAREP).
PMID 22719821 · PMC3374610 · PloS one · 2012 · 8 claims · 7 setups
METAREP version 1.3.1 is an open-source, scalable tool for querying, browsing and comparing extremely large volumes of metagenomic annotations, with an extended data model, dynamic weighting, distributed searches and advanced clustering.