Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Expression profiling of drug response--from genes to pathways.
PMID 17117610 · PMC3181826 · Dialogues in clinical neuroscience · 2006 · 8 claims · 8 setups
Understanding individual response to a drug (efficacy/tolerability) is the major bottleneck in current drug development and clinical trials.
-
Full-text index only
BTW: a web server for Boltzmann time warping of gene expression time series.
PMID 16845055 · PMC1538860 · Nucleic acids research · 2006 · 5 claims · 4 setups
Symmetric time warping distance is more flexible than Euclidean distance or correlation coefficient for identifying genes with similar temporal expression profiles, especially across sequences of different length.
-
Full-text index only
DAVID Bioinformatics Resources: expanded annotation database and novel algorithms to better extract biology from large gene lists.
PMID 17576678 · PMC1933169 · Nucleic acids research · 2007 · 8 claims · 4 setups
The DAVID Gene Concept uses a single-linkage method to agglomerate tens of millions of gene/protein identifiers from NCBI, PIR, UniProt and other resources into unified DAVID genes.
-
Full-text index only
Proteomic solutions for analytical challenges associated with alcohol research.
PMID 23584870 · PMC3860482 · Alcohol research & health : the journal of the National Institute on Alcohol Abuse and Alcoholism · 2008 · 7 claims · 4 setups
Protein-level meta-analyses analogous to the transcriptome meta-analysis by Mulligan et al. (2006) are not yet possible because proteins lack a uniform sample preparation/analysis method and span up to 8 orders of magnitude in abundance.
-
Full-text index only
MuPlex: multi-objective multiplex PCR assay design.
PMID 15980531 · PMC1160138 · Nucleic acids research · 2005 · 8 claims · 3 setups
MuPlex is a web-enabled system that designs multiplex PCR assays by selecting primer pairs for SNPs and partitioning them into multiplex-compatible tube sets.
-
Full-text index only
CGHPRO -- a comprehensive data analysis tool for array CGH.
PMID 15807904 · PMC1274268 · BMC bioinformatics · 2005 · 8 claims · 3 setups
CGHPRO is a user-friendly, versatile, stand-alone Java tool for normalization, visualization, breakpoint detection and comparative analysis of array-CGH data
-
Has reproduction · 76
Correcting scale distortion in RNA sequencing data.
PMID 39875825 · PMC11776150 · BMC bioinformatics · 2025 · 8 claims · 8 setups
Local averaging reveals expression-level-dependent biases that differ from sample to sample across all RNA-seq datasets studied, and are not corrected by conventional normalization (TPM/FPKM)
-
Has reproduction · 71
RNAmountAlign: Efficient software for local, global, semiglobal pairwise and multiple RNA sequence/structure alignment.
PMID 31978147 · PMC6980424 · PloS one · 2020 · 7 claims · 6 setups
RNAmountAlign performs pairwise local, global, and semiglobal (query search) alignment and progressive multiple alignment (global and local) using incremental ensemble mountain height, running in O(n^3) time and O(n^2) space for two sequences of length n
-
Has reproduction · 43
TransFlow: a Snakemake workflow for transmission analysis of Mycobacterium tuberculosis whole-genome sequencing data.
PMID 36469333 · PMC9825751 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 8 setups
TransFlow is a Snakemake- and Conda-based workflow that combines state-of-the-art tools into a single, fast, scalable pipeline for MTBC WGS-based transmission analysis.
-
Has reproduction · 100
Shiny-Calorie: a context-aware application for indirect calorimetry data analysis and visualization using R.
PMID 41640623 · PMC12867577 · Bioinformatics advances · 2026 · 8 claims · 8 setups
Shiny-Calorie is an open-source interactive application for transparent data and metadata integration, statistical analysis, and visualization of indirect calorimetry datasets.
-
Has reproduction · 43
StatsDB: platform-agnostic storage and understanding of next generation sequencing run metrics.
PMID 24627795 · PMC3938176 · F1000Research · 2013 · 8 claims · 6 setups
StatsDB is an open-source software package for storage and analysis of next generation sequencing run metrics, backed by an SQL (MySQL) database with Perl and Java APIs.
-
Full-text index only
BABELOMICS: a systems biology perspective in the functional annotation of genome-scale experiments.
PMID 16845052 · PMC1538844 · Nucleic acids research · 2006 · 8 claims · 8 setups
Babelomics is presented as an updated, complete suite of web tools for functional analysis of genome-scale experiments with new and improved modules
-
Full-text index only
Multiple whole genome alignments and novel biomedical applications at the VISTA portal.
PMID 17488840 · PMC1933192 · Nucleic acids research · 2007 · 8 claims · 4 setups
A novel multiple whole-genome alignment algorithm treats all genomes symmetrically, avoiding dependence on a single base/reference genome
-
Full-text index only
SpliceMiner: a high-throughput database implementation of the NCBI Evidence Viewer for microarray splice variant analysis.
PMID 17338820 · PMC1839109 · BMC bioinformatics · 2007 · 6 claims · 4 setups
EVDB is a comprehensive, non-redundant relational database of known human splice variants built from NCBI Entrez Gene and Evidence Viewer data
-
Full-text index only
An integrated database-pipeline system for studying single nucleotide polymorphisms and diseases.
PMID 19091018 · PMC2638159 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Existing SNP/disease databases are fragmented; no combined resource widely supports gene-, SNP-, and disease-related information together
-
Full-text index only
G2Cdb: the Genes to Cognition database.
PMID 18984621 · PMC2686544 · Nucleic acids research · 2009 · 7 claims · 7 setups
G2Cdb integrates experimentally validated synapse proteome datasets with mouse/human genomic annotation, phenotype, and human disease data in a gene-centric database.
-
Full-text index only
High fidelity of whole-genome amplified DNA on high-density single nucleotide polymorphism arrays.
PMID 18786630 · PMC2659594 · Genomics · 2008 · 8 claims · 7 setups
WGA product performs well on the Affymetrix 250K SNP array compared to genomic DNA, especially with the BRLMM calling algorithm.
-
Full-text index only
A general definition and nomenclature for alternative splicing events.
PMID 18688268 · PMC2467475 · PLoS computational biology · 2008 · 6 claims · 4 setups
Existing AS nomenclatures (Malko et al.'s 5-letter strings, Nagasaki et al.'s bit matrices, and the ASD/ATD/AEdb system) are redundant, ambiguous, or incapable of representing complex or large splicing variations.
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.