Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Update of the G2D tool for prioritization of gene candidates to inherited diseases.
PMID 17478516 · PMC1933178 · Nucleic acids research · 2007 · 8 claims · 4 setups
G2D is a web server that prioritizes candidate genes for inherited diseases using three distinct algorithms based on different input information.
-
Has reproduction · 74
SpaGene: A Deep Adversarial Framework for Spatial Gene Imputation.
PMID 42146899 · PMC13176606 · Computational and structural biotechnology journal · 2026 · 8 claims · 6 setups
SpaGene improves average PCC and SSIM and reduces RMSE compared to 6 baseline methods (SpaGE, gimVI, Tangram, VISTA, spRefine, stDiff) across 8 diverse ST-SC dataset pairs under gene-holdout evaluation.
-
Full-text index only
CONTRAST: a discriminative, phylogeny-free approach to multiple informant de novo gene prediction.
PMID 18096039 · PMC2246271 · Genome biology · 2007 · 8 claims · 5 setups
CONTRAST predicts exact coding region structures for 65% more human genes than the previous state-of-the-art de novo predictor (N-SCAN)
-
Has reproduction · 85
Digital sorting of complex tissues for cell type-specific gene expression profiles.
PMID 23497278 · PMC3626856 · BMC bioinformatics · 2013 · 8 claims · 8 setups
The Digital Sorting Algorithm (DSA) deconvolves mixed tissue expression into cell type-specific profiles using only marker genes, without requiring prior knowledge of cell type frequencies or in vitro pure-cell profiles.
-
Full-text index only
MatchMiner: a tool for batch navigation among gene and gene product identifiers.
PMID 12702208 · PMC154578 · Genome biology · 2003 · 8 claims · 3 setups
MatchMiner's LookUp function automates batch translation of an input list of gene identifiers into a matching list of a different identifier type.
-
Has reproduction · 65
FusionQ: a novel approach for gene fusion detection and quantification from paired-end RNA-Seq.
PMID 23768108 · PMC3691734 · BMC bioinformatics · 2013 · 8 claims · 8 setups
FusionQ is a novel tool that detects gene fusions, constructs chimerical transcript structures, and estimates their abundances from paired-end RNA-Seq data.
-
Full-text index only
PolySearch: a web-based text mining system for extracting relationships between human diseases, genes, mutations, drugs and metabolites.
PMID 18487273 · PMC2447794 · Nucleic acids research · 2008 · 8 claims · 7 setups
PolySearch supports more than 50 different classes of queries against nearly a dozen types of text, abstract, or bioinformatic databases
-
Full-text index only
Functional analysis of novel SNPs and mutations in human and mouse genomes.
PMID 19091009 · PMC2638150 · BMC bioinformatics · 2008 · 8 claims · 7 setups
FANS streamlines functional analysis of novel SNPs and mutations into a simplified, few-click, four-step procedure.
-
Full-text index only
Improved mutation tagging with gene identifiers applied to membrane protein stability prediction.
PMID 19758467 · PMC2745585 · BMC bioinformatics · 2009 · 8 claims · 4 setups
MutationTagger achieves 87% F-measure for the mutation retrieval task on a benchmark dataset
-
Has reproduction · 86
RNASEQR--a streamlined and accurate RNA-seq sequence analysis program.
PMID 22199257 · PMC3315322 · Nucleic acids research · 2012 · 8 claims · 7 setups
RNASEQR is a new RNA-seq mapper/aligner that combines a BWT-based (Bowtie) transcriptomic/genomic alignment with hash-based BLAT local alignment in three sequential steps: transcriptome mapping, novel exon detection, and anchor-and-align novel splice junction identification.
-
Has reproduction · 37
A Bayesian approach to accurate and robust signature detection on LINCS L1000 data.
PMID 32003771 · PMC7203754 · Bioinformatics (Oxford, England) · 2020 · 7 claims · 4 setups
A novel Bayesian peak deconvolution algorithm gives unbiased likelihood estimations for peak locations and derives probability-based z-scores.
-
Has reproduction · 67
binny: an automated binning algorithm to recover high-quality genomes from complex metagenomic datasets.
PMID 36239393 · PMC9677464 · Briefings in bioinformatics · 2022 · 8 claims · 8 setups
binny outperforms or is highly competitive with commonly used and state-of-the-art binning methods (MetaBAT2, MaxBin2, CONCOCT, VAMB, SemiBin, MetaDecoder)
-
Full-text index only
Integrating alternative splicing detection into gene prediction.
PMID 15705189 · PMC550657 · BMC bioinformatics · 2005 · 8 claims · 4 setups
An integrative intrinsic/extrinsic method was implemented in the gene finder EuGÈNE (as EuGÈNE-M) to detect AS evidence from aligned transcripts and generate alternative optimal gene predictions consistent with each detected AS event.
-
Has reproduction · 83
Macrel: antimicrobial peptide screening in genomes and metagenomes.
PMID 33384902 · PMC7751412 · PeerJ · 2020 · 8 claims · 8 setups
Macrel introduces a novel set of 22 peptide features (6 local, 16 global), including a new Free Energy Transition (FET) feature group, for AMP and hemolytic activity classification
-
Has reproduction · 50
Viewing RNA-seq data on the entire human genome.
PMID 28979763 · PMC5605993 · F1000Research · 2017 · 8 claims · 3 setups
RNA-Seq Viewer is a web application that visualizes genome-wide RNA-seq expression data pulled from NCBI's SRA and GEO databases using Ideogram.js
-
Has reproduction · 67
Leveraging RNA-seq deconvolution to improve complex in vitro model characterization.
PMID 40701251 · PMC12391696 · The Journal of biological chemistry · 2025 · 8 claims · 6 setups
RNA-seq deconvolution can predict cell type proportions from bulk RNA-seq using scRNA-seq references, offering a useful characterization tool for CIVMs where single-cell methods are impractical
-
Full-text index only
MILANO--custom annotation of microarray results using automatic literature searches.
PMID 15661078 · PMC547913 · BMC bioinformatics · 2005 · 7 claims · 4 setups
MILANO annotates microarray gene lists by counting literature co-occurrences of each gene with user-defined secondary terms
-
Full-text index only
A space-efficient and accurate method for mapping and aligning cDNA sequences onto genomic sequence.
PMID 18344523 · PMC2377433 · Nucleic acids research · 2008 · 7 claims · 6 setups
Spaln maps and aligns large cDNA sequence sets onto whole mammalian genomes using substantially less memory than comparable existing tools
-
Has reproduction · 77
SurvConvMixer: robust and interpretable cancer survival prediction based on ConvMixer using pathway-level gene expression images.
PMID 38539106 · PMC10967213 · BMC bioinformatics · 2024 · 6 claims · 5 setups
SurvConvMixer, using pathway-level gene expression images and ConvMixer, achieves strong internal validation AUC for overall survival prediction, especially on larger datasets like LUAD
-
Has reproduction · 78
Machine learning and free energy clustering reveal PAH protein binding linked to AD risk.
PMID 41953002 · PMC13053772 · iScience · 2026 · 8 claims · 8 setups
PARP1, PTPN1, and ITGA4 are core PAH protein targets identified via PPI network analysis and XGBoost feature selection from AD-associated DEGs