Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 81
Transcriptome analysis reveals differential splicing events in IPF lung tissue.
PMID 24647608 · PMC3960165 · PloS one · 2014 · 8 claims · 6 setups
873 genes are differentially expressed in IPF lung tissue versus healthy controls at FDR<5%, with more up-regulated than down-regulated genes.
-
Has reproduction · 87
Enhanced Generalizability of RNA Secondary Structure Prediction via Convolutional Block Attention Network and Ensemble Learning.
PMID 40871599 · PMC12388828 · Molecules (Basel, Switzerland) · 2025 · 8 claims · 8 setups
TrioFold integrates base-pairing clues from thermodynamic- and DL-based methods via ensemble learning and a convolutional block attention mechanism to enhance RSS prediction generalizability.
-
Has reproduction · 65
FusionQ: a novel approach for gene fusion detection and quantification from paired-end RNA-Seq.
PMID 23768108 · PMC3691734 · BMC bioinformatics · 2013 · 8 claims · 8 setups
FusionQ is a novel tool that detects gene fusions, constructs chimerical transcript structures, and estimates their abundances from paired-end RNA-Seq data.
-
Full-text index only
A comprehensive sensitivity analysis of microarray breast cancer classification under feature variability.
PMID 19941644 · PMC2789744 · BMC bioinformatics · 2009 · 7 claims · 4 setups
Feature variability strongly influences breast cancer signature composition even when array platform and patient stratification are identical.
-
Has reproduction · 67
binny: an automated binning algorithm to recover high-quality genomes from complex metagenomic datasets.
PMID 36239393 · PMC9677464 · Briefings in bioinformatics · 2022 · 8 claims · 8 setups
binny outperforms or is highly competitive with commonly used and state-of-the-art binning methods (MetaBAT2, MaxBin2, CONCOCT, VAMB, SemiBin, MetaDecoder)
-
Full-text index only
Computation of haplotypes on SNPs subsets: advantage of the "global method".
PMID 17067372 · PMC1636337 · BMC genetics · 2006 · 6 claims · 4 setups
The global method for subhaplotyping always yields a lower error rate than the direct method across datasets and SNP subset sizes
-
Full-text index only
A comparison of random sequence reads versus 16S rDNA sequences for estimating the biodiversity of a metagenomic library.
PMID 18682527 · PMC2532719 · Nucleic acids research · 2008 · 8 claims · 7 setups
Biodiversity observed by RSR analysis is consistent with that obtained by 16S rDNA analysis
-
Has reproduction · 89
Spatial information matters: are traditional imputation methods effective for spatial transcriptomics data?
PMID 41627342 · PMC12862982 · Briefings in bioinformatics · 2026 · 7 claims · 3 setups
No single existing SOTA imputation method consistently performs well across newer SRT platforms/datasets
-
Has reproduction · 81
Macrophages on the run: Exercise balances macrophage polarization for improved health.
PMID 39476967 · PMC11585839 · Molecular metabolism · 2024 · 8 claims · 7 setups
Immediate/acute exercise triggers an M1 (pro-inflammatory) macrophage polarization surge.
-
Has reproduction · 63
Creation of a Single Cell RNASeq Meta-Atlas to Define Human Liver Immune Homeostasis.
PMID 34335581 · PMC8322955 · Frontiers in immunology · 2021 · 7 claims · 7 setups
Independent human liver immune scRNA-seq datasets can be combined into an integrated meta-atlas in which all datasets co-cluster, despite differing cell-type proportions between studies.
-
Has reproduction · 90
Systematic Assessment of Small RNA Profiling in Human Extracellular Vesicles.
PMID 37444556 · PMC10340377 · Cancers · 2023 · 7 claims · 4 setups
Different EV isolation methods vary in reproducibility for isolating small RNAs and have characteristic effects on small RNA composition, with differential ultracentrifugation showing the highest variability/lowest replicability.
-
Has reproduction · 80
SLDMS: A Tool for Calculating the Overlapping Regions of Sequences.
PMID 35046988 · PMC8761809 · Frontiers in plant science · 2021 · 8 claims · 5 setups
SLDMS is a novel method for computing overlapping regions of sequencing reads using suffix array (SA), longest common prefix (LCP) array, document array (DA), and a monotonic stack.
-
Has reproduction · 84
AI-assisted discovery of an ethnicity-influenced driver of cell transformation in esophageal and gastroesophageal junction adenocarcinomas.
PMID 36134663 · PMC9675486 · JCI insight · 2022 · 8 claims · 8 setups
An AI-guided Boolean network approach (BoNE) models transcriptomic continuum states of normal esophagus, BE, and EAC to derive classifier gene signatures
-
Has reproduction · 68
Cell-type annotation with accurate unseen cell-type identification using multiple references.
PMID 37379341 · PMC10335708 · PLoS computational biology · 2023 · 8 claims · 4 setups
mtANN integrates multiple reference datasets and eight gene selection methods via ensemble learning (multiple deep classification models + majority voting) to improve cell-type annotation accuracy
-
Has reproduction · 58
HGA: de novo genome assembly method for bacterial genomes using high coverage short sequencing reads.
PMID 26945881 · PMC4779561 · BMC genomics · 2016 · 8 claims · 7 setups
HGA leads to significant improvement in assembly quality (N50 and corrected N50) for all 7 evaluated GAGE-B bacterial datasets using most of the 8 evaluated assemblers
-
Has reproduction · 49
EDGE COVID-19: a web platform to generate submission-ready genomes from SARS-CoV-2 sequencing efforts.
PMID 35561186 · PMC9113274 · Bioinformatics (Oxford, England) · 2022 · 7 claims · 5 setups
EDGE COVID-19 (EC-19) is a web-based platform that automates QC, reference-based variant/consensus calling, lineage determination, and submission of SARS-CoV-2 genomes and metadata to GenBank, GISAID and INSDC for both Illumina and ONT data.
-
Has reproduction · 50
Implementing the reuse of public DIA proteomics datasets: from the PRIDE database to Expression Atlas.
PMID 35701420 · PMC9197839 · Scientific data · 2022 · 8 claims · 5 setups
An open, containerised, Nextflow-orchestrated reanalysis pipeline combining metadata annotation, SWATH-MS analysis, statistical analysis, and Expression Atlas integration was developed for public DIA data.
-
Has reproduction · 43
Compression of structured high-throughput sequencing data.
PMID 24260313 · PMC3832420 · PloS one · 2013 · 8 claims · 7 setups
Leveraging an explicit data schema (separate field encoding, field modeling, template compression, domain modeling) enables stronger compression of HTS alignment data than general-purpose compression of serialized bytes.
-
Full-text index only
Evolutionary sequence analysis of complete eukaryote genomes.
PMID 15762985 · PMC1274250 · BMC bioinformatics · 2005 · 8 claims · 6 setups
A conservative genome-comparison method (MIA) identifies panorthologs — strict single-copy 1:1 orthologs containing only species divergences, no paralogy — to minimize errors from gene duplication in evolutionary sequence analysis.
-
Full-text index only
CLEAN: CLustering Enrichment ANalysis.
PMID 19640299 · PMC2734555 · BMC bioinformatics · 2009 · 8 claims · 4 setups
The gene-specific CLEAN score improves reproducibility of cluster analysis conclusions across independent datasets compared to the traditional cluster-wide score (cwCLEAN).