Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 87
Machine learning reveals microbial interactions driving plastic degradation across plastisphere environments.
PMID 41657981 · PMC12876002 · Frontiers in microbiology · 2025 · 6 claims · 7 setups
Wastewater plastispheres harbor the most diverse and compositionally even microbial communities among the three habitats.
-
Has reproduction
All of gene expression (AOE): An integrated index for public gene expression databases.
PMID 31978081 · PMC6980531 · PloS one · 2020 · 8 claims · 5 setups
AOE integrates publicly available gene expression data from GEO, ArrayExpress, and GEA into a single searchable index.
-
Has reproduction · 67
GEMmaker: process massive RNA-seq datasets on heterogeneous computational infrastructure.
PMID 35501696 · PMC9063052 · BMC bioinformatics · 2022 · 6 claims · 3 setups
GEMmaker, an nf-core compliant Nextflow workflow, can quantify gene expression from small to massive RNA-seq datasets while remaining reproducible via versioned containerized software.
-
Has reproduction · 50
Viewing RNA-seq data on the entire human genome.
PMID 28979763 · PMC5605993 · F1000Research · 2017 · 6 claims · 3 setups
RNA-Seq Viewer is a web application that visualizes genome-wide expression data from NCBI's SRA and GEO databases using an ideogram across the entire human genome.
-
Has reproduction · 65
FusionQ: a novel approach for gene fusion detection and quantification from paired-end RNA-Seq.
PMID 23768108 · PMC3691734 · BMC bioinformatics · 2013 · 8 claims · 8 setups
FusionQ is a novel tool that detects gene fusions, constructs chimerical transcript structures, and estimates their abundances from paired-end RNA-Seq data.
-
Full-text index only
C-terminal mutants of apolipoprotein L-I efficiently kill both Trypanosoma brucei brucei and Trypanosoma brucei rhodesiense.
PMID 19997494 · PMC2778949 · PLoS pathogens · 2009 · 8 claims · 8 setups
The C-terminal helix of apoL1 is entirely responsible for its interaction with SRA
-
Has reproduction · 90
Comparative Genomics Provides Insight into the Function of Broad-Host Range Sponge Symbionts.
PMID 34519538 · PMC8546597 · mBio · 2021 · 8 claims · 8 setups
Eleven new genomes were added to the Tethybacterales order and a novel family (Polydorabacteraceae) was identified
-
Has reproduction · 80
PanglaoDB: a web server for exploration of mouse and human single-cell RNA sequencing data.
PMID 30951143 · PMC6450036 · Database : the journal of biological databases and curation · 2019 · 7 claims · 7 setups
PanglaoDB is a web server providing pre-processed and pre-computed analyses of >1054 single-cell experiments (>4 million cells) from mouse and human across many tissues and platforms.
-
Has reproduction · 75
Identification of Novel Therapeutic Candidates Against SARS-CoV-2 Infections: An Application of RNA Sequencing Toward mRNA Based Nanotherapeutics.
PMID 35983322 · PMC9378778 · Frontiers in microbiology · 2022 · 6 claims · 7 setups
RPL29 (60S ribosomal protein L29) is highly/consistently expressed across all COVID-19 infected groups regardless of severity, suggesting it as a novel host therapeutic target for mRNA-based nanomedicines.
-
Has reproduction · 95
MetaMap: an atlas of metatranscriptomic reads in human disease-related RNA-seq data.
PMID 29901703 · PMC6025204 · GigaScience · 2018 · 6 claims · 7 setups
A two-step 'omni' RNA-seq pipeline (MetaMap) combining STAR human alignment with CLARK-S metagenomic classification can quantify archaeal, bacterial, and viral reads from the non-human read fraction of human RNA-seq data
-
Has reproduction · 78
QuasiFlow: a Nextflow pipeline for analysis of NGS-based HIV-1 drug resistance data.
PMID 36699347 · PMC9722223 · Bioinformatics advances · 2022 · 6 claims · 8 setups
QuasiFlow is a Nextflow pipeline that runs entirely locally via command-line tools and a local HIVdb database copy to analyze NGS-based HIV-1 drug resistance testing data.
-
Has reproduction · 54
Single-Cell Transcriptome Analysis Revealed Heterogeneity and Identified Novel Therapeutic Targets for Breast Cancer Subtypes.
PMID 37190091 · PMC10137100 · Cells · 2023 · 8 claims · 8 setups
Single-cell transcriptomic analysis of EPCAM+Lin- epithelial cells identified unique gene signatures/markers that distinguish ER+, HER2+, ER+HER2+, and TNBC molecular subtypes
-
Full-text index only
ChimerDB 2.0--a knowledgebase for fusion genes updated.
PMID 19906715 · PMC2808913 · Nucleic acids research · 2010 · 8 claims · 4 setups
ChimerDB 2.0 is an updated knowledgebase integrating fusion transcripts from GenBank transcriptome analysis with Sanger CGP, OMIM, PubMed, and Mitelman's database data.
-
Has reproduction · 99
getSequenceInfo: a suite of tools allowing to get genome sequence information from public repositories.
PMID 35804320 · PMC9264741 · BMC bioinformatics · 2022 · 8 claims · 8 setups
getSequenceInfo (gSeqI) allows programmatic (CLI) or GUI-based retrieval of sequence data and metadata from GenBank, RefSeq, and ENA across Linux, MacOS, and Windows.
-
Has reproduction · 90
Optimal Dual RNA-Seq Mapping for Accurate Pathogen Detection in Complex Eukaryotic Hosts.
PMID 39959292 · PMC11825298 · Bio-protocol · 2025 · 7 claims · 6 setups
Mapping adapter-trimmed reads first to the pathogen genome recovers more pathogen reads than the traditional host-first mapping approach.
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.
-
Has reproduction · 89
DFAST and DAGA: web-based integrated genome annotation tools and resources.
PMID 27867804 · PMC5107635 · Bioscience of microbiota, food and health · 2016 · 8 claims · 7 setups
DFAST is a web-based bacterial genome annotation and DDBJ submission pipeline with integrated CheckM quality assessment and ANI taxonomic assessment.
-
Has reproduction · 80
Curation of over 10 000 transcriptomic studies to enable data reuse.
PMID 33599246 · PMC7904053 · Database : the journal of biological databases and curation · 2021 · 8 claims · 6 setups
Gemma is a curated database and bioinformatics system that addresses metadata, probe annotation, and expression data inconsistencies in GEO to enable transcriptomic data reuse
-
Has reproduction · 100
The genome of the ant Tetramorium bicarinatum reveals a tandem organization of venom peptides genes allowing the prediction of their regulatory and evolutionary profiles.
PMID 38245722 · PMC10800049 · BMC genomics · 2024 · 8 claims · 8 setups
44 venom peptide genes were identified, distributed across four of the eleven chromosomes and organized in tandem repeat clusters.
-
Has reproduction · 50
rMAP: the Rapid Microbial Analysis Pipeline for ESKAPE bacterial group whole-genome sequence data.
PMID 34110280 · PMC8461470 · Microbial genomics · 2021 · 8 claims · 8 setups
rMAP is a pipeline capable of profiling the resistomes of ESKAPE pathogens using Illumina WGS data