Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Aerobic nonylphenol degradation and nitro-nonylphenol formation by microbial cultures from sediments.
PMID 20043151 · PMC2825322 · Applied microbiology and biotechnology · 2010 · 8 claims · 8 setups
Aerobic biodegradation of branched NP in polluted river sediment occurs within 8 days after a 2-day lag phase at 30°C
-
Has reproduction · 64
GeMI: interactive interface for transformer-based Genomic Metadata Integration.
PMID 35657113 · PMC9216561 · Database : the journal of biological databases and curation · 2022 · 8 claims · 5 setups
GeMI is a web tool that uses a fine-tuned GPT2 model to extract 15 structured key-value attributes from free-text GEO sample metadata.
-
Has reproduction
sRNAbench and sRNAtoolbox 2019: intuitive fast small RNA profiling and differential expression.
PMID 31114926 · PMC6602500 · Nucleic acids research · 2019 · 8 claims · 3 setups
sRNAtoolbox 2019 adds all major small RNA library preparation protocols (including UMI-based) to sRNAbench with automatic protocol-specific preprocessing.
-
Has reproduction · 89
HTSQualC is a flexible and one-step quality control software for high-throughput sequencing data analysis.
PMID 34548573 · PMC8455540 · Scientific reports · 2021 · 8 claims · 5 setups
HTSQualC is a standalone, one-step QC software that performs filtering and trimming of raw HTS data in a single run
-
Has reproduction · 92
Large-scale integration of single-cell transcriptomic data captures transitional progenitor states in mouse skeletal muscle regeneration.
PMID 34773081 · PMC8589952 · Communications biology · 2021 · 8 claims · 7 setups
Large-scale integration of 111 sc/snRNAseq datasets captures rare, transitional myogenic progenitor states (commitment and fusion) that are poorly represented in individual datasets.
-
Has reproduction · 76
Tracing human genetic histories and natural selection with precise local ancestry inference.
PMID 40379651 · PMC12084304 · Nature communications · 2025 · 7 claims · 7 setups
Orchestra, a two-stage LAI method combining a recombination-distance base layer with a deep learning (convolutional + attention) smoothing module, outperforms RFmix, FLARE and Gnomix in precision and recall across simulated admixture generations.
-
Full-text index only
Evaluating the performance of Affymetrix SNP Array 6.0 platform with 400 Japanese individuals.
PMID 18803882 · PMC2566316 · BMC genomics · 2008 · 8 claims · 5 setups
About 20% of the 909,622 SNPs on the SNP Array 6.0 are monomorphic in the Japanese population
-
Has reproduction · 100
Exploring Gene Expression Patterns in Alzheimer's Disease Using a Human Microarray Data Meta-Analysis.
PMID 41744654 · PMC12938635 · Biology · 2026 · 6 claims · 7 setups
AD brains show a distinct transcriptomic profile with up-regulation of immune/inflammation genes and down-regulation of synapse/neuronal-signaling genes
-
Has reproduction · 73
GREIN: An Interactive Web Platform for Re-analyzing GEO RNA-seq Data.
PMID 31110304 · PMC6527554 · Scientific reports · 2019 · 8 claims · 7 setups
GREIN is a web application providing user-friendly interfaces to manipulate, visualize, and analyze GEO RNA-seq data.
-
Has reproduction · 90
A Decentralized Kidney Transplant Biopsy Classifier for Transplant Rejection Developed Using Genes of the Banff-Human Organ Transplant Panel.
PMID 35619722 · PMC9128066 · Frontiers in immunology · 2022 · 6 claims · 6 setups
A random forest model trained solely on B-HOT panel genes (B-HOT Model) accurately classifies kidney transplant biopsies as NR, ABMR, or TCMR.
-
Full-text index only
High-throughput crystallography for structural genomics.
PMID 19765976 · PMC2764548 · Current opinion in structural biology · 2009 · 8 claims · 8 setups
SG programs use genomic sequence data to select structurally novel protein targets, avoiding proteins with known structural homologues