Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Full-text index only
MBGD update 2010: toward a comprehensive resource for exploring microbial genome diversity.
PMID 19906735 · PMC2808943 · Nucleic acids research · 2010 · 8 claims · 6 setups
MBGD allows users to create ortholog groups using a specified subgroup of organisms, distinguishing it from other comparative genomics resources
-
Has reproduction · 38
Genomic capacities for Reactive Oxygen Species metabolism across marine phytoplankton.
PMID 37098087 · PMC10128935 · PloS one · 2023 · 8 claims · 3 setups
Genes encoding superoxide (O2•−) scavenging are ubiquitous across phytoplankton, but their fractional gene allocation decreases with increasing cell radius, consistent with a nearly fixed core gene set.
-
Full-text index only
Fast and systematic genome-wide discovery of conserved regulatory elements using a non-alignment based approach.
PMID 15693947 · PMC551538 · Genome biology · 2005 · 7 claims · 8 setups
FastCompare, a non-alignment-based, linear-time algorithm, computes a genome-wide conservation score for all k-mers (7-9 nt) between two genomes to identify conserved regulatory elements
-
Full-text index only
Stable patterns of gene expression regulating carbohydrate metabolism determined by geographic ancestry.
PMID 20016837 · PMC2790609 · PloS one · 2009 · 8 claims · 6 setups
151 'geo-ancestral genes' were identified that are both differentially expressed between AA and CAU subjects and contain SNPs distinguishing YRI (African) from CEU (European) HapMap populations
-
Full-text index only
Dissecting microregulation of a master regulatory network.
PMID 18294391 · PMC2289817 · BMC genomics · 2008 · 8 claims · 6 setups
143 human miRNAs (termed p53-miRs) each contain at least one putative p53 binding site within 10 kb flanking sequence and are predicted to target at least one known gene
-
Has reproduction · 78
Fungal metabarcoding data integration framework for the MycoDiversity DataBase (MDDB).
PMID 32463383 · PMC7734503 · Journal of integrative bioinformatics · 2020 · 7 claims · 4 setups
Public fungal metabarcoding raw DNA data and their associated environmental metadata in sequence archives are heterogeneously annotated and lack a uniform processing pipeline, preventing large-scale biodiversity/distribution assessments.
-
Has reproduction · 68
Cell-type annotation with accurate unseen cell-type identification using multiple references.
PMID 37379341 · PMC10335708 · PLoS computational biology · 2023 · 8 claims · 4 setups
mtANN integrates multiple reference datasets and eight gene selection methods via ensemble learning (multiple deep classification models + majority voting) to improve cell-type annotation accuracy
-
Full-text index only
High resolution analysis of the human transcriptome: detection of extensive alternative splicing independent of transcriptional activity.
PMID 19804644 · PMC2768739 · BMC genetics · 2009 · 8 claims · 6 setups
The human GWSA uses exon body and exon-exon junction probes to directly measure over 280,000 known and predicted splicing events genome-wide.
-
Full-text index only
NGSTroubleFinder: a tool for detection and quantification of contamination and kinship across human NGS data.
PMID 41608734 · PMC12838523 · NAR genomics and bioinformatics · 2026 · 8 claims · 8 setups
NGSTroubleFinder detects cross-sample contamination, sample swaps, kinship, and sex mismatches from BAM/CRAM files without requiring additional variant-calling steps
-
Full-text index only
Automated recognition of retroviral sequences in genomic data--RetroTector.
PMID 17636050 · PMC1976444 · Nucleic acids research · 2007 · 8 claims · 8 setups
RetroTector uses 'fragment threading' (detection of chains of conserved retroviral motifs satisfying distance constraints) combined with LTR detection and protein reconstruction to identify ERVs in genomic sequences
-
Has reproduction · 88
AuPairWise: A Method to Estimate RNA-Seq Replicability through Co-expression.
PMID 27082953 · PMC4833304 · PLoS computational biology · 2016 · 7 claims · 6 setups
Sample-sample correlation of transcript abundances is a misleading measure of replicability for assessing differential expression, because it is dominated by gene-specific dynamic ranges rather than condition-dependent variation.
-
Has reproduction · 74
SpaGene: A Deep Adversarial Framework for Spatial Gene Imputation.
PMID 42146899 · PMC13176606 · Computational and structural biotechnology journal · 2026 · 8 claims · 6 setups
SpaGene improves average PCC and SSIM and reduces RMSE compared to 6 baseline methods (SpaGE, gimVI, Tangram, VISTA, spRefine, stDiff) across 8 diverse ST-SC dataset pairs under gene-holdout evaluation.
-
Has reproduction · 95
Utility of Triti-Map for bulk-segregated mapping of causal genes and regulatory elements in Triticeae.
PMID 35605195 · PMC9284283 · Plant communications · 2022 · 8 claims · 4 setups
Triti-Map is a computational package suite plus web interface specifically optimized for bulk-segregated gene mapping in Triticeae, accepting DNA-seq, RNA-seq/ChIP-seq, and traditional QTL data as input
-
Has reproduction · 61
TEMP: a computational method for analyzing transposable element polymorphism in populations.
PMID 24753423 · PMC4066757 · Nucleic acids research · 2014 · 8 claims · 8 setups
TEMP combines pair-end (discordant) read and split (soft-clipped) read information to identify both presence and absence of TE insertions in genomic DNA from heterogeneous/pooled samples.
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Full-text index only
HUPO Highlights.
PMID 19862759 · PMC4594800 · Proteomics · 2009 · 8 claims · 8 setups
Mass spectrometry analysis of human liver reference samples (French Reference liver + Huh7 hepatoma cells) achieves substantial human genome coverage via PeptideAtlas processing
-
Full-text index only
The Genographic Project public participation mitochondrial DNA database.
PMID 17604454 · PMC1904368 · PLoS genetics · 2007 · 7 claims · 4 setups
The Genographic Project created the largest standardized human mtDNA database to date, comprising 78,590 genotypes from the first 18 months of public participation.
-
Full-text index only
Variation analysis and gene annotation of eight MHC haplotypes: the MHC Haplotype Project.
PMID 18193213 · PMC2206249 · Immunogenetics · 2008 · 8 claims · 6 setups
Comparison of eight HLA-homozygous MHC haplotype sequences identified >44,000 variations (substitutions and indels), submitted to dbSNP
-
Full-text index only
Multi-species integration, alignment and annotation of single-cell RNA-seq data with CAMEX.
PMID 41723123 · PMC13035843 · Nature communications · 2026 · 8 claims · 6 setups
CAMEX outperforms state-of-the-art integration methods on cross-species scRNA-seq benchmarking datasets ranging from one to eleven species