Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins.
PMID 15608248 · PMC539979 · Nucleic acids research · 2005 · 7 claims · 5 setups
RefSeq provides a curated, non-redundant, explicitly linked collection of genomic, transcript and protein sequences spanning prokaryotes, eukaryotes and viruses.
-
Full-text index only
TRED: a Transcriptional Regulatory Element Database and a platform for in silico gene regulation studies.
PMID 15608156 · PMC539958 · Nucleic acids research · 2005 · 8 claims · 5 setups
TRED is a database collecting both cis-regulatory elements (promoters) and trans-regulatory elements (transcription factor binding/regulation data) with linked access.
-
Full-text index only
Large-scale structural analysis of the core promoter in mammalian and plant genomes.
PMID 16049029 · PMC1181242 · Nucleic acids research · 2005 · 8 claims · 7 setups
DNA encodes at least two independent levels of functional information: protein/TF-binding sequence information and physical/structural properties of the molecule itself.
-
Full-text index only
High-quality acinar cell isolation enables single-cell analysis of healthy and injured pancreas.
PMID 42013858 · PMC13198085 · Cell reports methods · 2026 · 8 claims · 7 setups
The DCTC protocol isolates up to 90% acinar cells from healthy wild-type pancreatic tissue without cell fixation or dead-cell removal kits
-
Full-text index only
TRED: a transcriptional regulatory element database, new entries and other development.
PMID 17202159 · PMC1899102 · Nucleic acids research · 2007 · 8 claims · 3 setups
TRED collects mammalian cis- and trans-regulatory elements together with experimental evidence, mapped onto assembled genomes
-
Has reproduction · 85
An extensive evaluation of read trimming effects on Illumina NGS data analysis.
PMID 24376861 · PMC3871669 · PloS one · 2013 · 8 claims · 8 setups
Read trimming increases the quality and reliability of downstream NGS analyses (RNA-Seq mapping, SNP identification, genome assembly) while reducing execution time and computational resources.
-
Has reproduction · 92
Large-scale integration of single-cell transcriptomic data captures transitional progenitor states in mouse skeletal muscle regeneration.
PMID 34773081 · PMC8589952 · Communications biology · 2021 · 8 claims · 7 setups
Large-scale integration of 111 sc/snRNAseq datasets captures rare, transitional myogenic progenitor states (commitment and fusion) that are poorly represented in individual datasets.
-
Full-text index only
Optimizing data-driven excellence: Canada's approach to using pathogen test datasets for quality control, pipeline development and training initiatives.
PMID 41591806 · PMC12847982 · Microbial genomics · 2026 · 8 claims · 5 setups
Standardized SARS-CoV-2 test datasets (Illumina and Nanopore) were developed as benchmarks for validating sequencing/bioinformatics pipelines across Canadian public health labs
-
Has reproduction · 80
VGEA: an RNA viral assembly toolkit.
PMID 34567846 · PMC8428259 · PeerJ · 2021 · 8 claims · 5 setups
VGEA is a Snakemake workflow that chains existing tools (fastp, BWA, SAMtools, IVA, shiver, SeqKit, QUAST, MultiQC) into an all-in-one RNA viral genome assembly pipeline
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
Single-molecule sequencing of an individual human genome.
PMID 19668243 · PMC4117198 · Nature biotechnology · 2009 · 8 claims · 7 setups
Single-molecule sequencing without cloning, amplification or ligation can sequence an individual human genome on one instrument by a single operator in four runs
-
Full-text index only
The genome sequence of the Woodland Grayling, Hipparchia fagi (Scopoli, 1763) (Lepidoptera: Nymphalidae).
PMID 42078576 · PMC13133624 · Wellcome open research · 2026 · 8 claims · 8 setups
A chromosome-level, haplotype-resolved genome assembly was produced for Hipparchia fagi (Woodland Grayling) as part of Project Psyche.
-
Has reproduction · 77
Accurate chromatin marks peak calling with Omnipeak.
PMID 41521664 · PMC12784980 · Nucleic acids research · 2026 · 8 claims · 6 setups
Omnipeak is a universal unsupervised peak-calling algorithm based on a constrained three-state hidden Markov model (zero, noise, signal states)
-
Has reproduction · 87
De Novo Transcriptome Meta-Assembly of the Mixotrophic Freshwater Microalga Euglena gracilis.
PMID 34072576 · PMC8227486 · Genes · 2021 · 7 claims · 8 setups
A new consensus transcriptome of E. gracilis was assembled by combining reads from five independent RNA-seq studies (23 samples)
-
Full-text index only
Adjustment of genomic waves in signal intensities from whole-genome SNP genotyping platforms.
PMID 18784189 · PMC2577347 · Nucleic acids research · 2008 · 8 claims · 6 setups
Genomic waves are present in both Illumina and Affymetrix SNP genotyping arrays, confirming they are not platform-specific
-
Full-text index only
StrainMake: reproducible hybrid metagenomics with MAG recovery and strain-level resolution.
PMID 42097292 · PMC13188985 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 5 setups
StrainMake is a Snakemake-based, Conda-managed workflow for de novo metagenomic analysis from short, long, or hybrid sequencing data.
-
Full-text index only
The genome sequence of the Eastern Rock Grayling, Hipparchia syriaca (Staudinger, 1871) (Lepidoptera: Nymphalidae).
PMID 41913757 · PMC13033134 · Wellcome open research · 2026 · 8 claims · 8 setups
A chromosome-level genome assembly was generated for Hipparchia syriaca (Eastern Rock Grayling) from a female specimen collected in Măcin, Romania
-
Full-text index only
The genome sequence of the Acorn Weevil, Curculio glandium (T.Marsham, 1802) (Coleoptera: Curculionidae).
PMID 42021765 · PMC13096787 · Wellcome open research · 2026 · 8 claims · 8 setups
A chromosomally complete genome assembly was produced for Curculio glandium (Acorn Weevil) as part of the Darwin Tree of Life project
-
Has reproduction · 71
Hyb: a bioinformatics pipeline for the analysis of CLASH (crosslinking, ligation and sequencing of hybrids) data.
PMID 24211736 · PMC3969109 · Methods (San Diego, Calif.) · 2014 · 8 claims · 6 setups
The 'hyb' pipeline detects, calls, folds and annotates chimeric reads from CLASH high-throughput sequencing data.