Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction
Inter- and Intraspecific Venom Variation in the Reclusive Rear-Fanged Black-Striped Snakes (Coniophanes).
PMID 41745774 · PMC12945099 · Toxins · 2026 · 8 claims · 4 setups
Coniophanes DVG transcriptomes contain 18 toxin families, with toxins comprising 38.8% to 66% of total transcriptome expression depending on species
-
Has reproduction · 99
Evaluation of taxonomic classification and profiling methods for long-read shotgun metagenomic sequencing datasets.
PMID 36513983 · PMC9749362 · BMC bioinformatics · 2022 · 8 claims · 7 setups
Long-read classifiers generally performed best among the 11 methods tested
-
Has reproduction · 95
transXpress: a Snakemake pipeline for streamlined de novo transcriptome assembly and annotation.
PMID 37016291 · PMC10074830 · BMC bioinformatics · 2023 · 6 claims · 7 setups
transXpress is a Snakemake pipeline that streamlines de novo transcriptome assembly, quantification, and annotation for non-model organisms
-
Has reproduction · 63
A worldwide map of swine short tandem repeats and their associations with evolutionary and environmental adaptations.
PMID 33892623 · PMC8063339 · Genetics, selection, evolution : GSE · 2021 · 8 claims · 8 setups
Identified 878,967 polymorphic STRs (pSTRs) from 394 deep-sequenced pig/Suidae genomes, the largest pSTR repository in pigs to date
-
Has reproduction · 50
BiRNA-BERT allows efficient RNA language modeling with adaptive tokenization.
PMID 41266599 · PMC12635123 · Communications biology · 2025 · 8 claims · 8 setups
BiRNA-BERT uses adaptive dual-tokenization that dynamically selects nucleotide-level (NUC) or byte-pair encoding (BPE) tokens based on input sequence length
-
Has reproduction · 74
ChIP-seq guidelines and practices of the ENCODE and modENCODE consortia.
PMID 22955991 · PMC3431496 · Genome research · 2012 · 8 claims · 8 setups
ENCODE/modENCODE define a set of working standards and guidelines for ChIP-seq covering antibody validation, experimental replication, sequencing depth, data/metadata reporting, and data quality assessment.
-
Has reproduction · 69
High-resolution transcriptome and genome-wide dynamics of RNA polymerase and NusA in Mycobacterium tuberculosis.
PMID 23222129 · PMC3553938 · Nucleic acids research · 2013 · 8 claims · 7 setups
NusA interacts with RNAP ubiquitously throughout the M. tuberculosis chromosome and its ChIP-seq profile mirrors RNAP distribution in both exponential and stationary phase, despite NusA not binding DNA directly.
-
Has reproduction · 66
RNAseq analysis of the parasitic nematode Strongyloides stercoralis reveals divergent regulation of canonical dauer pathways.
PMID 23145190 · PMC3493385 · PLoS neglected tropical diseases · 2012 · 8 claims · 8 setups
S. stercoralis possesses homologs of nearly all C. elegans dauer genes, but with significant differences in protein structure, developmental regulation, and gene family expansion.
-
Has reproduction · 82
Ordinal-level phylogenomics of the arthropod class Diplopoda (millipedes) based on an analysis of 221 nuclear protein-coding loci generated using next-generation sequence analyses.
PMID 24236165 · PMC3827447 · PloS one · 2013 · 8 claims · 8 setups
An ordinal-level phylogeny of Diplopoda reconstructed from 221 nuclear protein-coding loci (61,641 aligned amino acid columns) differs from existing classifications in fundamental ways.
-
Full-text index only
ECgene: genome annotation for alternative splicing.
PMID 15608289 · PMC540072 · Nucleic acids research · 2005 · 8 claims · 5 setups
ECgene combines genome-based EST clustering with a graph-theoretic transcript assembly procedure to predict gene models including alternative splicing events.
-
Full-text index only
Systematic identification of pseudogenes through whole genome expression evidence profiling.
PMID 16945953 · PMC1636364 · Nucleic acids research · 2006 · 8 claims · 8 setups
Developed a novel bioinformatics method that identifies pseudogenes by profiling whole-genome transcript and protein expression evidence
-
Has reproduction · 62
Metatranscriptomics of the human oral microbiome during health and disease.
PMID 24692635 · PMC3977359 · mBio · 2014 · 8 claims · 8 setups
Disease-associated periodontal communities display conserved community-level metabolic gene expression profiles between patients, whereas the metabolic gene expression of individual species is highly variable between patients.
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.
-
Full-text index only
The impact of peptide abundance and dynamic range on stable-isotope-based quantitative proteomic analyses.
PMID 18798661 · PMC2746028 · Journal of proteome research · 2008 · 8 claims · 7 setups
Over half of confidently identified peptides in complex mixtures have S/N ratios below 10 on both FT-ICR and Orbitrap instruments
-
Full-text index only
The Neandertal genome and ancient DNA authenticity.
PMID 19661919 · PMC2725275 · The EMBO journal · 2009 · 8 claims · 6 setups
Only direct assays of DNA sequence positions where Neandertals differ from all contemporary humans can reliably estimate human contamination.
-
Full-text index only
Functional restoration of BRCA2 protein by secondary BRCA2 mutations in BRCA2-mutated ovarian carcinoma.
PMID 19654294 · PMC2754824 · Cancer research · 2009 · 8 claims · 7 setups
Secondary BRCA2 mutations restore BRCA2 protein/reading frame and thereby cause acquired platinum and PARP inhibitor resistance in BRCA2-mutated ovarian carcinoma
-
Full-text index only
Genomic signatures of migratory preference and historical whaling in eastern South Pacific humpback whales.
PMID 41986456 · PMC13161203 · Communications biology · 2026 · 7 claims · 8 setups
Nuclear genomic data show no clear population structure among feeding grounds, indicating panmixia despite divergent migratory destinations
-
Full-text index only
Deciphering the RNA landscapes on mammalian cell surfaces.
PMID 40973153 · PMC12959771 · Protein & cell · 2026 · 7 claims · 8 setups
AMOUR, a T7 RNA polymerase-based in-situ linear amplification method, enables high-throughput profiling of outer membrane surface RNAs while preserving plasma membrane integrity, including in rare primary cell populations.
-
Full-text index only
SnakeAltPromoter Facilitates Differential Alternative Promoter Analysis.
PMID 41993886 · PMC13082578 · Computational and structural biotechnology journal · 2026 · 6 claims · 6 setups
SnakeAltPromoter is the first unified, reproducible Snakemake workflow that automates alternative promoter analysis from raw RNA-seq data using 3 complementary methods (ProActiv, Salmon, DEXSeq)
-
Has reproduction · 76
Transcriptional landscape of repetitive elements in normal and cancer human cells.
PMID 25012247 · PMC4122776 · BMC genomics · 2014 · 8 claims · 8 setups
RepEnrich, a computational method that uses all mapping reads (uniquely mapping plus multi-mapping reads assigned to repetitive element subfamily assemblies/pseudogenomes), quantifies genome-wide repetitive element enrichment