Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Has reproduction · 78
Enhancing chemotherapy response prediction via matched colorectal tumor-organoid gene expression analysis and network-based biomarker selection.
PMID 39754813 · PMC11754497 · Translational oncology · 2025 · 6 claims · 8 setups
A consensus WGCNA approach combining matched tumor-organoid and independent organoid drug-response expression data identifies gene modules and hub genes predictive of 5-FU chemotherapy response
-
Full-text index only
GeneTide--Terra Incognita Discovery Endeavor: a new transcriptome focused member of the GeneCards/GeneNote suite of databases.
PMID 15608261 · PMC540076 · Nucleic acids research · 2005 · 8 claims · 7 setups
GeneTide integrates UniGene, DoTS, AceView, BLAT/GeneLoc genomic alignment, and GeneAnnot probe-set data into a unified Consensus/Uniqueness/Score scheme to associate ESTs with GeneCards genes
-
Full-text index only
Ensembl's 10th year.
PMID 19906699 · PMC2808936 · Nucleic acids research · 2010 · 8 claims · 8 setups
Ensembl provides comprehensive gene annotation and integrated genomic resources (variation, regulation, comparative genomics) across a growing set of chordate genomes
-
Full-text index only
Correlating novel variable and conserved motifs in the Hemagglutinin protein with significant biological functions.
PMID 18681973 · PMC2553082 · Virology journal · 2008 · 8 claims · 6 setups
14 MEME blocks were identified in the HA protein of H3N2 strains (1968-1999), with blocks 1, 2, 3, and 7 correlating with several biological functions
-
Full-text index only
Candidate vaccine sequences to represent intra- and inter-clade HIV-1 variation.
PMID 19812689 · PMC2753653 · PloS one · 2009 · 7 claims · 5 setups
Natural CTL immunodominance toward variable proteome regions increases epitope mismatch with challenge strains and recapitulates the escape-driven CTL failure seen in natural infection, contributing to HIV vaccine failure
-
Full-text index only
Genome-wide identification of specific oligonucleotides using artificial neural network and computational genomic analysis.
PMID 17518996 · PMC1892811 · BMC bioinformatics · 2007 · 7 claims · 4 setups
The IAB algorithm (integration of ANN and BLAST) identifies genome-wide specific oligos much faster than pure BLAST search while maintaining comparable success rate and cross homology
-
Has reproduction · 49
EDGE COVID-19: a web platform to generate submission-ready genomes from SARS-CoV-2 sequencing efforts.
PMID 35561186 · PMC9113274 · Bioinformatics (Oxford, England) · 2022 · 7 claims · 5 setups
EDGE COVID-19 (EC-19) is a web-based platform that automates QC, reference-based variant/consensus calling, lineage determination, and submission of SARS-CoV-2 genomes and metadata to GenBank, GISAID and INSDC for both Illumina and ONT data.
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.
-
Full-text index only
RAId_DbS: mass-spectrometry based peptide identification web server with knowledge integration.
PMID 18954448 · PMC2605478 · BMC genomics · 2008 · 7 claims · 4 setups
Constructed enhanced protein databases integrating annotated SAPs, PTMs, and disease associations for 17 organisms.
-
Full-text index only
Machine-learning approaches for classifying haplogroup from Y chromosome STR data.
PMID 18551166 · PMC2396484 · PLoS computational biology · 2008 · 8 claims · 5 setups
Y-STR allelic variability is partitioned more by differences among haplogroups than by differences among populations, suggesting Y-STRs carry haplogroup information
-
Full-text index only
TEPEAK: A novel method for identifying and characterizing polymorphic transposable elements in non-model species populations.
PMID 41494038 · PMC12788660 · PLoS computational biology · 2026 · 8 claims · 6 setups
TEPEAK identifies and characterizes polymorphic TEs in populations without any prior TE sequence or loci information, using only a chromosome-level reference assembly.
-
Full-text index only
Uncovering Cas9 PAM diversity through metagenomic mining and machine learning.
PMID 41656299 · PMC12996302 · Nature communications · 2026 · 8 claims · 6 setups
CRISPR-PAMdb is a publicly accessible database compiling Cas9 protein sequences from 3.8 million bacterial/archaeal genomes and PAM profiles from 7.4 million phage/plasmid sequences
-
Has reproduction · 99
Systematic benchmarking of tools for CpG methylation detection from nanopore sequencing.
PMID 34103501 · PMC8187371 · Nature communications · 2021 · 8 claims · 6 setups
Existing Nanopore methylation detection tools present a tradeoff between false positives and false negatives and show high dispersion relative to expected methylation frequency values.
-
Has reproduction · 78
Single duplex DNA sequencing with CODEC detects mutations with high sensitivity.
PMID 37106072 · PMC10181940 · Nature genetics · 2023 · 8 claims · 8 setups
CODEC concatenates both strands of an original DNA duplex into a single NGS read pair via an adapter quadruplex and strand-displacing extension, enabling single-duplex resolution
-
Has reproduction · 84
Towards reliable whole genome sequencing for outbreak preparedness and response.
PMID 35945497 · PMC9361258 · BMC genomics · 2022 · 7 claims · 4 setups
Amplicon-based Nanopore sequencing can rapidly generate whole genome sequences in samples with viral load up to Ct 33.
-
Full-text index only
Finding signals that regulate alternative splicing in the post-genomic era.
PMID 12429065 · PMC244920 · Genome biology · 2002 · 8 claims · 8 setups
Alternative splicing generates protein and regulatory diversity from a limited number of genes and modulates isoform levels in a cell-context-specific manner
-
Full-text index only
Fast and systematic genome-wide discovery of conserved regulatory elements using a non-alignment based approach.
PMID 15693947 · PMC551538 · Genome biology · 2005 · 7 claims · 8 setups
FastCompare, a non-alignment-based, linear-time algorithm, computes a genome-wide conservation score for all k-mers (7-9 nt) between two genomes to identify conserved regulatory elements
-
Full-text index only
Manual annotation and analysis of the defensin gene cluster in the C57BL/6J mouse reference genome.
PMID 20003482 · PMC2807441 · BMC genomics · 2009 · 8 claims · 6 setups
Manual annotation of the mouse Chromosome 8 defensin region identifies 98 gene loci: 54 in the alpha-defensin cluster and 44 in the beta-defensin cluster
-
Has reproduction · 67
Satellitome Analysis and Transposable Elements Comparison in Geographically Distant Populations of Spodoptera frugiperda.
PMID 35455012 · PMC9026859 · Life (Basel, Switzerland) · 2022 · 8 claims · 5 setups
Most transposable elements are commonly shared across all eight geographically distant S. frugiperda samples, except Maverick and PIF/Harbinger elements which show divergent repeat copies