Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Population genomics: diversity and virulence in the Neisseria.
PMID 18822386 · PMC2612085 · Current opinion in microbiology · 2008 · 8 claims · 8 setups
Horizontal genetic exchange (recombination) plays a central role in shaping bacterial speciation and population structure in Neisseria
-
Full-text index only
Applying proteomics to the diagnosis and treatment of ALS and related diseases.
PMID 19670321 · PMC2836583 · Muscle & nerve · 2009 · 8 claims · 8 setups
Protein-based biomarkers for ALS/MND require further verification and large-scale validation/qualification studies, including disease mimics, before clinical use
-
Full-text index only
Variation analysis and gene annotation of eight MHC haplotypes: the MHC Haplotype Project.
PMID 18193213 · PMC2206249 · Immunogenetics · 2008 · 8 claims · 6 setups
Comparison of eight HLA-homozygous MHC haplotype sequences identified >44,000 variations (substitutions and indels), submitted to dbSNP
-
Has reproduction · 82
Landscape of allele-specific transcription factor binding in the human genome.
PMID 33980847 · PMC8115691 · Nature communications · 2021 · 8 claims · 6 setups
A novel statistical framework (ADASTRA) calls allele-specific TF binding from existing ChIP-Seq alignments by jointly correcting for background allelic dosage (BAD, from aneuploidy/CNVs) and reference mapping bias.
-
Full-text index only
Combining transcriptional profiling and genetic linkage analysis to uncover gene networks operating in hematopoietic stem cells and their progeny.
PMID 18560825 · PMC2493868 · Immunogenetics · 2008 · 8 claims · 8 setups
Neither transcriptional profiling alone nor genetic linkage analysis alone has been an effective approach to identify genes or gene networks that specify stemness or initiate differentiation/lineage specification.
-
Full-text index only
Relationships between emm and multilocus sequence types within a global collection of Streptococcus pyogenes.
PMID 18405369 · PMC2359762 · BMC microbiology · 2008 · 8 claims · 5 setups
emm type is often a poor marker for clonal genetic background across the global S. pyogenes collection
-
Has reproduction · 78
IsomiR_Window: a system for analyzing small-RNA-seq data in an integrative and user-friendly manner.
PMID 33522913 · PMC7852101 · BMC bioinformatics · 2021 · 8 claims · 2 setups
IsomiR Window is an integrated, user-friendly platform that systematically identifies, quantifies, and functionally explores isomiR expression in small-RNA-seq datasets without requiring computational skills
-
Has reproduction · 86
Plasmid transmission dynamics and evolution of partner quality in a natural population of Rhizobium leguminosarum.
PMID 41212030 · PMC12691615 · mBio · 2025 · 8 claims · 8 setups
Of the four most frequent plasmid types, types II and III have more stable size, larger core genomes, and track the chromosomal phylogeny (more vertical transmission), while types I and IV (pSym) vary in size and gene content with phylogenies consistent with frequent horizontal transmission.
-
Full-text index only
Genomic expression during human myelopoiesis.
PMID 17683550 · PMC2045681 · BMC genomics · 2007 · 8 claims · 5 setups
An integrated myelopoiesis expression dataset of 9,425 genes, each mapped to a unique genomic position, was generated from 24 microarray experiments across 8 myeloid cell types.
-
Full-text index only
Sys-BodyFluid: a systematical database for human body fluid proteome research.
PMID 18978022 · PMC2686600 · Nucleic acids research · 2009 · 6 claims · 4 setups
Sys-BodyFluid is a web-based database integrating proteomic data from 11 human body fluids (plasma/serum, urine, cerebrospinal fluid, saliva, bronchoalveolar lavage fluid, synovial fluid, nipple aspirate fluid, tear fluid, seminal fluid, milk, amniotic fluid), containing over 10,000 proteins
-
Full-text index only
Multilocus sequence typing of Cronobacter sakazakii and Cronobacter malonaticus reveals stable clonal structures with clinical significance which do not correlate with biotypes.
PMID 19852808 · PMC2770063 · BMC microbiology · 2009 · 8 claims · 6 setups
A seven-locus MLST scheme (atpD, fusA, glnS, gltB, gyrB, infB, pps) reliably identifies and discriminates C. sakazakii and C. malonaticus strains
-
Has reproduction · 69
Automatic discovery of 100-miRNA signature for cancer classification using ensemble feature selection.
PMID 31533612 · PMC6751684 · BMC bioinformatics · 2019 · 7 claims · 8 setups
An ensemble feature selection method based on classifier consensus identifies a robust 100-miRNA signature from TCGA data.
-
Full-text index only
SeqBuster, a bioinformatic tool for the processing and analysis of small RNAs datasets, reveals ubiquitous miRNA modifications in human embryonic cells.
PMID 20008100 · PMC2836562 · Nucleic acids research · 2010 · 8 claims · 6 setups
SeqBuster is a versatile web-based and stand-alone bioinformatic toolkit for processing and analyzing large-scale small RNA deep sequencing datasets.
-
Has reproduction · 68
Cell-type annotation with accurate unseen cell-type identification using multiple references.
PMID 37379341 · PMC10335708 · PLoS computational biology · 2023 · 8 claims · 4 setups
mtANN integrates multiple reference datasets and eight gene selection methods via ensemble learning (multiple deep classification models + majority voting) to improve cell-type annotation accuracy
-
Full-text index only
Oncogene mutations, copy number gains and mutant allele specific imbalance (MASI) frequently occur together in tumor cells.
PMID 19826477 · PMC2757721 · PloS one · 2009 · 8 claims · 8 setups
Homozygous mutations of oncogenes are frequent (20%) across 833 cancer cell lines of 12 tumor types in the Sanger database
-
Has reproduction · 50
Wx: a neural network-based feature selection algorithm for transcriptomic data.
PMID 31324856 · PMC6642261 · Scientific reports · 2019 · 8 claims · 8 setups
The Wx algorithm ranks genes by a discriminative index (DI) score representing classification power for distinguishing given groups, enabling intuitive selection of optimal biomarker genes.
-
Full-text index only
An integrated database of genes responsive to the Myc oncogenic transcription factor: identification of direct genomic targets.
PMID 14519204 · PMC328458 · Genome biology · 2003 · 8 claims · 6 setups
The Myc Target Gene database integrates literature evidence to prioritize candidate Myc-responsive genes and cluster them into functional groups
-
Full-text index only
In vitro identification and in silico utilization of interspecies sequence similarities using GeneChip technology.
PMID 15871745 · PMC1156887 · BMC genomics · 2005 · 7 claims · 6 setups
Only 14±2% of canine transcripts were detected by U133A probe sets versus 49±6% of human transcripts when hybridized to the same chip
-
Full-text index only
Non-EST-based prediction of novel alternatively spliced cassette exons with cell signaling function in Caenorhabditis elegans and human.
PMID 17452356 · PMC1904267 · Nucleic acids research · 2007 · 8 claims · 7 setups
PASE (Prediction of Alternative Signaling Exons) is a computational algorithm combining Markov splice-site models, a Bayesian classifier, species conservation, and Scansite motif scoring to identify novel alternative cassette exons involved in cell signaling.
-
Full-text index only
Candidate vaccine sequences to represent intra- and inter-clade HIV-1 variation.
PMID 19812689 · PMC2753653 · PloS one · 2009 · 7 claims · 5 setups
Natural CTL immunodominance toward variable proteome regions increases epitope mismatch with challenge strains and recapitulates the escape-driven CTL failure seen in natural infection, contributing to HIV vaccine failure