Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The past and future of tuberculosis research.
PMID 19855821 · PMC2745564 · PLoS pathogens · 2009 · 8 claims · 6 setups
Integrating systems biology with epidemiology ('systems epidemiology') will be required to better predict TB's trajectory and eliminate the disease
-
Full-text index only
Evolution of two distinct phylogenetic lineages of the emerging human pathogen Mycobacterium ulcerans.
PMID 17900363 · PMC2098775 · BMC evolutionary biology · 2007 · 8 claims · 4 setups
M. ulcerans has evolved into five InDel haplotypes that separate into two distinct phylogenetic lineages: a 'classical' lineage (Africa, Australia, South East Asia) and an 'ancestral' lineage (Asia, South America, Mexico)
-
Full-text index only
Genome sequences and great expectations.
PMID 11178275 · PMC150431 · Genome biology · 2001 · 8 claims · 3 setups
Function is known or can be predicted for an average of 62% of proteins across 31 analyzed genomes.
-
Has reproduction · 88
Comprehensive benchmarking of large language models for RNA secondary structure prediction.
PMID 40205851 · PMC11982019 · Briefings in bioinformatics · 2025 · 7 claims · 4 setups
Existing RNA-LLMs had not previously been evaluated for secondary structure prediction in a unified, fair experimental setup with the same datasets and prediction model.
-
Full-text index only
Compressing DNA sequence databases with coil.
PMID 18489794 · PMC2426707 · BMC bioinformatics · 2008 · 8 claims · 1 setups
coil achieves higher compression ratio than state-of-the-art general-purpose compression tools on a large GenBank EST database file
-
Full-text index only
Integration of cytogenetic landmarks into the draft sequence of the human genome.
PMID 11237021 · PMC7845515 · Nature · 2001 · 8 claims · 6 setups
7,600 cytogenetically defined landmarks (from a set of 8,877 clones) were placed on the draft sequence of the human genome as a public resource
-
Full-text index only
Predicting failure rate of PCR in large genomes.
PMID 18492719 · PMC2441781 · Nucleic acids research · 2008 · 7 claims · 8 setups
The number of predicted primer-binding sites in genomic DNA is the most important factor determining PCR failure.
-
Full-text index only
ORFer--retrieval of protein sequences and open reading frames from GenBank and storage into relational databases or text files.
PMID 12493080 · PMC139979 · BMC bioinformatics · 2002 · 6 claims · 6 setups
ORFer retrieves protein and nucleic acid sequences and annotations from NCBI GenBank using the XML sequence format
-
Full-text index only
Completing the map of human genetic variation.
PMID 17495918 · PMC2685471 · Nature · 2007 · 8 claims · 5 setups
A community resource initiative will sequence fosmid and BAC clone libraries from 62 HapMap individuals to systematically discover and resolve structural genetic variants at nucleotide resolution
-
Full-text index only
Integrative functional genomics.
PMID 15239826 · PMC463286 · Genome biology · 2004 · 8 claims · 8 setups
Ultra-conserved noncoding elements exist across human, mouse and rat genomes at very high sequence identity, often far from genes
-
Full-text index only
Human genome research in China.
PMID 15168679 · PMC7079922 · Journal of molecular medicine (Berlin, Germany) · 2004 · 8 claims · 8 setups
China completed its assigned 1% share of the international Human Genome Project sequencing effort and contributed ~10% of the HapMap effort
-
Full-text index only
Biologic diversity of polyomavirus BK genomic sequences: Implications for molecular diagnostic laboratories.
PMID 18712842 · PMC2906129 · Journal of medical virology · 2008 · 8 claims · 5 setups
Coverage of naturally occurring BKV strains varies substantially among current PCR diagnostic assays due to primer/probe mismatches
-
Full-text index only
Grammar-based distance in progressive multiple sequence alignment.
PMID 18616828 · PMC2478692 · BMC bioinformatics · 2008 · 7 claims · 3 setups
A grammar-based (LZ complexity) distance metric can be used to determine the order in which sequences are progressively pairwise aligned
-
Full-text index only
Physiology engages with functional genomics - at last.
PMID 16086845 · PMC1273626 · Genome biology · 2005 · 8 claims · 8 setups
Large-scale QTL phenotyping in rat strains reveals that most hypertension-related traits are sexually dimorphic
-
Full-text index only
High-accuracy proteome maps of human body fluids.
PMID 17140426 · PMC1794581 · Genome biology · 2006 · 8 claims · 5 setups
Large-scale, high-accuracy MS analyses of tear fluid, urine, and seminal plasma provide high-quality datasets useful for biomarker discovery
-
Full-text index only
Viruses take center stage in cellular evolution.
PMID 16787527 · PMC1779534 · Genome biology · 2006 · 8 claims · 6 setups
A large poxvirus-like dsDNA virus may have become the eukaryotic nucleus after being taken in by an ancestral cell (viral eukaryogenesis)
-
Full-text index only
Genetic analysis of completely sequenced disease-associated MHC haplotypes identifies shuffling of segments in recent human history.
PMID 16440057 · PMC1331980 · PLoS genetics · 2006 · 7 claims · 6 setups
Complete 4.25-Mb sequence of the QBL haplotype was determined by BAC shotgun sequencing and compared with PGF (reference) and COX haplotypes
-
Full-text index only
Complete genome sequence and comparative analysis of the wild-type commensal Escherichia coli strain SE11 isolated from a healthy adult.
PMID 18931093 · PMC2608844 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2008 · 8 claims · 6 setups
The SE11 genome comprises a 4.8 Mb chromosome encoding 4679 protein-coding genes and six plasmids encoding 323 protein-coding genes
-
Full-text index only
Screening of male breast cancer and of breast-ovarian cancer families for BRCA2 mutations using large bifluorescent amplicons.
PMID 11207042 · PMC2363770 · British journal of cancer · 2001 · 8 claims · 4 setups
FAMA using large bifluorescent amplicons (avg 1.2 kb) with chemical cleavage of mismatch, combined with DGGE for 9 small exons, allows sensitive, unbiased scanning of the entire BRCA2 coding sequence with few amplicons (15 FAMA + 9 DGGE)
-
Full-text index only
Bench-to-bedside review: fulfilling promises of the Human Genome Project.
PMID 12133180 · PMC137447 · Critical care (London, England) · 2002 · 8 claims · 4 setups
Most common diseases and many drug responses are influenced by inherited genetic variation