Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Host-pathogen studies in the post-genomic era.
PMID 11178231 · PMC138846 · Genome biology · 2000 · 7 claims · 7 setups
DNA arrays have been used to study host and/or pathogen gene expression for four viruses (HCMV, HHV8, HIV-1, HPV31) and two bacteria (Listeria monocytogenes, Salmonella)
-
Has reproduction · 90
Optimal Dual RNA-Seq Mapping for Accurate Pathogen Detection in Complex Eukaryotic Hosts.
PMID 39959292 · PMC11825298 · Bio-protocol · 2025 · 7 claims · 6 setups
Mapping adapter-trimmed reads first to the pathogen genome recovers more pathogen reads than the traditional host-first mapping approach.
-
Full-text index only
An Asilomar moment.
PMID 12372138 · PMC244904 · Genome biology · 2002 · 7 claims · 3 setups
The current push to restrict genomics/pathogen research funding and publication due to bioterrorism fears parallels the 1975 Asilomar Conference response to recombinant DNA risks.
-
Has reproduction · 85
Optimizing open data to support one health: best practices to ensure interoperability of genomic data from bacterial pathogens.
PMID 33103064 · PMC7568946 · One health outlook · 2020 · 8 claims · 3 setups
An open-access pathogen surveillance database (NCBI Pathogen Detection) plus contributor Best Practices enables FAIR, interoperable genomic data across human, animal, food, and environmental sources for One Health surveillance.
-
Has reproduction · 95
Reproducible, portable, and efficient ancient genome reconstruction with nf-core/eager.
PMID 33777521 · PMC7977378 · PeerJ · 2021 · 8 claims · 1 setups
nf-core/eager is a complete redesign and extension of the EAGER pipeline in Nextflow, built within the nf-core framework to ensure high-quality, sustainable software development.
-
Full-text index only
Dark matter in a deep-sea vent and in human mouth.
PMID 17803764 · PMC2040194 · Environmental microbiology · 2007 · 7 claims · 8 setups
The first genome sequence from the uncultured TM7 phylum was obtained by capturing and sequencing DNA from a single cell using a microfluidic device, yielding a 2.86 Mb assembly with 3245 predicted genes.
-
Full-text index only
More biology from the sequence.
PMID 11532209 · PMC138951 · Genome biology · 2001 · 8 claims · 8 setups
The Schizosaccharomyces pombe genome has been sequenced to completion with no gaps, telomere to telomere.
-
Has reproduction · 67
Essential Genes of Vibrio anguillarum and Other Vibrio spp. Guide the Development of New Drugs and Vaccines.
PMID 34745063 · PMC8564382 · Frontiers in microbiology · 2021 · 7 claims · 7 setups
Tn-seq using the TnSC189 mariner transposon identified 329 essential genes in V. anguillarum NB10Sm from a library of 52,662 insertion mutants.
-
Has reproduction
Human Retrotransposons and Effective Computational Detection Methods for Next-Generation Sequencing Data.
PMID 36295018 · PMC9605557 · Life (Basel, Switzerland) · 2022 · 8 claims · 7 setups
Transposable elements make up nearly 45% of the human genome, vastly exceeding the ~1.5% that is protein-coding.
-
Has reproduction · 54
Profiling chromatin accessibility responses in human neutrophils with sensitive pathogen detection.
PMID 34145026 · PMC8321655 · Life science alliance · 2021 · 6 claims · 6 setups
ATAC-seq reveals unique neutrophil chromatin accessibility changes in response to different stimuli before transcriptional activation, with most differential regions being challenge-specific in position, function, and motif.
-
Full-text index only
Thirty years into the genomics era: tumor viruses led the way.
PMID 17940630 · PMC1994808 · The Yale journal of biology and medicine · 2006 · 8 claims · 8 setups
Restriction endonuclease-based analysis and sequencing of small tumor virus genomes (SV40, phiX174) established the technical framework later used to sequence bacterial and human genomes
-
Has reproduction · 96
Deep learning based protocol to construct an immune-related gene network of host-pathogen interactions in plants.
PMID 36525344 · PMC9791427 · STAR protocols · 2023 · 6 claims · 6 setups
A deep-learning protocol (DLNet) ranks genes by their contribution to classifying treatment versus control expression data, identifying genes involved in host defense against pathogens.
-
Full-text index only
Massively parallel pyrosequencing in HIV research.
PMID 18614863 · PMC4221253 · AIDS (London, England) · 2008 · 8 claims · 8 setups
Massively parallel pyrosequencing platforms (454/Roche, Solexa/Illumina) enable very high-throughput DNA sequencing, up to ~1 billion bases per run
-
Full-text index only
An "omics" approach to uropathogenic Escherichia coli vaccinology.
PMID 19758805 · PMC2770165 · Trends in microbiology · 2009 · 8 claims · 8 setups
An 'omics'-based screening strategy integrating genomic, proteomic, and metabolomic data can identify PASivE UPEC proteins as vaccine candidates
-
Full-text index only
Molecular genomic approaches to infectious diseases in resource-limited settings.
PMID 19855820 · PMC2745561 · PLoS medicine · 2009 · 8 claims · 5 setups
Researchers in most developing countries lack the technology, resources, and capacity to participate fully in genomics research.
-
Has reproduction · 97
Invasive bacterial disease trends and characterization of group B streptococcal isolates among young infants in southern Mozambique, 2001-2015.
PMID 29351318 · PMC5774717 · PloS one · 2018 · 7 claims · 6 setups
A notable young infant GBS disease burden persisted during 2001–2015 despite significant declines in overall IBD, neonatal mortality, and stillbirth rates.
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.