Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 98
Large-scale quality assessment of prokaryotic genomes with metashot/prok-quality.
PMID 35136576 · PMC8804904 · F1000Research · 2021 · 8 claims · 6 setups
metashot/prok-quality is a container-enabled Nextflow pipeline for quality assessment and dereplication of draft prokaryotic genomes
-
Has reproduction · 89
Identification of genes influencing the evolution of Escherichia coli ST372 in dogs and humans.
PMID 36752777 · PMC9997745 · Microbial genomics · 2023 · 8 claims · 8 setups
Dogs are the dominant host of E. coli ST372, and clusters within the ST372 population structure exhibit distinctive O:H types.
-
Has reproduction · 95
A role for ColV plasmids in the evolution of pathogenic Escherichia coli ST58.
PMID 35115531 · PMC8813906 · Nature communications · 2022 · 8 claims · 8 setups
ST58 contains a major sub-lineage (BAP2, n=363) characterized by near-ubiquitous carriage of ColV plasmids
-
Full-text index only
Evidence for a preferential targeting of 3'-UTRs by cis-encoded natural antisense transcripts.
PMID 16204454 · PMC1243798 · Nucleic acids research · 2005 · 8 claims · 4 setups
Cis-encoded natural antisense RNAs show striking preferential complementarity to 3′-UTRs of their target genes in human and mouse genomes
-
Full-text index only
ECgene: genome annotation for alternative splicing.
PMID 15608289 · PMC540072 · Nucleic acids research · 2005 · 8 claims · 5 setups
ECgene combines genome-based EST clustering with a graph-theoretic transcript assembly procedure to predict gene models including alternative splicing events.
-
Full-text index only
Comprehensive genome analysis of 203 genomes provides structural genomics with new insights into protein family space.
PMID 16481312 · PMC1373602 · Nucleic acids research · 2006 · 8 claims · 7 setups
The number of protein families continues to expand steadily as more genomes are sequenced, showing no sign of saturation.
-
Full-text index only
High-resolution aCGH and expression profiling identifies a novel genomic subtype of ER negative breast cancer.
PMID 17925008 · PMC2246289 · Genome biology · 2007 · 7 claims · 8 setups
A novel subtype of high-grade ER-negative breast cancer exists, characterized by a low genomic instability index (GII)
-
Has reproduction · 92
Evaluation of core genome and whole genome multilocus sequence typing schemes for Campylobacter jejuni and Campylobacter coli outbreak detection in the USA.
PMID 37133905 · PMC10272873 · Microbial genomics · 2023 · 8 claims · 8 setups
cgMLST, wgMLST and hqSNP WGS-based analysis methods clustered C. jejuni and C. coli isolates in concordance with epidemiological data.
-
Full-text index only
Comparative genomics of emerging human ehrlichiosis agents.
PMID 16482227 · PMC1366493 · PLoS genetics · 2006 · 8 claims · 5 setups
Ehrlichia spp. and Anaplasma spp. display a unique large expansion of immunodominant outer membrane proteins (OMP-1/P44/Msp2 family) facilitating antigenic variation
-
Full-text index only
Complete genome of Phenylobacterium zucineum--a novel facultative intracellular bacterium isolated from human erythroleukemia cell line K562.
PMID 18700039 · PMC2529317 · BMC genomics · 2008 · 8 claims · 6 setups
Complete genome of P. zucineum HLK1T consists of a 3,996,255 bp circular chromosome and a 382,976 bp circular plasmid encoding 3,861 proteins, 42 tRNAs, and one 16S-23S-5S rRNA operon
-
Full-text index only
Comparative genomics of vertebrate Fox cluster loci.
PMID 17062144 · PMC1634998 · BMC genomics · 2006 · 8 claims · 3 setups
Two additional human paralogous Fox cluster regions exist, on chromosomes 14 and 20, beyond the previously known chromosome 6 and 16 loci
-
Full-text index only
The most frequent short sequences in non-coding DNA.
PMID 19966278 · PMC2831315 · Nucleic acids research · 2010 · 8 claims · 2 setups
Short frequent sequences (9-14 bases) in non-coding DNA may play a role in maintaining chromosome structure and function
-
Has reproduction · 55
Natural clines and human management impact the genetic structure of Algerian honey bee populations.
PMID 38114899 · PMC10729559 · Genetics, selection, evolution : GSE · 2023 · 7 claims · 8 setups
No significant admixture from European subspecies was detected in Algerian honey bees, suggesting large-scale queen imports have not occurred in Algeria.
-
Has reproduction · 43
TransFlow: a Snakemake workflow for transmission analysis of Mycobacterium tuberculosis whole-genome sequencing data.
PMID 36469333 · PMC9825751 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 8 setups
TransFlow is a Snakemake- and Conda-based workflow that combines state-of-the-art tools into a single, fast, scalable pipeline for MTBC WGS-based transmission analysis.
-
Full-text index only
Evolutionary sequence analysis of complete eukaryote genomes.
PMID 15762985 · PMC1274250 · BMC bioinformatics · 2005 · 8 claims · 6 setups
A conservative genome-comparison method (MIA) identifies panorthologs — strict single-copy 1:1 orthologs containing only species divergences, no paralogy — to minimize errors from gene duplication in evolutionary sequence analysis.
-
Full-text index only
The MAPPER database: a multi-genome catalog of putative transcription factor binding sites.
PMID 15608292 · PMC540057 · Nucleic acids research · 2005 · 8 claims · 6 setups
Built a library of 1134 HMM models (359 matrix-derived, 718 factor-derived, 57 JASPAR-derived), corresponding to 863 distinct TF names, from TRANSFAC and JASPAR binding site data
-
Full-text index only
Inparanoid: a comprehensive database of eukaryotic orthologs.
PMID 15608241 · PMC540061 · Nucleic acids research · 2005 · 8 claims · 4 setups
The Inparanoid algorithm identifies true ortholog clusters by seeding on reciprocal best-matching pairs, gathering inparalogs (post-speciation duplicates) while excluding outparalogs (pre-speciation duplicates)
-
Full-text index only
The diploid genome sequence of an individual human.
PMID 17803354 · PMC1964779 · PLoS biology · 2007 · 7 claims · 6 setups
Generated an independently assembled diploid human genome sequence (HuRef) from both chromosome sets of a single individual using whole-genome shotgun Sanger sequencing
-
Full-text index only
BIPASS: BioInformatics Pipeline Alternative Splicing Services.
PMID 17584795 · PMC1933140 · Nucleic acids research · 2007 · 8 claims · 4 setups
BIPASS offers two complementary services for alternative splicing (AS) research: BIPAS-SpliceDB, a queryable pre-computed AS data warehouse, and BIPAS-Align&Splice, an online pipeline for user-submitted sequences.
-
Has reproduction · 54
Population structure analysis of Salmonella serovar Muenchen to redefine geno-serotyping using genome indexing approaches.
PMID 41743541 · PMC12929376 · Frontiers in microbiology · 2025 · 6 claims · 6 setups
Integrating genome-indexing (bettercallsal, DNA sketching + genome proximity) with SeqSero2 yields complementary serovar calls that improve discrimination of genomically distinct but antigenically similar serovars while retaining historical nomenclature