Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 100
Analysis of the Taxonomy, Synteny, and Virulence Factors for Soft Rot Pathogen Pectobacterium aroidearum in Amorphophallus konjac Using Comparative Genomics.
PMID 35910650 · PMC9326479 · Frontiers in microbiology · 2022 · 8 claims · 8 setups
The causal agent of konjac soft rot in China is Pectobacterium aroidearum, confirmed via in vitro/in vivo pathogenicity tests, ANI, dDDH, and phylogenomic analysis.
-
Has reproduction · 43
TransFlow: a Snakemake workflow for transmission analysis of Mycobacterium tuberculosis whole-genome sequencing data.
PMID 36469333 · PMC9825751 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 8 setups
TransFlow is a Snakemake- and Conda-based workflow that combines state-of-the-art tools into a single, fast, scalable pipeline for MTBC WGS-based transmission analysis.
-
Has reproduction · 89
Improved eukaryotic detection compatible with large-scale automated analysis of metagenomes.
PMID 37032329 · PMC10084625 · Microbiome · 2023 · 8 claims · 7 setups
MAPQ ≥30 filtering improves precision but substantially reduces recall, especially for unrepresented/divergent eukaryotic taxa
-
Has reproduction · 71
Parsimonious Gene Correlation Network Analysis (PGCNA): a tool to define modular gene co-expression for refined molecular stratification in cancer.
PMID 30993001 · PMC6459838 · NPJ systems biology and applications · 2019 · 8 claims · 7 setups
Retaining only the top ~3 most correlated edges per gene (EPG3) combined with FastUnfold clustering (termed PGCNA) produces gene co-expression modules with significantly better separation and enrichment of known biology than using all edges or other clustering methods.
-
Has reproduction · 92
Evaluation of core genome and whole genome multilocus sequence typing schemes for Campylobacter jejuni and Campylobacter coli outbreak detection in the USA.
PMID 37133905 · PMC10272873 · Microbial genomics · 2023 · 8 claims · 8 setups
cgMLST, wgMLST and hqSNP WGS-based analysis methods clustered C. jejuni and C. coli isolates in concordance with epidemiological data.
-
Full-text index only
Constitutional genetic variation at the human aromatase gene (Cyp19) and breast cancer risk.
PMID 10027313 · PMC2362434 · British journal of cancer · 1999 · 7 claims · 5 setups
Allelic distribution of the Cyp19 intron 4 STRP differs significantly between breast cancer cases and controls
-
Full-text index only
An SVD-based comparison of nine whole eukaryotic genomes supports a coelomate rather than ecdysozoan lineage.
PMID 15606920 · PMC544558 · BMC bioinformatics · 2004 · 8 claims · 7 setups
SVD-based analysis of tetrapeptide frequency vectors can compare whole eukaryotic proteomes without pre-defining orthologs or aligning homologous sites
-
Has reproduction · 67
Optimal scaling of digital transcriptomes.
PMID 24223126 · PMC3819321 · PloS one · 2013 · 8 claims · 8 setups
Fifteen existing and novel transcript-count normalization algorithms can be compared with two novel, mutually independent metrics: the number of "uniform" genes (sufficiently low coefficient of variation after normalization) and low average Spearman correlation between normalized expression profiles of gene pairs.
-
Has reproduction · 49
oPOSSUM-3: advanced analysis of regulatory motif over-representation across genes or ChIP-Seq datasets.
PMID 22973536 · PMC3429929 · G3 (Bethesda, Md.) · 2012 · 8 claims · 6 setups
oPOSSUM-3 is a web-accessible system that identifies over-represented TFBS and TFBS families in DNA sequences of co-expressed genes or in sequences from high-throughput methods such as ChIP-Seq.
-
Full-text index only
Variation in conserved non-coding sequences on chromosome 5q and susceptibility to asthma and atopy.
PMID 16336695 · PMC1325232 · Respiratory research · 2005 · 6 claims · 8 setups
There is overall little sequence variation in the conserved non-coding elements (CNEs) on 5q31, including none detected in CNE-B/CNS-1
-
Full-text index only
Genetic analysis of the GLUT10 glucose transporter (SLC2A10) polymorphisms in Caucasian American type 2 diabetes.
PMID 16336637 · PMC1325051 · BMC medical genetics · 2005 · 7 claims · 5 setups
GLUT10 (SLC2A10) is a facilitative glucose transporter gene mapped within the T2DM-linked chromosome 20q12-13.1 region
-
Full-text index only
Leveraging human genomic information to identify nonhuman primate sequences for expression array development.
PMID 16288651 · PMC1314899 · BMC genomics · 2005 · 8 claims · 6 setups
Human genomic DNA sequence can be leveraged to obtain 3' end sequence of NHP orthologs, which can then be used to generate NHP oligonucleotide microarrays
-
Full-text index only
The distribution of SNPs in human gene regulatory regions.
PMID 16209714 · PMC1260019 · BMC genomics · 2005 · 8 claims · 6 setups
SNPs occur with higher density closer to the transcriptional start site within gene promoter regions than in further upstream regions
-
Full-text index only
Comprehensive search for intra- and inter-specific sequence polymorphisms among coding envelope genes of retroviral origin found in the human genome: genes and pseudogenes.
PMID 16150157 · PMC1236922 · BMC genomics · 2005 · 8 claims · 5 setups
HERV-W (envW) and HERV-FRD (envFRD) envelope genes, both specifically expressed in placenta, show strong sequence conservation with only two nonsynonymous SNPs identified across 91 individuals
-
Full-text index only
Coxiella burnetii genotyping.
PMID 16102309 · PMC3320512 · Emerging infectious diseases · 2005 · 8 claims · 5 setups
Multispacer sequence typing (MST) is the first reliable method for typing Coxiella burnetii isolates
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Full-text index only
Computational verification of protein-protein interactions by orthologous co-expression.
PMID 15740634 · PMC555590 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Co-expression of orthologous protein pairs across multiple species can verify/predict S. cerevisiae PPIs with better performance than S. cerevisiae co-expression alone.
-
Full-text index only
Variants in the vitamin D receptor gene and asthma.
PMID 15651992 · PMC546000 · BMC genetics · 2005 · 8 claims · 6 setups
VDR gene variants are candidate risk factors for asthma/allergy because vitamin D effects are mediated through the VDR and VDR maps to the 12q asthma linkage region
-
Full-text index only
Bioinformatic mapping of AlkB homology domains in viruses.
PMID 15627404 · PMC544882 · BMC genomics · 2005 · 8 claims · 8 setups
AlkB-like domains are found in at least 22 different single-stranded RNA positive-strand plant viruses, mainly within a subgroup of the Flexiviridae family.
-
Full-text index only
Coverage of whole proteome by structural genomics observed through protein homology modeling database.
PMID 17146617 · PMC1769342 · Journal of structural and functional genomics · 2006 · 8 claims · 7 setups
FAMSBASE, a homology-modeling database of whole-genome ORFs, currently covers about 50% of predicted ORFs (368,724 of 734,193) across 276 genomes with modeled 3D structures.