Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Next-generation high-density self-assembling functional protein arrays.
PMID 18469824 · PMC3070491 · Nature methods · 2008 · 8 claims · 7 setups
A next-generation NAPPA method produces high-density protein microarrays displaying over 1500 unique proteins with >90% expression success
-
Has reproduction · 90
The first single-stranded DNA virus targeting Pectobacterium belongs to the family Microviridae and demonstrates a broad host range to Pectobacterium brasiliense soft rot pathogens.
PMID 41739249 · PMC12935840 · Archives of virology · 2026 · 7 claims · 7 setups
Phage Mimer is the first described single-stranded DNA phage targeting Pectobacterium and belongs to the family Microviridae, subfamily Bullavirinae.
-
Full-text index only
Disturbed interaction of p21-rac with mutated p67-phox causes chronic granulomatous disease.
PMID 8879195 · PMC2192830 · The Journal of experimental medicine · 1996 · 6 claims · 8 setups
The patient is a compound heterozygote for a p67-phox gene mutation: an in-frame deletion of lysine 58 on one allele and an 11-13 kb genomic deletion on the other allele.
-
Full-text index only
A modified T-test feature selection method and its application on the HapMap genotype data.
PMID 18267305 · PMC5054219 · Genomics, proteomics & bioinformatics · 2007 · 7 claims · 4 setups
A modified t-test ranking measure, extended to handle nominal SNP genotype data via vector transformation, can effectively rank SNPs by their discriminative capability for population classification.
-
Full-text index only
Incorporation of genetic model parameters for cost-effective designs of genetic association studies using DNA pooling.
PMID 17634103 · PMC1947971 · BMC genomics · 2007 · 8 claims · 4 setups
A closed-form approximation to the F-test non-centrality parameter (NCP) incorporating genetic model parameters (disease allele frequency, marker allele frequency, prevalence, genotype relative risk, sample size, genetic model, number of pools/replicates, machine variability) can be used to compute power for DNA pooling association studies
-
Full-text index only
Long-term trends in evolution of indels in protein sequences.
PMID 17298668 · PMC1805498 · BMC evolutionary biology · 2007 · 8 claims · 5 setups
More than one third of protein domains show a statistically significant tendency to increase or decrease in size over evolutionary distance.
-
Has reproduction · 71
polishCLR: A Nextflow Workflow for Polishing PacBio CLR Genome Assemblies.
PMID 36792366 · PMC9985148 · Genome biology and evolution · 2023 · 8 claims · 8 setups
polishCLR is a reproducible, containerized Nextflow workflow that implements best practices for polishing PacBio CLR genome assemblies.
-
Has reproduction · 85
PowerBacGWAS: a computational pipeline to perform power calculations for bacterial genome-wide association studies.
PMID 35338232 · PMC8956664 · Communications biology · 2022 · 8 claims · 8 setups
Two computational approaches (sub-sampling and phenotype-simulation) can be implemented to perform power calculations for bacterial GWAS using existing genome collections, packaged as the PowerBacGWAS pipeline
-
Full-text index only
The use of edge-betweenness clustering to investigate biological function in protein interaction networks.
PMID 15740614 · PMC555937 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Edge-Betweenness clustering separates protein interaction graphs into subgraphs whose GO term distributions show significant correlations, revealing biologically meaningful functional modules.
-
Full-text index only
A Korean family with Arg1448Cys mutation of SCN4A channel causing paramyotonia congenita: electrophysiologic, histopathologic, and molecular genetic studies.
PMID 12483017 · PMC3054970 · Journal of Korean medical science · 2002 · 7 claims · 5 setups
A missense mutation (Arg1448Cys, R1448C) in SCN4A causes paramyotonia congenita in this Korean family
-
Full-text index only
GoMiner: a resource for biological interpretation of genomic and proteomic data.
PMID 12702209 · PMC154579 · Genome biology · 2003 · 8 claims · 4 setups
GoMiner organizes 'interesting' gene lists (e.g., differentially expressed genes) into the Gene Ontology hierarchy for biological interpretation, displaying results as both a tree and a directed acyclic graph (DAG).
-
Full-text index only
Molecular analysis of X-linked chronic granulomatous disease in five unrelated Korean patients.
PMID 15082894 · PMC2822302 · Journal of Korean medical science · 2004 · 8 claims · 4 setups
Five unrelated Korean X-linked CGD patients each carry a distinct CYBB gene mutation: c.1663insT, c.1111-1G>T, c.39_40insG, c.927delC, and c.434T>C
-
Has reproduction · 43
Compression of structured high-throughput sequencing data.
PMID 24260313 · PMC3832420 · PloS one · 2013 · 8 claims · 7 setups
Leveraging an explicit data schema (separate field encoding, field modeling, template compression, domain modeling) enables stronger compression of HTS alignment data than general-purpose compression of serialized bytes.
-
Full-text index only
Bias of selection on human copy-number variants.
PMID 16482228 · PMC1366494 · PLoS genetics · 2006 · 8 claims · 8 setups
Human CNVs are significantly overrepresented near telomeres and centromeres and enriched in simple tandem repeats relative to the genome as a whole
-
Full-text index only
Protein under-wrapping causes dosage sensitivity and decreases gene duplicability.
PMID 18208334 · PMC2211539 · PLoS genetics · 2008 · 7 claims · 6 setups
Protein under-wrapping extent is negatively correlated with gene duplicability (family size) across six organisms (E. coli, yeast, worm, fly, human, thale cress)
-
Full-text index only
Direct inference of SNP heterozygosity rates and resolution of LOH detection.
PMID 18052545 · PMC2098867 · PLoS computational biology · 2007 · 6 claims · 7 setups
A large proportion of SNPs in dbSNP have high-variance HET rate estimates, limiting their reliability for LOH study design.
-
Full-text index only
Genetical genomics: spotlight on QTL hotspots.
PMID 18949031 · PMC2563687 · PLoS genetics · 2008 · 8 claims · 4 setups
Distant eQTL hotspots are rare and difficult to reliably verify across published genetical genomics studies
-
Full-text index only
What can genome-wide association studies tell us about the genetics of common disease?
PMID 18454206 · PMC2323402 · PLoS genetics · 2008 · 8 claims · 4 setups
Apparent patterns of common, low-effect disease-associated alleles largely reflect statistical power of studies rather than the true underlying distribution of disease variants
-
Full-text index only
EPD in its twentieth year: towards complete promoter coverage of selected model organisms.
PMID 16381980 · PMC1347508 · Nucleic acids research · 2006 · 7 claims · 4 setups
EPD is an annotated, non-redundant collection of experimentally defined eukaryotic POL II promoters accessed via genome position pointers.
-
Full-text index only
Power analysis for genome-wide association studies.
PMID 17725844 · PMC2042984 · BMC genetics · 2007 · 8 claims · 6 setups
Developed a method to compute genome-wide association study power using tag SNPs and representative population genotype data (HapMap), equivalent to the cumulative r2-adjusted power of Jorgenson and Witte.