Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Sequence variation in the human transcription factor gene POU5F1.
PMID 18254969 · PMC2275747 · BMC genetics · 2008 · 7 claims · 5 setups
POU5F1 is highly polymorphic, with a higher polymorphism density than most genes
-
Has reproduction
Predicting favorable landing pads for targeted integrations in Chinese hamster ovary cell lines by learning stability characteristics from random transgene integrations.
PMID 33304461 · PMC7710658 · Computational and structural biotechnology journal · 2020 · 7 claims · 6 setups
Expression stability in CHO cell lines is controlled at three levels: choice of integration site, integrity/concatemerization pattern of the transgene, and stress-related cellular processes.
-
Full-text index only
Copy number variants and common disorders: filling the gaps and exploring complexity in genome-wide association studies.
PMID 17953491 · PMC2039766 · PLoS genetics · 2007 · 8 claims · 5 setups
CNVs are not easily tagged by SNPs and often fall in genomic regions poorly covered by whole-genome SNP arrays or not genotyped by HapMap, so current GWASs have largely missed their contribution to complex disorders.
-
Full-text index only
An integrated genomic analysis of human glioblastoma multiforme.
PMID 18772396 · PMC2820389 · Science (New York, N.Y.) · 2008 · 8 claims · 7 setups
IDH1 is recurrently mutated at its active site (R132) in 12% of GBM patients, a previously unrecognized alteration in GBM.
-
Full-text index only
Correlating novel variable and conserved motifs in the Hemagglutinin protein with significant biological functions.
PMID 18681973 · PMC2553082 · Virology journal · 2008 · 8 claims · 6 setups
14 MEME blocks were identified in the HA protein of H3N2 strains (1968-1999), with blocks 1, 2, 3, and 7 correlating with several biological functions
-
Full-text index only
Human and mouse introns are linked to the same processes and functions through each genome's most frequent non-conserved motifs.
PMID 18450818 · PMC2425492 · Nucleic acids research · 2008 · 8 claims · 5 setups
Pyknons (recurrent, genome-specific, ≥16nt motifs with ≥30 intact intergenic/intronic copies and ≥1 exonic copy) span a substantial fraction of previously uncharacterized intronic space (7.4% human, 4.4% mouse)
-
Full-text index only
Twin peaks: the draft human genome sequence.
PMID 11276423 · PMC138909 · Genome biology · 2001 · 8 claims · 8 setups
The predicted number of human genes (~26,000-40,000) is far lower than the widely assumed ~100,000, though downstream RNA/protein complexity can still generate substantial biological complexity.
-
Full-text index only
6th annual meeting of the Complex Trait Consortium.
PMID 17906895 · PMC2042027 · Mammalian genome : official journal of the International Mammalian Genome Society · 2007 · 8 claims · 7 setups
The NIEHS Perlegen/resequencing project has generated over 8.5 million SNPs from 15 inbred mouse strains but shows a high false-negative discovery rate, with an estimated 45 million SNPs actually present.
-
Full-text index only
The EH1 motif in metazoan transcription factors.
PMID 16309560 · PMC1310626 · BMC genomics · 2005 · 8 claims · 5 setups
There is a statistically significant association between EH1hox motif HMM score and transcription factor function across human, Drosophila and C. elegans proteomes.
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.
-
Full-text index only
A response to Yu et al. "A forward-backward fragment assembling algorithm for the identification of genomic amplification and deletion breakpoints using high-density single nucleotide polymorphism (SNP) array", BMC Bioinformatics 2007, 8: 145.
PMID 17939873 · PMC2222656 · BMC bioinformatics · 2007 · 8 claims · 4 setups
Yu et al.'s original comparison ran RJaCGH's MCMC sampler for a severely insufficient number of iterations (50 burn-in, 500 total)
-
Full-text index only
SNP selection for genes of iron metabolism in a study of genetic modifiers of hemochromatosis.
PMID 18366708 · PMC2289803 · BMC medical genetics · 2008 · 7 claims · 6 setups
Illumina validation/design scores above 0.6 are not strongly correlated with actual SNP genotyping performance (Gentrain score)
-
Full-text index only
CoMoDis: composite motif discovery in mammalian genomes.
PMID 17130158 · PMC1702496 · Nucleic acids research · 2007 · 7 claims · 4 setups
CoMoDis is a new bioinformatics tool that streamlines computational identification of novel regulatory modules starting from a single seed motif
-
Full-text index only
Assessment of algorithms for high throughput detection of genomic copy number variation in oligonucleotide microarray data.
PMID 17910767 · PMC2148068 · BMC bioinformatics · 2007 · 8 claims · 4 setups
Different CNV analysis software packages produce highly variable numbers and types of candidate CNVs from the same data
-
Full-text index only
Investigating hookworm genomes by comparative analysis of two Ancylostoma species.
PMID 15854223 · PMC1112591 · BMC genomics · 2005 · 8 claims · 8 setups
Nearly 20,000 ESTs from 7 cDNA libraries define nearly 7,000 hookworm genes across A. caninum and A. ceylanicum
-
Full-text index only
Probing the cancer genome.
PMID 18492227 · PMC2441462 · Genome biology · 2008 · 8 claims · 8 setups
Combined Sanger and 454 pyrosequencing of MCF-7 BAC clones identified 157 PCR-confirmed translocation breakpoint junctions, including 10 in-frame junctions confirmed at the transcript level
-
Full-text index only
Proteomic approaches to cancer biomarkers.
PMID 19931265 · PMC2873613 · Gastroenterology · 2010 · 8 claims · 8 setups
Combining abundant-protein depletion, offline fractionation, and subproteome (e.g., glycoproteome) enrichment with 2D LC-MS/MS increases the dynamic range and depth of blood proteome analysis for biomarker discovery.
-
Has reproduction
Genomic prediction based on selective linkage disequilibrium pruning of low-coverage whole-genome sequence variants in a pure Duroc population.
PMID 37853325 · PMC10583454 · Genetics, selection, evolution : GSE · 2023 · 8 claims · 6 setups
Selective linkage disequilibrium pruning (SLDP) refines whole-genome SNP sets using GWAS prior information to improve genomic prediction accuracy.
-
Full-text index only
Comparing protein abundance and mRNA expression levels on a genomic scale.
PMID 12952525 · PMC193646 · Genome biology · 2003 · 8 claims · 8 setups
Correlations between mRNA expression and protein abundance are generally poor or limited across most studies reviewed, including in yeast and human cancers
-
Full-text index only
PeroxisomeDB: a database for the peroxisomal proteome, functional genomics and disease.
PMID 17135190 · PMC1747181 · Nucleic acids research · 2007 · 8 claims · 6 setups
PeroxisomeDB integrates the complete peroxisomal proteome of Homo sapiens and Saccharomyces cerevisiae into interrelated 'Genes', 'Functions', 'Metabolic pathways' and 'Diseases' sections with links to NCBI, ENSEMBL and UCSC