Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A high throughput method for genome-wide analysis of retroviral integration.
PMID 17028098 · PMC1636494 · Nucleic acids research · 2006 · 8 claims · 8 setups
VITA uses MmeI to cleave DNA at a fixed distance from its recognition site, generating 21-22 bp genomic tags that serve as signatures of lentiviral integration sites.
-
Full-text index only
A model-based approach to selection of tag SNPs.
PMID 16776821 · PMC1525207 · BMC bioinformatics · 2006 · 7 claims · 5 setups
The Li and Stephens hidden Markov model outperforms other tested models (simple Markov, two-state HMM, HMM-4D, greedy GR-1/GR-2) in description code-length, tag set information content, and prediction of tagged SNPs.
-
Full-text index only
An evaluation of the performance of tag SNPs derived from HapMap in a Caucasian population.
PMID 16532062 · PMC1391920 · PLoS genetics · 2006 · 8 claims · 5 setups
CEU HapMap-derived tSNPs capture most of the genetic variation observed in the Estonian (EGP) population sample
-
Full-text index only
Selecting additional tag SNPs for tolerating missing data in genotyping.
PMID 16259642 · PMC1316880 · BMC bioinformatics · 2005 · 7 claims · 6 setups
There exists a subset of SNPs (robust tag SNPs) that can distinguish all distinct haplotypes even when up to m SNPs are missing
-
Full-text index only
Systematic analysis of human kinase genes: a large number of genes and alternative splicing events result in functional and structural diversity.
PMID 16351747 · PMC1866387 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Systematic in silico search identified 5 novel human kinase genes (on chromosomes 1, 11, 13, 15, 16) and 1 pseudogene (chromosome X) absent from KinBase
-
Full-text index only
Software for tag single nucleotide polymorphism selection.
PMID 16004730 · PMC3525260 · Human genomics · 2005 · 8 claims · 3 setups
Pairwise R2 methods tend to pick more tagging SNPs than strictly needed because they miss redundancy where two or more tag SNPs jointly predict an untagged SNP with no single direct surrogate.
-
Full-text index only
A genome annotation-driven approach to cloning the human ORFeome.
PMID 15461802 · PMC545604 · Genome biology · 2004 · 8 claims · 8 setups
Existing human cDNA clone collections together provide only 60% coverage of full-length chromosome 22 ORFs, with the best single collection (MGC) providing 48%
-
Full-text index only
Application of single molecule technology to rapidly map long DNA and study the conformation of stretched DNA.
PMID 16243782 · PMC1266062 · Nucleic acids research · 2005 · 7 claims · 4 setups
DLA can generate a high signal-to-noise bisPNA binding-site map of the 185.1 kb BAC 12M9 using as few as 200 molecule traces
-
Full-text index only
Phylogenomic approaches to common problems encountered in the analysis of low copy repeats: the sulfotransferase 1A gene family example.
PMID 15752422 · PMC555591 · BMC evolutionary biology · 2005 · 8 claims · 8 setups
A previously unidentified fourth human SULT1A gene (SULT1A4) exists on chromosome 16 and is transcriptionally active
-
Full-text index only
Mice have a transcribed L-threonine aldolase/GLY1 gene, but the human GLY1 gene is a non-processed pseudogene.
PMID 15757516 · PMC555945 · BMC genomics · 2005 · 8 claims · 8 setups
Mouse has a transcribed, 7-exon L-threonine aldolase (GLY1) gene on chromosome 11 encoding a 400-residue protein homologous to bacterial threonine aldolase
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.
-
Has reproduction · 64
A novel and dual digestive symbiosis scales up the nutrition and immune system of the holobiont Rimicaris exoculata.
PMID 36333777 · PMC9636832 · Microbiome · 2022 · 7 claims · 6 setups
Genome-resolved metagenomics of separated foregut and midgut reconstructed 20 MAGs including novel lineages of Hepatoplasmataceae (foregut) and Deferribacteres (midgut).
-
Full-text index only
Insights into vertebrate evolution from the chicken genome sequence.
PMID 15693954 · PMC551526 · Genome biology · 2005 · 8 claims · 6 setups
Chicken has expanded gene families involved in egg production (e.g., avidin) and feather/scale/claw formation (avian-specific keratins) not present or lost in mammals
-
Full-text index only
Current status and the future for the genetics of type I diabetes.
PMID 19956094 · PMC2805458 · Genes and immunity · 2009 · 8 claims · 7 setups
A T1DGC genome-wide association meta-analysis of >7500 cases and >9000 controls identified 42 distinct genomic locations associated with T1D at P<10^-6.
-
Full-text index only
Sequence variation and linkage disequilibrium in the GABA transporter-1 gene (SLC6A1) in five populations: implications for pharmacogenetic research.
PMID 17941974 · PMC2175509 · BMC genetics · 2007 · 8 claims · 7 setups
SLC6A1 shows low levels of LD and an absence of major LD blocks across all five populations studied
-
Full-text index only
Mechanisms of disease: genetic insights into the etiology of type 2 diabetes and obesity.
PMID 18212765 · PMC7116808 · Nature clinical practice. Endocrinology & metabolism · 2008 · 8 claims · 8 setups
Six high-density genome-wide association studies in over 19,000 individuals identified approximately ten T2D-susceptibility loci, including HHEX, IDE, SLC30A8, FTO, CDKAL1, CDKN2A/CDKN2B, and IGF2BP2.
-
Has reproduction · 73
Vespucci: a system for building annotated databases of nascent transcripts.
PMID 24304890 · PMC3936758 · Nucleic acids research · 2014 · 8 claims · 7 setups
Existing ChIP-seq and RNA-seq analysis platforms (e.g. Cufflinks, peak callers) are unsuited to GRO-seq because they assume spliced/exonic reads, uniform density and paired-end data, and cannot identify transcriptional units de novo across the whole genome.