Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Identification and functional analyses of 11,769 full-length human cDNAs focused on alternative splicing.
PMID 19880432 · PMC2780955 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2009 · 8 claims · 5 setups
Identified 23,241 human genes transcribed into protein-coding mRNAs using full-length cDNA and 5'-EST sequence data
-
Full-text index only
ATBF1 and NQO1 as candidate targets for allelic loss at chromosome arm 16q in breast cancer: absence of somatic ATBF1 mutations and no role for the C609T NQO1 polymorphism.
PMID 18416817 · PMC2377272 · BMC cancer · 2008 · 8 claims · 7 setups
Five genes (NQO1, ATBF1, DBNDD1, HSBP1, CGI-38) at 16q show significantly lower mRNA expression in breast tumors with LOH at 16q compared to tumors without LOH
-
Full-text index only
Genome-wide estimation of transcript concentrations from spotted cDNA microarray data.
PMID 16204447 · PMC1243803 · Nucleic acids research · 2005 · 8 claims · 3 setups
A Bayesian model incorporating experimental covariates (array, pen, probe, dye, scanning) can estimate absolute transcript concentrations from spotted microarray intensities without needing calibration of each sample or gene individually
-
Full-text index only
Annotation and analysis of 10,000 expressed sequence tags from developing mouse eye and adult retina.
PMID 14519200 · PMC328454 · Genome biology · 2003 · 8 claims · 5 setups
Annotation of 8,633 high-quality non-mitochondrial/non-ribosomal ESTs shows 57% represent known genes and 43% are unknown or novel, with M15E having the highest proportion of novel ESTs
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
Identification of "pathologs" (disease-related genes) from the RIKEN mouse cDNA dataset using human curation plus FACTS, a new biological information extraction system.
PMID 15115540 · PMC420239 · BMC genomics · 2004 · 6 claims · 3 setups
Bioinformatic sequence comparison of 60,770 RIKEN FANTOM2 mouse cDNA clones identified 2,578 sequences with 70-85% identity to known human disease genes/proteins
-
Full-text index only
Decreased expression of the Id3 gene at 1p36.1 in ovarian adenocarcinomas.
PMID 11161400 · PMC2363740 · British journal of cancer · 2001 · 7 claims · 7 setups
Id3 mRNA and protein expression are decreased in ovarian cancer cell lines compared to immortalized HOSE cells
-
Full-text index only
All systems GO for understanding mouse gene function.
PMID 15610553 · PMC549721 · Journal of biology · 2004 · 7 claims · 4 setups
Quantitative, multivariate cross-tissue expression measurements are powerfully predictive of gene function
-
Full-text index only
A catalog of human cDNA expression clones and its application to structural genomics.
PMID 15345055 · PMC522878 · Genome biology · 2004 · 8 claims · 7 setups
A high-throughput screening approach can identify human cDNA clones from the hEx1 library that express soluble protein in E. coli
-
Full-text index only
Toxicogenomics research consortium sails into uncharted waters.
PMID 12460811 · PMC1241122 · Environmental health perspectives · 2002 · 8 claims · 8 setups
The NIEHS-funded $37 million Toxicogenomics Research Consortium (TRC) combines the NIEHS Microarray Center with five academic institutions (UNC, Duke, Fred Hutchinson/UW, MIT, OHSU) to define genetic variability, set gene expression standards, and study environmental stress responses.
-
Full-text index only
A genome annotation-driven approach to cloning the human ORFeome.
PMID 15461802 · PMC545604 · Genome biology · 2004 · 8 claims · 8 setups
Existing human cDNA clone collections together provide only 60% coverage of full-length chromosome 22 ORFs, with the best single collection (MGC) providing 48%
-
Full-text index only
Analysis of the prostate cancer cell line LNCaP transcriptome using a sequencing-by-synthesis approach.
PMID 17010196 · PMC1592491 · BMC genomics · 2006 · 8 claims · 7 setups
High-throughput 454 sequencing-by-synthesis of LNCaP cDNA can profile transcript abundance across the transcriptome
-
Full-text index only
Large-scale analysis of Macaca fascicularis transcripts and inference of genetic divergence between M. fascicularis and M. mulatta.
PMID 18294402 · PMC2287170 · BMC genomics · 2008 · 8 claims · 6 setups
Constructed full-length-enriched cDNA libraries and determined 85,721 EST sequences and 9407 full-insert sequences from cynomolgus macaque brain (7 regions), testis, and liver
-
Full-text index only
cDNA sequencing improves the detection of P53 missense mutations in colorectal cancer.
PMID 19671129 · PMC2731783 · BMC cancer · 2009 · 8 claims · 6 setups
cDNA sequencing detects P53 missense mutations in colorectal cancer more frequently and reliably than DNA sequencing
-
Full-text index only
Comparative analysis of genome tiling array data reveals many novel primate-specific functional RNAs in human.
PMID 17288572 · PMC1796608 · BMC evolutionary biology · 2007 · 8 claims · 6 setups
Widespread transcription occurs across the human genome outside known gene annotations, and the bulk of TARs represent genuine transcripts rather than experimental artifacts
-
Full-text index only
SelenoDB 1.0 : a database of selenoprotein genes, proteins and SECIS elements.
PMID 18174224 · PMC2238826 · Nucleic acids research · 2008 · 6 claims · 5 setups
Standard genome annotation pipelines misannotate selenoprotein genes because they rely on UGA as a universal stop codon, failing to recognize its dual role as the selenocysteine-recoding codon.
-
Full-text index only
Iterative class discovery and feature selection using Minimal Spanning Trees.
PMID 15355552 · PMC520744 · BMC bioinformatics · 2004 · 7 claims · 5 setups
Iterating between MST-based clustering and t-statistic feature selection removes noise genes step-wise while sharpening the sample clustering
-
Full-text index only
Non-EST based prediction of exon skipping and intron retention events using Pfam information.
PMID 16204458 · PMC1243800 · Nucleic acids research · 2005 · 7 claims · 5 setups
A novel ab initio method predicts exon skipping and intron retention events using only Pfam domain annotation, via a Viterbi-like dynamic programming algorithm applied to the Pfam alignment.
-
Full-text index only
Predicting candidate genes for human deafness disorders: a bioinformatics approach.
PMID 16854223 · PMC1564145 · BMC genomics · 2006 · 8 claims · 4 setups
A bioinformatic approach combining expression databases and protein interaction data narrows ~2400 candidate genes across deafness loci to a manageable set of candidates.
-
Full-text index only
X:Map: annotation and visualization of genome structure for Affymetrix exon array analysis.
PMID 17932061 · PMC2238884 · Nucleic acids research · 2008 · 7 claims · 4 setups
X:Map is a genome annotation database that maps every Affymetrix exon array probeset to Ensembl genome features (genes, ESTs, GenScan predictions) and supports both high-throughput and gene-centric analysis.