Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Identification of "pathologs" (disease-related genes) from the RIKEN mouse cDNA dataset using human curation plus FACTS, a new biological information extraction system.
PMID 15115540 · PMC420239 · BMC genomics · 2004 · 6 claims · 3 setups
Bioinformatic sequence comparison of 60,770 RIKEN FANTOM2 mouse cDNA clones identified 2,578 sequences with 70-85% identity to known human disease genes/proteins
-
Full-text index only
Mice have a transcribed L-threonine aldolase/GLY1 gene, but the human GLY1 gene is a non-processed pseudogene.
PMID 15757516 · PMC555945 · BMC genomics · 2005 · 8 claims · 8 setups
Mouse has a transcribed, 7-exon L-threonine aldolase (GLY1) gene on chromosome 11 encoding a 400-residue protein homologous to bacterial threonine aldolase
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
Database resources of the National Center for Biotechnology Information.
PMID 17170002 · PMC1781113 · Nucleic acids research · 2007 · 8 claims · 8 setups
NCBI maintains an integrated suite of database resources (Entrez, PubMed, RefSeq, dbSNP, BLAST, etc.) for molecular biology data retrieval and analysis
-
Full-text index only
Mining the draft human genome.
PMID 11236999 · PMC2658632 · Nature · 2001 · 8 claims · 7 setups
Protein-coding exons account for only about 3% of the human genome DNA, with repeat sequences making up around 46%.
-
Full-text index only
Integrative annotation of 21,037 human genes validated by full-length cDNA clones.
PMID 15103394 · PMC393292 · PLoS biology · 2004 · 8 claims · 5 setups
41,118 full-length human cDNAs from six high-throughput sequencing projects were exhaustively integratively characterized
-
Full-text index only
Characterization of 954 bovine full-CDS cDNA sequences.
PMID 16305752 · PMC1314900 · BMC genomics · 2005 · 7 claims · 8 setups
954 bovine full-length insert cDNA (bFLIC) clones representing 762 distinct loci were sequenced and characterized
-
Full-text index only
Characterization of the Schistosoma transcriptome opens up the world of helminth genomics.
PMID 14709167 · PMC395727 · Genome biology · 2003 · 8 claims · 5 setups
Near-complete transcriptome complements have now been described for S. japonicum and S. mansoni
-
Full-text index only
Switches in expression of Plasmodium falciparum var genes correlate with changes in antigenic and cytoadherent phenotypes of infected erythrocytes.
PMID 7606775 · PMC3730239 · Cell · 1995 · 6 claims · 8 setups
Expression of a specific var gene (A4var) correlates with A4 antigenic type and with binding to ICAM-1
-
Full-text index only
The gene guessing game.
PMID 11025532 · PMC2448377 · Yeast (Chichester, England) · 2000 · 8 claims · 6 setups
Published methods for estimating human gene number diverge widely, from ~30,000 to over 140,000 genes.
-
Full-text index only
Comparative gene finding in chicken indicates that we are closing in on the set of multi-exonic widely expressed human genes.
PMID 15809229 · PMC1074396 · Nucleic acids research · 2005 · 8 claims · 6 setups
Comparative gene finding (SGP2) between human and chicken, followed by RT-PCR verification, adds at most ~0.2% new genes to the multi-exonic human gene catalog
-
Full-text index only
Diversity of preferred nucleotide sequences around the translation initiation codon in eukaryote genomes.
PMID 18086709 · PMC2241899 · Nucleic acids research · 2008 · 8 claims · 5 setups
Preferred nucleotide sequences around the initiation codon are diverse among eukaryote species, but differences roughly reflect evolutionary relationships between species
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.
-
Full-text index only
Dynamic Proteomics: a database for dynamics and localizations of endogenous fluorescently-tagged proteins in living human cells.
PMID 19820112 · PMC2808965 · Nucleic acids research · 2010 · 8 claims · 6 setups
The Dynamic Proteomics database compiles fluorescence dynamics and localization data for endogenously YFP/Venus-tagged human proteins from the LARC library studied by Cohen et al.
-
Full-text index only
Ensembl 2005.
PMID 15608235 · PMC540092 · Nucleic acids research · 2005 · 8 claims · 4 setups
Ensembl's automatic gene build system can flexibly and reliably annotate a wide variety of genomes with limited species-specific evidence.
-
Full-text index only
A cell biological perspective on genome research.
PMID 8522596 · PMC2120688 · The Journal of cell biology · 1995 · 7 claims · 7 setups
Genome sequencing represents a sixth stage in the historical progression of structural biology (comparative anatomy through crystallography), and will be similarly valuable once related to function.
-
Full-text index only
Toxicogenomics: an emerging discipline.
PMID 12460812 · PMC1241126 · Environmental health perspectives · 2002 · 8 claims · 6 setups
Toxicogenomics applies genomic tools (microarrays, proteomics, metabolomics) to characterize how cells and organisms respond to chemical/drug exposures.
-
Full-text index only
GENCODE: producing a reference annotation for ENCODE.
PMID 16925838 · PMC1810553 · Genome biology · 2006 · 8 claims · 8 setups
GENCODE annotation combines initial manual annotation by HAVANA, experimental validation, and refinement based on results to identify protein-coding genes in ENCODE regions