Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A cell biological perspective on genome research.
PMID 8522596 · PMC2120688 · The Journal of cell biology · 1995 · 7 claims · 7 setups
Genome sequencing represents a sixth stage in the historical progression of structural biology (comparative anatomy through crystallography), and will be similarly valuable once related to function.
-
Full-text index only
Phylogenomic approaches to common problems encountered in the analysis of low copy repeats: the sulfotransferase 1A gene family example.
PMID 15752422 · PMC555591 · BMC evolutionary biology · 2005 · 8 claims · 8 setups
A previously unidentified fourth human SULT1A gene (SULT1A4) exists on chromosome 16 and is transcriptionally active
-
Full-text index only
More biology from the sequence.
PMID 11532209 · PMC138951 · Genome biology · 2001 · 8 claims · 8 setups
The Schizosaccharomyces pombe genome has been sequenced to completion with no gaps, telomere to telomere.
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.
-
Full-text index only
Helminth genomics: The implications for human health.
PMID 19855829 · PMC2757907 · PLoS neglected tropical diseases · 2009 · 8 claims · 7 setups
More than two billion people (one-third of humanity) are infected with helminth parasites, causing major morbidity, mortality, and poverty maintenance
-
Full-text index only
DNA sequencing: bench to bedside and beyond.
PMID 17855400 · PMC2094077 · Nucleic acids research · 2007 · 8 claims · 7 setups
DNA sequencing methods derived from Sanger's 1977 dideoxy method have dominated sequencing for 30 years despite being only incrementally refined.
-
Full-text index only
ASPIC: a web resource for alternative splicing prediction and transcript isoforms characterization.
PMID 16845044 · PMC1538898 · Nucleic acids research · 2006 · 8 claims · 2 setups
The ASPIC algorithm, using an optimization procedure that minimizes splice site predictions and transcript isoforms from multiple EST-genome alignments, outperforms other similar AS-prediction tools in sensitivity and selectivity
-
Full-text index only
Gene prediction in eukaryotes with a generalized hidden Markov model that uses hints from external sources.
PMID 16469098 · PMC1409804 · BMC bioinformatics · 2006 · 7 claims · 3 setups
AUGUSTUS+ extends the AUGUSTUS GHMM by combining intrinsic sequence information with extrinsic hints via an extended emission alphabet, so the GHMM jointly models the DNA sequence, gene structure, and hint collection.
-
Full-text index only
Upgrades to StellaBase facilitate medical and genetic studies on the starlet sea anemone, Nematostella vectensis.
PMID 17982171 · PMC2238866 · Nucleic acids research · 2008 · 6 claims · 5 setups
StellaBase Disease houses homology data for 155,904 invertebrate isoforms of human disease genes across four model systems, including 14,874 predicted Nematostella genes
-
Full-text index only
Genome annotation of a 1.5 Mb region of human chromosome 6q23 encompassing a quantitative trait locus for fetal hemoglobin expression in adults.
PMID 15169551 · PMC441375 · BMC genomics · 2004 · 8 claims · 8 setups
A very large, previously uncharacterized gene, AHI1, containing WD40 and SH3 domains was discovered in the candidate interval
-
Full-text index only
Sequence, "subtle" alternative splicing and expression of the CYYR1 (cysteine/tyrosine-rich 1) mRNA in human neuroendocrine tumors.
PMID 17442112 · PMC1863428 · BMC cancer · 2007 · 7 claims · 5 setups
CYYR1 mRNA undergoes a 'subtle' alternative splicing event generating two isoforms (CAG- and CAG+) differing by a single alanine codon at the exon3/exon4 junction
-
Full-text index only
Compressing DNA sequence databases with coil.
PMID 18489794 · PMC2426707 · BMC bioinformatics · 2008 · 8 claims · 1 setups
coil achieves higher compression ratio than state-of-the-art general-purpose compression tools on a large GenBank EST database file
-
Full-text index only
Characterization of 954 bovine full-CDS cDNA sequences.
PMID 16305752 · PMC1314900 · BMC genomics · 2005 · 7 claims · 8 setups
954 bovine full-length insert cDNA (bFLIC) clones representing 762 distinct loci were sequenced and characterized
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
ChimerDB--a knowledgebase for fusion sequences.
PMID 16381848 · PMC1347382 · Nucleic acids research · 2006 · 8 claims · 6 setups
ChimerDB integrates bioinformatics analysis of mRNA/EST sequences, manually collected literature data, and OMIM translocation data into a single fusion sequence knowledgebase
-
Full-text index only
Genome wide identification of recessive cancer genes by combinatorial mutation analysis.
PMID 18846217 · PMC2557123 · PloS one · 2008 · 7 claims · 4 setups
A combinatorial mutation analysis identified 154 candidate recessive cancer genes (pRecessiveCancer<1.5x10-7, FDR=0.39)
-
Full-text index only
Genome-wide census and expression profiling of chicken neuropeptide and prohormone convertase genes.
PMID 20006904 · PMC2814002 · Neuropeptides · 2010 · 8 claims · 5 setups
Bioinformatic survey of chicken genome/EST/HTGS databases identifies previously unreported chicken neuropeptide genes
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
NEIBank: genomics and bioinformatics resources for vision research.
PMID 18648525 · PMC2480482 · Molecular vision · 2008 · 8 claims · 7 setups
NEIBank is an integrated genomics and bioinformatics resource for vision research, combining EST/cDNA clone data, SAGE expression data, and eye disease gene databases.
-
Full-text index only
Annotation and analysis of 10,000 expressed sequence tags from developing mouse eye and adult retina.
PMID 14519200 · PMC328454 · Genome biology · 2003 · 8 claims · 5 setups
Annotation of 8,633 high-quality non-mitochondrial/non-ribosomal ESTs shows 57% represent known genes and 43% are unknown or novel, with M15E having the highest proportion of novel ESTs