Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins.
PMID 15608248 · PMC539979 · Nucleic acids research · 2005 · 7 claims · 5 setups
RefSeq provides a curated, non-redundant, explicitly linked collection of genomic, transcript and protein sequences spanning prokaryotes, eukaryotes and viruses.
-
Full-text index only
The Functional RNA Database 3.0: databases to support mining and annotation of functional RNAs.
PMID 18948287 · PMC2686472 · Nucleic acids research · 2009 · 8 claims · 5 setups
fRNAdb 3.0 is a completely rebuilt sequence database hosting a much larger collection of known/predicted non-coding RNA sequences with improved search functionality
-
Full-text index only
Database resources of the National Center for Biotechnology Information.
PMID 17170002 · PMC1781113 · Nucleic acids research · 2007 · 8 claims · 8 setups
NCBI maintains an integrated suite of database resources (Entrez, PubMed, RefSeq, dbSNP, BLAST, etc.) for molecular biology data retrieval and analysis
-
Full-text index only
Genomic organization and single-nucleotide polymorphism map of desmuslin, a novel intermediate filament protein on chromosome 15q26.3.
PMID 11454237 · PMC34549 · BMC genetics · 2001 · 6 claims · 4 setups
The desmuslin (DMN) gene was localized to chromosome 15q26.3 via electronic screening of the human genome database, and its 5-exon genomic organization was determined.
-
Full-text index only
Prediction of missed cleavage sites in tryptic peptides aids protein identification in proteomics.
PMID 17203985 · PMC2664920 · Journal of proteome research · 2007 · 8 claims · 4 setups
An information-theoretic log-likelihood scoring method can predict experimentally observed missed cleavage sites from amino acid sequence alone with up to 90% accuracy.
-
Full-text index only
Natural variation of HIV-1 group M integrase: implications for a new class of antiretroviral inhibitors.
PMID 18687142 · PMC2546438 · Retrovirology · 2008 · 7 claims · 6 setups
Integrase displays significantly less inter- and intra-subtype amino acid diversity and lower Shannon's entropy than protease or RT.
-
Full-text index only
Compressing DNA sequence databases with coil.
PMID 18489794 · PMC2426707 · BMC bioinformatics · 2008 · 8 claims · 1 setups
coil achieves higher compression ratio than state-of-the-art general-purpose compression tools on a large GenBank EST database file
-
Full-text index only
GeneSeer: a sage for gene names and genomic resources.
PMID 16176584 · PMC1266031 · BMC genomics · 2005 · 7 claims · 4 setups
GeneSeer aggregates gene name synonyms from GenBank, FlyBase, ExPASy, HUGO, ENSEMBL, UCSC and Gene Ontology into a name-translation database that maps any familiar name to a reference (SOFAR) identifier.
-
Full-text index only
Human Lsg1 defines a family of essential GTPases that correlates with the evolution of compartmentalization.
PMID 16209721 · PMC1262696 · BMC biology · 2005 · 8 claims · 9 setups
hLsg1 is the human orthologue of yeast Lsg1p and defines a family of circularly permuted GTPases named YRG (YlqF Related GTPases)
-
Full-text index only
Bioinformatic mapping of AlkB homology domains in viruses.
PMID 15627404 · PMC544882 · BMC genomics · 2005 · 8 claims · 8 setups
AlkB-like domains are found in at least 22 different single-stranded RNA positive-strand plant viruses, mainly within a subgroup of the Flexiviridae family.
-
Full-text index only
Towards a comprehensive structural coverage of completed genomes: a structural genomics viewpoint.
PMID 17349043 · PMC1829165 · BMC bioinformatics · 2007 · 8 claims · 6 setups
A combined target-selection approach — pursuing both structurally uncharacterised domain families and additional targets from large structurally characterised superfamilies — is essential for comprehensive structural coverage of the genomes.
-
Full-text index only
Assignment of Streptococcus agalactiae isolates to clonal complexes using a small set of single nucleotide polymorphisms.
PMID 18710585 · PMC2533671 · BMC microbiology · 2008 · 7 claims · 6 setups
A four-SNP set (glnA36, glnA429, glcK180, adhP111) identified via the Not-N algorithm plus empirical testing divides GBS into 10 groups concordant with eBURST-defined population structure.
-
Full-text index only
A survey of integral alpha-helical membrane proteins.
PMID 19760129 · PMC2780624 · Journal of structural and functional genomics · 2009 · 8 claims · 8 setups
An automated annotation pipeline defines the integral membrane genome and family associations for 21,379 proteins from 34 genomes, most belonging to 598 Pfam-derived membrane protein families.
-
Full-text index only
Pegasys: software for executing and integrating analyses of biological sequences.
PMID 15096276 · PMC406494 · BMC bioinformatics · 2004 · 8 claims · 7 setups
Pegasys is a flexible, modular, customizable software system for executing and integrating heterogeneous biological sequence analysis tools
-
Full-text index only
Use of modified U1 snRNAs to inhibit HIV-1 replication.
PMID 17158512 · PMC1802557 · Nucleic acids research · 2007 · 7 claims · 6 setups
U1 snRNAs complementary to 5 of 15 targeted conserved regions in the HIV-1 terminal exon significantly suppress HIV-1 protein expression and viral replication, coincident with loss of viral RNA
-
Full-text index only
Distinctive pattern of sequence polymorphism in the NS3 protein of hepatitis C virus type 1b reflects conflicting evolutionary pressures.
PMID 18632963 · PMC2577380 · The Journal of general virology · 2008 · 7 claims · 6 setups
NS3 shows less evidence of purifying selection acting on its CTL epitopes than the other 9 HCV proteins, while outside the CTL epitopes NS3 is more conserved than the other proteins.
-
Full-text index only
Gene prediction in eukaryotes with a generalized hidden Markov model that uses hints from external sources.
PMID 16469098 · PMC1409804 · BMC bioinformatics · 2006 · 7 claims · 3 setups
AUGUSTUS+ extends the AUGUSTUS GHMM by combining intrinsic sequence information with extrinsic hints via an extended emission alphabet, so the GHMM jointly models the DNA sequence, gene structure, and hint collection.
-
Full-text index only
mtDB: Human Mitochondrial Genome Database, a resource for population genetics and medical sciences.
PMID 16381973 · PMC1347373 · Nucleic acids research · 2006 · 8 claims · 3 setups
mtDB is a comprehensive, actively maintained database of published human mitochondrial genome sequences, providing a common resource for population genetics and medical research
-
Full-text index only
Analysis of human sarcospan as a candidate gene for CFEOM1.
PMID 11180757 · PMC29083 · BMC genetics · 2001 · 7 claims · 5 setups
Sarcospan sequence is unmutated in all six CFEOM1 families studied
-
Full-text index only
LMPD: LIPID MAPS proteome database.
PMID 16381922 · PMC1347484 · Nucleic acids research · 2006 · 8 claims · 5 setups
LMPD is an object-relational database of lipid-associated protein sequences and annotations, publicly available from the LIPID MAPS Consortium website.