Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
DAVID Knowledgebase: a gene-centered database integrating heterogeneous gene annotation resources to facilitate high-throughput gene functional analysis.
PMID 17980028 · PMC2186358 · BMC bioinformatics · 2007 · 7 claims · 3 setups
The DAVID Gene Concept, a single-linkage algorithm, merges gene clusters from Entrez Gene, UniRef100, and PIR-NREF100 that share protein IDs and species into unified DAVID gene clusters, improving cross-referencing between NCBI and UniProt systems
-
Full-text index only
Clustering of phosphorylation site recognition motifs can be exploited to predict the targets of cyclin-dependent kinase.
PMID 17316440 · PMC1852407 · Genome biology · 2007 · 8 claims · 6 setups
CDK consensus motifs are frequently clustered (closely spaced) in known CDK substrate proteins rather than uniformly distributed
-
Full-text index only
The truth about mouse, human, worms and yeast.
PMID 15601543 · PMC3525071 · Human genomics · 2004 · 8 claims · 8 setups
Comparing genomes in pairs or larger sets (mouse-human, C. elegans-C. briggsae, multiple Saccharomyces, human-pufferfish, etc.) reveals unsuspected genes and helps eliminate false-positive gene predictions
-
Full-text index only
BIPASS: BioInformatics Pipeline Alternative Splicing Services.
PMID 17584795 · PMC1933140 · Nucleic acids research · 2007 · 8 claims · 4 setups
BIPASS offers two complementary services for alternative splicing (AS) research: BIPAS-SpliceDB, a queryable pre-computed AS data warehouse, and BIPAS-Align&Splice, an online pipeline for user-submitted sequences.
-
Full-text index only
GeneSeer: a sage for gene names and genomic resources.
PMID 16176584 · PMC1266031 · BMC genomics · 2005 · 7 claims · 4 setups
GeneSeer aggregates gene name synonyms from GenBank, FlyBase, ExPASy, HUGO, ENSEMBL, UCSC and Gene Ontology into a name-translation database that maps any familiar name to a reference (SOFAR) identifier.
-
Full-text index only
GenomeTrafac: a whole genome resource for the detection of transcription factor binding site clusters associated with conventional and microRNA encoding genes conserved between mouse and human gene orthologs.
PMID 17178752 · PMC1781107 · Nucleic acids research · 2007 · 8 claims · 5 setups
GenomeTrafac is a web-accessible database enabling genome-wide detection of conserved cis-element clusters in human-mouse gene orthologs, covering both conventional and microRNA genes
-
Full-text index only
Human-zebrafish non-coding conserved elements act in vivo to regulate transcription.
PMID 16179648 · PMC1236720 · Nucleic acids research · 2005 · 8 claims · 4 setups
Deeply conserved human-zebrafish non-coding elements are enriched for in vivo cis-acting transcriptional regulatory activity.
-
Full-text index only
Dcode.org anthology of comparative genomic tools.
PMID 15980535 · PMC1160116 · Nucleic acids research · 2005 · 8 claims · 7 setups
The dcode.org suite (zPicture, Mulan, eShadow, rVista 2.0, multiTF, Creme 2.0, ECR Browser) provides integrated tools for comparative genomic analysis and non-coding regulatory element discovery.
-
Full-text index only
Satellog: a database for the identification and prioritization of satellite repeats in disease association studies.
PMID 15949044 · PMC1181805 · BMC bioinformatics · 2005 · 7 claims · 6 setups
Satellog is a database cataloging all pure 1-16 unit satellite repeats in the human genome with supplementary polymorphism, gene-location, and expression data for prioritizing repeats in disease-association studies.
-
Full-text index only
Divergence of exonic splicing elements after gene duplication and the impact on gene structures.
PMID 19883501 · PMC3091315 · Genome biology · 2009 · 8 claims · 7 setups
ESEs and ESSs diverge especially fast shortly after gene duplication, correlating with time since duplication (Ks)
-
Full-text index only
Genome informatics: taming the avalanche of genomic data.
PMID 15642109 · PMC549058 · Genome biology · 2005 · 8 claims · 7 setups
Ultraconserved regions (>100 bp, 100% conserved among mammals) exist in the genome and their function remains unknown
-
Full-text index only
DNA sequence and analysis of human chromosome 9.
PMID 15164053 · PMC2734081 · Nature · 2004 · 8 claims · 8 setups
The finished euchromatic sequence of chromosome 9 comprises 109,044,351 base pairs, representing >99.6% of the region.
-
Full-text index only
Predictive screening for regulators of conserved functional gene modules (gene batteries) in mammals.
PMID 15882449 · PMC1134656 · BMC genomics · 2005 · 8 claims · 4 setups
A predictive computational screen covering ~40% of annotated protein-coding genes identified 21 co-expressed gene clusters with statistically supported sharing of cis-regulatory motifs.
-
Full-text index only
Genomic sequencing of the severe acute respiratory syndrome-coronavirus.
PMID 16916263 · PMC7121524 · Methods in molecular biology (Clifton, N.J.) · 2006 · 7 claims · 7 setups
PCR-based amplification and direct sequencing of SARS-CoV genome fragments is feasible from uncultured clinical specimens (serum, nasopharyngeal aspirate, stool), avoiding culture-derived artifacts and biohazard risk.
-
Full-text index only
ChimerDB--a knowledgebase for fusion sequences.
PMID 16381848 · PMC1347382 · Nucleic acids research · 2006 · 8 claims · 6 setups
ChimerDB integrates bioinformatics analysis of mRNA/EST sequences, manually collected literature data, and OMIM translocation data into a single fusion sequence knowledgebase
-
Full-text index only
piRNABank: a web resource on classified and clustered Piwi-interacting RNAs.
PMID 17881367 · PMC2238943 · Nucleic acids research · 2008 · 6 claims · 4 setups
piRNABank is a web-accessible database storing empirically known piRNA sequences and annotations for human, mouse and rat.
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
Multilocus sequence typing supports the hypothesis that Ochrobactrum anthropi displays a human-associated subpopulation.
PMID 20021660 · PMC2810298 · BMC microbiology · 2009 · 8 claims · 6 setups
A novel Multi-Locus Sequence Typing (MLST) scheme for O. anthropi was developed for the first time, based on 7 genes (3490 nucleotides) evolving mostly by neutral mutations
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
Widespread A-to-I RNA editing of Alu-containing mRNAs in the human transcriptome.
PMID 15534692 · PMC526178 · PLoS biology · 2004 · 8 claims · 6 setups
Intramolecular pairs of oppositely oriented Alu elements within the same pre-mRNA form dsRNA foldback structures that are major substrates for A-to-I RNA editing