Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
Prediction-based approaches to characterize bidirectional promoters in the mammalian genome.
PMID 18366609 · PMC2386062 · BMC genomics · 2008 · 8 claims · 7 setups
The mapping algorithm identified 5,647 candidate bidirectional promoter regions in the mouse genome, similar in number to those previously found in human.
-
Full-text index only
MatchMiner: a tool for batch navigation among gene and gene product identifiers.
PMID 12702208 · PMC154578 · Genome biology · 2003 · 8 claims · 3 setups
MatchMiner's LookUp function automates batch translation of an input list of gene identifiers into a matching list of a different identifier type.
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
ChimerDB 2.0--a knowledgebase for fusion genes updated.
PMID 19906715 · PMC2808913 · Nucleic acids research · 2010 · 8 claims · 4 setups
ChimerDB 2.0 is an updated knowledgebase integrating fusion transcripts from GenBank transcriptome analysis with Sanger CGP, OMIM, PubMed, and Mitelman's database data.
-
Full-text index only
Bases and spaces: resources on the web for accessing the draft human genome.
PMID 11178254 · PMC138875 · Genome biology · 2000 · 8 claims · 8 setups
By combining currently available genomic databases and mapping resources (GenBank/Entrez, UniGene, RH maps, BAC fingerprint maps, Ensembl, NIX), it is possible to devise strategies that fully exploit the fragmentary draft human genome sequence.
-
Full-text index only
Haplotype analysis of common variants in the BRCA1 gene and risk of sporadic breast cancer.
PMID 15743496 · PMC1064127 · Breast cancer research : BCR · 2005 · 7 claims · 5 setups
A common BRCA1 haplotype (haplotype 2, C A G G) is associated with a modest increase in sporadic breast cancer risk
-
Full-text index only
AceView: a comprehensive cDNA-supported gene and transcripts annotation.
PMID 16925834 · PMC1810549 · Genome biology · 2006 · 8 claims · 4 setups
At the mRNA level, AceView transcripts are the closest match to Gencode transcripts among all evaluated methods, including alternative splice variants
-
Full-text index only
AUGUSTUS at EGASP: using EST, protein and genomic alignments for improved gene prediction in the human genome.
PMID 16925833 · PMC1810548 · Genome biology · 2006 · 8 claims · 5 setups
AUGUSTUS predicted significantly more genes correctly than any other ab initio program in EGASP
-
Full-text index only
ChimerDB--a knowledgebase for fusion sequences.
PMID 16381848 · PMC1347382 · Nucleic acids research · 2006 · 8 claims · 6 setups
ChimerDB integrates bioinformatics analysis of mRNA/EST sequences, manually collected literature data, and OMIM translocation data into a single fusion sequence knowledgebase
-
Full-text index only
Database resources of the National Center for Biotechnology Information.
PMID 17170002 · PMC1781113 · Nucleic acids research · 2007 · 8 claims · 8 setups
NCBI maintains an integrated suite of database resources (Entrez, PubMed, RefSeq, dbSNP, BLAST, etc.) for molecular biology data retrieval and analysis
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.
-
Full-text index only
Atlas - a data warehouse for integrative bioinformatics.
PMID 15723693 · PMC554782 · BMC bioinformatics · 2005 · 8 claims · 3 setups
Atlas is a biological data warehouse that locally stores and integrates sequences, molecular interactions, homology information, functional annotations, and ontologies
-
Has reproduction · 84
Pharokka: a fast scalable bacteriophage annotation tool.
PMID 36453861 · PMC9805569 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 5 setups
Pharokka is a one-line, fast, scalable bacteriophage annotation tool producing standards-compliant outputs, installable via a two-line bioconda command
-
Full-text index only
Sequencing of 16S rRNA gene: a rapid tool for identification of Bacillus anthracis.
PMID 12396926 · PMC2730316 · Emerging infectious diseases · 2002 · 7 claims · 4 setups
All 86 B. anthracis isolates tested share an identical 16S rRNA gene sequence, designated 16S type 6
-
Full-text index only
Accelerated evolution of the ASPM gene controlling brain size begins prior to human brain expansion.
PMID 15045028 · PMC374243 · PLoS biology · 2004 · 8 claims · 6 setups
The ASPM gene shows accelerated (positively selected) evolution in the African hominoid clade, and this acceleration precedes hominid brain expansion by several million years.
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.
-
Full-text index only
Comparative genomics search for losses of long-established genes on the human lineage.
PMID 18085818 · PMC2134963 · PLoS computational biology · 2007 · 8 claims · 6 setups
A novel comparative genomics method (TransMap-based syntenic mapping of gene structures between human, mouse, and dog) can detect losses of well-established single-copy genes without relying on sequence homology to a parental gene, distinguishing them from typical duplication- or retrotransposition-derived pseudogenes.
-
Has reproduction · 99
getSequenceInfo: a suite of tools allowing to get genome sequence information from public repositories.
PMID 35804320 · PMC9264741 · BMC bioinformatics · 2022 · 8 claims · 8 setups
getSequenceInfo (gSeqI) allows programmatic (CLI) or GUI-based retrieval of sequence data and metadata from GenBank, RefSeq, and ENA across Linux, MacOS, and Windows.
-
Full-text index only
Molecular evolution and multilocus sequence typing of 145 strains of SARS-CoV.
PMID 16112670 · PMC7118731 · FEBS letters · 2005 · 8 claims · 7 setups
145 SARS-CoV genomes can be divided into three groups: animal-origin viruses, first-epidemic clinical viruses, and GD03T0013