Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genotyping, Orientalis-like Yersinia pestis, and plague pandemics.
PMID 15498160 · PMC3320270 · Emerging infectious diseases · 2004 · 6 claims · 6 setups
Multiple spacer typing (MST), based on intergenic spacer sequencing, differentiates the three Y. pestis biovars
-
Full-text index only
PA-GOSUB: a searchable database of model organism protein sequences with their predicted Gene Ontology molecular function and subcellular localization.
PMID 15608166 · PMC540074 · Nucleic acids research · 2005 · 7 claims · 4 setups
PA-GOSUB significantly extends the coverage of GO molecular function and subcellular localization annotations for 10 model organism proteomes compared with existing databases (GOA, Swiss-Prot).
-
Full-text index only
Comprehensive genome analysis of 203 genomes provides structural genomics with new insights into protein family space.
PMID 16481312 · PMC1373602 · Nucleic acids research · 2006 · 8 claims · 7 setups
The number of protein families continues to expand steadily as more genomes are sequenced, showing no sign of saturation.
-
Full-text index only
TPRpred: a tool for prediction of TPR-, PPR- and SEL1-like repeats from protein sequences.
PMID 17199898 · PMC1774580 · BMC bioinformatics · 2007 · 7 claims · 8 setups
TPRpred detects divergent/remote-homolog TPR repeat units that existing resources (Pfam, SMART, REP) fail to detect
-
Full-text index only
Coiled-coil protein composition of 22 proteomes--differences and common themes in subcellular infrastructure and traffic control.
PMID 16288662 · PMC1322226 · BMC evolutionary biology · 2005 · 7 claims · 5 setups
Proteins with extended coiled-coil domains (>250 amino acids) are largely absent from bacterial genomes but present in archaea and eukaryotes.
-
Full-text index only
Large-scale structural analysis of the core promoter in mammalian and plant genomes.
PMID 16049029 · PMC1181242 · Nucleic acids research · 2005 · 8 claims · 7 setups
DNA encodes at least two independent levels of functional information: protein/TF-binding sequence information and physical/structural properties of the molecule itself.
-
Full-text index only
Paircomp, FamilyRelationsII and Cartwheel: tools for interspecific sequence comparison.
PMID 15790396 · PMC1087472 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Paircomp, FamilyRelationsII, and Cartwheel together form an integrated system for comparing, viewing, and managing analyses of BAC-sized (~100 kb) genomic sequence pairs.
-
Full-text index only
External contamination in single cell mtDNA analysis.
PMID 17668059 · PMC1930155 · PloS one · 2007 · 8 claims · 6 setups
External DNA contamination is a real and non-negligible problem in single-cell mtDNA sequence analysis
-
Full-text index only
Local combinational variables: an approach used in DNA-binding helix-turn-helix motif prediction with sequence information.
PMID 19651875 · PMC2761287 · Nucleic acids research · 2009 · 8 claims · 7 setups
The LCV approach predicts HTH motifs with 93.29% accuracy, 93.93% sensitivity and 92.66% specificity using only primary sequence information
-
Full-text index only
The functional importance of disease-associated mutation.
PMID 12220483 · PMC128831 · BMC bioinformatics · 2002 · 6 claims · 1 setups
Disease-associated mutations occur in conserved regions of genes and can be used to identify likely disease-causing mutations
-
Has reproduction · 50
MEDUSA: A Pipeline for Sensitive Taxonomic Classification and Flexible Functional Annotation of Metagenomic Shotgun Sequences.
PMID 35330728 · PMC8940201 · Frontiers in genetics · 2022 · 6 claims · 6 setups
MEDUSA is an automated, Conda-installable and Snakemake-managed pipeline performing preprocessing, assembly, alignment, taxonomic classification, and functional annotation on shotgun data.
-
Full-text index only
Direct maximum parsimony phylogeny reconstruction from genotype data.
PMID 18053244 · PMC2222657 · BMC bioinformatics · 2007 · 6 claims · 4 setups
The paper presents the first practical method for computing maximum parsimony phylogenies directly from genotype data, using integer linear programming.
-
Full-text index only
Use of suppression subtractive hybridisation to extend our knowledge of genome diversity in Campylobacter jejuni.
PMID 17470265 · PMC1868759 · BMC genomics · 2007 · 8 claims · 6 setups
There is a clear correlation between MLST clonal complex and the distribution of metabolic genes involved in alternative terminal electron acceptor use.
-
Has reproduction · 93
Experimental identification and in silico prediction of bacterivory in green algae.
PMID 33649548 · PMC8245530 · The ISME journal · 2021 · 7 claims · 6 setups
Five prasinophyte strains (Pterosperma cristatum NIES626, Pyramimonas parkeae CCMP726, Pyramimonas parkeae NIES254, Nephroselmis pyriformis RCC618, Dolichomastix tenuilepis CCMP3274) ingest live fluorescently labeled bacteria, detected by microscopy and/or flow cytometry
-
Full-text index only
Backseat drivers take the wheel.
PMID 18068625 · PMC2705833 · Cancer cell · 2007 · 8 claims · 8 setups
Systematic resequencing combined with functional validation can distinguish rare driver FLT3 mutations from passenger mutations in AML patients negative for known activating mutations
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Full-text index only
Identification and characterisation of Pseudomonas 16S ribosomal DNA from ileal biopsies of children with Crohn's disease.
PMID 18974839 · PMC2572839 · PloS one · 2008 · 7 claims · 6 setups
Pseudomonas 16S rDNA is significantly more prevalent in ileal biopsies of CD patients than non-IBD patients
-
Full-text index only
A surrogate-based approach for post-genomic partner identification.
PMID 11602024 · PMC57814 · BMC biotechnology · 2001 · 8 claims · 5 setups
Peptide surrogates derived from random phage display libraries contain amino acid sequence information that identifies the natural biological partner of the panned target via database searching.
-
Full-text index only
Systematic identification of pseudogenes through whole genome expression evidence profiling.
PMID 16945953 · PMC1636364 · Nucleic acids research · 2006 · 8 claims · 8 setups
Developed a novel bioinformatics method that identifies pseudogenes by profiling whole-genome transcript and protein expression evidence
-
Full-text index only
Comparative analysis of genome tiling array data reveals many novel primate-specific functional RNAs in human.
PMID 17288572 · PMC1796608 · BMC evolutionary biology · 2007 · 8 claims · 6 setups
Widespread transcription occurs across the human genome outside known gene annotations, and the bulk of TARs represent genuine transcripts rather than experimental artifacts