Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
FeatureScan: revealing property-dependent similarity of nucleotide sequences.
PMID 16845077 · PMC1538849 · Nucleic acids research · 2006 · 6 claims · 5 setups
FeatureScan transforms nucleotide sequences into numerical signals of physico-chemical/conformational properties and compares them via a convolution/correlation (Fourier transform) method rather than comparing letters
-
Full-text index only
TFBScluster web server for the identification of mammalian composite regulatory elements.
PMID 16845063 · PMC1538905 · Nucleic acids research · 2006 · 7 claims · 5 setups
TFBScluster is a web server that identifies genome-wide clusters of TFBSs conserved in multiple mammalian species using human or mouse as the reference genome.
-
Full-text index only
BTW: a web server for Boltzmann time warping of gene expression time series.
PMID 16845055 · PMC1538860 · Nucleic acids research · 2006 · 5 claims · 4 setups
Symmetric time warping distance is more flexible than Euclidean distance or correlation coefficient for identifying genes with similar temporal expression profiles, especially across sequences of different length.
-
Full-text index only
GeneAlign: a coding exon prediction tool based on phylogenetical comparisons.
PMID 16845010 · PMC1538901 · Nucleic acids research · 2006 · 8 claims · 5 setups
GeneAlign predicts coding exons by using signal detection (GeneSplicer/WMM) combined with CORAL, a heuristic linear-time alignment tool, to align candidate signal-flanked regions against annotated exons of a homologous organism's genes
-
Full-text index only
Prediction of catalytic residues using Support Vector Machine with selected protein sequence and structural properties.
PMID 16790052 · PMC1534064 · BMC bioinformatics · 2006 · 8 claims · 7 setups
The Sequential Minimal Optimization (SMO) SVM algorithm was the best-performing classifier among 26 WEKA classifiers for predicting catalytic residues
-
Has reproduction · 68
LaSSO, a strategy for genome-wide mapping of intronic lariats and branch points using RNA-seq.
PMID 24709818 · PMC4079972 · Genome research · 2014 · 8 claims · 8 setups
LaSSO (Lariat Sequence Site Origin) identifies intronic lariat reads and pinpoints branch points genome-wide from RNA-seq data by considering every intronic base as a potential branch point and including all possible exon-skipping lariats.
-
Full-text index only
Genome-wide survey for biologically functional pseudogenes.
PMID 16680195 · PMC1456316 · PLoS computational biology · 2006 · 8 claims · 6 setups
A subset of ancient, cross-species-conserved pseudogenes (30 of 1,453 candidate quartets) show evidence consistent with retained biological function
-
Full-text index only
Novel gene and gene model detection using a whole genome open reading frame analysis in proteomics.
PMID 16646984 · PMC1557991 · Genome biology · 2006 · 8 claims · 4 setups
A six-frame genomic ORF translation used as an MS search database can detect novel peptides absent from standard protein databases, revealing incomplete genome annotation.
-
Full-text index only
Eight previously unidentified mutations found in the OA1 ocular albinism gene.
PMID 16646960 · PMC1468396 · BMC medical genetics · 2006 · 7 claims · 5 setups
Sequencing of the nine OA1 exons in 72 individuals identified ten different mutations across seven unrelated families and three sporadic cases.
-
Full-text index only
Predicting deleterious nsSNPs: an analysis of sequence and structural attributes.
PMID 16630345 · PMC1489951 · BMC bioinformatics · 2006 · 8 claims · 7 setups
Sequence conservation (PSIC score difference) at the nsSNP position is the single most useful attribute for predicting deleterious vs neutral status.
-
Full-text index only
DNA sequence of human chromosome 17 and analysis of rearrangement in the human lineage.
PMID 16625196 · PMC2610434 · Nature · 2006 · 8 claims · 7 setups
A finished sequence of human chromosome 17 (78,839,971 bases, ~2.8% of the euchromatic genome) was generated.
-
Full-text index only
Identifying repeat domains in large genomes.
PMID 16507140 · PMC1431705 · Genome biology · 2006 · 7 claims · 5 setups
A repeat domain graph, built using a modified A-Bruijn graph framework, decomposes a repeat library into shared repeat domains and reveals the mosaic structure of repeat families.
-
Full-text index only
The fragile breakage versus random breakage models of chromosome evolution.
PMID 16501665 · PMC1378107 · PLoS computational biology · 2006 · 8 claims · 6 setups
Sankoff and Trinh's synteny block identification algorithm (ST-Synteny) is flawed, producing erroneous block identifications even in small toy examples.
-
Full-text index only
Novel dengue virus type 1 from travelers to Yap State, Micronesia.
PMID 16494770 · PMC3373118 · Emerging infectious diseases · 2006 · 8 claims · 5 setups
DENV-1 responsible for the 2004 Yap State dengue outbreak was isolated from serum of 4 Japanese travelers returning from Yap
-
Full-text index only
Evolution of the NANOG pseudogene family in the human and chimpanzee genomes.
PMID 16469101 · PMC1457002 · BMC evolutionary biology · 2006 · 7 claims · 5 setups
The NANOG gene and all pseudogenes except NANOGP8 occupy orthologous chromosomal positions in the chimpanzee genome, indicating they originated before the human-chimpanzee divergence.
-
Full-text index only
MPromDb: an integrated resource for annotation and visualization of mammalian gene promoters and ChIP-chip experimental data.
PMID 16381984 · PMC1347458 · Nucleic acids research · 2006 · 8 claims · 5 setups
MPromDb is a novel database integrating experimentally supported gene promoters, TSS annotation, cis-regulatory elements, CpG islands, and ChIP-chip data with an integrated visualization interface.
-
Full-text index only
ABS: a database of Annotated regulatory Binding Sites from orthologous promoters.
PMID 16381947 · PMC1347478 · Nucleic acids research · 2006 · 7 claims · 6 setups
ABS is a public database of experimentally identified TF binding sites conserved in orthologous vertebrate gene promoters, manually curated from the literature.
-
Full-text index only
From genomics to chemical genomics: new developments in KEGG.
PMID 16381885 · PMC1347464 · Nucleic acids research · 2006 · 8 claims · 5 setups
KEGG BRITE has been formally added as a fourth main KEGG database to establish a logical foundation for functional interpretation and pathway reconstruction.
-
Full-text index only
TreeFam: a curated database of phylogenetic trees of animal gene families.
PMID 16381935 · PMC1347480 · Nucleic acids research · 2006 · 7 claims · 6 setups
Tree-based inference of orthologs and paralogs is more robust than BLAST-based methods because evolutionary rates (and thus pairwise BLAST scores) vary across gene family members
-
Has reproduction · 62
Metatranscriptomics of the human oral microbiome during health and disease.
PMID 24692635 · PMC3977359 · mBio · 2014 · 8 claims · 8 setups
Disease-associated periodontal communities display conserved community-level metabolic gene expression profiles between patients, whereas the metabolic gene expression of individual species is highly variable between patients.