Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Evaluation of six methods for estimating synonymous and nonsynonymous substitution rates.
PMID 17127215 · PMC5054070 · Genomics, proteomics & bioinformatics · 2006 · 8 claims · 4 setups
Incorporating more sequence evolution features (transition/transversion bias, nucleotide/codon frequency bias) into Ka/Ks estimation methods yields more accurate and reliable estimates.
-
Full-text index only
Vertebrate gene finding from multiple-species alignments using a two-level strategy.
PMID 16925840 · PMC1810555 · Genome biology · 2006 · 8 claims · 5 setups
DOGFISH cleanly separates a multi-species alignment classifier (RVM cascade) from an HMM-based structure predictor, avoiding tight coupling of alignment complexity with HMM formalism
-
Full-text index only
The vertebrate genome annotation (Vega) database.
PMID 18003653 · PMC2238886 · Nucleic acids research · 2008 · 8 claims · 8 setups
Vega is a database for viewing manual genome annotation of human, mouse and zebrafish genomic sequences produced at the Wellcome Trust Sanger Institute.
-
Full-text index only
Multiplexed discovery of sequence polymorphisms using base-specific cleavage and MALDI-TOF MS.
PMID 15731331 · PMC549577 · Nucleic acids research · 2005 · 8 claims · 7 setups
Multiplexed base-specific cleavage/MALDI-TOF MS (Multiplexed Comparative Sequence Analysis) enables simultaneous discovery of sequence polymorphisms across multiple target regions
-
Full-text index only
Computing Ka and Ks with a consideration of unequal transitional substitutions.
PMID 16740169 · PMC1552089 · BMC evolutionary biology · 2006 · 7 claims · 7 setups
MYN, a modified version of the Yang-Nielsen (YN) algorithm based on the Tamura-Nei Model, allows unequal transitional substitution rates between purines (κR) and pyrimidines (κY) plus codon frequency bias
-
Full-text index only
Ensembl 2007.
PMID 17148474 · PMC1761443 · Nucleic acids research · 2007 · 8 claims · 7 setups
Ensembl added 18 new chordate genomes this year, increasing total genomes available from 15 to 33, the largest yearly increase to date.
-
Full-text index only
Ensembl 2009.
PMID 19033362 · PMC2686571 · Nucleic acids research · 2009 · 8 claims · 6 setups
Ensembl provides comprehensive, consistently annotated genome information for chordate genomes with automatically generated genesets and comparative genomics data
-
Full-text index only
SVC: structured visualization of evolutionary sequence conservation.
PMID 15991338 · PMC1160265 · Nucleic acids research · 2005 · 7 claims · 5 setups
SVC aligns protein-coding sequences of orthologous gene pairs and maps them back onto their encoding exons/introns to generate a scaffold of conserved gene structure.
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs
-
Full-text index only
F-SNP: computationally predicted functional SNPs for disease association studies.
PMID 17986460 · PMC2238878 · Nucleic acids research · 2008 · 6 claims · 8 setups
F-SNP is a database integrating functional effect predictions for SNPs from 16 bioinformatics tools/databases across four categories: splicing, transcription, translation, and post-translation
-
Full-text index only
Frameshift mutations in coding repeats of protein tyrosine phosphatase genes in colorectal tumors with microsatellite instability.
PMID 19000305 · PMC2586028 · BMC cancer · 2008 · 7 claims · 6 setups
16 PTP candidate genes containing coding mononucleotide repeats (cMNR) of at least 7 units were identified via bioinformatic analysis and screened in MSI-H cell lines, cancers, and adenomas
-
Full-text index only
G-quadruplexes: the beginning and end of UTRs.
PMID 18832370 · PMC2577360 · Nucleic acids research · 2008 · 8 claims · 5 setups
UTRs show significant strand asymmetry with C-PQS more common than G-PQS, consistent with general depletion of G-quadruplex-forming RNA
-
Full-text index only
Functional analysis of novel SNPs and mutations in human and mouse genomes.
PMID 19091009 · PMC2638150 · BMC bioinformatics · 2008 · 8 claims · 7 setups
FANS streamlines functional analysis of novel SNPs and mutations into a simplified, few-click, four-step procedure.
-
Full-text index only
An optimized procedure for the design and evaluation of Ecotilling assays.
PMID 18973671 · PMC2586031 · BMC genomics · 2008 · 8 claims · 7 setups
An optimized procedure integrating Vector NTI, Ensembl, Genomatix Suite, GelBuddy, and sequencing/functional-prediction tools streamlines the design, evaluation and interpretation of human Ecotilling assays
-
Full-text index only
Mutation analysis of the MSMB gene in familial prostate cancer.
PMID 19997100 · PMC2816656 · British journal of cancer · 2010 · 8 claims · 5 setups
No deleterious mutations were found in the MSMB coding region in 192 familial prostate cancer cases
-
Full-text index only
PigGIS: Pig Genomic Informatics System.
PMID 17090590 · PMC1669765 · Nucleic acids research · 2007 · 7 claims · 7 setups
PigGIS identified 15,700 pig consensus sequences covering 18.5 Mb of homologous human exons
-
Full-text index only
Conserved elements with potential to form polymorphic G-quadruplex structures in the first intron of human genes.
PMID 18187510 · PMC2275096 · Nucleic acids research · 2008 · 8 claims · 6 setups
G-richness downstream of the TSS is strand-biased, concentrated on the nontemplate strand, with a peak at +200 to +300 bp
-
Has reproduction · 94
A Deluge of Complex Repeats: The Solanum Genome.
PMID 26241045 · PMC4524691 · PloS one · 2015 · 8 claims · 7 setups
~50–60% of the S. tuberosum and S. lycopersicum genomes are composed of repetitive elements
-
Full-text index only
Satellog: a database for the identification and prioritization of satellite repeats in disease association studies.
PMID 15949044 · PMC1181805 · BMC bioinformatics · 2005 · 7 claims · 6 setups
Satellog is a database cataloging all pure 1-16 unit satellite repeats in the human genome with supplementary polymorphism, gene-location, and expression data for prioritizing repeats in disease-association studies.
-
Full-text index only
Anopheles gambiae genome reannotation through synthesis of ab initio and comparative gene prediction algorithms.
PMID 16569258 · PMC1557760 · Genome biology · 2006 · 8 claims · 7 setups
An exon-gene-union (EGU) algorithm followed by an open-reading-frame-selection algorithm can synthesize ab initio (GENSCAN, GeneMark, SNAP) and comparative (Ensembl/Genewise) predictions into a single, more complete CDS set