Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 95
OptiType: precision HLA typing from next-generation sequencing data.
PMID 25143287 · PMC4441069 · Bioinformatics (Oxford, England) · 2014 · 8 claims · 8 setups
OptiType, an ILP-based HLA genotyping algorithm, produces accurate four-digit HLA-I predictions from NGS data not enriched for the HLA cluster.
-
Has reproduction · 74
Disome-seq reveals widespread ribosome collisions that promote cotranslational protein folding.
PMID 33402206 · PMC7784341 · Genome biology · 2021 · 8 claims · 8 setups
Disome-seq sequences mRNA fragments protected by two stacked (collided) ribosomes, detecting ribosome collisions at codon resolution.
-
Full-text index only
Metagenomic study of the oral microbiota by Illumina high-throughput sequencing.
PMID 19796657 · PMC3568755 · Journal of microbiological methods · 2009 · 8 claims · 6 setups
The 16S rRNA V5 hypervariable region, amplified as a short ~82-base segment, provides reliable taxonomic identification of oral bacteria against public databases like HOMD.
-
Full-text index only
NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins.
PMID 15608248 · PMC539979 · Nucleic acids research · 2005 · 7 claims · 5 setups
RefSeq provides a curated, non-redundant, explicitly linked collection of genomic, transcript and protein sequences spanning prokaryotes, eukaryotes and viruses.
-
Full-text index only
SNP500Cancer: a public resource for sequence validation, assay development, and frequency analysis for genetic variation in candidate genes.
PMID 16381944 · PMC1347513 · Nucleic acids research · 2006 · 7 claims · 4 setups
SNP500Cancer provides sequence and genotype assay information for candidate cancer-related SNPs to support molecular epidemiology and complex disease mapping studies
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Full-text index only
Genome-wide identification of specific oligonucleotides using artificial neural network and computational genomic analysis.
PMID 17518996 · PMC1892811 · BMC bioinformatics · 2007 · 7 claims · 4 setups
The IAB algorithm (integration of ANN and BLAST) identifies genome-wide specific oligos much faster than pure BLAST search while maintaining comparable success rate and cross homology
-
Has reproduction · 45
Accurate sequence variant genotyping in cattle using variation-aware genome graphs.
PMID 31092189 · PMC6521551 · Genetics, selection, evolution : GSE · 2019 · 8 claims · 7 setups
Graphtyper outperformed GATK and SAMtools in genotype concordance, non-reference sensitivity, and non-reference discrepancy compared to microarray genotypes
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
Large-scale molecular analysis of a 34 Mb interval on chromosome 6q: major refinement of the RP25 interval.
PMID 18510646 · PMC2689154 · Annals of human genetics · 2008 · 7 claims · 5 setups
Direct sequencing of 43 candidate genes in 7 Spanish arRP families identified 244 sequence variants (76 novel), none pathogenic, excluding these genes as disease-causing.
-
Full-text index only
SNAP: predict effect of non-synonymous polymorphisms on function.
PMID 17526529 · PMC1920242 · Nucleic acids research · 2007 · 7 claims · 8 setups
SNAP, a neural network-based method using sequence-derived information, predicts whether a non-synonymous SNP is neutral or non-neutral for protein function
-
Full-text index only
Identification of deleterious non-synonymous single nucleotide polymorphisms using sequence-derived information.
PMID 18588693 · PMC2446391 · BMC bioinformatics · 2008 · 8 claims · 5 setups
A decision tree built on 10 selected sequence-derived features classifies SAPs as Disease or Polymorphism with 82.6% accuracy and 0.607 MCC in cross-validation.
-
Full-text index only
Genomic views of distant-acting enhancers.
PMID 19741700 · PMC2923221 · Nature · 2009 · 8 claims · 8 setups
Meta-analysis of ~1200 top GWAS SNPs found that in 40% of cases (472/1170) no known exons overlap the linked SNP or its haplotype block, implying noncoding variation causally contributes to many traits.
-
Full-text index only
SOP3v2: web-based selection of oligonucleotide primer trios for genotyping of human and mouse polymorphisms.
PMID 15980532 · PMC1160243 · Nucleic acids research · 2005 · 7 claims · 3 setups
SOP 3 v2 is a web-based application that outputs recommended forward/reverse PCR primers plus a sequencing primer optimized for sequence-based genotyping of human and mouse polymorphisms
-
Full-text index only
Comparative genomics comes of age.
PMID 12186641 · PMC139393 · Genome biology · 2002 · 8 claims · 8 setups
Only about 50% of conserved sequence elements (exons+introns) in orthologous human-mouse genes correspond to exons, implying substantial non-exonic conservation
-
Full-text index only
HLA-A gene polymorphism defined by high-resolution sequence-based typing in 161 Northern Chinese Han people.
PMID 15629059 · PMC5172246 · Genomics, proteomics & bioinformatics · 2003 · 7 claims · 5 setups
HLA-A gene shows high polymorphism in the Northern Chinese Han population, with 74 gene types and 36 alleles detected in 161 individuals
-
Full-text index only
Genome-wide detection of segmental duplications and potential assembly errors in the human genome sequence.
PMID 12702206 · PMC154576 · Genome biology · 2003 · 8 claims · 6 setups
Segmental duplications comprise 3.53% (107.4/3,043.1 Mb) of the June 2002 human genome assembly
-
Full-text index only
A genome annotation-driven approach to cloning the human ORFeome.
PMID 15461802 · PMC545604 · Genome biology · 2004 · 8 claims · 8 setups
Existing human cDNA clone collections together provide only 60% coverage of full-length chromosome 22 ORFs, with the best single collection (MGC) providing 48%
-
Full-text index only
A genome-wide survey of Major Histocompatibility Complex (MHC) genes and their paralogues in zebrafish.
PMID 16271140 · PMC1309616 · BMC genomics · 2005 · 8 claims · 4 setups
149 putative MHC gene loci and their paralogues were identified in the zebrafish genome using sequence similarity searches against the Zv4 draft assembly.