Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A modified T-test feature selection method and its application on the HapMap genotype data.
PMID 18267305 · PMC5054219 · Genomics, proteomics & bioinformatics · 2007 · 7 claims · 4 setups
A modified t-test ranking measure, extended to handle nominal SNP genotype data via vector transformation, can effectively rank SNPs by their discriminative capability for population classification.
-
Full-text index only
Application of machine learning in SNP discovery.
PMID 16398931 · PMC1955739 · BMC bioinformatics · 2006 · 8 claims · 6 setups
PolyBayes produces high false-positive SNP predictions even with stringent parameters
-
Full-text index only
Supervised learning-based tagSNP selection for genome-wide disease classifications.
PMID 18366619 · PMC2386071 · BMC genomics · 2008 · 7 claims · 2 setups
SRFA (Supervised Recursive Feature Addition) is a novel feature selection method combining supervised learning and statistical redundancy measures for SNP selection
-
Full-text index only
SNP-RFLPing: restriction enzyme mining for SNPs in genomes.
PMID 16503968 · PMC1386656 · BMC genomics · 2006 · 8 claims · 2 setups
SNP-RFLPing accepts three flexible input types (dbSNP rs#/ss# IDs, HUGO gene name/Entrez gene ID, or free-form SNP-in-sequence including IUPAC or [dNTP1/dNTP2] formats) for human, rat, and mouse genomes
-
Full-text index only
A comparison of classification methods for predicting Chronic Fatigue Syndrome based on genetic data.
PMID 19772600 · PMC2765429 · Journal of translational medicine · 2009 · 7 claims · 3 setups
The naive Bayes model with the wrapper-based feature selection approach performed best among all predictive models tested for distinguishing CFS from controls.
-
Full-text index only
VarDetect: a nucleotide sequence variation exploratory tool.
PMID 19091032 · PMC2638149 · BMC bioinformatics · 2008 · 8 claims · 2 setups
VarDetect is a stand-alone software tool that automatically detects nucleotide variation (SNPs) from fluorescence-based chromatogram traces using pre-calculated peak content ratios and artifact-handling rules.
-
Full-text index only
BioHealthBase: informatics support in the elucidation of influenza virus host pathogen interactions and virulence.
PMID 17965094 · PMC2238987 · Nucleic acids research · 2008 · 7 claims · 5 setups
BioHealthBase BRC is a public integrated bioinformatics database and analysis resource for influenza virus, Francisella tularensis, Mycobacterium tuberculosis, Microsporidia species and ricin toxin.
-
Full-text index only
Natural selection of protein structural and functional properties: a single nucleotide polymorphism perspective.
PMID 18397526 · PMC2643940 · Genome biology · 2008 · 8 claims · 8 setups
The SNP A/S ratio is a robust measure of selective constraint, correlating with interspecies Ka/Ks ratios and with protein sequence conservation.
-
Full-text index only
Sequence variation in G-protein-coupled receptors: analysis of single nucleotide polymorphisms.
PMID 15784611 · PMC1069129 · Nucleic acids research · 2005 · 7 claims · 8 setups
Position-specific phylogenetic features describing evolutionary conservation at a site (e.g. SIFT score, normalized site entropy, residue frequency change) are the best individual discriminators of disease-causing versus neutral GPCR mutations.
-
Has reproduction · 40
On the holobiont 'predictome' of immunocompetence in pigs.
PMID 37127575 · PMC10150480 · Genetics, selection, evolution : GSE · 2023 · 8 claims · 8 setups
Holobiont (combined genotype + microbiome) models performed better than partial models (genotype-only or microbiome-only) overall
-
Full-text index only
Detecting purely epistatic multi-locus interactions by an omnibus permutation test on ensembles of two-locus analyses.
PMID 19761607 · PMC2759961 · BMC bioinformatics · 2009 · 8 claims · 5 setups
2LOmb performs an omnibus permutation test on ensembles of two-locus analyses via a four-step algorithm (two-locus analysis, permutation test, global p-value determination, progressive ensemble search)
-
Full-text index only
Fast-evolving noncoding sequences in the human genome.
PMID 17578567 · PMC2394770 · Genome biology · 2007 · 8 claims · 6 setups
1,356 conserved noncoding sequences show human-specific accelerated substitution rates (ANC sequences) relative to chimpanzee
-
Has reproduction · 83
Analyzing biomarker discovery: Estimating the reproducibility of biomarker sets.
PMID 35901020 · PMC9333302 · PloS one · 2022 · 7 claims · 3 setups
A Reproducibility Score, RS(D,BD), defined as the average Jaccard overlap between biomarker sets found by the same discovery process on comparable datasets from the same distribution, quantifies biomarker reproducibility on a 0-1 scale
-
Full-text index only
Analysis of expressed sequence tags from Actinidia: applications of a cross species EST database for gene discovery in the areas of flavor, health, color and ripening.
PMID 18655731 · PMC2515324 · BMC genomics · 2008 · 7 claims · 6 setups
A collection of 132,577 ESTs from four Actinidia species was generated and clustered into 41,858 non-redundant clusters (18,070 TCs and 23,788 singletons)
-
Full-text index only
Identification of deleterious non-synonymous single nucleotide polymorphisms using sequence-derived information.
PMID 18588693 · PMC2446391 · BMC bioinformatics · 2008 · 8 claims · 5 setups
A decision tree built on 10 selected sequence-derived features classifies SAPs as Disease or Polymorphism with 82.6% accuracy and 0.607 MCC in cross-validation.