Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 85
Predicting the pathogenicity of missense variants using features derived from AlphaFold2.
PMID 37084271 · PMC10203375 · Bioinformatics (Oxford, England) · 2023 · 6 claims · 8 setups
AlphaFold2-derived structural features (solvent accessibility, amino acid network features, physicochemical environment, pLDDT) can be used to train a random forest classifier (AlphScore) that distinguishes proxy-benign from proxy-pathogenic missense variants.
-
Full-text index only
Searching for interpretable rules for disease mutations: a simulated annealing bump hunting strategy.
PMID 16984653 · PMC1618409 · BMC bioinformatics · 2006 · 8 claims · 6 setups
The proposed feature set outperforms existing published feature sets for predicting effects of amino acid substitutions
-
Full-text index only
Reverse polarization in amino acid and nucleotide substitution patterns between human-mouse orthologs of two compositional extrema.
PMID 17895298 · PMC2533592 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2007 · 8 claims · 7 setups
Nucleotide and amino acid substitution trends between human-mouse orthologs are highly asymmetric and polarized in opposite directions for high-GC versus low-GC gene groups.
-
Full-text index only
Predicting the phenotypic effects of non-synonymous single nucleotide polymorphisms based on support vector machines.
PMID 18005451 · PMC2216041 · BMC bioinformatics · 2007 · 8 claims · 5 setups
Parepro, an SVM-based method integrating three attribute sets (RD, MI, IE) derived from evolutionary and residue-property information, predicts whether an nsSNP is deleterious or neutral.
-
Full-text index only
Exhaustive prediction of disease susceptibility to coding base changes in the human genome.
PMID 18793467 · PMC2537574 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Inter-species conservation is the strongest single predictor of disease-associated coding mutations among the factors tested.
-
Full-text index only
SNP-VISTA: an interactive SNP visualization tool.
PMID 16336665 · PMC1325058 · BMC bioinformatics · 2005 · 7 claims · 3 setups
SNP-VISTA is an interactive Java-based visualization tool with two versions, GeneSNP-VISTA and EcoSNP-VISTA, for exploring large-scale SNP datasets
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Full-text index only
SNAP: predict effect of non-synonymous polymorphisms on function.
PMID 17526529 · PMC1920242 · Nucleic acids research · 2007 · 7 claims · 8 setups
SNAP, a neural network-based method using sequence-derived information, predicts whether a non-synonymous SNP is neutral or non-neutral for protein function
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Full-text index only
Positive natural selection in the evolution of human metapneumovirus attachment glycoprotein.
PMID 17931731 · PMC7114232 · Virus research · 2008 · 7 claims · 5 setups
8 amino acid sites in the extracellular domain of hMPV lineage 1a show a higher rate of nonsynonymous than synonymous substitutions (posterior probability >0.95), indicating positive selection.
-
Full-text index only
In silico analysis of missense substitutions using sequence-alignment based methods.
PMID 18951440 · PMC3431198 · Human mutation · 2008 · 8 claims · 7 setups
Carefully validated PMSA-based computational algorithms can achieve predictive values of ~75-95% for classifying missense substitutions as pathogenic or neutral.
-
Full-text index only
Mutation spectrum of homogentisic acid oxidase (HGD) in alkaptonuria.
PMID 19862842 · PMC2830005 · Human mutation · 2009 · 8 claims · 6 setups
Mutations in HGD, which encodes homogentisate dioxygenase, cause AKU by blocking conversion of homogentisic acid to maleylacetoacetic acid in the tyrosine catabolic pathway
-
Full-text index only
Flux balance analysis of mycolic acid pathway: targets for anti-tubercular drugs.
PMID 16261191 · PMC1246807 · PLoS computational biology · 2005 · 7 claims · 7 setups
A comprehensive stoichiometric model of the MAP was built comprising 197 metabolites, 219 reactions, and 28 proteins
-
Full-text index only
The protein-phosphatome of the human malaria parasite Plasmodium falciparum.
PMID 18793411 · PMC2559854 · BMC genomics · 2008 · 8 claims · 8 setups
P. falciparum possesses 27 putative protein phosphatase sequences across the four major PP families (PPP, PPM, PTP, NIF), plus 7 additional sequences predicted to dephosphorylate non-protein substrates, totaling 34.
-
Full-text index only
Sequence variation in G-protein-coupled receptors: analysis of single nucleotide polymorphisms.
PMID 15784611 · PMC1069129 · Nucleic acids research · 2005 · 7 claims · 8 setups
Position-specific phylogenetic features describing evolutionary conservation at a site (e.g. SIFT score, normalized site entropy, residue frequency change) are the best individual discriminators of disease-causing versus neutral GPCR mutations.
-
Full-text index only
Genetic variation of St. Louis encephalitis virus.
PMID 18632961 · PMC2696384 · The Journal of general virology · 2008 · 8 claims · 4 setups
Phylogenetic analysis of 106 SLEV E gene sequences confirms seven major lineages (I-VII) and refines them into 13 clades (IA, IB, IIA, IIB, IIC, IID, IIG, III, IV, VA, VB, VI, VII)
-
Full-text index only
Using structural bioinformatics to investigate the impact of non synonymous SNPs and disease mutations: scope and limitations.
PMID 19758473 · PMC2745591 · BMC bioinformatics · 2009 · 8 claims · 8 setups
None of 39 tested structural properties can be used as a sole classification criterion to separate neutral SNPs from disease mutations.
-
Full-text index only
SNAP predicts effect of mutations on protein function.
PMID 18757876 · PMC2562009 · Bioinformatics (Oxford, England) · 2008 · 8 claims · 3 setups
SNAP is a publicly available web-server implementation predicting functional effects (neutral/non-neutral) of single amino acid substitutions.
-
Full-text index only
Hepatitis B virus genotypes circulating in Brazil: molecular characterization of genotype F isolates.
PMID 18036224 · PMC2231365 · BMC microbiology · 2007 · 8 claims · 4 setups
Genotypes A, D, and F co-circulate in each of the five Brazilian geographic regions, with no other genotypes identified among 303 isolates
-
Full-text index only
Predicting positive p53 cancer rescue regions using Most Informative Positive (MIP) active learning.
PMID 19756158 · PMC2742196 · PLoS computational biology · 2009 · 8 claims · 4 setups
MIP active learning is a novel active learning method that preferentially seeks informative Positive (functionally active) examples rather than only maximizing classifier accuracy.