Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Non-EST based prediction of exon skipping and intron retention events using Pfam information.
PMID 16204458 · PMC1243800 · Nucleic acids research · 2005 · 7 claims · 5 setups
A novel ab initio method predicts exon skipping and intron retention events using only Pfam domain annotation, via a Viterbi-like dynamic programming algorithm applied to the Pfam alignment.
-
Full-text index only
Computational approaches for predicting the biological effect of p53 missense mutations: a comparison of three sequence analysis based methods.
PMID 16522644 · PMC1390679 · Nucleic acids research · 2006 · 7 claims · 6 setups
Align-GVGD predicts loss of transactivation activity with high specificity (~88%) but lower sensitivity (67.9-71.2%) for neutral mutants
-
Full-text index only
Cancer-specific high-throughput annotation of somatic mutations: computational prediction of driver missense mutations.
PMID 19654296 · PMC2763410 · Cancer research · 2009 · 7 claims · 7 setups
CHASM, a Random Forest-based computational method, was developed to identify and prioritize missense mutations likely to be functional drivers of tumor cell proliferation.
-
Full-text index only
Endonuclease-independent insertion provides an alternative pathway for L1 retrotransposition in the human genome.
PMID 17517773 · PMC1920257 · Nucleic acids research · 2007 · 8 claims · 5 setups
An endonuclease-independent pathway (NCLI) for L1 insertion has been active in recent human genome evolution
-
Full-text index only
Prediction of specificity-determining residues for small-molecule kinase inhibitors.
PMID 19032760 · PMC2655090 · BMC bioinformatics · 2008 · 8 claims · 5 setups
S-Filter is a novel method combining sequence and structural information (within PFAAT) to predict specificity-determining residues and selectivity profiles for small-molecule kinase inhibitors
-
Full-text index only
Aberrant 5' splice sites in human disease genes: mutation pattern, nucleotide structure and comparison of computational tools that predict their utilization.
PMID 17576681 · PMC1934990 · Nucleic acids research · 2007 · 8 claims · 4 setups
Cryptic 5'ss are best predicted by computational algorithms that accommodate nucleotide dependencies (e.g., Markov model, maximum entropy, maximum dependence decomposition) rather than by weight-matrix models
-
Full-text index only
Systems biology of gene regulation fulfills its promise.
PMID 16719937 · PMC1779525 · Genome biology · 2006 · 8 claims · 8 setups
Suz12, a Polycomb Group complex component, has DNA targets identifiable by ChIP-chip and can silence large genomic regions in a cell-type-specific manner.
-
Has reproduction · 85
Predicting the pathogenicity of missense variants using features derived from AlphaFold2.
PMID 37084271 · PMC10203375 · Bioinformatics (Oxford, England) · 2023 · 6 claims · 8 setups
AlphaFold2-derived structural features (solvent accessibility, amino acid network features, physicochemical environment, pLDDT) can be used to train a random forest classifier (AlphScore) that distinguishes proxy-benign from proxy-pathogenic missense variants.
-
Full-text index only
The distribution of SNPs in human gene regulatory regions.
PMID 16209714 · PMC1260019 · BMC genomics · 2005 · 8 claims · 6 setups
SNPs occur with higher density closer to the transcriptional start site within gene promoter regions than in further upstream regions
-
Full-text index only
Predicting the phenotypic effects of non-synonymous single nucleotide polymorphisms based on support vector machines.
PMID 18005451 · PMC2216041 · BMC bioinformatics · 2007 · 8 claims · 5 setups
Parepro, an SVM-based method integrating three attribute sets (RD, MI, IE) derived from evolutionary and residue-property information, predicts whether an nsSNP is deleterious or neutral.
-
Full-text index only
CTCFBSDB: a CTCF-binding site database for characterization of vertebrate genomic insulators.
PMID 17981843 · PMC2238977 · Nucleic acids research · 2008 · 7 claims · 8 setups
CTCF is the only identified trans-acting factor in vertebrates that confers enhancer-blocking insulator activity
-
Full-text index only
Towards a comprehensive structural coverage of completed genomes: a structural genomics viewpoint.
PMID 17349043 · PMC1829165 · BMC bioinformatics · 2007 · 8 claims · 6 setups
A combined target-selection approach — pursuing both structurally uncharacterised domain families and additional targets from large structurally characterised superfamilies — is essential for comprehensive structural coverage of the genomes.
-
Full-text index only
Protective effect of paraoxonase 1 gene variant Gln192Arg in age-related macular degeneration.
PMID 20042177 · PMC3026437 · American journal of ophthalmology · 2010 · 6 claims · 4 setups
The Gln192Arg PON1 polymorphism is associated with decreased susceptibility to AMD, particularly wet AMD, indicating a protective effect
-
Full-text index only
The mitochondrial genome, a growing interest inside an organelle.
PMID 18488415 · PMC2526360 · International journal of nanomedicine · 2008 · 8 claims · 8 setups
mtDNA mutations are causally linked to a wide range of mitochondrial diseases, aging, and chronic degenerative diseases
-
Has reproduction · 83
Integrative transcriptomic and machine learning framework reveals candidate genes and potential mechanisms of aflatoxin B1 exposure in breast cancer.
PMID 41688730 · PMC12982753 · Scientific reports · 2026 · 7 claims · 8 setups
Twenty-two genes lie at the intersection of AFB1-predicted targets and breast cancer-associated co-expression modules/DEGs
-
Full-text index only
TRED: a Transcriptional Regulatory Element Database and a platform for in silico gene regulation studies.
PMID 15608156 · PMC539958 · Nucleic acids research · 2005 · 8 claims · 5 setups
TRED is a database collecting both cis-regulatory elements (promoters) and trans-regulatory elements (transcription factor binding/regulation data) with linked access.
-
Full-text index only
Automatic annotation of eukaryotic genes, pseudogenes and promoters.
PMID 16925832 · PMC1810547 · Genome biology · 2006 · 8 claims · 6 setups
Fgenesh++ gene prediction pipeline identifies 91% of coding nucleotides with 90% specificity
-
Full-text index only
EGASP: Introduction.
PMID 16925831 · PMC1810546 · Genome biology · 2006 · 8 claims · 5 setups
Computational gene finding methods, when compared to the GENCODE golden standard annotation, show that the human genome annotation is nearly complete in terms of novel protein-coding loci.
-
Full-text index only
miRGator: an integrated system for functional annotation of microRNAs.
PMID 17942429 · PMC2238850 · Nucleic acids research · 2008 · 8 claims · 8 setups
miRGator integrates target prediction, functional enrichment analysis (GO/pathway/disease), and expression data (miRNA/mRNA/protein) into one system for functional annotation of miRNAs
-
Full-text index only
TRED: a transcriptional regulatory element database, new entries and other development.
PMID 17202159 · PMC1899102 · Nucleic acids research · 2007 · 8 claims · 3 setups
TRED collects mammalian cis- and trans-regulatory elements together with experimental evidence, mapped onto assembled genomes