Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Preponderance of the oncogenic V599E and V599K mutations in B-raf kinase domain is enhanced in melanoma cutaneous/subcutaneous metastases.
PMID 15935100 · PMC1164406 · BMC cancer · 2005 · 8 claims · 4 setups
B-raf exon 15 somatic mutations occur in 24/60 (40%) of cutaneous/subcutaneous melanoma metastases, similar to the frequency in primary melanomas
-
Full-text index only
Predictive screening for regulators of conserved functional gene modules (gene batteries) in mammals.
PMID 15882449 · PMC1134656 · BMC genomics · 2005 · 8 claims · 4 setups
A predictive computational screen covering ~40% of annotated protein-coding genes identified 21 co-expressed gene clusters with statistically supported sharing of cis-regulatory motifs.
-
Full-text index only
Large-scale analysis of human alternative protein isoforms: pattern classification and correlation with subcellular localization signals.
PMID 15860772 · PMC1087780 · Nucleic acids research · 2005 · 8 claims · 8 setups
Constructed a large-scale dataset of 6876 human alternative protein isoforms from 2624 genes by combining H-Invitational full-length cDNA data and SwissProt VARSPLIC entries
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
TRED: a Transcriptional Regulatory Element Database and a platform for in silico gene regulation studies.
PMID 15608156 · PMC539958 · Nucleic acids research · 2005 · 8 claims · 5 setups
TRED is a database collecting both cis-regulatory elements (promoters) and trans-regulatory elements (transcription factor binding/regulation data) with linked access.
-
Full-text index only
Whole genome association mapping by incompatibilities and local perfect phylogenies.
PMID 17042942 · PMC1624851 · BMC bioinformatics · 2006 · 8 claims · 8 setups
Blossoc scores the perfect phylogenetic tree spanning the largest compatible region around each marker as a decision tree for case/control status to detect association
-
Full-text index only
Evaluation of NTHL1, NEIL1, NEIL2, MPG, TDG, UNG and SMUG1 genes in familial colorectal cancer predisposition.
PMID 17029639 · PMC1624846 · BMC cancer · 2006 · 6 claims · 4 setups
Coding sequences and intron-exon boundaries of NTHL1, NEIL1, NEIL2, MPG, TDG, UNG and SMUG1 were screened in 94 familial CRC cases with known genes excluded
-
Full-text index only
In silico and in vivo splicing analysis of MLH1 and MSH2 missense mutations shows exon- and tissue-specific effects.
PMID 16995940 · PMC1590028 · BMC genomics · 2006 · 8 claims · 6 setups
In silico ESE-prediction algorithms (ESEfinder, RescueESE, PESX) do not reliably predict actual in vivo splicing behavior of missense mutations
-
Full-text index only
Using several pair-wise informant sequences for de novo prediction of alternatively spliced transcripts.
PMID 16925842 · PMC1810557 · Genome biology · 2006 · 8 claims · 4 setups
MARS, an extension of the Twinscan algorithm, uses multiple pairwise informant genomes to predict human alternatively spliced transcripts de novo without expressed sequence information.
-
Full-text index only
Exogean: a framework for annotating protein-coding genes in eukaryotic genomic DNA.
PMID 16925841 · PMC1810556 · Genome biology · 2006 · 8 claims · 5 setups
Exogean is a framework using directed acyclic coloured multigraphs (DACMs) to represent biological objects (mRNA, ESTs, protein alignments, exons) and iteratively combine them into complex protein-coding transcript models.
-
Full-text index only
Automatic annotation of eukaryotic genes, pseudogenes and promoters.
PMID 16925832 · PMC1810547 · Genome biology · 2006 · 8 claims · 6 setups
Fgenesh++ gene prediction pipeline identifies 91% of coding nucleotides with 90% specificity
-
Full-text index only
EGASP: Introduction.
PMID 16925831 · PMC1810546 · Genome biology · 2006 · 8 claims · 5 setups
Computational gene finding methods, when compared to the GENCODE golden standard annotation, show that the human genome annotation is nearly complete in terms of novel protein-coding loci.
-
Full-text index only
Adaptively inferring human transcriptional subnetworks.
PMID 16760900 · PMC1681499 · Molecular systems biology · 2006 · 8 claims · 7 setups
A multivariate linear spline (MARS-based) model correlating PWM binding scores with log expression ratios can identify active cis-motif combinations in mammalian promoters without requiring gene clustering.
-
Full-text index only
Benchmarking ortholog identification methods using functional genomics data.
PMID 16613613 · PMC1557999 · Genome biology · 2006 · 8 claims · 7 setups
InParanoid is the best overall ortholog identification method for identifying functionally equivalent proteins when sensitivity and selectivity are combined into an overall score.
-
Full-text index only
Anopheles gambiae genome reannotation through synthesis of ab initio and comparative gene prediction algorithms.
PMID 16569258 · PMC1557760 · Genome biology · 2006 · 8 claims · 7 setups
An exon-gene-union (EGU) algorithm followed by an open-reading-frame-selection algorithm can synthesize ab initio (GENSCAN, GeneMark, SNAP) and comparative (Ensembl/Genewise) predictions into a single, more complete CDS set
-
Full-text index only
Identification and evolutionary analysis of novel exons and alternative splicing events using cross-species EST-to-genome comparisons in human, mouse and rat.
PMID 16536879 · PMC1479377 · BMC bioinformatics · 2006 · 8 claims · 6 setups
ENACE, a cross-species EST-to-genome comparison algorithm, can identify novel cassette-on exons and retained introns for EST-scanty species and distinguish conserved vs lineage-specific exons
-
Full-text index only
Prospective health care: the second transformation of medicine.
PMID 16522218 · PMC1431721 · Genome biology · 2006 · 8 claims · 7 setups
Predictive biomarkers (genomic, proteomic, metabolomic) will enable quantification of disease risk and anticipation of onset before pathological damage occurs.
-
Full-text index only
Functional genomics of early cortex patterning.
PMID 16515721 · PMC1431711 · Genome biology · 2006 · 7 claims · 6 setups
The neocortical protomap is established as continuous rostro-caudal gradients of gene expression in progenitor cells, not discrete compartments
-
Full-text index only
Statistical learning of peptide retention behavior in chromatographic separations: a new kernel-based approach for computational proteomics.
PMID 18053132 · PMC2254445 · BMC bioinformatics · 2007 · 6 claims · 5 setups
The paired oligo-border kernel (POBK) combined with SVMs predicts peptide adsorption/elution in SAX-SPE and retention time in IP-RP-HPLC more accurately than existing methods.