Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
High accuracy mass spectrometry analysis as a tool to verify and improve gene annotation using Mycobacterium tuberculosis as an example.
PMID 18597682 · PMC2483986 · BMC genomics · 2008 · 8 claims · 5 setups
High-accuracy MS proteomics (LTQ-Orbitrap) can be used to verify and improve gene annotation by identifying peptides specific to one of two competing annotation datasets.
-
Has reproduction · 78
Machine learning and free energy clustering reveal PAH protein binding linked to AD risk.
PMID 41953002 · PMC13053772 · iScience · 2026 · 7 claims · 8 setups
An integrated framework of bioinformatics, machine learning, and ΔG clustering can prioritize PAHs for AD-associated neurotoxicity.
-
Full-text index only
Inference of transcriptional regulation using gene expression data from the bovine and human genomes.
PMID 17683551 · PMC1978505 · BMC genomics · 2007 · 7 claims · 8 setups
Using human reference promoter sequences is a useful approach for studying gene expression regulation in species with limited or non-existing genomic sequence, such as cattle.
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
How to find soluble proteins: a comprehensive analysis of alpha/beta hydrolases for recombinant expression in E. coli.
PMID 15804363 · PMC1079826 · BMC genomics · 2005 · 7 claims · 7 setups
Predicted solubility in E. coli (via CV-CV') depends on hydrolase size, phylogenetic origin, homologous family, and superfamily
-
Full-text index only
TRED: a Transcriptional Regulatory Element Database and a platform for in silico gene regulation studies.
PMID 15608156 · PMC539958 · Nucleic acids research · 2005 · 8 claims · 5 setups
TRED is a database collecting both cis-regulatory elements (promoters) and trans-regulatory elements (transcription factor binding/regulation data) with linked access.
-
Full-text index only
A survey of integral alpha-helical membrane proteins.
PMID 19760129 · PMC2780624 · Journal of structural and functional genomics · 2009 · 8 claims · 8 setups
An automated annotation pipeline defines the integral membrane genome and family associations for 21,379 proteins from 34 genomes, most belonging to 598 Pfam-derived membrane protein families.
-
Full-text index only
TRED: a transcriptional regulatory element database, new entries and other development.
PMID 17202159 · PMC1899102 · Nucleic acids research · 2007 · 8 claims · 3 setups
TRED collects mammalian cis- and trans-regulatory elements together with experimental evidence, mapped onto assembled genomes
-
Full-text index only
SNAP predicts effect of mutations on protein function.
PMID 18757876 · PMC2562009 · Bioinformatics (Oxford, England) · 2008 · 8 claims · 3 setups
SNAP is a publicly available web-server implementation predicting functional effects (neutral/non-neutral) of single amino acid substitutions.
-
Full-text index only
In silico and in vivo splicing analysis of MLH1 and MSH2 missense mutations shows exon- and tissue-specific effects.
PMID 16995940 · PMC1590028 · BMC genomics · 2006 · 8 claims · 6 setups
In silico ESE-prediction algorithms (ESEfinder, RescueESE, PESX) do not reliably predict actual in vivo splicing behavior of missense mutations
-
Full-text index only
MBGD update 2010: toward a comprehensive resource for exploring microbial genome diversity.
PMID 19906735 · PMC2808943 · Nucleic acids research · 2010 · 8 claims · 6 setups
MBGD allows users to create ortholog groups using a specified subgroup of organisms, distinguishing it from other comparative genomics resources
-
Full-text index only
Pegasys: software for executing and integrating analyses of biological sequences.
PMID 15096276 · PMC406494 · BMC bioinformatics · 2004 · 8 claims · 7 setups
Pegasys is a flexible, modular, customizable software system for executing and integrating heterogeneous biological sequence analysis tools
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Full-text index only
CRSD: a comprehensive web server for composite regulatory signature discovery.
PMID 16845073 · PMC1538777 · Nucleic acids research · 2006 · 7 claims · 5 setups
CRSD is a comprehensive web server integrating six large-scale databases (UniGene, mature microRNAs, putative promoter, TRANSFAC, pathway, GO) plus two newly constructed genome-wide databases (MRS and TRS) for composite regulatory signature discovery
-
Full-text index only
PA-GOSUB: a searchable database of model organism protein sequences with their predicted Gene Ontology molecular function and subcellular localization.
PMID 15608166 · PMC540074 · Nucleic acids research · 2005 · 7 claims · 4 setups
PA-GOSUB significantly extends the coverage of GO molecular function and subcellular localization annotations for 10 model organism proteomes compared with existing databases (GOA, Swiss-Prot).
-
Full-text index only
Widespread A-to-I RNA editing of Alu-containing mRNAs in the human transcriptome.
PMID 15534692 · PMC526178 · PLoS biology · 2004 · 8 claims · 6 setups
Intramolecular pairs of oppositely oriented Alu elements within the same pre-mRNA form dsRNA foldback structures that are major substrates for A-to-I RNA editing
-
Full-text index only
SECIS elements in the coding regions of selenoprotein transcripts are functional in higher eukaryotes.
PMID 17169995 · PMC1802603 · Nucleic acids research · 2007 · 8 claims · 5 setups
SECIS elements located within coding regions of selenoprotein mRNAs support functional Sec insertion in mammalian cells
-
Full-text index only
QTL MatchMaker: a multi-species quantitative trait loci (QTL) database and query system for annotation of genes and QTL.
PMID 16381937 · PMC1347390 · Nucleic acids research · 2006 · 8 claims · 5 setups
QTL MatchMaker integrates QTL information with physical, genetic and cytogenetic maps across human, mouse and rat genomes
-
Full-text index only
Using structural bioinformatics to investigate the impact of non synonymous SNPs and disease mutations: scope and limitations.
PMID 19758473 · PMC2745591 · BMC bioinformatics · 2009 · 8 claims · 8 setups
None of 39 tested structural properties can be used as a sole classification criterion to separate neutral SNPs from disease mutations.