Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
miRGator: an integrated system for functional annotation of microRNAs.
PMID 17942429 · PMC2238850 · Nucleic acids research · 2008 · 8 claims · 8 setups
miRGator integrates target prediction, functional enrichment analysis (GO/pathway/disease), and expression data (miRNA/mRNA/protein) into one system for functional annotation of miRNAs
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Full-text index only
Benchmarking ortholog identification methods using functional genomics data.
PMID 16613613 · PMC1557999 · Genome biology · 2006 · 8 claims · 7 setups
InParanoid is the best overall ortholog identification method for identifying functionally equivalent proteins when sensitivity and selectivity are combined into an overall score.
-
Full-text index only
Human disease classification in the postgenomic era: a complex systems approach to human pathobiology.
PMID 17625512 · PMC1948102 · Molecular systems biology · 2007 · 8 claims · 5 setups
Current syndromic disease classification lacks specificity despite historically serving clinicians well
-
Full-text index only
Aggregation propensity of the human proteome.
PMID 18927604 · PMC2557143 · PLoS computational biology · 2008 · 8 claims · 7 setups
Long proteins have, on average, less intense/pronounced aggregation peaks than short proteins
-
Has reproduction · 84
Improving recombinant protein production by yeast through genome-scale modeling using proteome constraints.
PMID 35624178 · PMC9142503 · Nature communications · 2022 · 7 claims · 5 setups
pcSecYeast, a proteome-constrained genome-scale model integrating metabolism, translation, and detailed secretory pathway processing (translocation, PTMs, folding, misfolding, degradation), was constructed for S. cerevisiae
-
Full-text index only
Proteomics in alcohol research.
PMID 12875051 · PMC6683837 · Alcohol research & health : the journal of the National Institute on Alcohol Abuse and Alcoholism · 2002 · 7 claims · 8 setups
The proteome is larger and more complex than the genome due to differential splicing, post-translational modifications (PTMs), and protein-protein interactions.
-
Full-text index only
Protein length in eukaryotic and prokaryotic proteomes.
PMID 15951512 · PMC1150220 · Nucleic acids research · 2005 · 7 claims · 5 setups
Eukaryotic proteins are significantly longer than prokaryotic proteins across virtually all functional categories and the majority of protein families
-
Full-text index only
How to find soluble proteins: a comprehensive analysis of alpha/beta hydrolases for recombinant expression in E. coli.
PMID 15804363 · PMC1079826 · BMC genomics · 2005 · 7 claims · 7 setups
Predicted solubility in E. coli (via CV-CV') depends on hydrolase size, phylogenetic origin, homologous family, and superfamily
-
Full-text index only
Computational verification of protein-protein interactions by orthologous co-expression.
PMID 15740634 · PMC555590 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Co-expression of orthologous protein pairs across multiple species can verify/predict S. cerevisiae PPIs with better performance than S. cerevisiae co-expression alone.
-
Full-text index only
Identification of the proliferation/differentiation switch in the cellular network of multicellular organisms.
PMID 17166053 · PMC1664705 · PLoS computational biology · 2006 · 8 claims · 8 setups
Integrating interactome and transcriptome data reveals a pair of transcriptionally anticorrelated network modules (P and D) each comprising hundreds of genes, present across individuals and species.
-
Full-text index only
A protein interaction based model for schizophrenia study.
PMID 19091023 · PMC2638163 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Products of 36 schizophrenia candidate genes cluster together into a single connected component within a PPI sub-network of 831 proteins
-
Full-text index only
A map of human protein interactions derived from co-expression of human mRNAs and their orthologs.
PMID 18414481 · PMC2387231 · Molecular systems biology · 2008 · 8 claims · 6 setups
Comparing human mRNA co-expression with co-expression of orthologous gene pairs in five other organisms identifies proteins that physically associate
-
Full-text index only
A global view of protein expression in human cells, tissues, and organs.
PMID 20029370 · PMC2824494 · Molecular systems biology · 2009 · 7 claims · 6 setups
A high fraction (>65%) of proteins is expressed in most human cells and tissues, while very few proteins (<2%) are detected in any single cell type.
-
Full-text index only
Phenotypic categorization of genetic skin diseases reveals new relations between phenotypes, genes and pathways.
PMID 19744994 · PMC2773259 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 5 setups
560 genetic skin diseases can be decomposed into 71 elementary phenotypic features (42 dermatologic, 29 systemic) that combine to represent each disease as a point in a multidimensional phenotype space
-
Full-text index only
A rigorous method for multigenic families' functional annotation: the peptidyl arginine deiminase (PADs) proteins family example.
PMID 16271148 · PMC1310624 · BMC genomics · 2005 · 8 claims · 5 setups
Integrating EST-based expression data with phylogenetic analysis is a valid new method for functionally annotating multigenic protein families
-
Has reproduction · 83
Integrative transcriptomic and machine learning framework reveals candidate genes and potential mechanisms of aflatoxin B1 exposure in breast cancer.
PMID 41688730 · PMC12982753 · Scientific reports · 2026 · 7 claims · 8 setups
Twenty-two genes lie at the intersection of AFB1-predicted targets and breast cancer-associated co-expression modules/DEGs
-
Full-text index only
High resolution analysis of the human transcriptome: detection of extensive alternative splicing independent of transcriptional activity.
PMID 19804644 · PMC2768739 · BMC genetics · 2009 · 8 claims · 6 setups
The human GWSA uses exon body and exon-exon junction probes to directly measure over 280,000 known and predicted splicing events genome-wide.
-
Full-text index only
An SVM-based system for predicting protein subnuclear localizations.
PMID 16336650 · PMC1325059 · BMC bioinformatics · 2005 · 7 claims · 3 setups
New kernels defined on k-peptide vectors mapped by BLOSUM62-based high-scored pair matrices (D1, D2, D3) improve SVM discrimination of protein subnuclear localization compared to conventional k-peptide encodings.
-
Full-text index only
Rho GTPases in human breast tumours: expression and mutation analyses and correlation with clinical parameters.
PMID 12237774 · PMC2364248 · British journal of cancer · 2002 · 8 claims · 8 setups
RhoA, RhoB, Rac1 and Cdc42 protein levels are markedly overexpressed in breast tumours compared to matched normal tissue from the same patient