Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins.
PMID 15608248 · PMC539979 · Nucleic acids research · 2005 · 7 claims · 5 setups
RefSeq provides a curated, non-redundant, explicitly linked collection of genomic, transcript and protein sequences spanning prokaryotes, eukaryotes and viruses.
-
Full-text index only
Proteomic approaches to cancer biomarkers.
PMID 19931265 · PMC2873613 · Gastroenterology · 2010 · 8 claims · 8 setups
Combining abundant-protein depletion, offline fractionation, and subproteome (e.g., glycoproteome) enrichment with 2D LC-MS/MS increases the dynamic range and depth of blood proteome analysis for biomarker discovery.
-
Full-text index only
Interaction profile-based protein classification of death domain.
PMID 15189571 · PMC459208 · BMC bioinformatics · 2004 · 7 claims · 6 setups
An SVM-based classifier using Residue Pair Interaction Profiles (RPIPs) can classify death domain superfamily members into subfamilies with 89% average cross-validation accuracy
-
Full-text index only
Network-assisted protein identification and data interpretation in shotgun proteomics.
PMID 19690572 · PMC2736651 · Molecular systems biology · 2009 · 7 claims · 7 setups
Confidently identified proteins in a sample form tightly connected sub-networks in the protein interaction network, with significantly higher clustering coefficients than random or topology-matched random sub-networks.
-
Full-text index only
EpiToolKit--a web server for computational immunomics.
PMID 18440979 · PMC2447732 · Nucleic acids research · 2008 · 7 claims · 3 setups
EpiToolKit is a web server integrating five MHC class I and two MHC class II epitope prediction methods in a unified, user-friendly interface.
-
Full-text index only
SNAP: predict effect of non-synonymous polymorphisms on function.
PMID 17526529 · PMC1920242 · Nucleic acids research · 2007 · 7 claims · 8 setups
SNAP, a neural network-based method using sequence-derived information, predicts whether a non-synonymous SNP is neutral or non-neutral for protein function
-
Full-text index only
A dynamic range compression and three-dimensional peptide fractionation analysis platform expands proteome coverage and the diagnostic potential of whole saliva.
PMID 19813771 · PMC2789208 · Journal of proteome research · 2009 · 7 claims · 7 setups
Coupling DRC (hexapeptide libraries) with 3D peptide fractionation (IEF + SCX + µLC-MS/MS) substantially increases the number of proteins identified in whole saliva
-
Full-text index only
Synaptic proteins linked to HIV-1 infection and immunoproteasome induction: proteomic analysis of human synaptosomes.
PMID 19693676 · PMC2824116 · Journal of neuroimmune pharmacology : the official journal of the Society on NeuroImmune Pharmacology · 2010 · 8 claims · 7 setups
Proteomic screening of human synaptosomes identifies a set of proteins differentially expressed in HIV/AIDS brain
-
Has reproduction · 83
Integrative transcriptomic and machine learning framework reveals candidate genes and potential mechanisms of aflatoxin B1 exposure in breast cancer.
PMID 41688730 · PMC12982753 · Scientific reports · 2026 · 7 claims · 8 setups
Twenty-two genes lie at the intersection of AFB1-predicted targets and breast cancer-associated co-expression modules/DEGs
-
Has reproduction · 92
An integrative proteomics method identifies a regulator of translation during stem cell maintenance and differentiation.
PMID 34772928 · PMC8590018 · Nature communications · 2021 · 8 claims · 5 setups
PISA-Express is a method that simultaneously measures protein expression and thermal stability (solubility) changes using only two samples per replicate per cell type
-
Has reproduction · 63
Target identification for repurposed drugs active against SARS-CoV-2 via high-throughput inverse docking.
PMID 34825285 · PMC8616721 · Journal of computer-aided molecular design · 2022 · 8 claims · 6 setups
Combining Vinardo, Ledock, and Korp-PL scoring functions (via averaged Z-scores) improves correct target identification over any single scoring function.
-
Full-text index only
A global proteomics approach identifies novel phosphorylated signaling proteins in GPVI-activated platelets: involvement of G6f, a novel platelet Grb2-binding membrane adapter.
PMID 16941570 · PMC1869047 · Proteomics · 2006 · 8 claims · 7 setups
96 proteins undergo post-translational modification (phosphorylation) in response to CRP stimulation of human platelets, including 11 proteins not previously identified in platelets
-
Full-text index only
Comparative cytochrome P450 proteomics in the livers of immunodeficient mice using 18O stable isotope labeling.
PMID 17296599 · PMC2315784 · Molecular & cellular proteomics : MCP · 2007 · 8 claims · 5 setups
SDS-PAGE combined with post-digest 18O/16O labeling and LC-MS/MS enables relative quantification of multiple P450 proteins from liver microsomes, including highly homologous isoforms
-
Full-text index only
Identification of candidate disease genes by integrating Gene Ontologies and protein-interaction networks: case study of primary immunodeficiencies.
PMID 19073697 · PMC2632920 · Nucleic acids research · 2009 · 8 claims · 5 setups
Combining high protein-interaction network scores with significant PID-related GO terms identifies novel PID candidate genes
-
Has reproduction · 100
Gene signature discovery and systematic validation across diverse clinical cohorts for TB prognosis and response to treatment.
PMID 37471455 · PMC10393163 · PLoS computational biology · 2023 · 8 claims · 8 setups
A network-based meta-analysis across studies identifies a common 45-gene signature specific to active TB disease that accounts for cohort/population heterogeneity
-
Has reproduction · 58
Genome-wide identification and characterization of germin-like protein family in Brassica juncea reveals their role against biotic stress.
PMID 41327045 · PMC12763953 · BMC plant biology · 2025 · 8 claims · 8 setups
102 GLPs were identified in B. juncea, 51 in B. nigra, and 48 in B. rapa via genome-wide in-silico analysis
-
Full-text index only
ORFer--retrieval of protein sequences and open reading frames from GenBank and storage into relational databases or text files.
PMID 12493080 · PMC139979 · BMC bioinformatics · 2002 · 6 claims · 6 setups
ORFer retrieves protein and nucleic acid sequences and annotations from NCBI GenBank using the XML sequence format
-
Full-text index only
Molecular phylogeny of the kelch-repeat superfamily reveals an expansion of BTB/kelch proteins in animals.
PMID 13678422 · PMC222960 · BMC bioinformatics · 2003 · 8 claims · 8 setups
The human genome encodes at least 71 kelch-repeat proteins
-
Full-text index only
How to find soluble proteins: a comprehensive analysis of alpha/beta hydrolases for recombinant expression in E. coli.
PMID 15804363 · PMC1079826 · BMC genomics · 2005 · 7 claims · 7 setups
Predicted solubility in E. coli (via CV-CV') depends on hydrolase size, phylogenetic origin, homologous family, and superfamily
-
Full-text index only
Searching for interpretable rules for disease mutations: a simulated annealing bump hunting strategy.
PMID 16984653 · PMC1618409 · BMC bioinformatics · 2006 · 8 claims · 6 setups
The proposed feature set outperforms existing published feature sets for predicting effects of amino acid substitutions