Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Evolutionary trace annotation of protein function in the structural proteome.
PMID 20036248 · PMC2831211 · Journal of molecular biology · 2010 · 8 claims · 7 setups
ET-ranked residue clusters can be used to build 3D templates that predict GO function in enzymes and non-enzymes alike, without prior knowledge of functional mechanism.
-
Has reproduction · 62
Predicting Bone Metastasis Using Gene Expression-Based Machine Learning Models.
PMID 34858485 · PMC8631472 · Frontiers in genetics · 2021 · 8 claims · 5 setups
A DNN model built on 34 top-ranked hub genes achieves the highest prediction accuracy (AUC 92.11%) for distinguishing primary from bone-metastasized tumors
-
Full-text index only
A modified T-test feature selection method and its application on the HapMap genotype data.
PMID 18267305 · PMC5054219 · Genomics, proteomics & bioinformatics · 2007 · 7 claims · 4 setups
A modified t-test ranking measure, extended to handle nominal SNP genotype data via vector transformation, can effectively rank SNPs by their discriminative capability for population classification.
-
Full-text index only
Prioritization of candidate cancer genes--an aid to oncogenomic studies.
PMID 18710882 · PMC2566894 · Nucleic acids research · 2008 · 8 claims · 8 setups
Computational classifiers using combinations of protein conservation, gene structure, protein domains, protein interactions, and regulatory data can distinguish known cancer genes (CD/CR) from unlabelled human genes
-
Full-text index only
BLASTO: a tool for searching orthologous groups.
PMID 17483516 · PMC1933156 · Nucleic acids research · 2007 · 7 claims · 2 setups
BLASTO treats each orthologous group as a unit and outputs a ranked list of orthologous groups instead of single sequences
-
Has reproduction · 82
eVITTA: a web-based visualization and inference toolbox for transcriptome analysis.
PMID 34019643 · PMC8218201 · Nucleic acids research · 2021 · 8 claims · 7 setups
eVITTA is a web-based toolbox with three integrated modules (easyGEO, easyGSEA, easyVizR) for GEO data access, functional profiling, and multi-dataset comparison
-
Full-text index only
Function2Gene: a gene selection tool to increase the power of genetic association studies by utilizing public databases and expert knowledge.
PMID 18631403 · PMC2500032 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Function2Gene is a set of Perl programs that queries public databases (NCBI, GeneCards, Harvester, with Uniprot/Ensembl also supported) using expert-selected keywords to rank genes by prior probability of disease association.
-
Has reproduction · 71
Chemical genomics informs antibiotic and essential gene function in Acinetobacter baumannii.
PMID 40153700 · PMC11975115 · PLoS genetics · 2025 · 8 claims · 6 setups
The vast majority of A. baumannii essential genes show significant chemical-gene interactions upon knockdown
-
Has reproduction · 76
PI3K inhibitors protect against glucocorticoid-induced skin atrophy.
PMID 30737086 · PMC6441871 · EBioMedicine · 2019 · 8 claims · 8 setups
Drug repurposing screen of the LINCS transcriptional signature database identified PI3K/mTOR/Akt inhibitors as the most prominent pharmacological class capable of repressing REDD1/FKBP51 expression.
-
Has reproduction · 89
Graph Random Forest: A Graph Embedded Algorithm for Identifying Highly Connected Important Features.
PMID 37509188 · PMC10377046 · Biomolecules · 2023 · 8 claims · 6 setups
GRF identifies effective features that form highly connected sub-graphs on the underlying biological network
-
Has reproduction · 71
Newborn sex-specific transcriptome signatures and gestational exposure to fine particles: findings from the ENVIRONAGE birth cohort.
PMID 28583124 · PMC5458481 · Environmental health : a global access science source · 2017 · 8 claims · 6 setups
Gestational PM2.5 exposure is associated with sex-specific transcriptome signatures in cord blood, with a significant sex-by-exposure interaction for many genes.
-
Has reproduction · 44
Dynamic Gene Attention Focus (DyGAF): Enhancing Biomarker Identification Through Dual-Model Attention Networks.
PMID 40160891 · PMC11951896 · Bioinformatics and biology insights · 2025 · 6 claims · 5 setups
DyGAF, a dual-model attention neural network (independent Model A + dependent Model B), identifies and ranks genes by significance for COVID-19 biomarker discovery more effectively than differential expression analysis (DEA) and random forest (RF) feature selection
-
Has reproduction · 96
Deep learning based protocol to construct an immune-related gene network of host-pathogen interactions in plants.
PMID 36525344 · PMC9791427 · STAR protocols · 2023 · 7 claims · 6 setups
DLNet algorithm ranks genes based on their contribution to classifying treatment vs. control expression data
-
Has reproduction
D2H2: diabetes data and hypothesis hub.
PMID 38107655 · PMC10723036 · Bioinformatics advances · 2023 · 6 claims · 6 setups
D2H2 is a web-based portal integrating hundreds of curated diabetes-relevant transcriptomics datasets with bioinformatics tools for gene/gene set queries
-
Has reproduction · 32
Integrated analysis of post-transcriptional regulations reveals insights into acute myeloid leukemia.
PMID 41407883 · PMC12712020 · Communications biology · 2025 · 8 claims · 6 setups
PTRs computed from integrated transcriptomic/proteomic data of 44 AML samples are highly conserved across AML subtypes and with 29 other human tissues, indicating broadly conserved post-transcriptional mechanisms.
-
Full-text index only
Boosting accuracy of automated classification of fluorescence microscope images for location proteomics.
PMID 15207009 · PMC449699 · BMC bioinformatics · 2004 · 8 claims · 8 setups
New classifiers (SVMs, ensembles) and new wavelet-derived (Gabor, Daubechies) features improve recognition of protein subcellular location patterns over the previous neural network approach
-
Full-text index only
InSite: a computational method for identifying protein-protein interaction binding sites on a proteome-wide scale.
PMID 17868464 · PMC2375030 · Genome biology · 2007 · 8 claims · 8 setups
InSite predicts protein-pair-specific binding motifs ('Motif M on protein A binds to protein B') by integrating heterogeneous PPI and motif-motif interaction evidence within a Bayesian network trained by EM
-
Full-text index only
Gene-disease relationship discovery based on model-driven data integration and database view definition.
PMID 19042916 · PMC2639000 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 4 setups
Explicit gene–disease relationships can be formulated as candidate gene definitions (e.g., co-localization, dysregulation, functional similarity) that may include intermediary orthologous or interacting genes
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
Babelomics: advanced functional profiling of transcriptomics, proteomics and genomics experiments.
PMID 18515841 · PMC2447758 · Nucleic acids research · 2008 · 8 claims · 5 setups
Babelomics is a web suite offering both conventional functional enrichment methods and more advanced gene set analysis (GSA) methods, a combination offered by only one other tool (FuncAssociate) among competitors.