Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
GeneTide--Terra Incognita Discovery Endeavor: a new transcriptome focused member of the GeneCards/GeneNote suite of databases.
PMID 15608261 · PMC540076 · Nucleic acids research · 2005 · 8 claims · 7 setups
GeneTide integrates UniGene, DoTS, AceView, BLAT/GeneLoc genomic alignment, and GeneAnnot probe-set data into a unified Consensus/Uniqueness/Score scheme to associate ESTs with GeneCards genes
-
Full-text index only
VIRGO: computational prediction of gene functions.
PMID 16845022 · PMC1538839 · Nucleic acids research · 2006 · 8 claims · 6 setups
VIRGO constructs a functional linkage network (FLN) from gene expression and molecular interaction data, labels genes with GO annotations, and propagates these labels to predict functions of unlabelled genes
-
Full-text index only
Functional nsSNPs from carcinogenesis-related genes expressed in breast tissue: potential breast cancer risk alleles and their distribution across human populations.
PMID 16595073 · PMC3500178 · Human genomics · 2006 · 7 claims · 5 setups
A bioinformatics strategy cross-referencing carcinogenesis-related gene lists with breast-tissue expression data can identify candidate breast cancer risk nsSNPs.
-
Full-text index only
InSite: a computational method for identifying protein-protein interaction binding sites on a proteome-wide scale.
PMID 17868464 · PMC2375030 · Genome biology · 2007 · 8 claims · 8 setups
InSite predicts protein-pair-specific binding motifs ('Motif M on protein A binds to protein B') by integrating heterogeneous PPI and motif-motif interaction evidence within a Bayesian network trained by EM
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Full-text index only
SNAP: predict effect of non-synonymous polymorphisms on function.
PMID 17526529 · PMC1920242 · Nucleic acids research · 2007 · 7 claims · 8 setups
SNAP, a neural network-based method using sequence-derived information, predicts whether a non-synonymous SNP is neutral or non-neutral for protein function
-
Full-text index only
Alkyltransferase-mediated toxicity of 1,3-butadiene diepoxide.
PMID 18712882 · PMC2614079 · Chemical research in toxicology · 2008 · 8 claims · 6 setups
hAGT expression increases mutations and cytotoxicity following BDO exposure, and hAGT-DNA cross-links form in the presence of BDO
-
Full-text index only
Canine tumor cross-species genomics uncovers targets linked to osteosarcoma progression.
PMID 20028558 · PMC2803201 · BMC genomics · 2009 · 8 claims · 7 setups
High expression of IL-8 and SLC1A3, identified via cross-species mining, is associated with poor outcome in an independent population of human osteosarcoma patients
-
Full-text index only
Proteomic analysis of integrin-associated complexes identifies RCC2 as a dual regulator of Rac1 and Arf6.
PMID 19738201 · PMC2857963 · Science signaling · 2009 · 8 claims · 8 setups
A novel ligand-affinity/cross-linking proteomic methodology enables isolation of labile integrin-associated signaling complexes
-
Full-text index only
NGSTroubleFinder: a tool for detection and quantification of contamination and kinship across human NGS data.
PMID 41608734 · PMC12838523 · NAR genomics and bioinformatics · 2026 · 8 claims · 8 setups
NGSTroubleFinder detects cross-sample contamination, sample swaps, kinship, and sex mismatches from BAM/CRAM files without requiring additional variant-calling steps
-
Full-text index only
Genomic instability and mono-parental expression mitigate genomic shock in a cross-subgenus Leishmania hybrid.
PMID 41918820 · PMC13034039 · NAR molecular medicine · 2026 · 7 claims · 8 setups
An in vitro cross between L. infantum and L. tarentolae produced a viable inter-subgenus hybrid, demonstrating genomic compatibility between highly divergent Leishmania species.
-
Full-text index only
EpiXFormer: a cross-attention neural network for predicting cell type-specific transcription factor binding sites.
PMID 41527854 · PMC12796812 · Briefings in bioinformatics · 2026 · 8 claims · 8 setups
EpiXFormer achieves high accuracy (mean AUROC ~0.99) predicting binding sites of both TFs and non-sequence-specific DBPs across 199 DBP-cell type pairs
-
Full-text index only
Comprehensive analysis of the causal risk factor from hypertension associated with prognosis and therapeutic response in renal cell carcinoma by multi-omics analysis and validation.
PMID 41680825 · PMC12998095 · Biology direct · 2026 · 8 claims · 8 setups
A 48-gene cross-species hypertension (HTN) gene module identified from human and SHR rat scRNA-seq can classify ccRCC patients into two molecular subgroups with distinct survival and targeted therapy response
-
Has reproduction · 74
SpaGene: A Deep Adversarial Framework for Spatial Gene Imputation.
PMID 42146899 · PMC13176606 · Computational and structural biotechnology journal · 2026 · 8 claims · 6 setups
SpaGene improves average PCC and SSIM and reduces RMSE compared to 6 baseline methods (SpaGE, gimVI, Tangram, VISTA, spRefine, stDiff) across 8 diverse ST-SC dataset pairs under gene-holdout evaluation.
-
Has reproduction · 95
Utility of Triti-Map for bulk-segregated mapping of causal genes and regulatory elements in Triticeae.
PMID 35605195 · PMC9284283 · Plant communications · 2022 · 8 claims · 4 setups
Triti-Map is a computational package suite plus web interface specifically optimized for bulk-segregated gene mapping in Triticeae, accepting DNA-seq, RNA-seq/ChIP-seq, and traditional QTL data as input
-
Has reproduction · 90
A Decentralized Kidney Transplant Biopsy Classifier for Transplant Rejection Developed Using Genes of the Banff-Human Organ Transplant Panel.
PMID 35619722 · PMC9128066 · Frontiers in immunology · 2022 · 6 claims · 6 setups
A random forest model trained solely on B-HOT panel genes (B-HOT Model) accurately classifies kidney transplant biopsies as NR, ABMR, or TCMR.
-
Full-text index only
Science review: searching for gene candidates in acute lung injury.
PMID 15566614 · PMC1065043 · Critical care (London, England) · 2004 · 8 claims · 8 setups
The candidate gene approach combined with an ortholog gene database and gene ontology analysis identifies ALI candidate genes, with blood coagulation and inflammation ontologies most highly represented
-
Full-text index only
Interaction profile-based protein classification of death domain.
PMID 15189571 · PMC459208 · BMC bioinformatics · 2004 · 7 claims · 6 setups
An SVM-based classifier using Residue Pair Interaction Profiles (RPIPs) can classify death domain superfamily members into subfamilies with 89% average cross-validation accuracy
-
Full-text index only
Columba: an integrated database of proteins, structures, and annotations.
PMID 15801979 · PMC1087474 · BMC bioinformatics · 2005 · 8 claims · 6 setups
COLUMBA physically integrates data from twelve protein structure-related databases (PDB, KEGG, Swiss-Prot, CATH, SCOP, Gene Ontology, ENZYME, etc.) into a single PostgreSQL data warehouse.
-
Full-text index only
Searching for interpretable rules for disease mutations: a simulated annealing bump hunting strategy.
PMID 16984653 · PMC1618409 · BMC bioinformatics · 2006 · 8 claims · 6 setups
The proposed feature set outperforms existing published feature sets for predicting effects of amino acid substitutions