Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 62
Specific Marker Gene Analysis for Primary Central Nervous System Lymphoma Based on Methylation Difference and Development of Detection Primers.
PMID 41097902 · PMC12528802 · Brain and behavior · 2025 · 7 claims · 5 setups
450K microarray analysis identified 26 overlapping differential methylation sites distinguishing PCNSL from CNS and non-CNS groups
-
Has reproduction · 59
Advances in genomic and pharmacokinetic profiling for clinical stratification of metastatic breast cancer.
PMID 41369820 · PMC12799884 · Discover oncology · 2025 · 8 claims · 8 setups
Key genes AR, AKT1, UBC, CDH1, SMAD3, ROR1, and ROR2 are associated with chemotherapy resistance and poor prognosis in metastatic breast cancer
-
Has reproduction · 78
Machine learning and free energy clustering reveal PAH protein binding linked to AD risk.
PMID 41953002 · PMC13053772 · iScience · 2026 · 8 claims · 8 setups
PARP1, PTPN1, and ITGA4 are core PAH protein targets identified via PPI network analysis and XGBoost feature selection from AD-associated DEGs
-
Has reproduction · 78
Identification of succinylation-related genes in bladder cancer: integration of single-cell and transcriptomic data.
PMID 42220482 · PMC13218921 · Frontiers in immunology · 2026 · 8 claims · 8 setups
KCTD16, CD3D, and GSDMB were identified as prognostic succinylation-related genes in BLCA
-
Has reproduction · 95
nf-rnaSeqCount: A Nextflow pipeline for obtaining raw read counts from RNA-seq data.
PMID 35574063 · PMC9097006 · South African computer journal = Suid-Afrikaanse rekenaartydskrif · 2021 · 7 claims · 5 setups
nf-rnaSeqCount is a portable, reproducible Nextflow pipeline that maps RNA-seq reads to a reference genome and quantifies gene abundance for differential expression analysis
-
Has reproduction · 89
Identification of potential therapeutic targets for nonischemic cardiomyopathy in European ancestry: an integrated multiomics analysis.
PMID 39267096 · PMC11396958 · Cardiovascular diabetology · 2024 · 8 claims · 7 setups
Two-sample MR analysis identified 255 circulating plasma proteins associated with NISCM
-
Has reproduction · 40
Machine learning developed an intratumor heterogeneity signature for predicting clinical outcome and immunotherapy benefit in bladder cancer.
PMID 39100839 · PMC11291408 · Translational andrology and urology · 2024 · 8 claims · 8 setups
An integrative machine learning procedure (10 methods, 101 algorithm combinations) identified an Enet (alpha=0.2)-based intratumor heterogeneity-related signature (IRS) with the highest average C-index (0.69) across TCGA and GEO cohorts
-
Has reproduction · 53
Combining evidence of preferential gene-tissue relationships from multiple sources.
PMID 23950964 · PMC3741196 · PloS one · 2013 · 8 claims · 8 setups
A high-level integration approach combining three methods across four human microarray datasets, merged by consensus voting and a rule-based inner/total score, predicts preferentially expressed genes while reducing method- and study-specific bias.
-
Has reproduction
Using random walks to identify cancer-associated modules in expression data.
PMID 24128261 · PMC4015830 · BioData mining · 2013 · 8 claims · 8 setups
Walktrap-GM, a random-walk community detection algorithm adapted with stopping criteria (maximum modularity, maximum size, maximum module score), identifies modules significantly enriched with cancer genes in expression-weighted interaction networks.
-
Full-text index only
Discovering multiple transcripts of human hepatocytes using massively parallel signature sequencing (MPSS).
PMID 17601345 · PMC1929076 · BMC genomics · 2007 · 8 claims · 8 setups
MPSS detected 10,279 UniGene clusters, representing 7,475 known genes, in human hepatocytes
-
Full-text index only
g:Profiler--a web-based toolset for functional profiling of gene lists from large-scale experiments.
PMID 17478515 · PMC1933153 · Nucleic acids research · 2007 · 8 claims · 5 setups
g:Profiler integrates four modules (g:Profiler core, g:Convert, g:Orth, g:Sorter) into a single cross-linked web tool for gene list analysis
-
Full-text index only
DiRE: identifying distant regulatory elements of co-expressed genes.
PMID 18487623 · PMC2447744 · Nucleic acids research · 2008 · 8 claims · 4 setups
DiRE predicts distant regulatory elements by combining gene co-expression data, comparative genomics and TFBS profiles to determine TFBS-association signatures
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
Molecular interactions between HNF4a, FOXA2 and GABP identified at regulatory DNA elements through ChIP-sequencing.
PMID 19822575 · PMC2794179 · Nucleic acids research · 2009 · 8 claims · 6 setups
ChIP-seq identified 3064 GABP peaks, 7266 FOXA2 peaks and 18783 HNF4a peaks in HepG2 cells
-
Full-text index only
DNA methylation biomarkers-based pan-cancer classifier: predictive modeling for cancer classification.
PMID 42152108 · PMC13185202 · Genome medicine · 2026 · 8 claims · 5 setups
Relatively simple ML models (logistic regression) outperform complex algorithms such as deep neural networks for methylation-based cancer classification
-
Full-text index only
TF2TG: an online resource mining the potential gene targets of transcription factors in Drosophila.
PMID 40314147 · PMC12774851 · Genetics · 2026 · 8 claims · 8 setups
TF2TG is an online resource integrating motif scan data, ChIP-seq peaks (modENCODE/modERN), Hi-C (TADs), REDfly-curated CRMs, ATAC-seq, protein-protein interaction data, and tissue-specific expression to predict TF-target gene relationships in Drosophila
-
Full-text index only
TEDD: a comprehensive database for translation efficiency dynamics.
PMID 41217970 · PMC12807600 · Nucleic acids research · 2026 · 8 claims · 4 setups
TEDD integrates 1518 RNA-seq, Ribo-seq, and RNC-seq samples from 143 human projects (279 datasets) spanning 24 tissues/cell types, 74 cell lines, and 52 conditions.
-
Full-text index only
CellPredX, a computational framework for cross-data type, cross-sample, and cross-protocol cell type annotation through domain adaptation and deep metric learning.
PMID 41481570 · PMC12758788 · PLoS computational biology · 2026 · 8 claims · 7 setups
CellPredX is a unified semi-supervised framework integrating domain adaptation and deep metric learning to align heterogeneous embeddings for cross-modality cell type annotation.
-
Full-text index only
Leveraging the germ layer development patterns to predict prognosis and identify MEST as a novel therapeutic target in glioma.
PMID 41501725 · PMC12870398 · Cancer cell international · 2026 · 7 claims · 8 setups
MEST is a key oncogenic GLD-related gene and a novel therapeutic target in glioma, identified via a machine learning feature selection framework
-
Full-text index only
A latent activated olfactory stem cell state revealed by single-cell transcriptomic and epigenomic profiling.
PMID 41512864 · PMC12903091 · Stem cell reports · 2026 · 7 claims · 8 setups
HBC-derived regeneration proceeds via three distinct lineages (rHBC, Sus, mOSN) marked by sequential, lineage-specific transcription factor (TF) expression cascades.