Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Early feature extraction drives model performance in high-resolution chromatin accessibility prediction.
PMID 41526189 · PMC12951969 · Genome research · 2026 · 8 claims · 6 setups
Early feature extraction (via ConvNeXt V2 blocks), rather than downstream architecture type, is the primary determinant of prediction accuracy in high-resolution chromatin accessibility prediction.
-
Full-text index only
EpiXFormer: a cross-attention neural network for predicting cell type-specific transcription factor binding sites.
PMID 41527854 · PMC12796812 · Briefings in bioinformatics · 2026 · 8 claims · 8 setups
EpiXFormer achieves high accuracy (mean AUROC ~0.99) predicting binding sites of both TFs and non-sequence-specific DBPs across 199 DBP-cell type pairs
-
Full-text index only
CLAMP: predicting specific protein-mediated chromatin loops in diverse species with a chromatin accessibility language model.
PMID 41555433 · PMC12903630 · Genome biology · 2026 · 8 claims · 8 setups
CLAMP, a chromatin-accessibility language model, predicts protein-mediated chromatin loops across 10 species, 18 proteins, and 24 cell types with superior performance versus existing methods.
-
Full-text index only
Architectural and evolutionary features of TE-derived TSSs shape tissue-specific promoter activity in the human genome.
PMID 41620470 · PMC12963367 · Nature communications · 2026 · 8 claims · 8 setups
A three-step RAMPAGE-based pipeline can systematically identify TE-derived transcription start sites (TSSs) genome-wide, distinguishing them from autonomous TE transcription and background noise.
-
Full-text index only
A generic reference defined by consensus peaks for single-cell ATAC-seq data analysis.
PMID 41663439 · PMC12996591 · Nature communications · 2026 · 7 claims · 7 setups
Aggregating peaks from 624 high-quality bulk ATAC-seq datasets defines ~1.4 million observed consensus peaks (cPeaks) covering ~30% of the genome.
-
Has reproduction · 85
Ensembl 2013.
PMID 23203987 · PMC3531136 · Nucleic acids research · 2013 · 8 claims · 8 setups
Ensembl (http://www.ensembl.org) provides genome information for sequenced chordate genomes, currently supporting 70 species with a focus on human, mouse, zebrafish and rat.