Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Score Matching for Differential Abundance Testing of Compositional High-Throughput Sequencing Data.
PMID 41944570 · PMC13055433 · Statistics in medicine · 2026 · 8 claims · 3 setups
cosmoDA extends the a-b power interaction model by adding a linear covariate effect on the location vector, enabling differential abundance testing on compositional data with feature interactions.
-
Full-text index only
Geometry-aware graph attention networks to explain single-cell chromatin states and gene expression with SEAGALL.
PMID 42026624 · PMC13238118 · Genome biology · 2026 · 8 claims · 6 setups
SEAGALL combines a geometry-regularised autoencoder (GRAE) to embed cells and build a cell-cell graph with a graph attention network (GAT) classifier and GNNExplainer-based XAI to identify features driving cell type/phenotype.
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
AXOLOTL: an accurate method for detecting aberrant gene expression in rare diseases using coexpression constraints.
PMID 42083807 · PMC13198384 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 6 setups
AXOLOTL is a novel ensemble outlier detection method that incorporates coexpression constraints to detect aberrant gene expression events in RNA expression matrices.
-
Has reproduction · 72
Prediction of prognostic signatures in triple-negative breast cancer based on the differential expression analysis via NanoString nCounter immune panel.
PMID 33138797 · PMC7607642 · BMC cancer · 2020 · 8 claims · 8 setups
edgeR-based DEG selection is more appropriate for feature selection than Elastic Net when sample sizes are small.
-
Full-text index only
Boosting accuracy of automated classification of fluorescence microscope images for location proteomics.
PMID 15207009 · PMC449699 · BMC bioinformatics · 2004 · 8 claims · 8 setups
New classifiers (SVMs, ensembles) and new wavelet-derived (Gabor, Daubechies) features improve recognition of protein subcellular location patterns over the previous neural network approach
-
Full-text index only
Discovering cancer genes by integrating network and functional properties.
PMID 19765316 · PMC2758898 · BMC medical genomics · 2009 · 8 claims · 6 setups
Cancer genes have distinct PPI network topology (higher connectivity, higher clustering coefficient, shorter path length to known cancer genes) compared to non-cancer genes
-
Full-text index only
TSniffer: unbiased de novo identification of RNA editing sites and quantification of editing activity in RNA-seq data.
PMID 41549280 · PMC12838065 · Genome biology · 2026 · 8 claims · 6 setups
TSniffer is a novel tool that uses a rolling window Fisher's exact test approach to identify RNA editing sites (TsRegions) de novo in RNA-seq data without relying on editing databases or two-sample differential comparison.
-
Full-text index only
A novel deep learning-driven framework for improving lncRNA comprehensive annotation with LncADeep 2.0.
PMID 41923359 · PMC13090826 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 8 setups
LncADeep 2.0 outperforms LncADeep and other existing tools for lncRNA identification on both GENCODE annotated transcripts and independent RNA-seq data
-
Full-text index only
Ensembl's 10th year.
PMID 19906699 · PMC2808936 · Nucleic acids research · 2010 · 8 claims · 8 setups
Ensembl provides comprehensive gene annotation and integrated genomic resources (variation, regulation, comparative genomics) across a growing set of chordate genomes
-
Full-text index only
Anopheles gambiae genome reannotation through synthesis of ab initio and comparative gene prediction algorithms.
PMID 16569258 · PMC1557760 · Genome biology · 2006 · 8 claims · 7 setups
An exon-gene-union (EGU) algorithm followed by an open-reading-frame-selection algorithm can synthesize ab initio (GENSCAN, GeneMark, SNAP) and comparative (Ensembl/Genewise) predictions into a single, more complete CDS set
-
Full-text index only
Analysis of expressed sequence tags from Actinidia: applications of a cross species EST database for gene discovery in the areas of flavor, health, color and ripening.
PMID 18655731 · PMC2515324 · BMC genomics · 2008 · 7 claims · 6 setups
A collection of 132,577 ESTs from four Actinidia species was generated and clustered into 41,858 non-redundant clusters (18,070 TCs and 23,788 singletons)
-
Full-text index only
BaGPipe: an automated, reproducible, and flexible pipeline for bacterial genome-wide association studies.
PMID 41896736 · PMC13147680 · BMC microbiology · 2026 · 7 claims · 8 setups
BaGPipe is an automated, reproducible Nextflow pipeline that integrates pre-processing, Pyseer-based association analysis, and downstream visualisation into a unified bacterial GWAS workflow