Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 100
Betacoronavirus-specific alternate splicing.
PMID 35074468 · PMC8782732 · Genomics · 2022 · 8 claims · 8 setups
Genes showing differential alternative splicing in SARS-CoV-2 have a similar functional profile to those in SARS-CoV and MERS, affecting a diverse set of genes and biological functions related to virus biology.
-
Full-text index only
Improvements to cardiovascular gene ontology.
PMID 19046747 · PMC2706316 · Atherosclerosis · 2009 · 8 claims · 8 setups
Gene Ontology (GO) provides a controlled vocabulary that links current functional knowledge of genes to high-throughput genomic and proteomic datasets, aiding data interpretation.
-
Has reproduction · 51
SGCP: a spectral self-learning method for clustering genes in co-expression networks.
PMID 38956463 · PMC11221046 · BMC bioinformatics · 2024 · 7 claims · 4 setups
SGCP, a spectral self-learning method, yields gene co-expression modules with higher GO enrichment than WGCNA, CoExpNets, and CEMiTool across 12 real gene expression datasets.
-
Full-text index only
L2L: a simple tool for discovering the hidden significance in microarray expression data.
PMID 16168088 · PMC1242216 · Genome biology · 2005 · 8 claims · 4 setups
L2L systematically compares a user's differentially expressed gene list against a database of published differentially expressed gene lists to find statistically significant overlaps and generate hypotheses about shared mechanisms
-
Full-text index only
Functional annotation and identification of candidate disease genes by computational analysis of normal tissue gene expression data.
PMID 18560577 · PMC2409962 · PloS one · 2008 · 7 claims · 5 setups
Ranked Coexpression Groups (RCG) built from k=6 nearest coexpressed genes, combined with a majority-rule functional characterization, integrate multiple datasets/coexpression measures to generate high-confidence functional annotation predictions
-
Has reproduction · 50
Genetic parallels in biomineralization of the calcareous sponge Sycon ciliatum and stony corals.
PMID 40922549 · PMC12419799 · eLife · 2025 · 8 claims · 8 setups
829 genes are overexpressed in the oscular region of increased calcite spicule formation in S. ciliatum
-
Has reproduction · 86
Molecular Classification Models for Triple Negative Breast Cancer Subtype Using Machine Learning.
PMID 34575658 · PMC8472680 · Journal of personalized medicine · 2021 · 6 claims · 4 setups
A training gene set of 719 unique upregulated DEGs (subtype-specific) can be used to build ML models that classify TNBC into BLIA, BLIS, MES, and LAR subtypes.
-
Full-text index only
Independent component analysis reveals new and biologically significant structures in micro array data.
PMID 16762055 · PMC1557674 · BMC bioinformatics · 2006 · 7 claims · 8 setups
ICA applied to three microarray datasets reveals many biologically significant components, including low-ranking ones not obvious by rank alone
-
Full-text index only
Identifying alternative hyper-splicing signatures in MG-thymoma by exon arrays.
PMID 18545673 · PMC2409220 · PloS one · 2008 · 8 claims · 6 setups
An integrative ad-hoc functional GO analysis combining threshold-based (Fisher exact/hypergeometric) and threshold-free (Kolmogorov-Smirnov) statistics, plus term-to-parent comparisons, detects disease-relevant splicing events from exon array data.
-
Has reproduction · 57
Diapause vs. reproductive programs: transcriptional phenotypes in a keystone copepod.
PMID 33782539 · PMC8007741 · Communications biology · 2021 · 8 claims · 7 setups
t-SNE clustering of all-gene expression data groups field-collected (diapause program) samples into one cluster while early and late culture (reproductive program) samples separate into two distinct phenotypes
-
Full-text index only
Inferring combinatorial regulation of transcription in silico.
PMID 15647509 · PMC546154 · Nucleic acids research · 2005 · 8 claims · 5 setups
Combining Cluster-Buster (TFBS cluster prediction) with GOSSIP (rigorous GO enrichment statistics with multiple-testing/FDR correction) predicts biological functions controlled by combinatorial transcription factor action, without prior knowledge of factor targets
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
Gene-disease relationship discovery based on model-driven data integration and database view definition.
PMID 19042916 · PMC2639000 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 4 setups
Explicit gene–disease relationships can be formulated as candidate gene definitions (e.g., co-localization, dysregulation, functional similarity) that may include intermediary orthologous or interacting genes
-
Has reproduction · 90
A2TEA: Identifying trait-specific evolutionary adaptations.
PMID 37224329 · PMC10186066 · F1000Research · 2022 · 8 claims · 7 setups
A2TEA integrates gene family expansion analysis with differential expression data across species to identify genes that were targets of evolutionary adaptation to a given stress/treatment
-
Has reproduction · 83
Functional module detection through integration of single-cell RNA sequencing data with protein-protein interaction networks.
PMID 33138772 · PMC7607865 · BMC genomics · 2020 · 8 claims · 6 setups
scPPIN integrates scRNA-seq-derived p-values with PPINs to detect maximum-weight connected subgraphs (active/functional modules) via an exact Steiner-tree approach
-
Full-text index only
A statistical framework for consolidating "sibling" probe sets for Affymetrix GeneChip data.
PMID 18435860 · PMC2397416 · BMC genomics · 2008 · 7 claims · 4 setups
A two-way ANOVA model with a treatment x probe-set interaction term can automatically determine whether sibling probe sets for a gene behave similarly (non-significant interaction, consolidate) or differently (significant interaction, treat as independent)
-
Has reproduction · 82
Reusable building blocks in biological systems.
PMID 30958230 · PMC6303794 · Journal of the Royal Society, Interface · 2018 · 8 claims · 4 setups
Biological systems can be decomposed into phenotypic building blocks (PBBs) via k-maximally reusable decompositions (k-MRD) that maximize average reusability across conditions.
-
Has reproduction · 46
De novo transcriptome assembly and comprehensive assessment provide insight into fruiting body formation of Sparassis latifolia.
PMID 35773379 · PMC9247108 · Scientific reports · 2022 · 6 claims · 7 setups
De novo transcriptome assembly of S. latifolia produced 48,549 unigenes, 71.53% (34,728) of which were annotated against KEGG, GO, and/or KOG databases
-
Has reproduction · 73
Transcriptome assembly, profiling and differential gene expression analysis of the halophyte Suaeda fruticosa provides insights into salt tolerance.
PMID 25943316 · PMC4422317 · BMC genomics · 2015 · 7 claims · 6 setups
De novo assembly of the S. fruticosa transcriptome (Velvet/Oases k-45, CDHIT-EST) produced 54,526 high-quality unigenes with N50 of 957 bp
-
Full-text index only
Gene- and evidence-based candidate gene selection for schizophrenia and gene feature analysis.
PMID 19944577 · PMC2826526 · Artificial intelligence in medicine · 2010 · 8 claims · 5 setups
The SCOR method outperforms the CCOR method for prioritizing schizophrenia candidate genes