Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
L2L: a simple tool for discovering the hidden significance in microarray expression data.
PMID 16168088 · PMC1242216 · Genome biology · 2005 · 8 claims · 4 setups
L2L systematically compares a user's differentially expressed gene list against a database of published differentially expressed gene lists to find statistically significant overlaps and generate hypotheses about shared mechanisms
-
Full-text index only
KEGG spider: interpretation of genomics data in the context of the global gene metabolic network.
PMID 19094223 · PMC2646283 · Genome biology · 2008 · 8 claims · 8 setups
KEGG spider, using a global 'pathway-free' metabolic network framework, provides deeper insight into metabolism variations than existing enrichment-based methods.
-
Full-text index only
Divergence of exonic splicing elements after gene duplication and the impact on gene structures.
PMID 19883501 · PMC3091315 · Genome biology · 2009 · 8 claims · 7 setups
ESEs and ESSs diverge especially fast shortly after gene duplication, correlating with time since duplication (Ks)
-
Full-text index only
MatchMiner: a tool for batch navigation among gene and gene product identifiers.
PMID 12702208 · PMC154578 · Genome biology · 2003 · 8 claims · 3 setups
MatchMiner's LookUp function automates batch translation of an input list of gene identifiers into a matching list of a different identifier type.
-
Has reproduction · 75
Why an integrated view of gene expression studies on hematopoiesis in mouse aging is better than the sum of their parts.
PMID 38627103 · PMC11586588 · FEBS letters · 2024 · 7 claims · 4 setups
Combining differentially expressed (DE) gene lists from multiple publications into a unified 'aging list' (AL) with citation counts, and deriving a shorter high-confidence 'aging signature' (AS, genes cited in >3 publications, ~200 genes), produces a more reliable and consistent picture of hematopoietic aging than any single study.
-
Has reproduction · 57
KARAJ: An Efficient Adaptive Multi-Processor Tool to Streamline Genomic and Transcriptomic Sequence Data Acquisition.
PMID 36430895 · PMC9694301 · International journal of molecular sciences · 2022 · 8 claims · 6 setups
KARAJ automates end-to-end querying and downloading of genomic/transcriptomic sequence data from a list of PMCIDs, URLs, or accession numbers
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Full-text index only
Functional nsSNPs from carcinogenesis-related genes expressed in breast tissue: potential breast cancer risk alleles and their distribution across human populations.
PMID 16595073 · PMC3500178 · Human genomics · 2006 · 7 claims · 5 setups
A bioinformatics strategy cross-referencing carcinogenesis-related gene lists with breast-tissue expression data can identify candidate breast cancer risk nsSNPs.
-
Has reproduction · 63
Target identification for repurposed drugs active against SARS-CoV-2 via high-throughput inverse docking.
PMID 34825285 · PMC8616721 · Journal of computer-aided molecular design · 2022 · 8 claims · 6 setups
Combining Vinardo, Ledock, and Korp-PL scoring functions (via averaged Z-scores) improves correct target identification over any single scoring function.
-
Full-text index only
MILANO--custom annotation of microarray results using automatic literature searches.
PMID 15661078 · PMC547913 · BMC bioinformatics · 2005 · 7 claims · 4 setups
MILANO annotates microarray gene lists by counting literature co-occurrences of each gene with user-defined secondary terms
-
Full-text index only
Rapid identification of PAX2/5/8 direct downstream targets in the otic vesicle by combinatorial use of bioinformatics tools.
PMID 18828907 · PMC2760872 · Genome biology · 2008 · 8 claims · 8 setups
A combinatorial bioinformatics pipeline (evolutionary double filtering comparative genomics, GXD/ZFIN database queries, MEDLINE text mining) can rapidly and specifically identify PAX2/5/8 direct downstream targets in the otic vesicle
-
Full-text index only
Identification and analysis of co-occurrence networks with NetCutter.
PMID 18781200 · PMC2526157 · PloS one · 2008 · 8 claims · 4 setups
Random sampling from a complete permutation set of the bipartite graph permits co-occurrence analysis with optimal stringency, and the edge-swapping (ES) model closely approximates this and is the preferred null-model among six tested.
-
Full-text index only
Involvement of potential pathways in malignant transformation from oral leukoplakia to oral squamous cell carcinoma revealed by proteomic analysis.
PMID 19691830 · PMC2746235 · BMC genomics · 2009 · 7 claims · 6 setups
85 proteins are differentially and consistently expressed (>2-fold change, P<0.05) between paired OLK and OSCC tissues, including 52 up-regulated and 33 down-regulated proteins
-
Full-text index only
Inherited disorder phenotypes: controlled annotation and statistical analysis for knowledge mining from gene lists.
PMID 16351744 · PMC1866390 · BMC bioinformatics · 2005 · 5 claims · 3 setups
OMIM Clinical Synopsis free-text phenotype and location names can be normalized and hierarchically structured into a controlled vocabulary suitable for computational analysis
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Has reproduction · 85
Integrated Analysis of Multiple Microarrays Based on Raw Data Identified Novel Gene Signatures in Recurrent Implantation Failure.
PMID 35197930 · PMC8859149 · Frontiers in endocrinology · 2022 · 6 claims · 7 setups
Robust Rank Aggregation (RRA) can integrate DEG lists from multiple independent RIF microarray datasets to identify robust DEGs.
-
Has reproduction · 78
SPRR1B+ keratinocytes prime oral mucosa for rapid wound healing via STAT3 activation.
PMID 39300285 · PMC11413210 · Communications biology · 2024 · 8 claims · 8 setups
A shared wound healing gene set (P2-WHGs, 146 genes) is constitutively expressed in uninjured oral mucosa but not in uninjured skin
-
Has reproduction · 44
Dynamic Gene Attention Focus (DyGAF): Enhancing Biomarker Identification Through Dual-Model Attention Networks.
PMID 40160891 · PMC11951896 · Bioinformatics and biology insights · 2025 · 6 claims · 5 setups
DyGAF, a dual-model attention neural network (independent Model A + dependent Model B), identifies and ranks genes by significance for COVID-19 biomarker discovery more effectively than differential expression analysis (DEA) and random forest (RF) feature selection
-
Has reproduction · 73
Identification of intrinsic genes across general hypertension, hypertension with left ventricular remodeling, and uncontrolled hypertension.
PMID 36277786 · PMC9582241 · Frontiers in cardiovascular medicine · 2022 · 7 claims · 8 setups
FBXW4 and 13 other genes are uniquely enriched in the general hypertension group
-
Has reproduction · 54
Gene-Expression Profiling Suggests Impaired Signaling via the Interferon Pathway in Cstb-/- Microglia.
PMID 27355630 · PMC4927094 · PloS one · 2016 · 8 claims · 8 setups
In Cstb-/- microglia, 184 genes were differentially expressed relative to control, of which 33 were identified by both microarray and RNA-seq.