Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
L2L: a simple tool for discovering the hidden significance in microarray expression data.
PMID 16168088 · PMC1242216 · Genome biology · 2005 · 8 claims · 4 setups
L2L systematically compares a user's differentially expressed gene list against a database of published differentially expressed gene lists to find statistically significant overlaps and generate hypotheses about shared mechanisms
-
Full-text index only
MILANO--custom annotation of microarray results using automatic literature searches.
PMID 15661078 · PMC547913 · BMC bioinformatics · 2005 · 7 claims · 4 setups
MILANO annotates microarray gene lists by counting literature co-occurrences of each gene with user-defined secondary terms
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Has reproduction · 77
Comparison of RNA-Seq by poly (A) capture, ribosomal RNA depletion, and DNA microarray for expression profiling.
PMID 24888378 · PMC4070569 · BMC genomics · 2014 · 8 claims · 8 setups
Ribo-Zero-Seq removes rRNA with efficiency comparable to poly(A)-based mRNA-Seq in both FF and FFPE RNA, whereas DSN-Seq leaves significantly more rRNA and shows greater variation.
-
Full-text index only
Getting positive about selection.
PMID 12914654 · PMC193638 · Genome biology · 2003 · 8 claims · 4 setups
Purifying selection is the predominant form of molecular evolution, preserving fitness by eliminating deleterious mutations, while positive selection is rare but critical for adaptation.
-
Has reproduction · 83
Hobbes: optimized gram-based methods for efficient read alignment.
PMID 22199254 · PMC3315303 · Nucleic acids research · 2012 · 8 claims · 4 setups
Hobbes, a gram-based short-read mapper supporting Hamming and edit distance, is faster than all other read-mapping programs tested while maintaining high mapping quality.
-
Full-text index only
KEGG spider: interpretation of genomics data in the context of the global gene metabolic network.
PMID 19094223 · PMC2646283 · Genome biology · 2008 · 8 claims · 8 setups
KEGG spider, using a global 'pathway-free' metabolic network framework, provides deeper insight into metabolism variations than existing enrichment-based methods.
-
Full-text index only
Identification and analysis of co-occurrence networks with NetCutter.
PMID 18781200 · PMC2526157 · PloS one · 2008 · 8 claims · 4 setups
Random sampling from a complete permutation set of the bipartite graph permits co-occurrence analysis with optimal stringency, and the edge-swapping (ES) model closely approximates this and is the preferred null-model among six tested.
-
Full-text index only
Divergence of exonic splicing elements after gene duplication and the impact on gene structures.
PMID 19883501 · PMC3091315 · Genome biology · 2009 · 8 claims · 7 setups
ESEs and ESSs diverge especially fast shortly after gene duplication, correlating with time since duplication (Ks)
-
Full-text index only
BABELOMICS: a systems biology perspective in the functional annotation of genome-scale experiments.
PMID 16845052 · PMC1538844 · Nucleic acids research · 2006 · 8 claims · 8 setups
Babelomics is presented as an updated, complete suite of web tools for functional analysis of genome-scale experiments with new and improved modules
-
Full-text index only
Involvement of potential pathways in malignant transformation from oral leukoplakia to oral squamous cell carcinoma revealed by proteomic analysis.
PMID 19691830 · PMC2746235 · BMC genomics · 2009 · 7 claims · 6 setups
85 proteins are differentially and consistently expressed (>2-fold change, P<0.05) between paired OLK and OSCC tissues, including 52 up-regulated and 33 down-regulated proteins
-
Has reproduction · 85
Integrated Analysis of Multiple Microarrays Based on Raw Data Identified Novel Gene Signatures in Recurrent Implantation Failure.
PMID 35197930 · PMC8859149 · Frontiers in endocrinology · 2022 · 6 claims · 7 setups
Robust Rank Aggregation (RRA) can integrate DEG lists from multiple independent RIF microarray datasets to identify robust DEGs.
-
Has reproduction · 71
Chemical genomics informs antibiotic and essential gene function in Acinetobacter baumannii.
PMID 40153700 · PMC11975115 · PLoS genetics · 2025 · 8 claims · 6 setups
The vast majority of A. baumannii essential genes show significant chemical-gene interactions upon knockdown
-
Full-text index only
LMPD: LIPID MAPS proteome database.
PMID 16381922 · PMC1347484 · Nucleic acids research · 2006 · 8 claims · 5 setups
LMPD is an object-relational database of lipid-associated protein sequences and annotations, publicly available from the LIPID MAPS Consortium website.
-
Has reproduction · 55
Gene module regulation in dilated cardiomyopathy and the role of Na/K-ATPase.
PMID 35901050 · PMC9333241 · PloS one · 2022 · 5 claims · 8 setups
Several co-expressed gene modules are significantly associated with left ventricle ejection fraction (LVEF) and the DCM phenotype, enriched in fibrosis-related, small molecule transporting-related, and immune response-related pathways.
-
Has reproduction · 44
Dynamic Gene Attention Focus (DyGAF): Enhancing Biomarker Identification Through Dual-Model Attention Networks.
PMID 40160891 · PMC11951896 · Bioinformatics and biology insights · 2025 · 6 claims · 5 setups
DyGAF, a dual-model attention neural network (independent Model A + dependent Model B), identifies and ranks genes by significance for COVID-19 biomarker discovery more effectively than differential expression analysis (DEA) and random forest (RF) feature selection
-
Has reproduction
Using random walks to identify cancer-associated modules in expression data.
PMID 24128261 · PMC4015830 · BioData mining · 2013 · 8 claims · 8 setups
Walktrap-GM, a random-walk community detection algorithm adapted with stopping criteria (maximum modularity, maximum size, maximum module score), identifies modules significantly enriched with cancer genes in expression-weighted interaction networks.
-
Full-text index only
The DAVID Gene Functional Classification Tool: a novel biological module-centric algorithm to functionally analyze large gene lists.
PMID 17784955 · PMC2375021 · Genome biology · 2007 · 8 claims · 6 setups
Gene-gene functional similarity can be measured using kappa statistics applied to a binary gene-annotation-term matrix built from 14 annotation categories.
-
Full-text index only
LIMPIC: a computational method for the separation of protein MALDI-TOF-MS signals from noise.
PMID 17386085 · PMC1847688 · BMC bioinformatics · 2007 · 7 claims · 4 setups
LIMPIC is a computational method for detecting protein peaks from linear-mode MALDI-TOF-MS data using background noise reduction and baseline removal followed by non-uniform threshold peak detection and multi-spectra detection-rate classification.
-
Full-text index only
Versatile online-offline engine for automated acquisition of high-resolution tandem mass spectra.
PMID 18841935 · PMC2716176 · Analytical chemistry · 2008 · 7 claims · 6 setups
An integrated hardware/software online-offline platform (TriVersa NanoMate + 12T LTQ FT Ultra + AUTOMATION WAREHOUSE) automates collection of high-resolution MS/MS data for polypeptides >3 kDa that cannot be acquired online on a chromatographic timescale