Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
g:Profiler--a web-based toolset for functional profiling of gene lists from large-scale experiments.
PMID 17478515 · PMC1933153 · Nucleic acids research · 2007 · 8 claims · 5 setups
g:Profiler integrates four modules (g:Profiler core, g:Convert, g:Orth, g:Sorter) into a single cross-linked web tool for gene list analysis
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Full-text index only
CRSD: a comprehensive web server for composite regulatory signature discovery.
PMID 16845073 · PMC1538777 · Nucleic acids research · 2006 · 7 claims · 5 setups
CRSD is a comprehensive web server integrating six large-scale databases (UniGene, mature microRNAs, putative promoter, TRANSFAC, pathway, GO) plus two newly constructed genome-wide databases (MRS and TRS) for composite regulatory signature discovery
-
Full-text index only
Onto-Tools: new additions and improvements in 2006.
PMID 17584796 · PMC1933142 · Nucleic acids research · 2007 · 8 claims · 3 setups
OE2GO enables functional profiling for organisms lacking public-domain annotations by allowing users to supply custom GO-format annotation files and OBO-format ontology files
-
Full-text index only
GeneKeyDB: a lightweight, gene-centric, relational database to support data mining environments.
PMID 15790402 · PMC1274265 · BMC bioinformatics · 2005 · 8 claims · 6 setups
GeneKeyDB is a lightweight, gene-centric relational database that supports data mining and integration with computational analysis tools.
-
Full-text index only
KEGG spider: interpretation of genomics data in the context of the global gene metabolic network.
PMID 19094223 · PMC2646283 · Genome biology · 2008 · 8 claims · 8 setups
KEGG spider, using a global 'pathway-free' metabolic network framework, provides deeper insight into metabolism variations than existing enrichment-based methods.
-
Has reproduction · 76
GeneSetCart: assembling, augmenting, combining, visualizing, and analyzing gene sets.
PMID 40208796 · PMC11984350 · GigaScience · 2025 · 8 claims · 8 setups
GeneSetCart is a web-based platform that lets users assemble, augment, combine, visualize, and analyze gene sets from multiple sources in one place
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Full-text index only
Seeded Bayesian Networks: constructing genetic networks from microarray data.
PMID 18601736 · PMC2474592 · BMC systems biology · 2008 · 8 claims · 4 setups
Seeding Bayesian Network analysis with prior networks derived from literature and/or PPI data improves recovery of known gene-gene interactions compared to BN analysis without a seed
-
Full-text index only
Pegasys: software for executing and integrating analyses of biological sequences.
PMID 15096276 · PMC406494 · BMC bioinformatics · 2004 · 8 claims · 7 setups
Pegasys is a flexible, modular, customizable software system for executing and integrating heterogeneous biological sequence analysis tools
-
Full-text index only
Gene-disease relationship discovery based on model-driven data integration and database view definition.
PMID 19042916 · PMC2639000 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 4 setups
Explicit gene–disease relationships can be formulated as candidate gene definitions (e.g., co-localization, dysregulation, functional similarity) that may include intermediary orthologous or interacting genes
-
Has reproduction · 68
Targeting the epigenome and the integrated stress response to normalize colorectal cancer subclonal plasticity and progression.
PMID 41963303 · PMC13181133 · Cell death & disease · 2026 · 7 claims · 8 setups
The integrated stress response induces colorectal cancer cell plasticity, subclonal diversity, and tumor progression in stress-surviving cells
-
Full-text index only
CEAS: cis-regulatory element annotation system.
PMID 16845068 · PMC1538818 · Nucleic acids research · 2006 · 7 claims · 5 setups
CEAS is the first web server to streamline genome-scale ChIP-chip downstream analyses for biologists without strong bioinformatics support
-
Has reproduction · 88
nf-core/isoseq: simple gene and isoform annotation with PacBio Iso-Seq long-read sequencing.
PMID 36961337 · PMC10199315 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 4 setups
nf-core/isoseq is a new automated Nextflow-based pipeline that processes raw Iso-Seq subreads through to genome annotation (BED format) without requiring transcriptome assembly.
-
Full-text index only
SNP-RFLPing: restriction enzyme mining for SNPs in genomes.
PMID 16503968 · PMC1386656 · BMC genomics · 2006 · 8 claims · 2 setups
SNP-RFLPing accepts three flexible input types (dbSNP rs#/ss# IDs, HUGO gene name/Entrez gene ID, or free-form SNP-in-sequence including IUPAC or [dNTP1/dNTP2] formats) for human, rat, and mouse genomes
-
Has reproduction · 90
A2TEA: Identifying trait-specific evolutionary adaptations.
PMID 37224329 · PMC10186066 · F1000Research · 2022 · 8 claims · 7 setups
A2TEA integrates gene family expansion analysis with differential expression data across species to identify genes that were targets of evolutionary adaptation to a given stress/treatment
-
Full-text index only
Using several pair-wise informant sequences for de novo prediction of alternatively spliced transcripts.
PMID 16925842 · PMC1810557 · Genome biology · 2006 · 8 claims · 4 setups
MARS, an extension of the Twinscan algorithm, uses multiple pairwise informant genomes to predict human alternatively spliced transcripts de novo without expressed sequence information.
-
Full-text index only
DiRE: identifying distant regulatory elements of co-expressed genes.
PMID 18487623 · PMC2447744 · Nucleic acids research · 2008 · 8 claims · 4 setups
DiRE predicts distant regulatory elements by combining gene co-expression data, comparative genomics and TFBS profiles to determine TFBS-association signatures
-
Has reproduction · 56
Analysis of subcellular transcriptomes by RNA proximity labeling with Halo-seq.
PMID 34875090 · PMC8887463 · Nucleic acids research · 2022 · 6 claims · 8 setups
Halo-seq pairs a light-activatable Halo-DBF ligand with Click chemistry to label and purify spatially defined RNA populations in living cells with high spatial specificity (~100 nm radius)
-
Full-text index only
Shaken not stirred: a global research cocktail served in Hinxton.
PMID 18036269 · PMC2258181 · Genome biology · 2007 · 8 claims · 8 setups
Network-guided reverse genetics using probabilistic functional gene networks (e.g. YeastNet, WormNet) reduces the search space for identifying genes in a given biological process