Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Organization of physical interactomes as uncovered by network schemas.
PMID 18949022 · PMC2561054 · PLoS computational biology · 2008 · 7 claims · 5 setups
A computational procedure can systematically identify 'emergent' network schemas that are both recurrent and over-represented relative to randomized networks preserving lower-order subschema distributions
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
DAVID Knowledgebase: a gene-centered database integrating heterogeneous gene annotation resources to facilitate high-throughput gene functional analysis.
PMID 17980028 · PMC2186358 · BMC bioinformatics · 2007 · 7 claims · 3 setups
The DAVID Gene Concept, a single-linkage algorithm, merges gene clusters from Entrez Gene, UniRef100, and PIR-NREF100 that share protein IDs and species into unified DAVID gene clusters, improving cross-referencing between NCBI and UniProt systems
-
Full-text index only
Better smelling through genetics: mammalian odor perception.
PMID 18938244 · PMC2590501 · Current opinion in neurobiology · 2008 · 8 claims · 8 setups
Odorant receptor (OR) gene repertoire size and pseudogene fraction vary dramatically across mammalian species
-
Full-text index only
The complete genome sequence of Vibrio cholerae: a tale of two chromosomes and of two lifestyles.
PMID 11178241 · PMC138858 · Genome biology · 2000 · 8 claims · 4 setups
The V. cholerae O1 (El Tor) genome consists of two chromosomes with asymmetrically distributed gene functions
-
Has reproduction · 56
Identification of key genes in chickpea transcriptomics and the development of ChickpeaOmicsR as a comprehensive resource to advance breeding and genomic studies.
PMID 41909810 · PMC13022592 · Frontiers in bioinformatics · 2026 · 8 claims · 4 setups
ChickpeaOmicsR is the first comprehensive/specialized R package integrating transcriptomic, genomic, and proteomic (RNA-seq, GWAS, PPI) data within a unified, reproducible framework and standardizing fragmented chickpea gene nomenclature.
-
Full-text index only
BioGPS: an extensible and customizable portal for querying and organizing gene annotation resources.
PMID 19919682 · PMC3091323 · Genome biology · 2009 · 8 claims · 4 setups
BioGPS aggregates distributed, third-party gene annotation resources into a single customizable portal for human, mouse, and rat genes.
-
Has reproduction · 69
Meta-analysis of six dairy cattle breeds reveals biologically relevant candidate genes for mastitis resistance.
PMID 39009986 · PMC11247842 · Genetics, selection, evolution : GSE · 2024 · 6 claims · 8 setups
Meta-analysis of GWAS across multiple breeds for CM and SCS identified 58 lead markers associated with mastitis incidence, including 16 loci not overlapping previously identified QTL in AnimalQTLdb.
-
Has reproduction
Genome-wide associations of aortic distensibility suggest causality for aortic aneurysms and brain white matter hyperintensities.
PMID 35922433 · PMC9349177 · Nature communications · 2022 · 8 claims · 7 setups
Genome-wide association of six CMR-derived aortic traits in up to 32,590 UK Biobank participants identifies 102 loci (including 27 novel associations) for aortic distensibility and area.
-
Full-text index only
Columba: an integrated database of proteins, structures, and annotations.
PMID 15801979 · PMC1087474 · BMC bioinformatics · 2005 · 8 claims · 6 setups
COLUMBA physically integrates data from twelve protein structure-related databases (PDB, KEGG, Swiss-Prot, CATH, SCOP, Gene Ontology, ENZYME, etc.) into a single PostgreSQL data warehouse.
-
Full-text index only
L1Base: from functional annotation to prediction of active LINE-1 elements.
PMID 15608246 · PMC539998 · Nucleic acids research · 2005 · 7 claims · 6 setups
L1Base is a database of putatively active LINE-1 insertions in human, mouse and rat genomes, containing FLI-L1s (intact in both ORFs), ORF2-L1s (intact ORF2, disrupted ORF1), and FLnI-L1s (full-length, >6000 bp, non-intact)
-
Full-text index only
Multi-organ expression profiling uncovers a gene module in coronary artery disease involving transendothelial migration of leukocytes and LIM domain binding 2: the Stockholm Atherosclerosis Gene Expression (STAGE) study.
PMID 19997623 · PMC2780352 · PLoS genetics · 2009 · 8 claims · 6 setups
Functionally associated gene modules, not individual genes, underlie CAD development and can be identified via multi-organ expression clustering
-
Full-text index only
Pathway analysis for intracellular Porphyromonas gingivalis using a strain ATCC 33277 specific database.
PMID 19723305 · PMC2753363 · BMC microbiology · 2009 · 8 claims · 5 setups
Using the ATCC 33277-specific genome annotation improves proteome coverage (more proteins identified and more abundance ratios calculated) compared to the W83 annotation
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.
-
Has reproduction · 74
Discovery of a novel filamentous prophage in the genome of the Mimosa pudica microsymbiont Cupriavidus taiwanensis STM 6018.
PMID 36925474 · PMC10011098 · Frontiers in microbiology · 2023 · 8 claims · 8 setups
The STM 6018 genome contains two prophages: a complete Mu-like capsular phage and a filamentous phage that integrates into a putative dif site.
-
Full-text index only
Discovery of protein-protein interactions using a combination of linguistic, statistical and graphical information.
PMID 15941473 · PMC1164402 · BMC bioinformatics · 2005 · 8 claims · 5 setups
A combined linguistic+statistical+rule-based method achieves precision 0.61 and recall 0.97 (f=0.74) detecting yeast protein-protein interactions across 12,300 Medline abstracts.
-
Has reproduction · 88
nf-core/isoseq: simple gene and isoform annotation with PacBio Iso-Seq long-read sequencing.
PMID 36961337 · PMC10199315 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 4 setups
nf-core/isoseq is a new automated Nextflow-based pipeline that processes raw Iso-Seq subreads through to genome annotation (BED format) without requiring transcriptome assembly.
-
Full-text index only
MACSIMS: multiple alignment of complete sequences information management system.
PMID 16792820 · PMC1539025 · BMC bioinformatics · 2006 · 8 claims · 5 setups
MACSIMS is a multiple alignment-based information management system combining knowledge-based database mining with ab initio sequence predictions
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Full-text index only
Integrating alternative splicing detection into gene prediction.
PMID 15705189 · PMC550657 · BMC bioinformatics · 2005 · 8 claims · 4 setups
An integrative intrinsic/extrinsic method was implemented in the gene finder EuGÈNE (as EuGÈNE-M) to detect AS evidence from aligned transcripts and generate alternative optimal gene predictions consistent with each detected AS event.