Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Sequence similarity network reveals common ancestry of multidomain proteins.
PMID 18475320 · PMC2377100 · PLoS computational biology · 2008 · 8 claims · 6 setups
Traditional homology definitions do not capture multidomain evolution; the authors extend the definition to include domain insertion via a common ancestral locus model.
-
Full-text index only
The other side of comparative genomics: genes with no orthologs between the cow and other mammalian species.
PMID 20003425 · PMC2808326 · BMC genomics · 2009 · 7 claims · 4 setups
3,801 bovine genes have no orthologs in human, mouse and dog, and 1,010 human genes have no orthologs in cow despite having orthologs in mouse and dog
-
Full-text index only
SuperCYP: a comprehensive database on Cytochrome P450 enzymes including a tool for analysis of CYP-drug interactions.
PMID 19934256 · PMC2808967 · Nucleic acids research · 2010 · 8 claims · 6 setups
SuperCYP is a comprehensive relational database aggregating CYP enzyme, drug metabolism, SNP/mutation, and structural information from literature and web resources.
-
Full-text index only
A new procedure for determining the genetic basis of a physiological process in a non-model species, illustrated by cold induced angiogenesis in the carp.
PMID 19852815 · PMC2771047 · BMC genomics · 2009 · 8 claims · 5 setups
The Conditional Stepped Reciprocal Best Hit (CSRBH) approach, combining direct RBH and zebrafish-stepped RBH (SRBH), outperformed other ortholog assignment methods and attained 8,726 carp-human functional homolog relationships for 16,650 carp contigs
-
Full-text index only
Extraction of human kinase mutations from literature, databases and genotyping studies.
PMID 19758464 · PMC2745582 · BMC bioinformatics · 2009 · 7 claims · 6 setups
A literature mining pipeline combining MutationFinder, false-positive filtering, and SVM-based classification can extract and disambiguate single-point mutation mentions from abstracts and full text
-
Full-text index only
IDEAL-Q, an automated tool for label-free quantitation analysis using an efficient peptide alignment approach and spectral data validation.
PMID 19752006 · PMC2808259 · Molecular & cellular proteomics : MCP · 2010 · 6 claims · 5 setups
IDEAL-Q predicts the elution time of peptides unidentified in a given LC-MS/MS run (but identified in others) using a computation-efficient linear regression plus fragmental refining function, avoiding costly whole-dataset pattern recognition
-
Full-text index only
Automated mapping of DNA replication fork progression in human cells with ForkML.
PMID 41577668 · PMC12932727 · Nature communications · 2026 · 8 claims · 8 setups
ForkML uses double BrdU pulse-labelling and nanopore sequencing with a machine-learning fork detection/orientation pipeline to automatically map thousands of individual replication fork velocities.
-
Full-text index only
The genome sequence of Menetries' Clouded Yellow, Colias thisoa Ménétriès, 1832 (Lepidoptera: Pieridae).
PMID 41716938 · PMC12914173 · Wellcome open research · 2026 · 8 claims · 7 setups
A chromosome-level, haplotype-resolved genome assembly was generated for Colias thisoa as part of Project Psyche
-
Full-text index only
MIMIC: a flexible pipeline to register and summarize IMC-MSI experiments.
PMID 41917425 · PMC13201759 · Communications biology · 2026 · 7 claims · 6 setups
MIMIC is a reproducible, semi-automated workflow that co-registers and jointly analyzes MALDI-MSI and IMC data using a chain of before/after microscopy images.
-
Full-text index only
scSNViz: visualization and analysis of cell-specific expressed SNVs.
PMID 41533688 · PMC12866635 · Bioinformatics (Oxford, England) · 2026 · 7 claims · 7 setups
scSNViz is an R package for exploration, quantification, and visualization of expressed SNVs from cell-barcoded scRNA-seq data, supporting VAF estimation, SNV clustering, and 2D/3D visualization.
-
Full-text index only
Performance assessment of promoter predictions on ENCODE regions in the EGASP experiment.
PMID 16925837 · PMC1810552 · Genome biology · 2006 · 6 claims · 3 setups
Promoter predictors that combine promoter prediction with gene prediction (N-SCAN, Fprom) achieve better performance than pure ab initio promoter predictors, mainly by reducing the promoter search space and false positives
-
Full-text index only
DDBJ in collaboration with mass-sequencing teams on annotation.
PMID 15608189 · PMC539974 · Nucleic acids research · 2005 · 7 claims · 5 setups
DDBJ collected and released 1,066,084 entries (718,072,425 bases) in the past year, including the complete chimpanzee chromosome 22 sequence and silkworm whole-genome shotgun data
-
Full-text index only
Construction and use of spotted large-insert clone DNA microarrays for the detection of genomic copy number changes.
PMID 17406619 · PMC2688820 · Nature protocols · 2007 · 8 claims · 7 setups
Combining three human-optimized DOP-PCR primers before a secondary amino-labeled PCR increases array hybridization sensitivity and reproducibility sixfold compared to the standard 6MW DOP-PCR primer
-
Full-text index only
CYCLONET--an integrated database on cell cycle regulation and carcinogenesis.
PMID 17202170 · PMC1899094 · Nucleic acids research · 2007 · 7 claims · 4 setups
Cyclonet is a web-based integrated database combining 'omics' and chemoinformatics data on mammalian cell cycle regulation in normal and pathological (cancer) states, built on a systems biology approach.
-
Full-text index only
BRENDA, AMENDA and FRENDA: the enzyme information system in 2007.
PMID 17202167 · PMC1899097 · Nucleic acids research · 2007 · 7 claims · 6 setups
BRENDA is the largest publicly available enzyme information system worldwide, manually curated from primary literature and covering all identified enzymes regardless of source.
-
Full-text index only
SUPERFAMILY--sophisticated comparative genomics, data mining, visualization and phylogeny.
PMID 19036790 · PMC2686452 · Nucleic acids research · 2009 · 7 claims · 6 setups
SUPERFAMILY provides structural, functional and evolutionary annotation for proteins from all completely sequenced genomes using SCOP-based hidden Markov models
-
Full-text index only
The genome sequence of the fin whale, Balaenoptera physalus (Linnaeus, 1758) (Artiodactyla: Balaenopteridae).
PMID 41657664 · PMC12877476 · Wellcome open research · 2026 · 8 claims · 8 setups
A chromosome-level genome assembly was produced for Balaenoptera physalus (fin whale) from a male specimen collected in Orkney, Scotland, as part of the Darwin Tree of Life project
-
Full-text index only
The genome sequence of a rove beetle, Tachyporus hypnorum (Fabricius, 1775) (Coleoptera: Staphylinidae).
PMID 41993728 · PMC13080330 · Wellcome open research · 2026 · 6 claims · 8 setups
A chromosome-level genome assembly was generated for Tachyporus hypnorum, the first high-quality genome for the genus Tachyporus.
-
Has reproduction · 93
Firefly genomes illuminate parallel origins of bioluminescence in beetles.
PMID 30324905 · PMC6191289 · eLife · 2018 · 7 claims · 8 setups
Bioluminescence arose independently (parallel/convergent origins) in fireflies and click beetles rather than from a single common ancestral origin.
-
Has reproduction · 73
Proteogenomic analysis prioritises functional single nucleotide variants in cancer samples.
PMID 29221171 · PMC5707065 · Oncotarget · 2017 · 8 claims · 6 setups
A customised SAAV peptide database built from RNA-seq/WGS variant calls can be used to search proteomics data and detect single amino acid variant (SAAV)-containing peptides at the protein level