Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Integration with the human genome of peptide sequences obtained by high-throughput mass spectrometry.
PMID 15642101 · PMC549070 · Genome biology · 2005 · 8 claims · 4 setups
PeptideAtlas, a public database integrating MS/MS-derived peptide identifications with the human genome, was built as an expandable resource for proteomic data.
-
Full-text index only
Novel gene and gene model detection using a whole genome open reading frame analysis in proteomics.
PMID 16646984 · PMC1557991 · Genome biology · 2006 · 8 claims · 4 setups
A six-frame genomic ORF translation used as an MS search database can detect novel peptides absent from standard protein databases, revealing incomplete genome annotation.
-
Full-text index only
MEROPS: the peptidase database.
PMID 19892822 · PMC2808883 · Nucleic acids research · 2010 · 8 claims · 5 setups
MEROPS is a manually curated hierarchical classification of peptidases and protein inhibitors organized into protein species, families, and clans based on sequence and structural homology.
-
Full-text index only
SVNeoPP: A Workflow for Structural-Variant-Derived Neoantigen Prediction and Prioritization Using Multi-Omics Data.
PMID 41892252 · PMC13024079 · Biology · 2026 · 8 claims · 7 setups
SVNeoPP is an end-to-end Snakemake workflow that takes WGS and RNA-seq as input to call/annotate SVs, reconstruct altered transcripts and coding sequences in an isoform-aware, traceable manner, and generate candidate peptides.
-
Full-text index only
Systematic analysis of human kinase genes: a large number of genes and alternative splicing events result in functional and structural diversity.
PMID 16351747 · PMC1866387 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Systematic in silico search identified 5 novel human kinase genes (on chromosomes 1, 11, 13, 15, 16) and 1 pseudogene (chromosome X) absent from KinBase
-
Has reproduction · 76
The genome and development-dependent transcriptomes of Pyronema confluens: a window into fungal evolution.
PMID 24068976 · PMC3778014 · PLoS genetics · 2013 · 8 claims · 8 setups
The 50 Mb P. confluens genome with 13,369 predicted protein-coding genes is more characteristic of higher filamentous ascomycetes than of the large, repeat-rich Tuber melanosporum genome, showing that the truffle's expanded genome is not typical of the Pezizales.
-
Full-text index only
The Proteomic Code: a molecular recognition code for proteins.
PMID 17999762 · PMC2206014 · Theoretical biology & medical modelling · 2007 · 8 claims · 8 setups
The Proteomic Code is a set of rules by which genetic information is transferred into the physico-chemical properties of amino acids, determining protein-protein interactions and folding; it is part of the redundant Genetic Code.
-
Full-text index only
A compatible exon-exon junction database for the identification of exon skipping events using tandem mass spectrum data.
PMID 19087293 · PMC2636810 · BMC bioinformatics · 2008 · 6 claims · 6 setups
A theoretical exon-exon junction protein database accounting for all in-phase (frame-preserving) exon combinations can be built from the Ensembl Core Database using Perl/Bioperl/MySQL/Ensembl API.
-
Full-text index only
Evidence for a novel gene associated with human influenza A viruses.
PMID 19917120 · PMC2780412 · Virology journal · 2009 · 8 claims · 8 setups
A 167-codon ORF (NEG8) on the negative-sense genomic strand of segment 8 is associated with early-20th-century human influenza A isolates
-
Full-text index only
The gentle art of gene arrangement: the meaning of gene clusters.
PMID 11897017 · PMC139018 · Genome biology · 2002 · 8 claims · 7 setups
Gene order in eukaryotic genomes is likely optimized by natural selection rather than arising purely by chance reshuffling.
-
Full-text index only
Comparing protein abundance and mRNA expression levels on a genomic scale.
PMID 12952525 · PMC193646 · Genome biology · 2003 · 8 claims · 8 setups
Correlations between mRNA expression and protein abundance are generally poor or limited across most studies reviewed, including in yeast and human cancers
-
Full-text index only
Anopheles gambiae genome reannotation through synthesis of ab initio and comparative gene prediction algorithms.
PMID 16569258 · PMC1557760 · Genome biology · 2006 · 8 claims · 7 setups
An exon-gene-union (EGU) algorithm followed by an open-reading-frame-selection algorithm can synthesize ab initio (GENSCAN, GeneMark, SNAP) and comparative (Ensembl/Genewise) predictions into a single, more complete CDS set
-
Full-text index only
HIV-1 evolution following transmission to an HLA-B*5801-positive patient.
PMID 19909081 · PMC2779566 · The Journal of infectious diseases · 2009 · 8 claims · 8 setups
Multiple escape mutations developed rapidly in HLA-B*5801-restricted epitopes in Gag, Nef, and Pol following transmission
-
Full-text index only
Exploring the regulatory potential of RNA structures in 202 cyanobacterial genomes.
PMID 41641705 · PMC12873609 · Nucleic acids research · 2026 · 7 claims · 8 setups
Screening 202 cyanobacterial genomes identified 402 CRSs matching known RNA families (Rfam and Rho-independent terminators) and 409 novel CRSs.
-
Full-text index only
High accuracy mass spectrometry analysis as a tool to verify and improve gene annotation using Mycobacterium tuberculosis as an example.
PMID 18597682 · PMC2483986 · BMC genomics · 2008 · 8 claims · 5 setups
High-accuracy MS proteomics (LTQ-Orbitrap) can be used to verify and improve gene annotation by identifying peptides specific to one of two competing annotation datasets.
-
Full-text index only
The Proteomics Identifications database: 2010 update.
PMID 19906717 · PMC2808904 · Nucleic acids research · 2010 · 8 claims · 6 setups
PRIDE has become one of the main repositories for MS-based proteomics data, with substantial growth in data holdings over the last two years.
-
Has reproduction · 67
Essential Genes of Vibrio anguillarum and Other Vibrio spp. Guide the Development of New Drugs and Vaccines.
PMID 34745063 · PMC8564382 · Frontiers in microbiology · 2021 · 7 claims · 7 setups
Tn-seq using the TnSC189 mariner transposon identified 329 essential genes in V. anguillarum NB10Sm from a library of 52,662 insertion mutants.
-
Has reproduction · 93
Characterization of protein isoform diversity in human umbilical vein endothelial cells via long-read proteogenomics.
PMID 36457147 · PMC9721438 · RNA biology · 2022 · 8 claims · 7 setups
Long-read RNA-seq detected 53,863 transcript isoforms from 10,426 genes in HUVECs, of which 22,195 were novel
-
Full-text index only
Searching for new clues about the molecular cause of endomyocardial fibrosis by way of in silico proteomics and analytical chemistry.
PMID 19823676 · PMC2757908 · PloS one · 2009 · 8 claims · 4 setups
Cross-reactivity of antibodies against C-terminal sequences of ribosomal P proteins from several animals, plants and protozoa with heart tissue may mediate EMF similarly to how T. cruzi C-termini mediate Chaga's disease
-
Full-text index only
Proteomics data repositories.
PMID 19795424 · PMC2908408 · Proteomics · 2009 · 5 claims · 5 setups
The YRC Public Data Repository (YRC PDR) provides a single unified interface disseminating multi-technology proteomics data (mass spectrometry, yeast two-hybrid, fluorescence microscopy, structure prediction) linked to protein annotations from many source databases.