Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.
-
Full-text index only
MitoP2: the mitochondrial proteome database--now including mouse data.
PMID 16381964 · PMC1347489 · Nucleic acids research · 2006 · 8 claims · 8 setups
MitoP2 is a database integrating manually annotated mitochondrial reference proteins, functions, and disease associations for yeast, human, and mouse, with cross-species orthologue mapping
-
Full-text index only
Advancement of biomarker discovery and validation through the HUPO plasma proteome project.
PMID 15502245 · PMC3839274 · Disease markers · 2004 · 7 claims · 3 setups
Standardization of specimen collection, handling, storage, and choice of serum vs. plasma/anticoagulant is essential for comparable proteomic biomarker discovery.
-
Full-text index only
Computational verification of protein-protein interactions by orthologous co-expression.
PMID 15740634 · PMC555590 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Co-expression of orthologous protein pairs across multiple species can verify/predict S. cerevisiae PPIs with better performance than S. cerevisiae co-expression alone.
-
Full-text index only
A mouse plasma peptide atlas as a resource for disease proteomics.
PMID 18522751 · PMC2481425 · Genome biology · 2008 · 8 claims · 6 setups
A publicly available, high-quality mouse plasma peptide/protein repository (mouse PeptideAtlas) was built from 568 LC-MS/MS runs on four reference plasma pools.
-
Full-text index only
NCBI Reference Sequences: current status, policy and new initiatives.
PMID 18927115 · PMC2686572 · Nucleic acids research · 2009 · 7 claims · 5 setups
RefSeq is a curated, non-redundant, explicitly linked database of nucleotide and protein sequences spanning genomes, transcripts and proteins across prokaryotes, eukaryotes and viruses
-
Has reproduction · 77
Representing and querying disease networks using graph databases.
PMID 27462371 · PMC4960687 · BioData mining · 2016 · 7 claims · 8 setups
Graph databases are well suited for representing biological information that is highly connected, semi-structured, and unpredictable.
-
Full-text index only
Codon usage comparison of novel genes in clinical isolates of Haemophilus influenzae.
PMID 15983137 · PMC1160521 · Nucleic acids research · 2005 · 8 claims · 4 setups
A codon usage similarity statistic (ε, based on squared/absolute differences of codon frequencies with an optimized amino acid usage factor) was developed to compare ORFs against a set of 80 reference genomes.
-
Full-text index only
A computational screen for type I polyketide synthases in metagenomics shotgun data.
PMID 18953415 · PMC2568958 · PloS one · 2008 · 8 claims · 6 setups
Combining HMM domain searches with maximum-likelihood phylogenetic trees can discriminate true PKS I sequences from evolutionarily related but functionally different enzymes (e.g., FAS I) in metagenomic data.
-
Full-text index only
Aggregation propensity of the human proteome.
PMID 18927604 · PMC2557143 · PLoS computational biology · 2008 · 8 claims · 7 setups
Long proteins have, on average, less intense/pronounced aggregation peaks than short proteins
-
Full-text index only
Comparing cellular proteomes by mass spectrometry.
PMID 19886975 · PMC2784314 · Genome biology · 2009 · 8 claims · 6 setups
MS-based proteomics combined with cryo-electron tomography (CET) enables determination of absolute and relative protein abundances and localization
-
Full-text index only
High resolution analysis of the human transcriptome: detection of extensive alternative splicing independent of transcriptional activity.
PMID 19804644 · PMC2768739 · BMC genetics · 2009 · 8 claims · 6 setups
The human GWSA uses exon body and exon-exon junction probes to directly measure over 280,000 known and predicted splicing events genome-wide.
-
Full-text index only
Pathway analysis for intracellular Porphyromonas gingivalis using a strain ATCC 33277 specific database.
PMID 19723305 · PMC2753363 · BMC microbiology · 2009 · 8 claims · 5 setups
Using the ATCC 33277-specific genome annotation improves proteome coverage (more proteins identified and more abundance ratios calculated) compared to the W83 annotation
-
Full-text index only
Disease-specific proteins from rheumatoid arthritis patients.
PMID 16778393 · PMC2729955 · Journal of Korean medical science · 2006 · 8 claims · 6 setups
Fibronectin, semaphorin 7A precursor, growth factor receptor-bound protein 7 (GRB7), and immunoglobulin µ chain specifically associate with antibodies purified from RA synovial fluid
-
Full-text index only
HUPO Highlights.
PMID 19862759 · PMC4594800 · Proteomics · 2009 · 8 claims · 8 setups
Mass spectrometry analysis of human liver reference samples (French Reference liver + Huh7 hepatoma cells) achieves substantial human genome coverage via PeptideAtlas processing
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Has reproduction · 95
In vivo structural characterization of the SARS-CoV-2 RNA genome identifies host proteins vulnerable to repurposed drugs.
PMID 33636127 · PMC7871767 · Cell · 2021 · 8 claims · 8 setups
icSHAPE was used to determine the in vivo and in vitro structural landscape of the SARS-CoV-2 RNA genome in infected Huh7.5.1 cells, plus UTR structures of six other coronaviruses
-
Full-text index only
A pharmacogenetics study of the human glucuronosyltransferase UGT1A4.
PMID 19890225 · PMC6177227 · Pharmacogenetics and genomics · 2009 · 7 claims · 6 setups
Extensive sequencing of UGT1A4 (promoter to exon 1+2000bp) identified numerous novel polymorphisms: 13 intronic, 39 promoter, and 14 exonic variants (10 causing amino acid changes)
-
Full-text index only
Benchmarking ortholog identification methods using functional genomics data.
PMID 16613613 · PMC1557999 · Genome biology · 2006 · 8 claims · 7 setups
InParanoid is the best overall ortholog identification method for identifying functionally equivalent proteins when sensitivity and selectivity are combined into an overall score.
-
Has reproduction · 80
Transcriptome-Proteome Profiling in Burkholderia thailandensis during the Transition from Exponential to Stationary Phase.
PMID 40680064 · PMC12322963 · Journal of proteome research · 2025 · 8 claims · 7 setups
928 differentially accumulating mRNAs (564 up, 364 down) were identified between exponential and stationary phase