Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
A systematic comparative and structural analysis of protein phosphorylation sites based on the mtcPTM database.
PMID 17521420 · PMC1929158 · Genome biology · 2007 · 7 claims · 6 setups
mtcPTM is a hierarchically organized database of human and mouse phosphosites that preserves experimental context, enabling comparison of phosphorylation patterns across conditions
-
Full-text index only
MAZIE: a mass and charge inference engine to enhance database searching of tandem mass spectra.
PMID 19850495 · PMC2818324 · Journal of the American Society for Mass Spectrometry · 2010 · 7 claims · 4 setups
MAZIE is a post-acquisition Perl algorithm that determines precursor ion monoisotopic mass and charge (+1 to +4) from MS1 zoom scan isotopic distributions on a Thermo LTQ-XL
-
Full-text index only
Prediction of missed cleavage sites in tryptic peptides aids protein identification in proteomics.
PMID 17203985 · PMC2664920 · Journal of proteome research · 2007 · 8 claims · 4 setups
An information-theoretic log-likelihood scoring method can predict experimentally observed missed cleavage sites from amino acid sequence alone with up to 90% accuracy.
-
Has reproduction · 65
Estimating biodiversity across the tree of life on Mount Everest's southern flank with environmental DNA.
PMID 36148432 · PMC9486557 · iScience · 2022 · 8 claims · 6 setups
eDNA from ten high-alpine ponds and streams (4,500-5,500 m) on Mt. Everest's southern flank revealed 187 potential orders from 36 phyla across the Tree of Life.
-
Full-text index only
An integrated database of genes responsive to the Myc oncogenic transcription factor: identification of direct genomic targets.
PMID 14519204 · PMC328458 · Genome biology · 2003 · 8 claims · 6 setups
The Myc Target Gene database integrates literature evidence to prioritize candidate Myc-responsive genes and cluster them into functional groups
-
Full-text index only
Towards a comprehensive structural coverage of completed genomes: a structural genomics viewpoint.
PMID 17349043 · PMC1829165 · BMC bioinformatics · 2007 · 8 claims · 6 setups
A combined target-selection approach — pursuing both structurally uncharacterised domain families and additional targets from large structurally characterised superfamilies — is essential for comprehensive structural coverage of the genomes.
-
Full-text index only
The mammalian phenotype ontology: enabling robust annotation and comparative analysis.
PMID 20052305 · PMC2801442 · Wiley interdisciplinary reviews. Systems biology and medicine · 2009 · 8 claims · 6 setups
The Mammalian Phenotype (MP) Ontology enables classification and organization of phenotypic data for mouse and other mammalian species in a computationally useful, standardized manner.
-
Full-text index only
Structure of protein interaction networks and their implications on drug design.
PMID 19876376 · PMC2760708 · PLoS computational biology · 2009 · 8 claims · 6 setups
Budding yeast and human PINs are scale-rich and configured as highly optimized tolerance (HOT) networks similar to Internet router-level topology, rather than scale-free networks formed by preferential attachment.
-
Full-text index only
Cancer-specific high-throughput annotation of somatic mutations: computational prediction of driver missense mutations.
PMID 19654296 · PMC2763410 · Cancer research · 2009 · 7 claims · 7 setups
CHASM, a Random Forest-based computational method, was developed to identify and prioritize missense mutations likely to be functional drivers of tumor cell proliferation.
-
Full-text index only
PA-GOSUB: a searchable database of model organism protein sequences with their predicted Gene Ontology molecular function and subcellular localization.
PMID 15608166 · PMC540074 · Nucleic acids research · 2005 · 7 claims · 4 setups
PA-GOSUB significantly extends the coverage of GO molecular function and subcellular localization annotations for 10 model organism proteomes compared with existing databases (GOA, Swiss-Prot).
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.
-
Full-text index only
Pigs in sequence space: a 0.66X coverage pig genome survey based on shotgun sequencing.
PMID 15885146 · PMC1142312 · BMC genomics · 2005 · 8 claims · 7 setups
Pig sequence is closer to human than mouse is, across exons, UTRs, introns, intergenic regions, ultra-conserved elements, and miRNAs
-
Full-text index only
Expansion of the Bactericidal/Permeability Increasing-like (BPI-like) protein locus in cattle.
PMID 17362520 · PMC1839098 · BMC genomics · 2007 · 8 claims · 8 setups
The bovine BPI-like locus spans 470 kbp and contains 14 contiguous genes (13 intact + 1 pseudogene); 9 are orthologous to human/mouse BPI-like genes and 4 (named BSP30A, BSP30B, BSP30C, BSP30D) arose through cattle-specific duplication of the PSP gene
-
Full-text index only
Adaptive discriminant function analysis and reranking of MS/MS database search results for improved peptide identification in shotgun proteomics.
PMID 18788775 · PMC3744223 · Journal of proteome research · 2008 · 7 claims · 4 setups
PeptideProphet's fixed LDA coefficients for combining search scores (Xcorr', ΔCn, SpRank) may not be optimal under all search/instrument conditions.
-
Full-text index only
In silico promoters: modelling of cis-regulatory context facilitates target predictio.
PMID 18505473 · PMC3823354 · Journal of cellular and molecular medicine · 2009 · 8 claims · 8 setups
An integrated 'profiling of transcriptional targets' (PTT) strategy by Freebern et al. identified IGF-1 as a co-modulator of immune cell function genes in mitogen/drug-activated T cells.
-
Full-text index only
Metagenomic study of the oral microbiota by Illumina high-throughput sequencing.
PMID 19796657 · PMC3568755 · Journal of microbiological methods · 2009 · 8 claims · 6 setups
The 16S rRNA V5 hypervariable region, amplified as a short ~82-base segment, provides reliable taxonomic identification of oral bacteria against public databases like HOMD.
-
Has reproduction · 66
HTSstation: a web application and open-access libraries for high-throughput sequencing data analysis.
PMID 24475057 · PMC3903476 · PloS one · 2014 · 8 claims · 5 setups
HTSstation is a web application suite coupling simple web forms to modular analysis pipelines for ChIP-seq, RNA-seq, 4C-seq and re-sequencing HTS applications, accessible at http://htsstation.epfl.ch.
-
Has reproduction · 59
Downregulation of Splicing Factor PTBP1 Curtails FBXO5 Expression to Promote Cellular Senescence in Lung Adenocarcinoma.
PMID 39057099 · PMC11276454 · Current issues in molecular biology · 2024 · 8 claims · 8 setups
PTBP1 is significantly upregulated across multiple cancer types including LUAD, and higher PTBP1 levels are associated with worse LUAD patient survival
-
Has reproduction · 85
Exploring microproteins from various model organisms using the mip-mining database.
PMID 37919660 · PMC10623795 · BMC genomics · 2023 · 5 claims · 4 setups
Mip-mining is a database of 336 curated RNA-seq datasets from 8626 samples across nine species, built specifically to explore microprotein functions under stress and disease conditions