Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 83
Analyzing biomarker discovery: Estimating the reproducibility of biomarker sets.
PMID 35901020 · PMC9333302 · PloS one · 2022 · 8 claims · 6 setups
Non-reproducibility is a common problem in biomarker discovery: independently derived biomarker sets for the same phenotype frequently share very few features across studies
-
Full-text index only
CLEAN: CLustering Enrichment ANalysis.
PMID 19640299 · PMC2734555 · BMC bioinformatics · 2009 · 8 claims · 4 setups
The gene-specific CLEAN score improves reproducibility of cluster analysis conclusions across independent datasets compared to the traditional cluster-wide score (cwCLEAN).
-
Has reproduction · 50
Implementing the reuse of public DIA proteomics datasets: from the PRIDE database to Expression Atlas.
PMID 35701420 · PMC9197839 · Scientific data · 2022 · 8 claims · 5 setups
An open, containerised, Nextflow-orchestrated reanalysis pipeline combining metadata annotation, SWATH-MS analysis, statistical analysis, and Expression Atlas integration was developed for public DIA data.
-
Full-text index only
Development of proteomic patterns for detecting lung cancer.
PMID 14757945 · PMC3851077 · Disease markers · 2003 · 8 claims · 3 setups
A decision tree classification algorithm built on three serum protein mass peaks (8122Da, 1452Da, 1610Da) can discriminate lung cancer patients from healthy controls
-
Full-text index only
Comparison of multidimensional shotgun technologies targeting tissue proteomics.
PMID 19960471 · PMC3465977 · Electrophoresis · 2009 · 6 claims · 4 setups
CITP-based multidimensional separation achieves superior overall proteome performance (more total peptide, distinct peptide, and distinct protein identifications) than SCX/MuDPIT under matched conditions
-
Full-text index only
A novel wavelet-based thresholding method for the pre-processing of mass spectrometry data that accounts for heterogeneous noise.
PMID 18615428 · PMC2855839 · Proteomics · 2008 · 6 claims · 4 setups
Noise in SELDI-TOF/MALDI-TOF mass spectrometry data is heteroscedastic across the m/z range, with larger variance at lower m/z values, contrary to the homogeneous noise assumption of existing wavelet denoising methods.
-
Full-text index only
Human and mouse oligonucleotide-based array CGH.
PMID 16361265 · PMC1316119 · Nucleic acids research · 2005 · 8 claims · 8 setups
Oligo array CGH detects single copy gains, multi-copy amplifications, and homozygous/heterozygous deletions as small as 100 kb
-
Has reproduction · 89
Graph Random Forest: A Graph Embedded Algorithm for Identifying Highly Connected Important Features.
PMID 37509188 · PMC10377046 · Biomolecules · 2023 · 8 claims · 6 setups
GRF identifies effective features that form highly connected sub-graphs on the underlying biological network
-
Full-text index only
Sample preparation for serum/plasma profiling and biomarker identification by mass spectrometry.
PMID 17166507 · PMC7094463 · Journal of chromatography. A · 2007 · 8 claims · 8 setups
Standardizing sample preparation procedures for serum/plasma profiling is critical for obtaining reliable biomarkers, since slight procedural changes can produce very different protein profiles.