Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Identification of diagnostic markers for tuberculosis by proteomic fingerprinting of serum.
PMID 16980117 · PMC7159276 · Lancet (London, England) · 2006 · 8 claims · 5 setups
An SVM classifier trained on serum proteomic profiles discriminated patients with active tuberculosis from controls with clinically overlapping conditions
-
Full-text index only
Vertebrate gene finding from multiple-species alignments using a two-level strategy.
PMID 16925840 · PMC1810555 · Genome biology · 2006 · 8 claims · 5 setups
DOGFISH cleanly separates a multi-species alignment classifier (RVM cascade) from an HMM-based structure predictor, avoiding tight coupling of alignment complexity with HMM formalism
-
Full-text index only
CARAT: a novel method for allelic detection of DNA copy number changes using high density oligonucleotide arrays.
PMID 16504045 · PMC1402331 · BMC bioinformatics · 2006 · 8 claims · 5 setups
CARAT is a novel algorithm that uses SNP probe intensity and genotype-based allelic dosage response in a regression framework to estimate allele-specific copy number genome-wide.
-
Full-text index only
LIMPIC: a computational method for the separation of protein MALDI-TOF-MS signals from noise.
PMID 17386085 · PMC1847688 · BMC bioinformatics · 2007 · 7 claims · 4 setups
LIMPIC is a computational method for detecting protein peaks from linear-mode MALDI-TOF-MS data using background noise reduction and baseline removal followed by non-uniform threshold peak detection and multi-spectra detection-rate classification.
-
Full-text index only
Clustering of phosphorylation site recognition motifs can be exploited to predict the targets of cyclin-dependent kinase.
PMID 17316440 · PMC1852407 · Genome biology · 2007 · 8 claims · 6 setups
CDK consensus motifs are frequently clustered (closely spaced) in known CDK substrate proteins rather than uniformly distributed
-
Full-text index only
High-throughput chromatin information enables accurate tissue-specific prediction of transcription factor binding sites.
PMID 18988630 · PMC2662491 · Nucleic acids research · 2009 · 8 claims · 8 setups
Incorporating H3K4me3 chromatin modification estimates greatly improves the accuracy of in silico prediction of in vivo TF binding for a wide range of TFs in human and mouse
-
Full-text index only
Proteomic profiling of urine identifies specific fragments of SERPINA1 and albumin as biomarkers of preeclampsia.
PMID 18984079 · PMC2679897 · American journal of obstetrics and gynecology · 2008 · 8 claims · 7 setups
Women with severe preeclampsia present a unique urinary proteomic fingerprint distinguishable from controls.
-
Full-text index only
Exhaustive prediction of disease susceptibility to coding base changes in the human genome.
PMID 18793467 · PMC2537574 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Inter-species conservation is the strongest single predictor of disease-associated coding mutations among the factors tested.
-
Full-text index only
Adaptive discriminant function analysis and reranking of MS/MS database search results for improved peptide identification in shotgun proteomics.
PMID 18788775 · PMC3744223 · Journal of proteome research · 2008 · 7 claims · 4 setups
PeptideProphet's fixed LDA coefficients for combining search scores (Xcorr', ΔCn, SpRank) may not be optimal under all search/instrument conditions.
-
Full-text index only
Global sequencing of proteolytic cleavage sites in apoptosis by specific labeling of protein N termini.
PMID 18722006 · PMC2566540 · Cell · 2008 · 7 claims · 8 setups
A subtiligase-based N-terminal biotinylation and enrichment method enables global identification and sequencing of protease cleavage sites in complex mixtures
-
Full-text index only
A biomarker panel for peripheral arterial disease.
PMID 18687758 · PMC3133945 · Vascular medicine (London, England) · 2008 · 7 claims · 6 setups
β2-microglobulin (β2M) and cystatin C had the highest correlation with ABI among all plasma markers tested, higher than age, smoking, or diabetes status
-
Full-text index only
Identification of deleterious non-synonymous single nucleotide polymorphisms using sequence-derived information.
PMID 18588693 · PMC2446391 · BMC bioinformatics · 2008 · 8 claims · 5 setups
A decision tree built on 10 selected sequence-derived features classifies SAPs as Disease or Polymorphism with 82.6% accuracy and 0.607 MCC in cross-validation.
-
Has reproduction · 58
A comparative study of techniques for differential expression analysis on RNA-Seq data.
PMID 25119138 · PMC4132098 · PloS one · 2014 · 8 claims · 8 setups
edgeR performs slightly better than DESeq and Cuffdiff2 in terms of the ability to uncover true positives.
-
Full-text index only
Prediction-based approaches to characterize bidirectional promoters in the mammalian genome.
PMID 18366609 · PMC2386062 · BMC genomics · 2008 · 8 claims · 7 setups
The mapping algorithm identified 5,647 candidate bidirectional promoter regions in the mouse genome, similar in number to those previously found in human.
-
Full-text index only
Discovery and identification of potential biomarkers of papillary thyroid carcinoma.
PMID 19785722 · PMC2761863 · Molecular cancer · 2009 · 8 claims · 7 setups
A 3-peak (m/z 9190, 6631, 8697 Da) SVM classification model discriminates PTC from non-cancer controls with high sensitivity and specificity
-
Full-text index only
Mining novel biomarkers for prognosis of gastric cancer with serum proteomics.
PMID 19740432 · PMC2753349 · Journal of experimental & clinical cancer research : CR · 2009 · 7 claims · 4 setups
A 5-peak prognosis pattern (4474, 4542, 6443/6643, 4988, 6685 Da) predicts poor vs good prognosis in GC with higher sensitivity/specificity than CEA and TNM stage
-
Full-text index only
Network-assisted protein identification and data interpretation in shotgun proteomics.
PMID 19690572 · PMC2736651 · Molecular systems biology · 2009 · 7 claims · 7 setups
Confidently identified proteins in a sample form tightly connected sub-networks in the protein interaction network, with significantly higher clustering coefficients than random or topology-matched random sub-networks.
-
Full-text index only
A proteomic strategy to identify novel serum biomarkers for liver cirrhosis and hepatocellular cancer in individuals with fatty liver disease.
PMID 19656391 · PMC2729079 · BMC cancer · 2009 · 8 claims · 4 setups
2D gel proteomic profiling of immunodepleted serum reveals protein spot patterns that differentiate pre-cirrhotic NAFLD, cirrhotic NAFLD, and cirrhotic NAFLD with HCC
-
Full-text index only
Revealing potential interfering genes between abdominal aortic aneurysm and periodontitis through machine learning and bioinformatics analysis.
PMID 40857330 · PMC12380325 · PloS one · 2025 · 6 claims · 8 setups
90 interacting genes were identified at the intersection of AAA DEGs and periodontitis WGCNA module genes
-
Full-text index only
CIRCE: a scalable Python package to predict cis-regulatory DNA interactions from single-cell chromatin accessibility data.
PMID 41734268 · PMC12987762 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 5 setups
CIRCE re-implements the Cicero co-accessibility algorithm in Python, producing near-identical results while running much faster and using far less memory