Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Finding disease candidate genes by liquid association.
PMID 17915034 · PMC2246280 · Genome biology · 2007 · 7 claims · 6 setups
LA can detect functionally associated genes that are not directly co-expressed by identifying a mediating gene Z whose expression level changes the correlation between X and Y.
-
Full-text index only
An integrated database of genes responsive to the Myc oncogenic transcription factor: identification of direct genomic targets.
PMID 14519204 · PMC328458 · Genome biology · 2003 · 8 claims · 6 setups
The Myc Target Gene database integrates literature evidence to prioritize candidate Myc-responsive genes and cluster them into functional groups
-
Full-text index only
Dynamic Proteomics: a database for dynamics and localizations of endogenous fluorescently-tagged proteins in living human cells.
PMID 19820112 · PMC2808965 · Nucleic acids research · 2010 · 8 claims · 6 setups
The Dynamic Proteomics database compiles fluorescence dynamics and localization data for endogenously YFP/Venus-tagged human proteins from the LARC library studied by Cohen et al.
-
Has reproduction
DAGFormer: A graph-based domain adaptation approach for single-cell cancer drug response prediction.
PMID 41417875 · PMC12795466 · PLoS computational biology · 2025 · 7 claims · 4 setups
DAGFormer, a graph-based domain adaptation framework integrating bulk and scRNA-seq data, predicts single-cell drug responses more accurately than existing methods.
-
Has reproduction · 81
Enabling Single-Cell Drug Response Annotations from Bulk RNA-Seq Using SCAD.
PMID 36762572 · PMC10104628 · Advanced science (Weinheim, Baden-Wurttemberg, Germany) · 2023 · 7 claims · 7 setups
SCAD, a transfer learning framework integrating adversarial discriminative domain adaptation (ADDA), can infer single-cell drug sensitivities by transferring knowledge from bulk RNA-seq pharmacogenomic data (GDSC) to scRNA-seq target domains
-
Full-text index only
Integration with the human genome of peptide sequences obtained by high-throughput mass spectrometry.
PMID 15642101 · PMC549070 · Genome biology · 2005 · 8 claims · 4 setups
PeptideAtlas, a public database integrating MS/MS-derived peptide identifications with the human genome, was built as an expandable resource for proteomic data.
-
Full-text index only
CancerGenes: a gene selection resource for cancer genome projects.
PMID 17088289 · PMC1781153 · Nucleic acids research · 2007 · 6 claims · 4 setups
CancerGenes is a gene list-centric web resource that combines expert-annotated gene lists with data from public databases (Entrez Gene, Ensembl BioMart, Kim et al. promoter data, Sanger COSMIC) to support gene selection for cancer re-sequencing projects.
-
Full-text index only
PubMatrix: a tool for multiplex literature mining.
PMID 14667255 · PMC317283 · BMC bioinformatics · 2003 · 8 claims · 3 setups
PubMatrix is a web-based CGI tool that queries PubMed with two lists of terms (search terms vs modifier terms) and returns a matrix of pairwise co-occurrence frequency counts
-
Has reproduction · 53
PulmonDB: a curated lung disease gene expression database.
PMID 31949184 · PMC6965635 · Scientific reports · 2020 · 6 claims · 6 setups
PulmonDB is a curated, web-based gene expression database and R package integrating microarray and RNA-seq data for COPD and IPF with manually curated controlled-vocabulary annotation.
-
Has reproduction · 50
Exploiting convergent phenotypes to derive a pan-cancer cisplatin response gene expression signature.
PMID 37076665 · PMC10115855 · NPJ precision oncology · 2023 · 8 claims · 8 setups
A convergent-phenotype-based seed gene/co-expression method can extract consensus gene expression signatures predictive of response to chemotherapeutic drugs in the GDSC database
-
Has reproduction · 82
Landscape of allele-specific transcription factor binding in the human genome.
PMID 33980847 · PMC8115691 · Nature communications · 2021 · 8 claims · 6 setups
A novel statistical framework (ADASTRA) calls allele-specific TF binding from existing ChIP-Seq alignments by jointly correcting for background allelic dosage (BAD, from aneuploidy/CNVs) and reference mapping bias.
-
Full-text index only
Bases and spaces: resources on the web for accessing the draft human genome.
PMID 11178254 · PMC138875 · Genome biology · 2000 · 8 claims · 8 setups
By combining currently available genomic databases and mapping resources (GenBank/Entrez, UniGene, RH maps, BAC fingerprint maps, Ensembl, NIX), it is possible to devise strategies that fully exploit the fragmentary draft human genome sequence.
-
Full-text index only
Oncogene mutations, copy number gains and mutant allele specific imbalance (MASI) frequently occur together in tumor cells.
PMID 19826477 · PMC2757721 · PloS one · 2009 · 8 claims · 8 setups
Homozygous mutations of oncogenes are frequent (20%) across 833 cancer cell lines of 12 tumor types in the Sanger database
-
Has reproduction · 90
pysradb: A Python package to query next-generation sequencing metadata and data from NCBI Sequence Read Archive.
PMID 31114675 · PMC6505635 · F1000Research · 2019 · 6 claims · 7 setups
pysradb provides a simple, user-friendly command-line interface for querying metadata and downloading datasets from SRA without requiring knowledge of a programming language.
-
Has reproduction · 81
Identification of Proteins Deregulated by Platinum-Based Chemotherapy as Novel Biomarkers and Therapeutic Targets in Non-Small Cell Lung Cancer.
PMID 33777753 · PMC7991912 · Frontiers in oncology · 2021 · 7 claims · 8 setups
Cisplatin exposure induces significant deregulation of protein expression networks in NSCLC cells
-
Full-text index only
In silico and in vivo splicing analysis of MLH1 and MSH2 missense mutations shows exon- and tissue-specific effects.
PMID 16995940 · PMC1590028 · BMC genomics · 2006 · 8 claims · 6 setups
In silico ESE-prediction algorithms (ESEfinder, RescueESE, PESX) do not reliably predict actual in vivo splicing behavior of missense mutations
-
Full-text index only
X:Map: annotation and visualization of genome structure for Affymetrix exon array analysis.
PMID 17932061 · PMC2238884 · Nucleic acids research · 2008 · 7 claims · 4 setups
X:Map is a genome annotation database that maps every Affymetrix exon array probeset to Ensembl genome features (genes, ESTs, GenScan predictions) and supports both high-throughput and gene-centric analysis.
-
Full-text index only
Frameshift mutations in coding repeats of protein tyrosine phosphatase genes in colorectal tumors with microsatellite instability.
PMID 19000305 · PMC2586028 · BMC cancer · 2008 · 7 claims · 6 setups
16 PTP candidate genes containing coding mononucleotide repeats (cMNR) of at least 7 units were identified via bioinformatic analysis and screened in MSI-H cell lines, cancers, and adenomas
-
Full-text index only
Using proteomic approach to identify tumor-associated antigens as markers in hepatocellular carcinoma.
PMID 18672925 · PMC2680441 · Journal of proteome research · 2008 · 8 claims · 6 setups
34 immunoreactive protein spots were detected in 2DE Western blots probed with HCC patient sera but not normal sera
-
Full-text index only
Characterization of the human DYRK1A promoter and its regulation by the transcription factor E2F1.
PMID 18366763 · PMC2292204 · BMC molecular biology · 2008 · 8 claims · 8 setups
Transcription start sites of human DYRK1A are distributed over an 800 bp region within an unmethylated CpG island