Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 90
pysradb: A Python package to query next-generation sequencing metadata and data from NCBI Sequence Read Archive.
PMID 31114675 · PMC6505635 · F1000Research · 2019 · 6 claims · 7 setups
pysradb provides a simple, user-friendly command-line interface for querying metadata and downloading datasets from SRA without requiring knowledge of a programming language.
-
Full-text index only
Gene-disease relationship discovery based on model-driven data integration and database view definition.
PMID 19042916 · PMC2639000 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 4 setups
Explicit gene–disease relationships can be formulated as candidate gene definitions (e.g., co-localization, dysregulation, functional similarity) that may include intermediary orthologous or interacting genes
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Has reproduction · 62
scATD: a high-throughput and interpretable framework for single-cell cancer drug resistance prediction and biomarker identification.
PMID 40501071 · PMC12159290 · Briefings in bioinformatics · 2025 · 8 claims · 6 setups
scATD enables high-throughput single-cell drug sensitivity prediction for new patients without model parameter retraining via bidirectional Bi-AdaIN style transfer
-
Has reproduction · 64
GeMI: interactive interface for transformer-based Genomic Metadata Integration.
PMID 35657113 · PMC9216561 · Database : the journal of biological databases and curation · 2022 · 8 claims · 5 setups
GeMI is a web tool that uses a fine-tuned GPT2 model to extract 15 structured key-value attributes from free-text GEO sample metadata.
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Full-text index only
Minisequencing mitochondrial DNA pathogenic mutations.
PMID 18402672 · PMC2377236 · BMC medical genetics · 2008 · 7 claims · 7 setups
A minisequencing multiplex assay can interrogate 25 pathogenic mtDNA mutations across the whole mtDNA genome in a single reaction using 13 amplicons.
-
Full-text index only
Structural genomics: a new era for pharmaceutical research.
PMID 11864367 · PMC139010 · Genome biology · 2002 · 8 claims · 6 setups
Automation of crystal mounting, diffraction data collection, and structure determination/refinement is a critical factor enabling large-scale structural genomics projects.
-
Full-text index only
The PeptideAtlas project.
PMID 16381952 · PMC1347403 · Nucleic acids research · 2006 · 8 claims · 5 setups
PeptideAtlas provides an automated repository that identifies peptides by MS/MS, statistically validates identifications, and maps them to eukaryotic genomes to enable data exchange and integration with genomic data.