Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Metappuccino: large language model-driven reconstruction of sequence read archive metadata for cancer research.
PMID 42057294 · PMC13148957 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 3 setups
Metappuccino reconstructs 19 metadata classes by combining deterministic rule-based extraction/normalization (for explicit context) with LoRA-specialized Mistral-7B-Instruct completion (for missing/ambiguous fields)
-
Full-text index only
Genome assembly comparison identifies structural variants in the human genome.
PMID 17115057 · PMC2674632 · Nature genetics · 2006 · 7 claims · 7 setups
Genome assembly comparison is a robust approach for identifying all classes of genetic variation, with no lower size limit.
-
Full-text index only
An integrated database-pipeline system for studying single nucleotide polymorphisms and diseases.
PMID 19091018 · PMC2638159 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Existing SNP/disease databases are fragmented; no combined resource widely supports gene-, SNP-, and disease-related information together
-
Full-text index only
CaHoT-GRN: context-aware high-order topology learning for robust single-cell gene regulatory network inference.
PMID 42059479 · PMC13130071 · Briefings in bioinformatics · 2026 · 7 claims · 5 setups
CaHoT-GRN integrates pretrained biological language model embeddings (DNABERT for DNA, ESM for protein) with scRNA-seq expression data to improve GRN inference
-
Has reproduction · 59
De novo assembly of a transcriptome for Calanus finmarchicus (Crustacea, Copepoda)--the dominant zooplankter of the North Atlantic Ocean.
PMID 24586345 · PMC3929608 · PloS one · 2014 · 8 claims · 8 setups
A de novo transcriptome for Calanus finmarchicus was assembled from six developmental-stage libraries, yielding 206,041 contigs and a reference set of 96,090 unique comps, representing a new molecular resource for this species.
-
Full-text index only
EGenBio: a data management system for evolutionary genomics and biodiversity.
PMID 17118150 · PMC1683573 · BMC bioinformatics · 2006 · 7 claims · 7 setups
EGenBio is a web-based system for integrated management, filtering, curation, and visualization of large-scale genomic sequences, alignments, and phylogenetic trees for evolutionary genomics and biodiversity research.