Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
SETDB1 enables development beyond cleavage stages by extinguishing the MERVL-driven two-cell totipotency transcriptional program in the mouse embryo.
PMID 41697236 · PMC12908936 · eLife · 2026 · 8 claims · 5 setups
Maternal SETDB1 is required for mouse preimplantation development beyond the eight-cell stage
-
Full-text index only
Distinct radial glia subtypes regulate midbrain dopaminergic neuron development.
PMID 41699318 · PMC13061605 · Nature neuroscience · 2026 · 8 claims · 8 setups
Rgl1 is the progenitor of the mesDA neuronal lineage
-
Full-text index only
Single-cell epigenetic profiling reveals a tumor-intrinsic interferon response program in ccRCC tied to poor prognosis and BAP1 loss.
PMID 41719400 · PMC12922754 · Science advances · 2026 · 8 claims · 8 setups
Subclustering of ccRCC tumor cells reveals four shared epigenetic programs (C0-C3) recurrent across patients, cohorts, and disease stages
-
Has reproduction · 71
polishCLR: A Nextflow Workflow for Polishing PacBio CLR Genome Assemblies.
PMID 36792366 · PMC9985148 · Genome biology and evolution · 2023 · 8 claims · 8 setups
polishCLR is a reproducible, containerized Nextflow workflow that implements best practices for polishing PacBio CLR genome assemblies.
-
Full-text index only
SGCRNA: spectral clustering-guided co-expression network analysis without scale-free constraints for multi-omic data.
PMID 41615289 · PMC12856952 · Briefings in bioinformatics · 2026 · 8 claims · 8 setups
WGCNA's reliance on a scale-free topology assumption is problematic because real co-expression networks do not consistently exhibit scale-free properties
-
Has reproduction · 98
Uncertainty in the mating strategy of honeybees causes bias and unreliability in the estimates of genetic parameters.
PMID 38632535 · PMC11022492 · Genetics, selection, evolution : GSE · 2024 · 7 claims · 3 setups
The most precise estimates of genetic parameters and genetic trends are obtained when breeding queens are mated with drones of a single DPQ that is correctly assigned in the pedigree (SS mating).
-
Full-text index only
Advancement of biomarker discovery and validation through the HUPO plasma proteome project.
PMID 15502245 · PMC3839274 · Disease markers · 2004 · 7 claims · 3 setups
Standardization of specimen collection, handling, storage, and choice of serum vs. plasma/anticoagulant is essential for comparable proteomic biomarker discovery.
-
Full-text index only
An SVM-based system for predicting protein subnuclear localizations.
PMID 16336650 · PMC1325059 · BMC bioinformatics · 2005 · 7 claims · 3 setups
New kernels defined on k-peptide vectors mapped by BLOSUM62-based high-scored pair matrices (D1, D2, D3) improve SVM discrimination of protein subnuclear localization compared to conventional k-peptide encodings.
-
Full-text index only
Prediction of candidate primary immunodeficiency disease genes using a support vector machine learning approach.
PMID 19801557 · PMC2780952 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2009 · 6 claims · 3 setups
An SVM trained on 69 binary features of known PID genes can accurately classify PID vs non-PID genes and predict novel candidate PID genes
-
Full-text index only
Limitations in SELDI-TOF MS whole serum proteomic profiling with IMAC surface to specifically detect colorectal cancer.
PMID 19689818 · PMC2743709 · BMC cancer · 2009 · 7 claims · 3 setups
The previously reported classifier (m/z 8,132 and 4,002) failed to discriminate CRC patients from healthy volunteers in this independent validation cohort
-
Has reproduction · 83
Accurate prediction of metagenome-assembled genome completeness by MAGISTA, a random forest model built on alignment-free intra-bin statistics.
PMID 35248155 · PMC8898458 · Environmental microbiome · 2022 · 7 claims · 7 setups
MAGISTA, a random forest model built on alignment-free intra-bin distance-distribution statistics, can estimate MAG completeness and purity without relying on reference marker genes.
-
Full-text index only
Toward stem cell systems biology: from molecules to networks and landscapes.
PMID 19329576 · PMC2738746 · Cold Spring Harbor symposia on quantitative biology · 2008 · 7 claims · 6 setups
Stem-cell-fate specification is an extremely complex process regulated by multiple mutually interacting molecular mechanisms with numerous regulatory feedback loops.
-
Full-text index only
Testing groups of genomic locations for enrichment in disease loci using linkage scan data: a method for hypothesis testing.
PMID 16848972 · PMC3525155 · Human genomics · 2006 · 8 claims · 2 setups
A method testing enrichment of a group of genomic locations for disease loci by comparing the average NPL score of the group to a null distribution from randomly drawn groups of equal size
-
Has reproduction · 89
MirDIP 5.2: tissue context annotation and novel microRNA curation.
PMID 36453996 · PMC9825511 · Nucleic acids research · 2023 · 7 claims · 6 setups
mirDIP 5.2 removed eight outdated resources, added miRNATIP, and ran five prediction algorithms against miRBase and mirGeneDB miRNAs to expand and improve interaction coverage
-
Full-text index only
Splicing bioinformatics to biology.
PMID 16732900 · PMC1779529 · Genome biology · 2006 · 8 claims · 8 setups
Mutually exclusive selection of Dscam exon 6 variants is governed by base pairing between a conserved intronic docking site and selector sequences adjacent to each alternative exon.
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Full-text index only
The UCSC Genome Browser Database: 2008 update.
PMID 18086701 · PMC2238835 · Nucleic acids research · 2008 · 8 claims · 8 setups
The UCSC Genome Browser Database (GBD) provides integrated sequence and annotation data for a large collection of vertebrate and model organism genomes.
-
Full-text index only
COSMIC (the Catalogue of Somatic Mutations in Cancer): a resource to investigate acquired mutations in human cancer.
PMID 19906727 · PMC2808858 · Nucleic acids research · 2010 · 8 claims · 6 setups
COSMIC is the largest public resource for information on somatically acquired mutations in human cancer, freely available without restriction
-
Has reproduction · 90
A2TEA: Identifying trait-specific evolutionary adaptations.
PMID 37224329 · PMC10186066 · F1000Research · 2022 · 8 claims · 7 setups
A2TEA integrates gene family expansion analysis with differential expression data across species to identify genes that were targets of evolutionary adaptation to a given stress/treatment
-
Has reproduction · 80
PanglaoDB: a web server for exploration of mouse and human single-cell RNA sequencing data.
PMID 30951143 · PMC6450036 · Database : the journal of biological databases and curation · 2019 · 7 claims · 7 setups
PanglaoDB is a web server providing pre-processed and pre-computed analyses of published mouse and human scRNA-seq experiments through a user-friendly interface.