Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 84
Pharokka: a fast scalable bacteriophage annotation tool.
PMID 36453861 · PMC9805569 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 5 setups
Pharokka is a one-line, fast, scalable bacteriophage annotation tool producing standards-compliant outputs, installable via a two-line bioconda command
-
Full-text index only
Dynamic Proteomics: a database for dynamics and localizations of endogenous fluorescently-tagged proteins in living human cells.
PMID 19820112 · PMC2808965 · Nucleic acids research · 2010 · 8 claims · 6 setups
The Dynamic Proteomics database compiles fluorescence dynamics and localization data for endogenously YFP/Venus-tagged human proteins from the LARC library studied by Cohen et al.
-
Full-text index only
An integrated database of genes responsive to the Myc oncogenic transcription factor: identification of direct genomic targets.
PMID 14519204 · PMC328458 · Genome biology · 2003 · 8 claims · 6 setups
The Myc Target Gene database integrates literature evidence to prioritize candidate Myc-responsive genes and cluster them into functional groups
-
Full-text index only
Grammar-based distance in progressive multiple sequence alignment.
PMID 18616828 · PMC2478692 · BMC bioinformatics · 2008 · 7 claims · 3 setups
A grammar-based (LZ complexity) distance metric can be used to determine the order in which sequences are progressively pairwise aligned
-
Full-text index only
GeneKeyDB: a lightweight, gene-centric, relational database to support data mining environments.
PMID 15790402 · PMC1274265 · BMC bioinformatics · 2005 · 8 claims · 6 setups
GeneKeyDB is a lightweight, gene-centric relational database that supports data mining and integration with computational analysis tools.
-
Full-text index only
SpliceMiner: a high-throughput database implementation of the NCBI Evidence Viewer for microarray splice variant analysis.
PMID 17338820 · PMC1839109 · BMC bioinformatics · 2007 · 6 claims · 4 setups
EVDB is a comprehensive, non-redundant relational database of known human splice variants built from NCBI Entrez Gene and Evidence Viewer data
-
Full-text index only
CYCLONET--an integrated database on cell cycle regulation and carcinogenesis.
PMID 17202170 · PMC1899094 · Nucleic acids research · 2007 · 7 claims · 4 setups
Cyclonet is a web-based integrated database combining 'omics' and chemoinformatics data on mammalian cell cycle regulation in normal and pathological (cancer) states, built on a systems biology approach.
-
Has reproduction · 74
Evaluation of classification and forecasting methods on time series gene expression data.
PMID 33156855 · PMC7647064 · PloS one · 2020 · 8 claims · 4 setups
Deep learning based methods generally outperform traditional approaches for time series gene expression classification.
-
Full-text index only
Comparison of complete nuclear receptor sets from the human, Caenorhabditis elegans and Drosophila genomes.
PMID 11532213 · PMC55326 · Genome biology · 2001 · 7 claims · 5 setups
The human genome contains fewer than 50 functional nuclear receptors, far fewer than C. elegans and about twice as many as Drosophila
-
Full-text index only
In silico and in vitro comparative analysis to select, validate and test SNPs for human identification.
PMID 18076761 · PMC2222643 · BMC genomics · 2007 · 8 claims · 7 setups
A panel of 24 SNPs was selected and validated for human identification using 1,040 unrelated samples from three populations (Italian, Benin Gulf, Mongolian)
-
Full-text index only
Assignment of Streptococcus agalactiae isolates to clonal complexes using a small set of single nucleotide polymorphisms.
PMID 18710585 · PMC2533671 · BMC microbiology · 2008 · 7 claims · 6 setups
A four-SNP set (glnA36, glnA429, glcK180, adhP111) identified via the Not-N algorithm plus empirical testing divides GBS into 10 groups concordant with eBURST-defined population structure.
-
Full-text index only
The revolution of the biology of the genome.
PMID 15040884 · PMC7091781 · Cell research · 2004 · 8 claims · 6 setups
Polyploidization and gene duplication are the major mechanisms increasing eukaryotic genome size.
-
Full-text index only
Human Lsg1 defines a family of essential GTPases that correlates with the evolution of compartmentalization.
PMID 16209721 · PMC1262696 · BMC biology · 2005 · 8 claims · 9 setups
hLsg1 is the human orthologue of yeast Lsg1p and defines a family of circularly permuted GTPases named YRG (YlqF Related GTPases)
-
Full-text index only
Ab initio identification of putative human transcription factor binding sites by comparative genomics.
PMID 15865625 · PMC1097714 · BMC bioinformatics · 2005 · 8 claims · 5 setups
An integrated algorithm combining human-mouse genomic comparison, motif overrepresentation, and coregulation filters (GO annotation and microarray coexpression) can identify candidate transcription factor binding sites genome-wide
-
Full-text index only
The 10 sea urchin receptor for egg jelly proteins (SpREJ) are members of the polycystic kidney disease-1 (PKD1) family.
PMID 17629917 · PMC1934368 · BMC genomics · 2007 · 8 claims · 5 setups
Sea urchins possess 10 SpREJ (PKD1 family) genes, compared to five in humans, all defined by possession of a ~600 residue REJ domain
-
Has reproduction · 59
Downregulation of Splicing Factor PTBP1 Curtails FBXO5 Expression to Promote Cellular Senescence in Lung Adenocarcinoma.
PMID 39057099 · PMC11276454 · Current issues in molecular biology · 2024 · 8 claims · 8 setups
PTBP1 is significantly upregulated across multiple cancer types including LUAD, and higher PTBP1 levels are associated with worse LUAD patient survival
-
Has reproduction · 74
Wide-Open: Accelerating public data release by automating detection of overdue datasets.
PMID 28594819 · PMC5464523 · PLoS biology · 2017 · 6 claims · 5 setups
A general text-mining + API-query approach (Wide-Open) can automatically identify datasets that are overdue for public release in a repository
-
Full-text index only
Integrated proteomic and transcriptomic profiling of mouse lung development and Nmyc target genes.
PMID 17486137 · PMC2673710 · Molecular systems biology · 2007 · 8 claims · 7 setups
Global MudPIT-based proteomic profiling across six mouse lung developmental time points (E13.5–P56) identifies thousands of proteins and captures developmental/cell-biological expression patterns.
-
Full-text index only
Oncogene mutations, copy number gains and mutant allele specific imbalance (MASI) frequently occur together in tumor cells.
PMID 19826477 · PMC2757721 · PloS one · 2009 · 8 claims · 8 setups
Homozygous mutations of oncogenes are frequent (20%) across 833 cancer cell lines of 12 tumor types in the Sanger database
-
Full-text index only
Predicting the protein interaction landscape of a free-living bacterium with pooled-AlphaFold3.
PMID 41559189 · PMC13047044 · Molecular systems biology · 2026 · 8 claims · 6 setups
Pooled-AlphaFold3 prediction improves accuracy of genome-scale PPI screens compared to a paired approach while reducing inference time (~2-fold) and job count (~100-fold)