Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Functional nsSNPs from carcinogenesis-related genes expressed in breast tissue: potential breast cancer risk alleles and their distribution across human populations.
PMID 16595073 · PMC3500178 · Human genomics · 2006 · 7 claims · 5 setups
A bioinformatics strategy cross-referencing carcinogenesis-related gene lists with breast-tissue expression data can identify candidate breast cancer risk nsSNPs.
-
Full-text index only
Benchmarking methods for genome annotation using nanopore direct RNA in a non-model crop plant.
PMID 41800382 · PMC12967217 · Bioinformatics advances · 2026 · 6 claims · 8 setups
Annotation tools show substantial variation in isoform detection, structural completeness, splicing classification, and handling of 5' read truncation when applied to plant dRNA-seq data.
-
Full-text index only
MetaPepticon: automated prediction of anticancer peptides from microbial genomes and metagenomes.
PMID 41918857 · PMC13034871 · PeerJ · 2026 · 7 claims · 6 setups
MetaPepticon is a modular, end-to-end Snakemake pipeline that predicts ACP candidates directly from raw genomic, metagenomic, transcriptomic, metatranscriptomic reads, assembled contigs, or peptide sequences.
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
The use of edge-betweenness clustering to investigate biological function in protein interaction networks.
PMID 15740614 · PMC555937 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Edge-Betweenness clustering separates protein interaction graphs into subgraphs whose GO term distributions show significant correlations, revealing biologically meaningful functional modules.
-
Full-text index only
A DNA microarray survey of gene expression in normal human tissues.
PMID 15774023 · PMC1088941 · Genome biology · 2005 · 6 claims · 6 setups
Unsupervised hierarchical clustering of gene expression groups normal tissue samples largely according to anatomic location, cellular composition, or physiologic function.
-
Full-text index only
Dissecting microregulation of a master regulatory network.
PMID 18294391 · PMC2289817 · BMC genomics · 2008 · 8 claims · 6 setups
143 human miRNAs (termed p53-miRs) each contain at least one putative p53 binding site within 10 kb flanking sequence and are predicted to target at least one known gene
-
Full-text index only
MODBASE: a database of annotated comparative protein structure models and associated resources.
PMID 16381869 · PMC1347422 · Nucleic acids research · 2006 · 8 claims · 7 setups
MODBASE is a database of automatically calculated comparative protein structure models covering all UniProt sequences matchable to a known structure
-
Has reproduction · 49
Aberration in DNA methylation in B-cell lymphomas has a complex origin and increases with disease severity.
PMID 23326238 · PMC3542081 · PLoS genetics · 2013 · 8 claims · 8 setups
B-cell non-Hodgkin lymphomas display striking intra-tumor (intra-sample) and inter-patient (inter-sample) cytosine methylation heterogeneity that increases progressively with disease aggressiveness (NBC<NGC<FL<GCB<ABC).
-
Full-text index only
Predicting failure rate of PCR in large genomes.
PMID 18492719 · PMC2441781 · Nucleic acids research · 2008 · 7 claims · 8 setups
The number of predicted primer-binding sites in genomic DNA is the most important factor determining PCR failure.
-
Full-text index only
Uncovering Cas9 PAM diversity through metagenomic mining and machine learning.
PMID 41656299 · PMC12996302 · Nature communications · 2026 · 8 claims · 6 setups
CRISPR-PAMdb is a publicly accessible database compiling Cas9 protein sequences from 3.8 million bacterial/archaeal genomes and PAM profiles from 7.4 million phage/plasmid sequences
-
Has reproduction · 71
Interpretable artificial intelligence based on immunoregulation-related genes predicts prognosis and immunotherapy response in lung adenocarcinoma.
PMID 41048340 · PMC12491262 · Frontiers in bioinformatics · 2025 · 8 claims · 8 setups
IRG expression pattern clusters LUAD patients into groups with significantly different survival outcomes and immune cell infiltration
-
Has reproduction · 38
RNA-Seq transcriptome profiling of upland cotton (Gossypium hirsutum L.) root tissue under water-deficit stress.
PMID 24324815 · PMC3855774 · PloS one · 2013 · 8 claims · 8 setups
A total of 1,530 transcripts were differentially expressed between well-watered and water-deficit stressed field-grown upland cotton root tissues (913 up-regulated, 617 down-regulated).