Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Motif discovery in promoters of genes co-localized and co-expressed during myeloid cells differentiation.
PMID 19059999 · PMC2632922 · Nucleic acids research · 2009 · 6 claims · 8 setups
A novel multi-step computational method (built on approximate pattern enumeration, binomial over-representation scoring with FDR correction, and k-medoids clustering) can identify over-represented motifs in a selected set of promoters relative to a background promoter set.
-
Full-text index only
Short tandem repeats in human exons: a target for disease mutations.
PMID 18789129 · PMC2543027 · BMC genomics · 2008 · 8 claims · 6 setups
STRs are present in exons of 92% of known human genes, unlike longer tandem repeats which are rare in exons
-
Full-text index only
Identification and analysis of co-occurrence networks with NetCutter.
PMID 18781200 · PMC2526157 · PloS one · 2008 · 8 claims · 4 setups
Random sampling from a complete permutation set of the bipartite graph permits co-occurrence analysis with optimal stringency, and the edge-swapping (ES) model closely approximates this and is the preferred null-model among six tested.
-
Has reproduction · 90
PrimerSeq: Design and visualization of RT-PCR primers for alternative splicing using RNA-seq data.
PMID 24747190 · PMC4411361 · Genomics, proteomics & bioinformatics · 2014 · 8 claims · 3 setups
PrimerSeq is a user-friendly stand-alone software with a GUI for systematic design and visualization of RT-PCR primers for alternative splicing analysis using user-provided RNA-seq data.
-
Full-text index only
QuadBase: genome-wide database of G4 DNA--occurrence and conservation in human, chimpanzee, mouse and rat promoters and 146 microbes.
PMID 17962308 · PMC2238983 · Nucleic acids research · 2008 · 8 claims · 3 setups
QuadBase is a compendium of G4 DNA (quadruplex) motifs focused on their occurrence and conservation in promoters, composed of EuQuad and ProQuad
-
Full-text index only
Development of lead hammerhead ribozyme candidates against human rod opsin mRNA for retinal degeneration therapy.
PMID 19094986 · PMC3388947 · Experimental eye research · 2009 · 8 claims · 3 setups
Three lead hhRz candidates (CUC↓266, CUC↓1411, AUA↓1414) significantly knock down human RHO protein expression relative to control (p<0.05)
-
Full-text index only
Functional copy-number alterations in cancer.
PMID 18784837 · PMC2527508 · PloS one · 2008 · 8 claims · 3 setups
RAE is a comprehensive computational framework that robustly maps chromosomal alterations in tumor samples and statistically assesses their functional importance in cancer.
-
Full-text index only
PBAT: a comprehensive software package for genome-wide association analysis of complex family-based studies.
PMID 15814068 · PMC3525120 · Human genomics · 2005 · 8 claims · 1 setups
PBAT provides comprehensive tools for family-based association analysis, including nuclear families with missing parental genotypes, extended pedigrees, SNP and haplotype analysis, quantitative/qualitative/multivariate/longitudinal traits and time-to-onset phenotypes
-
Full-text index only
The relationship of potential G-quadruplex sequences in cis-upstream regions of the human genome to SP1-binding elements.
PMID 18353860 · PMC2377421 · Nucleic acids research · 2008 · 7 claims · 1 setups
A large number of upstream PQSSs incorporate the SP1-binding element, establishing a clear link between PQSS occurrence and SP1 elements
-
Full-text index only
A unique, consistent identifier for alternatively spliced transcript variants.
PMID 19865484 · PMC2765725 · PloS one · 2009 · 6 claims · 1 setups
Existing transcript identifiers (NM_ accessions, ENST identifiers) are unsuitable for uniquely identifying isoform structure across databases, methods, or organisms
-
Has reproduction · 62
E3RC: A step-by-step computational protocol for exploring enhancer RNA expression and regulation using conventional RNA-seq data.
PMID 40716058 · PMC12318280 · STAR protocols · 2025 · 6 claims · 3 setups
E3RC is a computational framework for identifying and quantifying eRNAs and characterizing their expression and transcriptional regulation using conventional RNA-seq data.
-
Full-text index only
Biocomputing enters its adolescence.
PMID 15960815 · PMC1175967 · Genome biology · 2005 · 8 claims · 8 setups
A 'match augmentation' algorithm efficiently matches structural motifs by prioritizing functionally significant residues, enabling function prediction between evolutionarily unrelated proteins
-
Has reproduction · 84
Fractional ridge regression: a fast, interpretable reparameterization of ridge regression.
PMID 33252656 · PMC7702219 · GigaScience · 2020 · 8 claims · 2 setups
Ridge regression can be reparameterized in terms of γ, the ratio between the L2-norms of the regularized and unregularized (OLS) coefficient solutions, defining 'fractional ridge regression' (FRR)
-
Has reproduction · 49
Numb prevents a complete epithelial-mesenchymal transition by modulating Notch signalling.
PMID 29187638 · PMC5721160 · Journal of the Royal Society, Interface · 2017 · 8 claims · 8 setups
Numb/Numbl acts as a 'phenotypic stability factor' (PSF) that inhibits a complete EMT by stabilizing the hybrid E/M phenotype
-
Full-text index only
Evola: Ortholog database of all human genes in H-InvDB with manual curation of phylogenetic trees.
PMID 17982176 · PMC2238928 · Nucleic acids research · 2008 · 6 claims · 7 setups
Evola combines genome synteny-based computational ortholog detection with manual curation of phylogenetic trees by experts to yield more reliable orthologs than automated pairwise methods
-
Full-text index only
Aberrant 5' splice sites in human disease genes: mutation pattern, nucleotide structure and comparison of computational tools that predict their utilization.
PMID 17576681 · PMC1934990 · Nucleic acids research · 2007 · 8 claims · 4 setups
Cryptic 5'ss are best predicted by computational algorithms that accommodate nucleotide dependencies (e.g., Markov model, maximum entropy, maximum dependence decomposition) rather than by weight-matrix models
-
Full-text index only
Cancer-specific high-throughput annotation of somatic mutations: computational prediction of driver missense mutations.
PMID 19654296 · PMC2763410 · Cancer research · 2009 · 7 claims · 7 setups
CHASM, a Random Forest-based computational method, was developed to identify and prioritize missense mutations likely to be functional drivers of tumor cell proliferation.
-
Full-text index only
Systems biology of gene regulation fulfills its promise.
PMID 16719937 · PMC1779525 · Genome biology · 2006 · 8 claims · 8 setups
Suz12, a Polycomb Group complex component, has DNA targets identifiable by ChIP-chip and can silence large genomic regions in a cell-type-specific manner.
-
Has reproduction · 84
An NMF-Based Methodology for Selecting Biomarkers in the Landscape of Genes of Heterogeneous Cancer-Associated Fibroblast Populations.
PMID 32425511 · PMC7218276 · Bioinformatics and biology insights · 2020 · 8 claims · 6 setups
An integrated methodology combining nonnegative matrix factorization (NMF), the WebGestalt functional enrichment tool, and an ad hoc gene extraction procedure automatically identifies informative gene subsets from microarray data matrices that differ in number of genes (rows) and patients (columns)
-
Has reproduction · 85
scSAMAC: saliency-adjusted masking induced attention contrastive learning for single-cell clustering.
PMID 40131310 · PMC11934584 · Briefings in bioinformatics · 2025 · 8 claims · 1 setups
scSAMAC integrates contrastive learning and negative binomial (NB) losses into a VAE, extracting features via contrastive unit similarity while preserving intrinsic data characteristics to enhance robustness and generalization in clustering.