Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A network model for the correlation between epistasis and genomic complexity.
PMID 18648534 · PMC2481279 · PloS one · 2008 · 8 claims · 5 setups
In small networks with multifunctional nodes, lack of redundancy, and absence of alternative pathways, epistasis is antagonistic on average.
-
Full-text index only
In silico promoters: modelling of cis-regulatory context facilitates target predictio.
PMID 18505473 · PMC3823354 · Journal of cellular and molecular medicine · 2009 · 8 claims · 8 setups
An integrated 'profiling of transcriptional targets' (PTT) strategy by Freebern et al. identified IGF-1 as a co-modulator of immune cell function genes in mitogen/drug-activated T cells.
-
Full-text index only
DiRE: identifying distant regulatory elements of co-expressed genes.
PMID 18487623 · PMC2447744 · Nucleic acids research · 2008 · 8 claims · 4 setups
DiRE predicts distant regulatory elements by combining gene co-expression data, comparative genomics and TFBS profiles to determine TFBS-association signatures
-
Full-text index only
In silico discovery of transcription regulatory elements in Plasmodium falciparum.
PMID 18257930 · PMC2268928 · BMC genomics · 2008 · 7 claims · 8 setups
GEMS, using hypergeometric scoring and PWM parameter optimization, reliably identifies high-confidence cis-regulatory elements in the AT-rich, repeat-rich P. falciparum genome
-
Has reproduction · 68
Bayesian transcriptome assembly.
PMID 25367074 · PMC4397945 · Genome biology · 2014 · 8 claims · 8 setups
Bayesembler, a probabilistic transcriptome assembler built on a Bayesian model of the RNA sequencing process with Gibbs sampling over expressed candidates, abundances and read assignments, is introduced.
-
Full-text index only
Searching for SNPs with cloud computing.
PMID 19930550 · PMC3091327 · Genome biology · 2009 · 8 claims · 4 setups
Crossbow combines the Bowtie short-read aligner and SOAPsnp SNP caller into a seamless, automatic Hadoop/MapReduce pipeline for whole-genome resequencing analysis
-
Full-text index only
Enrichment of sequencing targets from the human genome by solution hybridization.
PMID 19835619 · PMC2784331 · Genome biology · 2009 · 8 claims · 5 setups
Solution hybridization with 120-mer capture probes efficiently enriches targeted genomic sequences for next-generation sequencing
-
Full-text index only
BreakDancer: an algorithm for high-resolution mapping of genomic structural variation.
PMID 19668202 · PMC3661775 · Nature methods · 2009 · 8 claims · 8 setups
BreakDancer (BreakDancerMax + BreakDancerMini) is a software package that predicts a wide variety of structural variants including deletions, insertions, inversions, and intra/inter-chromosomal translocations from paired-end short-insert sequencing reads.
-
Full-text index only
Cancer-specific high-throughput annotation of somatic mutations: computational prediction of driver missense mutations.
PMID 19654296 · PMC2763410 · Cancer research · 2009 · 7 claims · 7 setups
CHASM, a Random Forest-based computational method, was developed to identify and prioritize missense mutations likely to be functional drivers of tumor cell proliferation.
-
Has reproduction · 50
Quality control method for RNA-seq using single nucleotide polymorphism allele frequency.
PMID 25243705 · PMC4231238 · Genes to cells : devoted to molecular & cellular mechanisms · 2014 · 8 claims · 8 setups
SNP allele frequency distributions from RNA-seq reads can detect contaminating cells whose genomic background differs from the target cells; the mode of the distribution reflects the cellular composition while its variance reflects PCR bias.
-
Full-text index only
bakR: uncovering differential RNA synthesis and degradation kinetics transcriptome-wide with Bayesian hierarchical modeling.
PMID 37028916 · PMC10275263 · RNA (New York, N.Y.) · 2023 · 8 claims · 4 setups
bakR uses Bayesian hierarchical modeling to share information (specifically a replicate variability vs. read count trend) across transcripts, increasing statistical power for differential kinetic analysis
-
Full-text index only
Large-scale estimation of bacterial and archaeal DNA prevalence in metagenomes reveals biome-specific patterns.
PMID 41854267 · PMC13098197 · mSystems · 2026 · 8 claims · 6 setups
SPF scalably and robustly estimates the fraction of bacterial and archaeal reads in a metagenome using detection of prokaryotic single-copy marker genes, without requiring eukaryotic or viral reference genomes
-
Full-text index only
MetaPepticon: automated prediction of anticancer peptides from microbial genomes and metagenomes.
PMID 41918857 · PMC13034871 · PeerJ · 2026 · 7 claims · 6 setups
MetaPepticon is a modular, end-to-end Snakemake pipeline that predicts ACP candidates directly from raw genomic, metagenomic, transcriptomic, metatranscriptomic reads, assembled contigs, or peptide sequences.
-
Has reproduction · 76
Transcriptional landscape of repetitive elements in normal and cancer human cells.
PMID 25012247 · PMC4122776 · BMC genomics · 2014 · 8 claims · 8 setups
RepEnrich, a computational method that uses all mapping reads (uniquely mapping plus multi-mapping reads assigned to repetitive element subfamily assemblies/pseudogenomes), quantifies genome-wide repetitive element enrichment
-
Full-text index only
isoSeQL: comparing long-read isoforms across multiple datasets.
PMID 41452740 · PMC12790818 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 4 setups
isoSeQL enables comparison of long-read isoform profiles across multiple datasets by consolidating SQANTI3-annotated samples into a unified SQLite database with consistent isoform IDs
-
Full-text index only
RUMINA: high-throughput deduplication of unique molecular identifiers for amplicon and whole-genome sequencing with enhanced error correction.
PMID 41734278 · PMC12975283 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 4 setups
RUMINA improves detection accuracy of ultra-low frequency SNVs (0.01%-1%) compared to UMI-tools and UMICollapse
-
Full-text index only
Comparative analysis of eccDNA and circRNA tools shows increased accuracy of tool combination.
PMID 41738836 · PMC13154841 · GigaScience · 2026 · 8 claims · 6 setups
Detection accuracy of eccDNA/circRNA tools is highly influenced by sequencing depth, alignment algorithm, and experimental enrichment protocol
-
Full-text index only
Extending differential gene expression testing to handle genome aneuploidy in cancer.
PMID 41894415 · PMC13061324 · PLoS computational biology · 2026 · 8 claims · 4 setups
DeConveil integrates CNV data into DGE analysis using a GLM with negative binomial distribution to correct for CN-driven gene dosage effects
-
Full-text index only
Machine learning-predicted chromatin organization landscape across pediatric tumors.
PMID 41904260 · PMC13039956 · Scientific reports · 2026 · 8 claims · 5 setups
SuPreMo-Akita (built on the Akita CNN) enables systematic in silico prediction of somatic SV effects on 3D genome folding across large SV cohorts where experimental testing is infeasible
-
Full-text index only
Benchmarking tools for deciphering cellular crosstalk in spatially-resolved transcriptomics.
PMID 41952215 · PMC13174004 · Genome biology · 2026 · 8 claims · 5 setups
No prior systematic, quantitative benchmark exists for CCI inference methods specifically developed for spatial transcriptomics across multiple platforms