Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
EGASP: the human ENCODE Genome Annotation Assessment Project.
PMID 16925836 · PMC1810551 · Genome biology · 2006 · 8 claims · 6 setups
Best-performing computational gene prediction methods correctly predict at least one transcript for close to 70% of annotated genes in the ENCODE regions.
-
Full-text index only
MultiPhyl: a high-throughput phylogenomics webserver using distributed computing.
PMID 17553837 · PMC1933173 · Nucleic acids research · 2007 · 8 claims · 8 setups
MultiPhyl is the first high-throughput distributed phylogenetics platform capable of using idle computational resources of many heterogeneous non-dedicated machines to form a phylogenetics supercomputer
-
Full-text index only
Pairagon+N-SCAN_EST: a model-based gene annotation pipeline.
PMID 16925839 · PMC1810554 · Genome biology · 2006 · 7 claims · 5 setups
Pairagon+N-SCAN_EST, using only native alignments, was as accurate as ENSEMBL and ExoGean in the EGASP mRNA/EST evidence assessment
-
Full-text index only
Benchmarking tools for the alignment of functional noncoding DNA.
PMID 14736341 · PMC344529 · BMC bioinformatics · 2004 · 8 claims · 4 setups
Global alignment tools (Avid, ClustalW, Lagan, Needle, DiAlign-G) typically have higher sensitivity over entire noncoding sequences and within constrained blocks than local tools
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
metaFun: An analysis pipeline for metagenomic big data with fast and unified functional searches.
PMID 41530917 · PMC12818822 · Gut microbes · 2026 · 8 claims · 8 setups
metaFun is an open-source, end-to-end Nextflow/Apptainer pipeline integrating quality control, taxonomic profiling, functional profiling, de novo assembly, binning, genome assessment, comparative genomics, network analysis, and strain-level microdiversity analysis into a unified framework
-
Has reproduction · 78
Metavisitor, a Suite of Galaxy Tools for Simple and Rapid Detection and Discovery of Viruses in Deep Sequence Data.
PMID 28045932 · PMC5207757 · PloS one · 2017 · 7 claims · 5 setups
Metavisitor is an open-source suite of modular Galaxy tools and preset workflows for detecting and assembling viral genomes from deep sequencing data.
-
Full-text index only
bakR: uncovering differential RNA synthesis and degradation kinetics transcriptome-wide with Bayesian hierarchical modeling.
PMID 37028916 · PMC10275263 · RNA (New York, N.Y.) · 2023 · 8 claims · 4 setups
bakR uses Bayesian hierarchical modeling to share information (specifically a replicate variability vs. read count trend) across transcripts, increasing statistical power for differential kinetic analysis
-
Has reproduction · 61
TEMP: a computational method for analyzing transposable element polymorphism in populations.
PMID 24753423 · PMC4066757 · Nucleic acids research · 2014 · 8 claims · 8 setups
TEMP combines pair-end (discordant) read and split (soft-clipped) read information to identify both presence and absence of TE insertions in genomic DNA from heterogeneous/pooled samples.
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs
-
Has reproduction · 71
RNAmountAlign: Efficient software for local, global, semiglobal pairwise and multiple RNA sequence/structure alignment.
PMID 31978147 · PMC6980424 · PloS one · 2020 · 7 claims · 6 setups
RNAmountAlign performs pairwise local, global, and semiglobal (query search) alignment and progressive multiple alignment (global and local) using incremental ensemble mountain height, running in O(n^3) time and O(n^2) space for two sequences of length n
-
Has reproduction · 87
R2DT is a framework for predicting and visualising RNA secondary structure using templates.
PMID 34108470 · PMC8190129 · Nature communications · 2021 · 8 claims · 6 setups
R2DT is a template-based computational framework/pipeline that predicts and visualises RNA 2D structure in standardised, community-accepted layouts
-
Full-text index only
Variant-resolved prediction of context-specific isoform variation with a graph-based attention model.
PMID 41547351 · PMC13069856 · Cell genomics · 2026 · 8 claims · 8 setups
Otari, an attention-based graph neural network trained on long-read transcriptomes across 30 tissues/brain regions, predicts tissue-specific differential isoform abundance
-
Full-text index only
PolyAseqTrap: a universal tool for genome-wide identification and quantification of polyadenylation sites from different 3' end sequencing data.
PMID 41620776 · PMC12947541 · Genome biology · 2026 · 6 claims · 7 setups
PolyAseqTrap is a universal R package for identifying and quantifying polyA sites from diverse 3' end sequencing data