Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 80
SLDMS: A Tool for Calculating the Overlapping Regions of Sequences.
PMID 35046988 · PMC8761809 · Frontiers in plant science · 2021 · 8 claims · 5 setups
SLDMS is a novel method for computing overlapping regions of sequencing reads using suffix array (SA), longest common prefix (LCP) array, document array (DA), and a monotonic stack.
-
Full-text index only
Genomics--from Neanderthals to high-throughput sequencing.
PMID 16934106 · PMC1779599 · Genome biology · 2006 · 8 claims · 8 setups
Next-generation sequencing platforms (GS20/454 and Solexa) can deliver the throughput and cost reductions needed for population-scale and medical resequencing.
-
Has reproduction · 76
Tracing human genetic histories and natural selection with precise local ancestry inference.
PMID 40379651 · PMC12084304 · Nature communications · 2025 · 7 claims · 7 setups
Orchestra, a two-stage LAI method combining a recombination-distance base layer with a deep learning (convolutional + attention) smoothing module, outperforms RFmix, FLARE and Gnomix in precision and recall across simulated admixture generations.
-
Full-text index only
SNPmasker: automatic masking of SNPs and repeats across eukaryotic genomes.
PMID 16845091 · PMC1538889 · Nucleic acids research · 2006 · 8 claims · 4 setups
SNPmasker is a web service combining SNP masking and repeat masking, supporting both coordinate-defined and homology-search-defined input regions, a combination not offered by prior tools
-
Full-text index only
EGASP: the human ENCODE Genome Annotation Assessment Project.
PMID 16925836 · PMC1810551 · Genome biology · 2006 · 8 claims · 6 setups
Best-performing computational gene prediction methods correctly predict at least one transcript for close to 70% of annotated genes in the ENCODE regions.
-
Full-text index only
Alignoth: portable and interactive visualization of read alignments.
PMID 41392197 · PMC12777968 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 3 setups
Alignoth is a command-line tool that generates self-contained, portable HTML reports of DNA sequencing read alignment pileups, with export to PNG, SVG, PDF, and JSON.
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Has reproduction · 30
IsoSCM: improved and alternative 3' UTR annotation using multiple change-point inference.
PMID 25406361 · PMC4274634 · RNA (New York, N.Y.) · 2015 · 8 claims · 6 setups
Existing ab initio assemblers (Cufflinks, Scripture) annotate at most one 3' boundary per terminal exon and therefore cannot assemble coexpressed tandem 3' UTR isoforms.
-
Has reproduction · 91
Whole genome and transcriptome maps of the entirely black native Korean chicken breed Yeonsan Ogye.
PMID 30010758 · PMC6065499 · GigaScience · 2018 · 8 claims · 6 setups
A draft genome (Ogye_1.1) was assembled using a hybrid de novo method combining high-depth Illumina short reads (376.6X) and low-depth PacBio long reads (9.7X)
-
Full-text index only
SeqBuster, a bioinformatic tool for the processing and analysis of small RNAs datasets, reveals ubiquitous miRNA modifications in human embryonic cells.
PMID 20008100 · PMC2836562 · Nucleic acids research · 2010 · 8 claims · 6 setups
SeqBuster is a versatile web-based and stand-alone bioinformatic toolkit for processing and analyzing large-scale small RNA deep sequencing datasets.
-
Full-text index only
EpiXFormer: a cross-attention neural network for predicting cell type-specific transcription factor binding sites.
PMID 41527854 · PMC12796812 · Briefings in bioinformatics · 2026 · 8 claims · 8 setups
EpiXFormer achieves high accuracy (mean AUROC ~0.99) predicting binding sites of both TFs and non-sequence-specific DBPs across 199 DBP-cell type pairs
-
Has reproduction · 79
RetroSnake: A modular pipeline to detect human endogenous retroviruses in genome sequencing data.
PMID 36339261 · PMC9626663 · iScience · 2022 · 8 claims · 4 setups
RetroSnake is an end-to-end, modular, computationally efficient Snakemake pipeline for detecting HERV-K insertions in short-read NGS data, from raw alignment files to an annotated interactive HTML report
-
Full-text index only
Metapipeline-DNA: A comprehensive germline and somatic genomics Nextflow pipeline.
PMID 41850291 · PMC13030954 · Cell reports methods · 2026 · 8 claims · 7 setups
Metapipeline-DNA automates germline and somatic DNA sequencing analysis end-to-end, from raw reads through preprocessing, feature detection, QC, and visualization.
-
Full-text index only
OTMODE: an optimal transport theory-based framework for identifying differential features in single-cell multi-omics data.
PMID 41335419 · PMC12766913 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 8 setups
OTMODE, using an unbalanced Sinkhorn algorithm and Wald test, improves differential feature identification in single-cell multi-omics data
-
Full-text index only
Analyses and comparison of accuracy of different genotype imputation methods.
PMID 18958166 · PMC2569208 · PloS one · 2008 · 8 claims · 3 setups
Stronger LD produces higher imputation accuracy rates for all five methods
-
Has reproduction · 89
Statistical framework for calling allelic imbalance in high-throughput sequencing data.
PMID 39966391 · PMC11836314 · Nature communications · 2025 · 8 claims · 6 setups
MIXALIME is a versatile computational framework for calling allele-specific variants (ASVs) from diverse high-throughput omics data
-
Has reproduction · 80
DMN-seq enriches DNA hypomethylated regions for biomarker discovery using 5-methylcytosine glycosylase.
PMID 41673887 · PMC13097799 · Genome biology · 2026 · 8 claims · 9 setups
DMN-seq (DMN+) uses DME to nick DNA specifically at 5mC sites, enabling 5mC detection at single-base resolution via selective adaptor ligation
-
Full-text index only
Identification of novel DNA sequence motifs that modulate transcription in T cells.
PMID 41514212 · PMC12879379 · BMC genomics · 2026 · 8 claims · 8 setups
Identified 2,036 novel DNA motifs enriched in regulatory regions of T-cell-specific genes