Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Whole genome amplification and de novo assembly of single bacterial cells.
PMID 19724646 · PMC2731171 · PloS one · 2009 · 8 claims · 6 setups
FACS-based single-cell isolation combined with strict handling procedures virtually eliminates contaminating DNA from single-cell MDA reactions
-
Full-text index only
QuantiSNP: an Objective Bayes Hidden-Markov Model to detect and accurately map copy number variation using SNP genotyping data.
PMID 17341461 · PMC1874617 · Nucleic acids research · 2007 · 8 claims · 7 setups
QuantiSNP (OB-HMM) provides probabilistic quantification of copy number states and significantly improves accuracy of segmental aneuploidy identification and breakpoint mapping relative to existing tools (BeadStudio/Illumina)
-
Has reproduction · 95
In vivo structural characterization of the SARS-CoV-2 RNA genome identifies host proteins vulnerable to repurposed drugs.
PMID 33636127 · PMC7871767 · Cell · 2021 · 8 claims · 8 setups
icSHAPE was used to determine the in vivo and in vitro structural landscape of the SARS-CoV-2 RNA genome in infected Huh7.5.1 cells, plus UTR structures of six other coronaviruses
-
Has reproduction · 83
Accurate prediction of metagenome-assembled genome completeness by MAGISTA, a random forest model built on alignment-free intra-bin statistics.
PMID 35248155 · PMC8898458 · Environmental microbiome · 2022 · 7 claims · 7 setups
MAGISTA, a random forest model built on alignment-free intra-bin distance-distribution statistics, can estimate MAG completeness and purity without relying on reference marker genes.
-
Has reproduction · 79
RetroSnake: A modular pipeline to detect human endogenous retroviruses in genome sequencing data.
PMID 36339261 · PMC9626663 · iScience · 2022 · 8 claims · 4 setups
RetroSnake is an end-to-end, modular, computationally efficient Snakemake pipeline for detecting HERV-K insertions in short-read NGS data, from raw alignment files to an annotated interactive HTML report
-
Full-text index only
Short tandem repeats in human exons: a target for disease mutations.
PMID 18789129 · PMC2543027 · BMC genomics · 2008 · 8 claims · 6 setups
STRs are present in exons of 92% of known human genes, unlike longer tandem repeats which are rare in exons
-
Full-text index only
SeqDoC: rapid SNP and mutation detection by direct comparison of DNA sequence chromatograms.
PMID 15927052 · PMC1156871 · BMC bioinformatics · 2005 · 8 claims · 6 setups
SeqDoC generates a subtracted difference trace between a reference and test chromatogram that highlights single base changes
-
Full-text index only
Satellog: a database for the identification and prioritization of satellite repeats in disease association studies.
PMID 15949044 · PMC1181805 · BMC bioinformatics · 2005 · 7 claims · 6 setups
Satellog is a database cataloging all pure 1-16 unit satellite repeats in the human genome with supplementary polymorphism, gene-location, and expression data for prioritizing repeats in disease-association studies.
-
Full-text index only
Fast and systematic genome-wide discovery of conserved regulatory elements using a non-alignment based approach.
PMID 15693947 · PMC551538 · Genome biology · 2005 · 7 claims · 8 setups
FastCompare, a non-alignment-based, linear-time algorithm, computes a genome-wide conservation score for all k-mers (7-9 nt) between two genomes to identify conserved regulatory elements
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
Effect of read-mapping biases on detecting allele-specific expression from RNA-sequencing data.
PMID 19808877 · PMC2788925 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 6 setups
Reads mapped to the reference genome show a significant bias toward the reference allele at heterozygous SNPs
-
Has reproduction · 67
Leveraging RNA-seq deconvolution to improve complex in vitro model characterization.
PMID 40701251 · PMC12391696 · The Journal of biological chemistry · 2025 · 8 claims · 6 setups
RNA-seq deconvolution can predict cell type proportions from bulk RNA-seq using scRNA-seq references, offering a useful characterization tool for CIVMs where single-cell methods are impractical
-
Has reproduction · 58
A comparative study of techniques for differential expression analysis on RNA-Seq data.
PMID 25119138 · PMC4132098 · PloS one · 2014 · 8 claims · 8 setups
edgeR performs slightly better than DESeq and Cuffdiff2 in terms of the ability to uncover true positives.
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs
-
Full-text index only
Searching for SNPs with cloud computing.
PMID 19930550 · PMC3091327 · Genome biology · 2009 · 8 claims · 4 setups
Crossbow combines the Bowtie short-read aligner and SOAPsnp SNP caller into a seamless, automatic Hadoop/MapReduce pipeline for whole-genome resequencing analysis
-
Full-text index only
COMUS: Clinician-Oriented locus-specific MUtation detection and deposition System.
PMID 19958500 · PMC2788389 · BMC genomics · 2009 · 8 claims · 6 setups
COMUS is a bioinformatics system for detecting and depositing new mutations from patient DNA with a clinician-friendly interface
-
Full-text index only
Allelotyping of pooled DNA with 250 K SNP microarrays.
PMID 17367522 · PMC1839100 · BMC genomics · 2007 · 8 claims · 5 setups
The polynomial based probe specific correction (PPC) algorithm is the most accurate method for estimating allele frequency from pooled DNA.
-
Full-text index only
Analysis of copy number variation using quantitative interspecies competitive PCR.
PMID 18697816 · PMC2553599 · Nucleic acids research · 2008 · 7 claims · 6 setups
qicPCR uses the entire genome of a single chimpanzee as a competitor, requiring only one reference sample for all assays and enabling large-scale multiplexing
-
Has reproduction · 86
The selection of software and database for metagenomics sequence analysis impacts the outcome of microbial profiling and pathogen detection.
PMID 37027361 · PMC10081788 · PloS one · 2023 · 7 claims · 7 setups
Obtaining an accurate species-level microbial profile using current direct-read metagenomics profiling software is still a challenging task.
-
Has reproduction · 85
Digital sorting of complex tissues for cell type-specific gene expression profiles.
PMID 23497278 · PMC3626856 · BMC bioinformatics · 2013 · 8 claims · 8 setups
The Digital Sorting Algorithm (DSA) deconvolves mixed tissue expression into cell type-specific profiles using only marker genes, without requiring prior knowledge of cell type frequencies or in vitro pure-cell profiles.