Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
BreakDancer: an algorithm for high-resolution mapping of genomic structural variation.
PMID 19668202 · PMC3661775 · Nature methods · 2009 · 8 claims · 8 setups
BreakDancer (BreakDancerMax + BreakDancerMini) is a software package that predicts a wide variety of structural variants including deletions, insertions, inversions, and intra/inter-chromosomal translocations from paired-end short-insert sequencing reads.
-
Has reproduction · 68
Bayesian transcriptome assembly.
PMID 25367074 · PMC4397945 · Genome biology · 2014 · 8 claims · 8 setups
Bayesembler, a probabilistic transcriptome assembler built on a Bayesian model of the RNA sequencing process with Gibbs sampling over expressed candidates, abundances and read assignments, is introduced.
-
Full-text index only
An integrative approach to reveal driver gene fusions from paired-end sequencing data in cancer.
PMID 19881495 · PMC3086882 · Nature biotechnology · 2009 · 8 claims · 8 setups
A 'concept signature' (ConSig) score algorithm ranks genes by association with molecular concepts characteristic of fusion or mutation cancer genes, nominating biologically important fusions from large candidate sets.
-
Has reproduction · 71
Systematic and computational identification of Androctonus crassicauda long non-coding RNAs.
PMID 33633149 · PMC7907363 · Scientific reports · 2021 · 7 claims · 7 setups
A custom ECF pipeline identified 13,401 lncRNAs in the A. crassicauda transcriptome (12,642 novel, 759 known).
-
Has reproduction · 65
FusionQ: a novel approach for gene fusion detection and quantification from paired-end RNA-Seq.
PMID 23768108 · PMC3691734 · BMC bioinformatics · 2013 · 8 claims · 8 setups
FusionQ is a novel tool that detects gene fusions, constructs chimerical transcript structures, and estimates their abundances from paired-end RNA-Seq data.
-
Full-text index only
Searching for SNPs with cloud computing.
PMID 19930550 · PMC3091327 · Genome biology · 2009 · 8 claims · 4 setups
Crossbow combines the Bowtie short-read aligner and SOAPsnp SNP caller into a seamless, automatic Hadoop/MapReduce pipeline for whole-genome resequencing analysis
-
Full-text index only
Bases and spaces: resources on the web for accessing the draft human genome.
PMID 11178254 · PMC138875 · Genome biology · 2000 · 8 claims · 8 setups
By combining currently available genomic databases and mapping resources (GenBank/Entrez, UniGene, RH maps, BAC fingerprint maps, Ensembl, NIX), it is possible to devise strategies that fully exploit the fragmentary draft human genome sequence.
-
Has reproduction · 58
A comparative study of techniques for differential expression analysis on RNA-Seq data.
PMID 25119138 · PMC4132098 · PloS one · 2014 · 8 claims · 8 setups
edgeR performs slightly better than DESeq and Cuffdiff2 in terms of the ability to uncover true positives.
-
Has reproduction · 100
Genomic insight into the influence of selection, crossbreeding, and geography on population structure in poultry.
PMID 36670351 · PMC9854048 · Genetics, selection, evolution : GSE · 2023 · 6 claims · 8 setups
Dutch traditional chicken breeds show a complex and admixed subdivided population structure that partly matches historical management-based clustering (past-productive, ornamental, country fowl, Lakenvelder).
-
Has reproduction · 69
Manual curation for improved genome annotation of the functionally extinct northern white rhinoceros (Ceratotherium simum cottoni).
PMID 41490125 · PMC12768360 · PloS one · 2026 · 6 claims · 5 setups
The original BRAKER3-based NWR annotation was of poor quality: only 51% of transcripts were correctly called, many were assigned uninformative protein names, and some were misassigned to incorrect or bacterial sequences.
-
Full-text index only
Discovery of novel human transcript variants by analysis of intronic single-block EST with polyadenylation site.
PMID 19906316 · PMC2784480 · BMC genomics · 2009 · 8 claims · 7 setups
Intronic single-block ESTs with poly(A/T) tails reveal previously unidentified novel transcript variants missed by existing databases.
-
Full-text index only
Identification and functional analyses of 11,769 full-length human cDNAs focused on alternative splicing.
PMID 19880432 · PMC2780955 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2009 · 8 claims · 5 setups
Identified 23,241 human genes transcribed into protein-coding mRNAs using full-length cDNA and 5'-EST sequence data
-
Full-text index only
A catalog of human cDNA expression clones and its application to structural genomics.
PMID 15345055 · PMC522878 · Genome biology · 2004 · 8 claims · 7 setups
A high-throughput screening approach can identify human cDNA clones from the hEx1 library that express soluble protein in E. coli
-
Has reproduction · 80
Bisulfite sequencing of chromatin immunoprecipitated DNA (BisChIP-seq) directly informs methylation status of histone-modified DNA.
PMID 22466171 · PMC3371705 · Genome research · 2012 · 8 claims · 8 setups
BisChIP-seq — bisulfite sequencing of chromatin immunoprecipitated DNA — enables direct genome-wide, base-resolution interrogation of DNA methylation on histone-modified DNA molecules
-
Has reproduction · 86
The selection of software and database for metagenomics sequence analysis impacts the outcome of microbial profiling and pathogen detection.
PMID 37027361 · PMC10081788 · PloS one · 2023 · 7 claims · 7 setups
Obtaining an accurate species-level microbial profile using current direct-read metagenomics profiling software is still a challenging task.