Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 63
hgtseq: A Standard Pipeline to Study Horizontal Gene Transfer.
PMID 36498841 · PMC9738810 · International journal of molecular sciences · 2022 · 8 claims · 8 setups
hgtseq is a fully automated, portable, and scalable Nextflow/nf-core pipeline for detecting horizontal gene transfer signatures from unmapped sequencing reads.
-
Full-text index only
BIPASS: BioInformatics Pipeline Alternative Splicing Services.
PMID 17584795 · PMC1933140 · Nucleic acids research · 2007 · 8 claims · 4 setups
BIPASS offers two complementary services for alternative splicing (AS) research: BIPAS-SpliceDB, a queryable pre-computed AS data warehouse, and BIPAS-Align&Splice, an online pipeline for user-submitted sequences.
-
Full-text index only
TassDB: a database of alternative tandem splice sites.
PMID 17142241 · PMC1669710 · Nucleic acids research · 2007 · 7 claims · 3 setups
TassDB is a relational database storing GYNGYN donor and NAGNAG acceptor tandem splice sites across eight species
-
Full-text index only
Stable in a genome of instability: an interview with Evan Eichler. Interview by Jane Gitschier.
PMID 18654618 · PMC2442658 · PLoS genetics · 2008 · 8 claims · 5 setups
Loss of AGG interruptions in CGG repeat tracts predisposes FMR1 alleles to instability and faster progression toward premutation/disease state
-
Full-text index only
Nucleotide-resolution analysis of structural variants using BreakSeq and a breakpoint library.
PMID 20037582 · PMC2951730 · Nature biotechnology · 2010 · 8 claims · 7 setups
A standardized, non-redundant library of 1,889 breakpoint-resolved SVs was assembled from eight published surveys
-
Has reproduction · 65
SPEAQeasy: a scalable pipeline for expression analysis and quantification for R/bioconductor-powered RNA-seq analyses.
PMID 33932985 · PMC8088074 · BMC bioinformatics · 2021 · 8 claims · 5 setups
SPEAQeasy is a portable, easy-to-install, Nextflow-powered RNA-seq processing pipeline that lowers the computational entry barrier for biologists/clinicians
-
Has reproduction · 95
nf-rnaSeqCount: A Nextflow pipeline for obtaining raw read counts from RNA-seq data.
PMID 35574063 · PMC9097006 · South African computer journal = Suid-Afrikaanse rekenaartydskrif · 2021 · 7 claims · 5 setups
nf-rnaSeqCount is a portable, reproducible Nextflow pipeline that maps RNA-seq reads to a reference genome and quantifies gene abundance for differential expression analysis
-
Full-text index only
Discovery of novel human transcript variants by analysis of intronic single-block EST with polyadenylation site.
PMID 19906316 · PMC2784480 · BMC genomics · 2009 · 8 claims · 7 setups
Intronic single-block ESTs with poly(A/T) tails reveal previously unidentified novel transcript variants missed by existing databases.
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.
-
Full-text index only
Ensembl 2009.
PMID 19033362 · PMC2686571 · Nucleic acids research · 2009 · 8 claims · 6 setups
Ensembl provides comprehensive, consistently annotated genome information for chordate genomes with automatically generated genesets and comparative genomics data
-
Full-text index only
ChimerDB 2.0--a knowledgebase for fusion genes updated.
PMID 19906715 · PMC2808913 · Nucleic acids research · 2010 · 8 claims · 4 setups
ChimerDB 2.0 is an updated knowledgebase integrating fusion transcripts from GenBank transcriptome analysis with Sanger CGP, OMIM, PubMed, and Mitelman's database data.
-
Has reproduction · 88
nf-core/isoseq: simple gene and isoform annotation with PacBio Iso-Seq long-read sequencing.
PMID 36961337 · PMC10199315 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 4 setups
nf-core/isoseq is a new automated Nextflow-based pipeline that processes raw Iso-Seq subreads through to genome annotation (BED format) without requiring transcriptome assembly.
-
Full-text index only
The other side of comparative genomics: genes with no orthologs between the cow and other mammalian species.
PMID 20003425 · PMC2808326 · BMC genomics · 2009 · 7 claims · 4 setups
3,801 bovine genes have no orthologs in human, mouse and dog, and 1,010 human genes have no orthologs in cow despite having orthologs in mouse and dog
-
Full-text index only
Target SNP selection in complex disease association studies.
PMID 15248903 · PMC487897 · BMC bioinformatics · 2004 · 7 claims · 3 setups
A computational pipeline can retrieve gene sequence, collect SNP variation data, and annotate SNPs falling in functional motifs (promoter, exon-intron structure, AU-rich elements, TF binding sites, splice sites) with expression in target tissue
-
Full-text index only
Variation resources at UC Santa Cruz.
PMID 17151077 · PMC1781230 · Nucleic acids research · 2007 · 8 claims · 8 setups
The UCSC Genome Browser variation resources integrate polymorphism data from public collections (dbSNP, HapMap, Affymetrix, Perlegen, SeattleSNPs) into a common format with additional annotations and genomic context.
-
Has reproduction
Using prototyping to choose a bioinformatics workflow management system.
PMID 33630841 · PMC7906312 · PLoS computational biology · 2021 · 8 claims · 6 setups
Prototyping a subset of a project's actual workflow in candidate software offers a low-cost (time and effort) way to make a more informed software selection than relying solely on reviews and recommendations.
-
Has reproduction · 79
RetroSnake: A modular pipeline to detect human endogenous retroviruses in genome sequencing data.
PMID 36339261 · PMC9626663 · iScience · 2022 · 8 claims · 4 setups
RetroSnake is an end-to-end, modular, computationally efficient Snakemake pipeline for detecting HERV-K insertions in short-read NGS data, from raw alignment files to an annotated interactive HTML report
-
Full-text index only
Systematic identification of pseudogenes through whole genome expression evidence profiling.
PMID 16945953 · PMC1636364 · Nucleic acids research · 2006 · 8 claims · 8 setups
Developed a novel bioinformatics method that identifies pseudogenes by profiling whole-genome transcript and protein expression evidence
-
Full-text index only
SpliceMiner: a high-throughput database implementation of the NCBI Evidence Viewer for microarray splice variant analysis.
PMID 17338820 · PMC1839109 · BMC bioinformatics · 2007 · 6 claims · 4 setups
EVDB is a comprehensive, non-redundant relational database of known human splice variants built from NCBI Entrez Gene and Evidence Viewer data
-
Full-text index only
Searching for SNPs with cloud computing.
PMID 19930550 · PMC3091327 · Genome biology · 2009 · 8 claims · 4 setups
Crossbow combines the Bowtie short-read aligner and SOAPsnp SNP caller into a seamless, automatic Hadoop/MapReduce pipeline for whole-genome resequencing analysis