Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Slider--maximum use of probability information for alignment of short sequence reads and SNP detection.
PMID 18974170 · PMC2638935 · Bioinformatics (Oxford, England) · 2009 · 7 claims · 3 setups
Slider aligns reads using all bases above a probability threshold (baseMinPrb) from prb files, generating all possible read sequences above a read probability threshold (read_0_MinPrb), rather than only the most probable sequence
-
Full-text index only
pmid-42057295
PMID 42057295 · PMC13141149 · 8 claims · 5 setups
nf-core/viralmetagenome is a Nextflow pipeline that automates untargeted reconstruction and variant analysis of eukaryotic DNA and RNA viruses from short-read metagenomic or hybridisation-capture data.
-
Full-text index only
TEPEAK: A novel method for identifying and characterizing polymorphic transposable elements in non-model species populations.
PMID 41494038 · PMC12788660 · PLoS computational biology · 2026 · 8 claims · 6 setups
TEPEAK identifies and characterizes polymorphic TEs in populations without any prior TE sequence or loci information, using only a chromosome-level reference assembly.
-
Has reproduction
Genome-wide signatures of convergent evolution in echolocating mammals.
PMID 24005325 · PMC3836225 · Nature · 2013 · 8 claims · 8 setups
Genome-wide convergent sequence evolution between echolocating lineages is not rare but widespread and continuously distributed, with signatures consistent with convergence in nearly 200 loci out of 2,326 examined.
-
Has reproduction · 100
Integrative transcriptome sequencing identifies trans-splicing events with important roles in human embryonic stem cell pluripotency.
PMID 24131564 · PMC3875859 · Genome research · 2014 · 8 claims · 8 setups
TSscan, a computational pipeline integrating long- and short-read transcriptome sequencing from multiple hESC lines, can detect trans-splicing while minimizing false positives from experimental artifacts and genetic rearrangements.
-
Full-text index only
High throughput sequencing and proteomics to identify immunogenic proteins of a new pathogen: the dirty genome approach.
PMID 20037647 · PMC2793016 · PloS one · 2009 · 7 claims · 7 setups
A dirty genome approach using unfinished, unclosed genome sequences combined with proteomics can rapidly identify immunogenic proteins useful for diagnostic tool development
-
Has reproduction · 83
Gene Expression Atlas update--a value-added database of microarray and sequencing-based functional genomics experiments.
PMID 22064864 · PMC3245177 · Nucleic acids research · 2012 · 8 claims · 5 setups
Gene Expression Atlas is an added-value database providing curated, re-annotated and statistically analysed gene expression data across cell types, organism parts, developmental stages, disease states and other biological/experimental conditions, derived from ArrayExpress Archive and the European Nucleotide Archive.
-
Has reproduction · 78
Metavisitor, a Suite of Galaxy Tools for Simple and Rapid Detection and Discovery of Viruses in Deep Sequence Data.
PMID 28045932 · PMC5207757 · PloS one · 2017 · 7 claims · 5 setups
Metavisitor is an open-source suite of modular Galaxy tools and preset workflows for detecting and assembling viral genomes from deep sequencing data.
-
Has reproduction · 73
Vespucci: a system for building annotated databases of nascent transcripts.
PMID 24304890 · PMC3936758 · Nucleic acids research · 2014 · 8 claims · 7 setups
Existing ChIP-seq and RNA-seq analysis platforms (e.g. Cufflinks, peak callers) are unsuited to GRO-seq because they assume spliced/exonic reads, uniform density and paired-end data, and cannot identify transcriptional units de novo across the whole genome.
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.
-
Has reproduction · 83
Macrel: antimicrobial peptide screening in genomes and metagenomes.
PMID 33384902 · PMC7751412 · PeerJ · 2020 · 8 claims · 8 setups
Macrel introduces a novel set of 22 peptide features (6 local, 16 global), including a new Free Energy Transition (FET) feature group, for AMP and hemolytic activity classification
-
Has reproduction · 74
Evaluation of classification and forecasting methods on time series gene expression data.
PMID 33156855 · PMC7647064 · PloS one · 2020 · 8 claims · 4 setups
Deep learning based methods generally outperform traditional approaches for time series gene expression classification.
-
Has reproduction · 30
SMRT and Illumina RNA sequencing reveal novel insights into the heat stress response and crosstalk with leaf senescence in tall fescue.
PMID 32746857 · PMC7397585 · BMC plant biology · 2020 · 8 claims · 7 setups
Combined PacBio SMRT and Illumina RNA sequencing generated a full-length reference transcriptome for tall fescue in the absence of a genome sequence.
-
Full-text index only
ChimerDB 2.0--a knowledgebase for fusion genes updated.
PMID 19906715 · PMC2808913 · Nucleic acids research · 2010 · 8 claims · 4 setups
ChimerDB 2.0 is an updated knowledgebase integrating fusion transcripts from GenBank transcriptome analysis with Sanger CGP, OMIM, PubMed, and Mitelman's database data.
-
Full-text index only
Water mass specific genes dominate the Southern Ocean microbiome.
PMID 41803086 · PMC12972064 · Nature communications · 2026 · 8 claims · 8 setups
The Southern Ocean microbial gene catalog is highly original and largely distinct from existing marine gene catalogs
-
Has reproduction · 90
Gap-free telomere-to-telomere haplotype assembly of the tomato hind (Cephalopholis sonnerati).
PMID 39578472 · PMC11584678 · Scientific data · 2024 · 8 claims · 8 setups
Two T2T gap-free haplotype assemblies of C. sonnerati (YSFRI_Csonn_HA_1.0 and YSFRI_Csonn_HB_1.0) were successfully generated, each spanning 24 chromosomes with no gaps.
-
Full-text index only
Proteomic analysis of differential proteins in pancreatic carcinomas: Effects of MBD1 knock-down by stable RNA interference.
PMID 18445260 · PMC2386481 · BMC cancer · 2008 · 8 claims · 5 setups
Stable RNAi-mediated MBD1 knock-down was successfully established in the BxPC-3 pancreatic cancer cell line using a recombinant siRNA plasmid
-
Full-text index only
Ensembl's 10th year.
PMID 19906699 · PMC2808936 · Nucleic acids research · 2010 · 8 claims · 8 setups
Ensembl provides comprehensive gene annotation and integrated genomic resources (variation, regulation, comparative genomics) across a growing set of chordate genomes