Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A space-efficient and accurate method for mapping and aligning cDNA sequences onto genomic sequence.
PMID 18344523 · PMC2377433 · Nucleic acids research · 2008 · 7 claims · 6 setups
Spaln maps and aligns large cDNA sequence sets onto whole mammalian genomes using substantially less memory than comparable existing tools
-
Full-text index only
The UCSC genome browser database: update 2007.
PMID 17142222 · PMC1669757 · Nucleic acids research · 2007 · 8 claims · 8 setups
The UCSC Genome Browser Database provides sequence and annotation data for 13 vertebrate and 19 invertebrate species as of September 2006.
-
Has reproduction · 30
IsoSCM: improved and alternative 3' UTR annotation using multiple change-point inference.
PMID 25406361 · PMC4274634 · RNA (New York, N.Y.) · 2015 · 8 claims · 6 setups
Existing ab initio assemblers (Cufflinks, Scripture) annotate at most one 3' boundary per terminal exon and therefore cannot assemble coexpressed tandem 3' UTR isoforms.
-
Full-text index only
Anopheles gambiae genome reannotation through synthesis of ab initio and comparative gene prediction algorithms.
PMID 16569258 · PMC1557760 · Genome biology · 2006 · 8 claims · 7 setups
An exon-gene-union (EGU) algorithm followed by an open-reading-frame-selection algorithm can synthesize ab initio (GENSCAN, GeneMark, SNAP) and comparative (Ensembl/Genewise) predictions into a single, more complete CDS set
-
Full-text index only
X:Map: annotation and visualization of genome structure for Affymetrix exon array analysis.
PMID 17932061 · PMC2238884 · Nucleic acids research · 2008 · 7 claims · 4 setups
X:Map is a genome annotation database that maps every Affymetrix exon array probeset to Ensembl genome features (genes, ESTs, GenScan predictions) and supports both high-throughput and gene-centric analysis.
-
Has reproduction · 86
RNASEQR--a streamlined and accurate RNA-seq sequence analysis program.
PMID 22199257 · PMC3315322 · Nucleic acids research · 2012 · 8 claims · 7 setups
RNASEQR is a new RNA-seq mapper/aligner that combines a BWT-based (Bowtie) transcriptomic/genomic alignment with hash-based BLAT local alignment in three sequential steps: transcriptome mapping, novel exon detection, and anchor-and-align novel splice junction identification.
-
Full-text index only
SpliceMiner: a high-throughput database implementation of the NCBI Evidence Viewer for microarray splice variant analysis.
PMID 17338820 · PMC1839109 · BMC bioinformatics · 2007 · 6 claims · 4 setups
EVDB is a comprehensive, non-redundant relational database of known human splice variants built from NCBI Entrez Gene and Evidence Viewer data
-
Has reproduction · 66
HTSstation: a web application and open-access libraries for high-throughput sequencing data analysis.
PMID 24475057 · PMC3903476 · PloS one · 2014 · 8 claims · 5 setups
HTSstation is a web application suite coupling simple web forms to modular analysis pipelines for ChIP-seq, RNA-seq, 4C-seq and re-sequencing HTS applications, accessible at http://htsstation.epfl.ch.
-
Full-text index only
Visualizing the genome: techniques for presenting human genome data and annotations.
PMID 12149135 · PMC119855 · BMC bioinformatics · 2002 · 8 claims · 4 setups
Web-based client-server genome browsers (e.g., LocusLink evidence viewer, UCSC genome browser) are limited by lack of true interactivity, requiring server round-trips for navigation
-
Has reproduction · 65
FusionQ: a novel approach for gene fusion detection and quantification from paired-end RNA-Seq.
PMID 23768108 · PMC3691734 · BMC bioinformatics · 2013 · 8 claims · 8 setups
FusionQ is a novel tool that detects gene fusions, constructs chimerical transcript structures, and estimates their abundances from paired-end RNA-Seq data.
-
Full-text index only
Ultra high throughput sequencing excludes MDH1 as candidate gene for RP28-linked retinitis pigmentosa.
PMID 20011630 · PMC2790479 · Molecular vision · 2009 · 8 claims · 5 setups
MDH1 is not the causative gene for RP28-linked autosomal recessive retinitis pigmentosa
-
Full-text index only
Eukan: a fully automated nuclear genome annotation pipeline for less studied and divergent eukaryotes.
PMID 41567515 · PMC12817076 · NAR genomics and bioinformatics · 2026 · 8 claims · 7 setups
Eukan automatically leverages RNA-Seq coverage to inform generalized Hidden Markov Model gene prediction and intron lengths to inform protein sequence alignments
-
Has reproduction · 94
A Deluge of Complex Repeats: The Solanum Genome.
PMID 26241045 · PMC4524691 · PloS one · 2015 · 8 claims · 8 setups
~50–60% of the genomes of S. tuberosum and S. lycopersicum are composed of repetitive elements
-
Full-text index only
Genome informatics: taming the avalanche of genomic data.
PMID 15642109 · PMC549058 · Genome biology · 2005 · 8 claims · 7 setups
Ultraconserved regions (>100 bp, 100% conserved among mammals) exist in the genome and their function remains unknown
-
Has reproduction · 66
RNAseq analysis of the parasitic nematode Strongyloides stercoralis reveals divergent regulation of canonical dauer pathways.
PMID 23145190 · PMC3493385 · PLoS neglected tropical diseases · 2012 · 8 claims · 8 setups
S. stercoralis possesses homologs of nearly all C. elegans dauer genes, but with significant differences in protein structure, developmental regulation, and gene family expansion.
-
Full-text index only
Expoldb: expression linked polymorphism database with inbuilt tools for analysis of expression and simple repeats.
PMID 17038195 · PMC1618849 · BMC genomics · 2006 · 8 claims · 6 setups
EXPOLDB is a novel database integrating human gene expression variability data (including monozygotic twin comparisons) with (TG/CA)n repeat polymorphism information
-
Full-text index only
Multiplex sequencing of paired-end ditags (MS-PET): a strategy for the ultra-high-throughput analysis of transcriptomes and genomes.
PMID 16840528 · PMC1524903 · Nucleic acids research · 2006 · 7 claims · 5 setups
MS-PET, which dimerizes PETs prior to 454 multiplex sequencing, achieves an approximate 100-fold efficiency increase over standard Sanger-based PET analysis
-
Has reproduction · 57
Analysis and comprehensive comparison of PacBio and nanopore-based RNA sequencing of the Arabidopsis transcriptome.
PMID 32536962 · PMC7291481 · Plant methods · 2020 · 8 claims · 8 setups
ONT Pc produces higher raw data quality (higher alignment rate, lower error rate) than ONT Dc, while PacBio generates the longest reads