Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Mapping proteins to disease terminologies: from UniProt to MeSH.
PMID 18460185 · PMC2367626 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Developed a three-step procedure (disease name extraction, exact matching, partial/similarity-based matching) to map UniProtKB/Swiss-Prot disease names to MeSH terms
-
Has reproduction · 75
Genomic regions and candidate genes selected during the breeding of rice in Vietnam.
PMID 35899250 · PMC9309459 · Evolutionary applications · 2022 · 8 claims · 7 setups
XP-CLR and FST scans identify genomic regions with distorted allele frequency/differentiation patterns resulting from differential selective pressures between Vietnamese rice subpopulations
-
Has reproduction · 67
Optimal scaling of digital transcriptomes.
PMID 24223126 · PMC3819321 · PloS one · 2013 · 8 claims · 8 setups
Fifteen existing and novel transcript-count normalization algorithms can be compared with two novel, mutually independent metrics: the number of "uniform" genes (sufficiently low coefficient of variation after normalization) and low average Spearman correlation between normalized expression profiles of gene pairs.
-
Has reproduction · 67
A consensus approach to vertebrate de novo transcriptome assembly from RNA-seq data: assembly of the duck (Anas platyrhynchos) transcriptome.
PMID 25009556 · PMC4070175 · Frontiers in genetics · 2014 · 8 claims · 8 setups
Multiple k-mer (MK) assemblies are more complete than single k-mer (SK) assemblies, showing higher reads-mapped-back-to-transcripts (RMBT) and higher CEGMA complete-gene percentages for all three tools.
-
Has reproduction · 74
Transcriptome profiling of Giardia intestinalis using strand-specific RNA-seq.
PMID 23555231 · PMC3610916 · PLoS computational biology · 2013 · 8 claims · 8 setups
Most of the G. intestinalis genome is transcribed in in vitro-grown trophozoites, but at vastly different expression levels.
-
Has reproduction · 86
RNASEQR--a streamlined and accurate RNA-seq sequence analysis program.
PMID 22199257 · PMC3315322 · Nucleic acids research · 2012 · 8 claims · 7 setups
RNASEQR is a new RNA-seq mapper/aligner that combines a BWT-based (Bowtie) transcriptomic/genomic alignment with hash-based BLAT local alignment in three sequential steps: transcriptome mapping, novel exon detection, and anchor-and-align novel splice junction identification.
-
Full-text index only
Interactome-transcriptome analysis reveals the high centrality of genes differentially expressed in lung cancer tissues.
PMID 16188928 · PMC4631381 · Bioinformatics (Oxford, England) · 2005 · 7 claims · 4 setups
Genes upregulated in squamous cell lung cancer are highly connected (well-connected) nodes in the protein interactome
-
Full-text index only
Multiplex sequencing of paired-end ditags (MS-PET): a strategy for the ultra-high-throughput analysis of transcriptomes and genomes.
PMID 16840528 · PMC1524903 · Nucleic acids research · 2006 · 7 claims · 5 setups
MS-PET, which dimerizes PETs prior to 454 multiplex sequencing, achieves an approximate 100-fold efficiency increase over standard Sanger-based PET analysis
-
Full-text index only
Conserved positive selection signals in gp41 across multiple subtypes and difference in selection signals detectable in gp41 sequences sampled during acute and chronic HIV-1 subtype C infection.
PMID 19025632 · PMC2630941 · Virology journal · 2008 · 8 claims · 4 setups
Twelve gp41 sites (outside the overlapping rev exon2 reading frame) show positive selection conserved across multiple HIV-1 M subtypes/CRFs, making them candidate targets for broadly protective vaccines.
-
Has reproduction · 57
Analysis and comprehensive comparison of PacBio and nanopore-based RNA sequencing of the Arabidopsis transcriptome.
PMID 32536962 · PMC7291481 · Plant methods · 2020 · 8 claims · 8 setups
ONT Pc produces higher raw data quality (higher alignment rate, lower error rate) than ONT Dc, while PacBio generates the longest reads
-
Has reproduction · 67
Satellitome Analysis and Transposable Elements Comparison in Geographically Distant Populations of Spodoptera frugiperda.
PMID 35455012 · PMC9026859 · Life (Basel, Switzerland) · 2022 · 8 claims · 5 setups
Most transposable elements are commonly shared across all eight geographically distant S. frugiperda samples, except Maverick and PIF/Harbinger elements which show divergent repeat copies
-
Has reproduction · 65
FusionQ: a novel approach for gene fusion detection and quantification from paired-end RNA-Seq.
PMID 23768108 · PMC3691734 · BMC bioinformatics · 2013 · 8 claims · 8 setups
FusionQ is a novel tool that detects gene fusions, constructs chimerical transcript structures, and estimates their abundances from paired-end RNA-Seq data.
-
Has reproduction · 74
Exploring candidate genes for pericarp russet pigmentation of sand pear (Pyrus pyrifolia) via RNA-Seq data in two genotypes contrasting for pericarp color.
PMID 24400075 · PMC3882208 · PloS one · 2014 · 8 claims · 5 setups
RNA-seq-based bulked segregant analysis of russet- vs green-pericarp F1 pools identified 29,100 unigenes, 206 of which were significantly differentially expressed (|log2 fold change| > 1).
-
Full-text index only
Genomic diversity and evolution of Mycobacterium ulcerans revealed by next-generation sequencing.
PMID 19806175 · PMC2736377 · PLoS pathogens · 2009 · 8 claims · 6 setups
Genome sequencing of three M. ulcerans strains (NM20/02, NM31/04, Jp8756) identified thousands of SNPs relative to reference strain Agy99
-
Full-text index only
The UCSC Proteome Browser.
PMID 15608236 · PMC540054 · Nucleic acids research · 2005 · 8 claims · 5 setups
The UCSC Proteome Browser is tightly integrated with the UCSC Genome Browser, giving users simultaneous access to genome and proteome data.
-
Full-text index only
FatiGO +: a functional profiling tool for genomic data. Integration of functional annotation, regulatory motifs and interaction data with microarray experiments.
PMID 17478504 · PMC1933151 · Nucleic acids research · 2007 · 8 claims · 8 setups
FatiGO+ is a web-based tool for functional profiling of genome-scale experiments that integrates functional annotation, regulatory motifs and interaction data