Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 75
Identification of Novel Therapeutic Candidates Against SARS-CoV-2 Infections: An Application of RNA Sequencing Toward mRNA Based Nanotherapeutics.
PMID 35983322 · PMC9378778 · Frontiers in microbiology · 2022 · 6 claims · 7 setups
RPL29 (60S ribosomal protein L29) is highly/consistently expressed across all COVID-19 infected groups regardless of severity, suggesting it as a novel host therapeutic target for mRNA-based nanomedicines.
-
Has reproduction · 98
Massively parallel genomic perturbations with multi-target CRISPR interrogates Cas9 activity and DNA repair at endogenous sites.
PMID 36064968 · PMC9481459 · Nature cell biology · 2022 · 8 claims · 6 setups
Multi-target gRNAs (mgRNAs) can direct Cas9 to over a hundred well-mapped endogenous genomic sites simultaneously, enabling massively parallel, high-throughput interrogation of Cas9 activity via short-read sequencing
-
Has reproduction · 64
A novel and dual digestive symbiosis scales up the nutrition and immune system of the holobiont Rimicaris exoculata.
PMID 36333777 · PMC9636832 · Microbiome · 2022 · 7 claims · 6 setups
Genome-resolved metagenomics of separated foregut and midgut reconstructed 20 MAGs including novel lineages of Hepatoplasmataceae (foregut) and Deferribacteres (midgut).
-
Full-text index only
Molecular cloning, genomic characterization and over-expression of a novel gene, XRRA1, identified from human colorectal cancer cell HCT116Clone2_XRR and macaque testis.
PMID 12908878 · PMC194569 · BMC genomics · 2003 · 8 claims · 7 setups
XRRA1 is a novel gene down-regulated ~2-fold in XR-resistant HCT116 Clone2_XRR relative to HCT116 Clone10, identified via cDNA microarray
-
Has reproduction · 66
RNAseq analysis of the parasitic nematode Strongyloides stercoralis reveals divergent regulation of canonical dauer pathways.
PMID 23145190 · PMC3493385 · PLoS neglected tropical diseases · 2012 · 8 claims · 8 setups
S. stercoralis possesses homologs of nearly all C. elegans dauer genes, but with significant differences in protein structure, developmental regulation, and gene family expansion.
-
Has reproduction · 65
FusionQ: a novel approach for gene fusion detection and quantification from paired-end RNA-Seq.
PMID 23768108 · PMC3691734 · BMC bioinformatics · 2013 · 8 claims · 8 setups
FusionQ is a novel tool that detects gene fusions, constructs chimerical transcript structures, and estimates their abundances from paired-end RNA-Seq data.
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.
-
Full-text index only
Shotgun haplotyping: a novel method for surveying allelic sequence variation.
PMID 16221968 · PMC1253838 · Nucleic acids research · 2005 · 8 claims · 7 setups
A novel shotgun haplotyping method generates haplotypic sequences from long PCR products by shotgun sequencing both alleles concurrently and using read-pair information to separate alleles during assembly
-
Has reproduction · 82
Ordinal-level phylogenomics of the arthropod class Diplopoda (millipedes) based on an analysis of 221 nuclear protein-coding loci generated using next-generation sequence analyses.
PMID 24236165 · PMC3827447 · PloS one · 2013 · 8 claims · 8 setups
An ordinal-level phylogeny of Diplopoda reconstructed from 221 nuclear protein-coding loci (61,641 aligned amino acid columns) differs from existing classifications in fundamental ways.
-
Full-text index only
Thirty years into the genomics era: tumor viruses led the way.
PMID 17940630 · PMC1994808 · The Yale journal of biology and medicine · 2006 · 8 claims · 8 setups
Restriction endonuclease-based analysis and sequencing of small tumor virus genomes (SV40, phiX174) established the technical framework later used to sequence bacterial and human genomes
-
Full-text index only
PigGIS: Pig Genomic Informatics System.
PMID 17090590 · PMC1669765 · Nucleic acids research · 2007 · 7 claims · 7 setups
PigGIS identified 15,700 pig consensus sequences covering 18.5 Mb of homologous human exons
-
Has reproduction · 74
Exploring candidate genes for pericarp russet pigmentation of sand pear (Pyrus pyrifolia) via RNA-Seq data in two genotypes contrasting for pericarp color.
PMID 24400075 · PMC3882208 · PloS one · 2014 · 8 claims · 5 setups
RNA-seq-based bulked segregant analysis of russet- vs green-pericarp F1 pools identified 29,100 unigenes, 206 of which were significantly differentially expressed (|log2 fold change| > 1).
-
Full-text index only
Upgrades to StellaBase facilitate medical and genetic studies on the starlet sea anemone, Nematostella vectensis.
PMID 17982171 · PMC2238866 · Nucleic acids research · 2008 · 6 claims · 5 setups
StellaBase Disease houses homology data for 155,904 invertebrate isoforms of human disease genes across four model systems, including 14,874 predicted Nematostella genes
-
Full-text index only
Complete genome sequence and comparative analysis of the wild-type commensal Escherichia coli strain SE11 isolated from a healthy adult.
PMID 18931093 · PMC2608844 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2008 · 8 claims · 6 setups
The SE11 genome comprises a 4.8 Mb chromosome encoding 4679 protein-coding genes and six plasmids encoding 323 protein-coding genes
-
Full-text index only
DDBJ dealing with mass data produced by the second generation sequencer.
PMID 18927114 · PMC2686496 · Nucleic acids research · 2009 · 8 claims · 7 setups
DDBJ collected and released 2,368,110 entries (1,415,106,598 bases) of original DNA sequence data from July 2007 to June 2008.
-
Full-text index only
Rise of the machines.
PMID 18670625 · PMC2467494 · PLoS genetics · 2008 · 8 claims · 4 setups
New short-read sequencing platforms (Illumina Genome Analyzer, 454 FLX, ABI SOLiD) enable rapid, scalable whole-genome resequencing that was previously restricted to dedicated sequencing centers using Sanger methods.
-
Has reproduction · 61
Comprehensive transcriptome study to develop molecular resources of the copepod Calanus sinicus for their potential ecological applications.
PMID 24982883 · PMC4055022 · BioMed research international · 2014 · 8 claims · 8 setups
Illumina RNA-Seq with Trinity de novo assembly produced a C. sinicus transcriptome of 69,751 contigs (average 928.8 bp, N50 1,127 bp) from 58.9 million reads.
-
Has reproduction · 87
A target enrichment method for gathering phylogenetic information from hundreds of loci: An example from the Compositae.
PMID 25202605 · PMC4103609 · Applications in plant sciences · 2014 · 8 claims · 8 setups
A custom sequence capture probe set (9678 baits targeting 1061 orthologous genes) was designed to enrich COS loci across the Compositae.
-
Has reproduction · 67
A consensus approach to vertebrate de novo transcriptome assembly from RNA-seq data: assembly of the duck (Anas platyrhynchos) transcriptome.
PMID 25009556 · PMC4070175 · Frontiers in genetics · 2014 · 8 claims · 8 setups
Multiple k-mer (MK) assemblies are more complete than single k-mer (SK) assemblies, showing higher reads-mapped-back-to-transcripts (RMBT) and higher CEGMA complete-gene percentages for all three tools.
-
Full-text index only
Genome-wide prioritization of disease genes and identification of disease-disease associations from an integrated human functional linkage network.
PMID 19728866 · PMC2768980 · Genome biology · 2009 · 6 claims · 6 setups
Integrating 16 genomic features (32 sub-features) via a naïve Bayes classifier produces a genome-scale FLN of 21,657 human genes and 22,388,609 weighted links that outperforms any individual data source for inferring functional linkages.