Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The human L-threonine 3-dehydrogenase gene is an expressed pseudogene.
PMID 12361482 · PMC131051 · BMC genetics · 2002 · 8 claims · 7 setups
The human TDH gene is located at chromosome 8p23-22, spans 10 kb, and has 8 exons that would be expected to encode a 369-residue ORF.
-
Full-text index only
Identification and functional analyses of 11,769 full-length human cDNAs focused on alternative splicing.
PMID 19880432 · PMC2780955 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2009 · 8 claims · 5 setups
Identified 23,241 human genes transcribed into protein-coding mRNAs using full-length cDNA and 5'-EST sequence data
-
Has reproduction · 60
TRAPID 2.0: a web application for taxonomic and functional analysis of de novo transcriptomes.
PMID 34197621 · PMC8464036 · Nucleic acids research · 2021 · 8 claims · 8 setups
TRAPID 2.0 is a web application performing global characterization of de novo transcriptomes via structural, functional, and taxonomic annotation in an initial processing phase, followed by an exploratory phase of downstream analyses.
-
Has reproduction · 61
Comprehensive transcriptome study to develop molecular resources of the copepod Calanus sinicus for their potential ecological applications.
PMID 24982883 · PMC4055022 · BioMed research international · 2014 · 8 claims · 8 setups
Illumina RNA-Seq with Trinity de novo assembly produced a C. sinicus transcriptome of 69,751 contigs (average 928.8 bp, N50 1,127 bp) from 58.9 million reads.
-
Has reproduction · 95
transXpress: a Snakemake pipeline for streamlined de novo transcriptome assembly and annotation.
PMID 37016291 · PMC10074830 · BMC bioinformatics · 2023 · 6 claims · 7 setups
transXpress is a Snakemake pipeline that streamlines de novo transcriptome assembly, quantification, and annotation for non-model organisms
-
Full-text index only
Species-specific protein sequence and fold optimizations.
PMID 12487631 · PMC139977 · BMC bioinformatics · 2002 · 7 claims · 7 setups
Environmental niche is a significant factor explaining variability in amino acid composition across 100 complete genomes
-
Full-text index only
Coverage of whole proteome by structural genomics observed through protein homology modeling database.
PMID 17146617 · PMC1769342 · Journal of structural and functional genomics · 2006 · 8 claims · 7 setups
FAMSBASE, a homology-modeling database of whole-genome ORFs, currently covers about 50% of predicted ORFs (368,724 of 734,193) across 276 genomes with modeled 3D structures.
-
Full-text index only
Evidence for a novel gene associated with human influenza A viruses.
PMID 19917120 · PMC2780412 · Virology journal · 2009 · 8 claims · 8 setups
A 167-codon ORF (NEG8) on the negative-sense genomic strand of segment 8 is associated with early-20th-century human influenza A isolates
-
Full-text index only
Interactome networks: the state of the science.
PMID 16515723 · PMC1431712 · Genome biology · 2006 · 8 claims · 8 setups
Spastin interacts with CHMP1B, an ESCRT-III-associated protein, supporting a role for spastin in intracellular membrane trafficking relevant to hereditary spastic paraplegia
-
Full-text index only
hORFeome v3.1: a resource of human open reading frames representing over 10,000 human genes.
PMID 17207965 · PMC4647941 · Genomics · 2007 · 8 claims · 7 setups
hORFeome v3.1 is a resource of 12,212 cloned human ORFs representing 10,214 genes, a 51% expansion over hORFeome v1.1
-
Has reproduction · 76
What the Phage: a scalable workflow for the identification and analysis of phage sequences.
PMID 36399058 · PMC9673492 · GigaScience · 2022 · 8 claims · 7 setups
WtP combines 11 tools (14 approaches) for phage prediction in a parallel, containerized Nextflow workflow
-
Full-text index only
Using DNA microarrays to study host-microbe interactions.
PMID 10998383 · PMC2627958 · Emerging infectious diseases · 2000 · 8 claims · 8 setups
DNA microarrays can measure transcript levels and detect sequence polymorphisms for every gene simultaneously in microbial genomes
-
Full-text index only
Genetic diversity among five T4-like bacteriophages.
PMID 16716236 · PMC1524935 · Virology journal · 2006 · 8 claims · 8 setups
A core set of 82 conserved genes (T4-like genes) is present in all five genomes analyzed, clustered in large collinear blocks.
-
Has reproduction · 50
Comparative analysis of circular RNAs between soybean cytoplasmic male-sterile line NJCMS1A and its maintainer NJCMS1B by high-throughput sequencing.
PMID 30208848 · PMC6134632 · BMC genomics · 2018 · 8 claims · 7 setups
2867 circRNAs were identified in soybean flower buds via high-throughput sequencing with RNase R enrichment, of which 1009 were differentially expressed between NJCMS1A and NJCMS1B
-
Has reproduction · 24
MiGPC: a comprehensive catalog of enzybiotics from environmental metagenomes.
PMID 41888223 · PMC13172421 · Scientific reports · 2026 · 8 claims · 8 setups
MiGPC is the first genome-resolved metagenomic gene and protein catalog specifically targeted to enzybiotics
-
Full-text index only
Characterization of 954 bovine full-CDS cDNA sequences.
PMID 16305752 · PMC1314900 · BMC genomics · 2005 · 7 claims · 8 setups
954 bovine full-length insert cDNA (bFLIC) clones representing 762 distinct loci were sequenced and characterized
-
Full-text index only
Distribution and effects of nonsense polymorphisms in human genes.
PMID 18852891 · PMC2561068 · PloS one · 2008 · 8 claims · 8 setups
Nonsense SNPs occur at a lower density than nonsynonymous SNPs, indicating stronger purifying selection against premature stop codons than amino acid changes.
-
Full-text index only
Genomic and bioinformatics analysis of human adenovirus type 37: new insights into corneal tropism.
PMID 18471294 · PMC2397415 · BMC genomics · 2008 · 7 claims · 7 setups
The complete genome of HAdV-37 was sequenced and annotated (35,213 bp, 56.6% GC content, 35 predicted coding sequences plus 8 hypothetical ORFs)
-
Full-text index only
The Bifidobacterium dentium Bd1 genome sequence reflects its genetic adaptation to the human oral cavity.
PMID 20041198 · PMC2788695 · PLoS genetics · 2009 · 8 claims · 8 setups
The B. dentium Bd1 genome was sequenced to completion, revealing a single circular 2,636,368 bp chromosome with 2,143 predicted ORFs
-
Has reproduction · 62
Bayesian prediction of RNA translation from ribosome profiling.
PMID 28126919 · PMC5389577 · Nucleic acids research · 2017 · 8 claims · 4 setups
Rp-Bp is an unsupervised Bayesian approach that uses a two-component 'high-low-low' mixture model to predict translated ORFs from ribosome profiles