Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Large-scale trends in the evolution of gene structures within 11 animal genomes.
PMID 16518452 · PMC1386723 · PLoS computational biology · 2006 · 8 claims · 5 setups
Change in intron–exon gene structure is gradual, clock-like, and largely independent of coding-sequence (protein) evolution
-
Full-text index only
Exonic enhancers are a widespread class of dual-function regulatory elements.
PMID 41927541 · PMC13216554 · Nature communications · 2026 · 8 claims · 8 setups
Many protein-coding exons possess enhancer activity across species, forming a class of candidate Exonic Enhancers (cEEs)
-
Full-text index only
DDBJ in collaboration with mass-sequencing teams on annotation.
PMID 15608189 · PMC539974 · Nucleic acids research · 2005 · 7 claims · 5 setups
DDBJ collected and released 1,066,084 entries (718,072,425 bases) in the past year, including the complete chimpanzee chromosome 22 sequence and silkworm whole-genome shotgun data
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.
-
Full-text index only
Improved reconstruction of transcripts and coding sequences from RNA-seq data.
PMID 41700087 · PMC12910111 · Nucleic acids research · 2026 · 7 claims · 3 setups
GeMoSeq combines combinatorial enumeration of candidate transcripts, splitting heuristics, and likelihood-based (EM) quantification for transcript reconstruction from RNA-seq data
-
Full-text index only
Phylogenetic variation and polymorphism at the toll-like receptor 4 locus (TLR4).
PMID 11104518 · PMC31919 · Genome biology · 2000 · 7 claims · 7 setups
The Tlr4 extracellular domain is far more variable than the cytoplasmic domain, both among mouse strains and among species
-
Full-text index only
The truth about mouse, human, worms and yeast.
PMID 15601543 · PMC3525071 · Human genomics · 2004 · 8 claims · 8 setups
Comparing genomes in pairs or larger sets (mouse-human, C. elegans-C. briggsae, multiple Saccharomyces, human-pufferfish, etc.) reveals unsuspected genes and helps eliminate false-positive gene predictions
-
Full-text index only
Classification of real and pseudo microRNA precursors using local structure-sequence features and support vector machine.
PMID 16381612 · PMC1360673 · BMC bioinformatics · 2005 · 7 claims · 7 setups
A 32-dimensional triplet structure-sequence feature vector combined with SVM (triplet-SVM) can distinguish real human pre-miRNAs from pseudo pre-miRNA hairpins with ~90% accuracy.
-
Full-text index only
A novel nonsense mutation in CRYBB1 associated with autosomal dominant congenital cataract.
PMID 18432316 · PMC2324115 · Molecular vision · 2008 · 7 claims · 5 setups
A novel heterozygous nonsense mutation (c.C737T, p.Q223X) in CRYBB1 is responsible for autosomal dominant congenital nuclear cataract in this family.
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.