Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Evidence for a preferential targeting of 3'-UTRs by cis-encoded natural antisense transcripts.
PMID 16204454 · PMC1243798 · Nucleic acids research · 2005 · 8 claims · 4 setups
Cis-encoded natural antisense RNAs show striking preferential complementarity to 3′-UTRs of their target genes in human and mouse genomes
-
Full-text index only
Comparative analysis of plant genomes allows the definition of the "Phytolongins": a novel non-SNARE longin domain protein family.
PMID 19889231 · PMC2779197 · BMC genomics · 2009 · 8 claims · 6 setups
A novel, plant-specific family of longin-related proteins, the 'Phytolongins', was identified in land plant genomes.
-
Has reproduction · 76
Organelle Genomes and Transcriptomes of Nymphaea Reveal the Interplay between Intron Splicing and RNA Editing.
PMID 34576004 · PMC8466565 · International journal of molecular sciences · 2021 · 8 claims · 7 setups
Multiple partially or fully intron-spliced intermediates co-exist within an organelle, and both cis- and trans-splicing introns are spliced randomly (no fixed order), generating diverse intermediates.
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.
-
Full-text index only
The truth about mouse, human, worms and yeast.
PMID 15601543 · PMC3525071 · Human genomics · 2004 · 8 claims · 8 setups
Comparing genomes in pairs or larger sets (mouse-human, C. elegans-C. briggsae, multiple Saccharomyces, human-pufferfish, etc.) reveals unsuspected genes and helps eliminate false-positive gene predictions
-
Full-text index only
Genome-wide survey for biologically functional pseudogenes.
PMID 16680195 · PMC1456316 · PLoS computational biology · 2006 · 8 claims · 6 setups
A subset of ancient, cross-species-conserved pseudogenes (30 of 1,453 candidate quartets) show evidence consistent with retained biological function
-
Has reproduction · 38
Genomic capacities for Reactive Oxygen Species metabolism across marine phytoplankton.
PMID 37098087 · PMC10128935 · PloS one · 2023 · 8 claims · 3 setups
Genes encoding superoxide (O2•−) scavenging are ubiquitous across phytoplankton, but their fractional gene allocation decreases with increasing cell radius, consistent with a nearly fixed core gene set.
-
Full-text index only
Ensembl 2007.
PMID 17148474 · PMC1761443 · Nucleic acids research · 2007 · 8 claims · 7 setups
Ensembl added 18 new chordate genomes this year, increasing total genomes available from 15 to 33, the largest yearly increase to date.
-
Full-text index only
Gene duplication: the genomic trade in spare parts.
PMID 15252449 · PMC449868 · PLoS biology · 2004 · 8 claims · 7 setups
Gene duplication relaxes selective constraint on one copy, allowing exploration of evolutionary space that is otherwise forbidden in single-copy genes, making duplication the major opportunity for new gene function evolution.
-
Full-text index only
ECgene: genome annotation for alternative splicing.
PMID 15608289 · PMC540072 · Nucleic acids research · 2005 · 8 claims · 5 setups
ECgene combines genome-based EST clustering with a graph-theoretic transcript assembly procedure to predict gene models including alternative splicing events.
-
Full-text index only
A new family of giardial cysteine-rich non-VSP protein genes and a novel cyst protein.
PMID 17183673 · PMC1762436 · PloS one · 2006 · 8 claims · 7 setups
HCNCp is a novel invariant (non-variant) cyst protein belonging to a new family of high-cysteine membrane proteins (HCMp) abundant in the Giardia genome
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
miRBase: tools for microRNA genomics.
PMID 17991681 · PMC2238936 · Nucleic acids research · 2008 · 8 claims · 6 setups
miRBase release 10.0 contains 5071 miRNA hairpin loci from 58 species, expressing 5922 distinct mature miRNA sequences, a growth of over 2000 sequences in 2 years
-
Full-text index only
SelenoDB 1.0 : a database of selenoprotein genes, proteins and SECIS elements.
PMID 18174224 · PMC2238826 · Nucleic acids research · 2008 · 6 claims · 5 setups
Standard genome annotation pipelines misannotate selenoprotein genes because they rely on UGA as a universal stop codon, failing to recognize its dual role as the selenocysteine-recoding codon.
-
Full-text index only
From microarrays to genome duplications.
PMID 12914655 · PMC193639 · Genome biology · 2003 · 8 claims · 8 setups
Gene3D shows that most genes across sequenced genomes can be assigned to known structural domain families, many of which are shared across kingdoms of life
-
Full-text index only
NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins.
PMID 15608248 · PMC539979 · Nucleic acids research · 2005 · 7 claims · 5 setups
RefSeq provides a curated, non-redundant, explicitly linked collection of genomic, transcript and protein sequences spanning prokaryotes, eukaryotes and viruses.
-
Full-text index only
Kinetoplastid genomics: the thin end of the wedge.
PMID 18675383 · PMC2676795 · Infection, genetics and evolution : journal of molecular epidemiology and evolutionary genetics in infectious diseases · 2008 · 8 claims · 8 setups
Completion of the T. brucei, T. cruzi, and L. major genome sequencing projects enabled numerous studies that would otherwise have been difficult or impossible.
-
Full-text index only
Evola: Ortholog database of all human genes in H-InvDB with manual curation of phylogenetic trees.
PMID 17982176 · PMC2238928 · Nucleic acids research · 2008 · 6 claims · 7 setups
Evola combines genome synteny-based computational ortholog detection with manual curation of phylogenetic trees by experts to yield more reliable orthologs than automated pairwise methods
-
Has reproduction · 83
MetaGT: A pipeline for de novo assembly of metatranscriptomes with the aid of metagenomic data.
PMID 36386613 · PMC9651917 · Frontiers in microbiology · 2022 · 7 claims · 4 setups
MetaGT is a pipeline that combines metatranscriptomic and metagenomic data from the same sample to assemble complete transcript sequences
-
Has reproduction · 74
Transcriptome profiling of Giardia intestinalis using strand-specific RNA-seq.
PMID 23555231 · PMC3610916 · PLoS computational biology · 2013 · 8 claims · 8 setups
Most of the G. intestinalis genome is transcribed in in vitro-grown trophozoites, but at vastly different expression levels.