Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A unique, consistent identifier for alternatively spliced transcript variants.
PMID 19865484 · PMC2765725 · PloS one · 2009 · 6 claims · 1 setups
Existing transcript identifiers (NM_ accessions, ENST identifiers) are unsuitable for uniquely identifying isoform structure across databases, methods, or organisms
-
Full-text index only
Widespread A-to-I RNA editing of Alu-containing mRNAs in the human transcriptome.
PMID 15534692 · PMC526178 · PLoS biology · 2004 · 8 claims · 6 setups
Intramolecular pairs of oppositely oriented Alu elements within the same pre-mRNA form dsRNA foldback structures that are major substrates for A-to-I RNA editing
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences
-
Full-text index only
TRED: a Transcriptional Regulatory Element Database and a platform for in silico gene regulation studies.
PMID 15608156 · PMC539958 · Nucleic acids research · 2005 · 8 claims · 5 setups
TRED is a database collecting both cis-regulatory elements (promoters) and trans-regulatory elements (transcription factor binding/regulation data) with linked access.
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
TRED: a transcriptional regulatory element database, new entries and other development.
PMID 17202159 · PMC1899102 · Nucleic acids research · 2007 · 8 claims · 3 setups
TRED collects mammalian cis- and trans-regulatory elements together with experimental evidence, mapped onto assembled genomes
-
Full-text index only
Aberrant 5' splice sites in human disease genes: mutation pattern, nucleotide structure and comparison of computational tools that predict their utilization.
PMID 17576681 · PMC1934990 · Nucleic acids research · 2007 · 8 claims · 4 setups
Cryptic 5'ss are best predicted by computational algorithms that accommodate nucleotide dependencies (e.g., Markov model, maximum entropy, maximum dependence decomposition) rather than by weight-matrix models
-
Full-text index only
NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins.
PMID 15608248 · PMC539979 · Nucleic acids research · 2005 · 7 claims · 5 setups
RefSeq provides a curated, non-redundant, explicitly linked collection of genomic, transcript and protein sequences spanning prokaryotes, eukaryotes and viruses.
-
Full-text index only
SECIS elements in the coding regions of selenoprotein transcripts are functional in higher eukaryotes.
PMID 17169995 · PMC1802603 · Nucleic acids research · 2007 · 8 claims · 5 setups
SECIS elements located within coding regions of selenoprotein mRNAs support functional Sec insertion in mammalian cells
-
Full-text index only
Genome-wide analyses of retrogenes derived from the human box H/ACA snoRNAs.
PMID 17175533 · PMC1802619 · Nucleic acids research · 2007 · 8 claims · 6 setups
202 novel box H/ACA RNA-related sequences were identified in the human genome
-
Full-text index only
Identification of the proliferation/differentiation switch in the cellular network of multicellular organisms.
PMID 17166053 · PMC1664705 · PLoS computational biology · 2006 · 8 claims · 8 setups
Integrating interactome and transcriptome data reveals a pair of transcriptionally anticorrelated network modules (P and D) each comprising hundreds of genes, present across individuals and species.
-
Full-text index only
Computational identification of transcriptional regulatory elements in DNA sequence.
PMID 16855295 · PMC1524905 · Nucleic acids research · 2006 · 8 claims · 3 setups
Weight matrix (PWM/PSSM) models of TF binding sites are grounded in biophysical theory of protein-DNA interactions, with position weights corresponding to log-odds contributions to binding free energy
-
Full-text index only
The relationship of potential G-quadruplex sequences in cis-upstream regions of the human genome to SP1-binding elements.
PMID 18353860 · PMC2377421 · Nucleic acids research · 2008 · 7 claims · 1 setups
A large number of upstream PQSSs incorporate the SP1-binding element, establishing a clear link between PQSS occurrence and SP1 elements
-
Full-text index only
Expansion of the BioCyc collection of pathway/genome databases to 160 genomes.
PMID 16246909 · PMC1266070 · Nucleic acids research · 2005 · 8 claims · 6 setups
The BioCyc collection has been expanded to 160 pathway/genome databases (PGDBs) organized into three curation tiers.
-
Full-text index only
Identifying the important HIV-1 recombination breakpoints.
PMID 18787691 · PMC2522274 · PLoS computational biology · 2008 · 8 claims · 3 setups
Local sequence identity between co-packaged parental RNAs strongly influences the probability of strand-transfer/breakpoint location, with fewer breakpoints occurring near mismatches
-
Has reproduction · 57
Data-driven projections of candidate enhancer-activating SNPs in immune regulation.
PMID 40011812 · PMC11863423 · BMC genomics · 2025 · 7 claims · 7 setups
A data-driven computational protocol combining motif scanning, open-chromatin filtering, gene proximity, dbSNP validation, spacing, and cross-species conservation can prioritize SNPs likely to create functional GAS motifs.
-
Full-text index only
Biased exon/intron distribution of cryptic and de novo 3' splice sites.
PMID 16141195 · PMC1197134 · Nucleic acids research · 2005 · 7 claims · 5 setups
Cryptic 3'ss (from 3'YAG consensus mutations) are significantly more frequent in exons than in introns
-
Full-text index only
Computer identification of snoRNA genes using a Mammalian Orthologous Intron Database.
PMID 16093549 · PMC1184218 · Nucleic acids research · 2005 · 8 claims · 5 setups
Created the Mammalian Orthologous Intron Database (MOID) containing orthologous introns of human, mouse and rat identified via conserved reading-frame position
-
Full-text index only
Towards a comprehensive structural coverage of completed genomes: a structural genomics viewpoint.
PMID 17349043 · PMC1829165 · BMC bioinformatics · 2007 · 8 claims · 6 setups
A combined target-selection approach — pursuing both structurally uncharacterised domain families and additional targets from large structurally characterised superfamilies — is essential for comprehensive structural coverage of the genomes.
-
Full-text index only
Comparative Toxicogenomics Database: a knowledgebase and discovery tool for chemical-gene-disease networks.
PMID 18782832 · PMC2686584 · Nucleic acids research · 2009 · 8 claims · 5 setups
CTD is a manually curated knowledgebase that integrates chemical-gene interactions, chemical-disease relationships, and gene-disease relationships into a chemical-gene-disease triad