Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Correlation of microsynteny conservation and disease gene distribution in mammalian genomes.
PMID 19909546 · PMC2779822 · BMC genomics · 2009 · 7 claims · 8 setups
Density of mouse orthologs of human disease genes correlates with regions of conserved microsynteny in the mouse genome
-
Full-text index only
miRNAMap 2.0: genomic maps of microRNAs in metazoan genomes.
PMID 18029362 · PMC2238982 · Nucleic acids research · 2008 · 8 claims · 6 setups
miRNAMap 2.0 is a resource collecting experimentally verified miRNAs and experimentally verified miRNA target genes in human, mouse, rat and other metazoan genomes
-
Full-text index only
MtSNPscore: a combined evidence approach for assessing cumulative impact of mitochondrial variations in disease.
PMID 19758471 · PMC2745589 · BMC bioinformatics · 2009 · 8 claims · 5 setups
MtSNPscore, a weighted scoring pipeline combining literature evidence, in silico predictions, and case/control frequency, can prioritize likely pathogenic mtDNA variations
-
Full-text index only
Expansion of the BioCyc collection of pathway/genome databases to 160 genomes.
PMID 16246909 · PMC1266070 · Nucleic acids research · 2005 · 8 claims · 6 setups
The BioCyc collection has been expanded to 160 pathway/genome databases (PGDBs) organized into three curation tiers.
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
Comparative genomic study reveals a transition from TA richness in invertebrates to GC richness in vertebrates at CpG flanking sites: an indication for context-dependent mutagenicity of methylated CpG sites.
PMID 19329065 · PMC5054122 · Genomics, proteomics & bioinformatics · 2008 · 8 claims · 8 setups
Nucleotide preference at CpG flanking sites transitions from 5' T (invertebrates) to 5' A (vertebrates) at the invertebrate-vertebrate boundary
-
Full-text index only
Pseudofam: the pseudogene families database.
PMID 18957444 · PMC2686518 · Nucleic acids research · 2009 · 8 claims · 7 setups
Pseudofam is an online database of pseudogene families built by mapping pseudogenes to Pfam protein families, providing query tools, statistics, and sequence alignments
-
Has reproduction · 38
Genomic capacities for Reactive Oxygen Species metabolism across marine phytoplankton.
PMID 37098087 · PMC10128935 · PloS one · 2023 · 8 claims · 3 setups
Genes encoding superoxide (O2•−) scavenging are ubiquitous across phytoplankton, but their fractional gene allocation decreases with increasing cell radius, consistent with a nearly fixed core gene set.
-
Full-text index only
Insights into the coupling of duplication events and macroevolution from an age profile of animal transmembrane gene families.
PMID 16895434 · PMC1534073 · PLoS computational biology · 2006 · 8 claims · 7 setups
The density of transmembrane gene duplicates positively correlates with the estimated maximum number of cell types of common ancestors
-
Full-text index only
EPGD: a comprehensive web resource for integrating and displaying eukaryotic paralog/paralogon information.
PMID 17984073 · PMC2238967 · Nucleic acids research · 2008 · 8 claims · 8 setups
EPGD is a gene-centered, internet-accessible database integrating paralog family and paralogon information for 26 eukaryotic genomes.
-
Full-text index only
A pharmacogene database enhanced by the 1000 Genomes Project.
PMID 19745786 · PMC2935084 · Pharmacogenetics and genomics · 2009 · 7 claims · 4 setups
The database provides a convenient portal for immediate utilization of newly released 1000 Genomes Project (KGP) data in pharmacogenetic studies
-
Full-text index only
From microarrays to genome duplications.
PMID 12914655 · PMC193639 · Genome biology · 2003 · 8 claims · 8 setups
Gene3D shows that most genes across sequenced genomes can be assigned to known structural domain families, many of which are shared across kingdoms of life
-
Has reproduction · 80
Recombination Facilitates Adaptive Evolution in Rhizobial Soil Bacteria.
PMID 34410427 · PMC8662638 · Molecular biology and evolution · 2021 · 8 claims · 7 setups
α varies from 0.07 to 0.39 across five Rhizobium species and is positively correlated with the level of recombination
-
Full-text index only
Does distance matter? Variations in alternative 3' splicing regulation.
PMID 17704130 · PMC2018619 · Nucleic acids research · 2007 · 8 claims · 7 setups
Alternative 3' splice sites can be distinguished from constitutive splice sites by a combination of sequence/conservation properties that vary depending on the distance between the splice sites.
-
Full-text index only
CpG_MI: a novel approach for identifying functional CpG islands in mammalian genomes.
PMID 19854943 · PMC2800233 · Nucleic acids research · 2010 · 8 claims · 6 setups
Functional ('bona fide') CGIs show distinct average/cumulative mutual information (AMI/CMI) distributions of neighboring CpG distances compared to non-functional CGIs and random genome segments
-
Full-text index only
Gene loss rate: a probabilistic measure for the conservation of eukaryotic genes.
PMID 17158152 · PMC1802574 · Nucleic acids research · 2007 · 8 claims · 8 setups
GLR is a novel maximum-likelihood measure of gene loss rate that probabilistically weighs all possible ancestral phyletic patterns rather than relying on a single parsimonious reconstruction.
-
Full-text index only
Protein co-evolution, co-adaptation and interactions.
PMID 18818697 · PMC2556093 · The EMBO journal · 2008 · 8 claims · 6 setups
The mirrortree method predicts protein-protein interactions by detecting pairs of protein families with similar phylogenetic trees (quantified as Pearson correlation of sequence similarity matrices).
-
Full-text index only
The integrated world of functional genomics.
PMID 12537543 · PMC151279 · Genome biology · 2003 · 8 claims · 8 setups
Integrating chromatin immunoprecipitation (promoter-binding) data with expression data reveals the yeast cell-cycle transcriptional regulatory network, including network motifs such as autoregulation, multi-component loops, and feedforward loops.
-
Full-text index only
The Origin at 150: is a new evolutionary synthesis in sight?
PMID 19836100 · PMC2784144 · Trends in genetics : TIG · 2009 · 8 claims · 4 setups
The Modern Synthesis (neo-Darwinism) has crumbled and its central tenets are overturned or radically revised in the post-genomic era
-
Has reproduction · 50
MEDUSA: A Pipeline for Sensitive Taxonomic Classification and Flexible Functional Annotation of Metagenomic Shotgun Sequences.
PMID 35330728 · PMC8940201 · Frontiers in genetics · 2022 · 6 claims · 6 setups
MEDUSA is an automated, Conda-installable and Snakemake-managed pipeline performing preprocessing, assembly, alignment, taxonomic classification, and functional annotation on shotgun data.