Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
Computational verification of protein-protein interactions by orthologous co-expression.
PMID 15740634 · PMC555590 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Co-expression of orthologous protein pairs across multiple species can verify/predict S. cerevisiae PPIs with better performance than S. cerevisiae co-expression alone.
-
Has reproduction · 95
Quantitative epigenetic co-variation in CpG islands and co-regulation of developmental genes.
PMID 23999385 · PMC6505400 · Scientific reports · 2013 · 8 claims · 8 setups
Four epigenetic modifications (DNA methylation, H3K4me2, H3K4me3, H3K27me3) in mouse CGIs undergo combinatorial variation (co-variation) across ESCs, NPCs and adult brain during neuron differentiation.
-
Full-text index only
Inferring combinatorial regulation of transcription in silico.
PMID 15647509 · PMC546154 · Nucleic acids research · 2005 · 8 claims · 5 setups
Combining Cluster-Buster (TFBS cluster prediction) with GOSSIP (rigorous GO enrichment statistics with multiple-testing/FDR correction) predicts biological functions controlled by combinatorial transcription factor action, without prior knowledge of factor targets
-
Full-text index only
PeroxisomeDB: a database for the peroxisomal proteome, functional genomics and disease.
PMID 17135190 · PMC1747181 · Nucleic acids research · 2007 · 8 claims · 6 setups
PeroxisomeDB integrates the complete peroxisomal proteome of Homo sapiens and Saccharomyces cerevisiae into interrelated 'Genes', 'Functions', 'Metabolic pathways' and 'Diseases' sections with links to NCBI, ENSEMBL and UCSC
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
Using several pair-wise informant sequences for de novo prediction of alternatively spliced transcripts.
PMID 16925842 · PMC1810557 · Genome biology · 2006 · 8 claims · 4 setups
MARS, an extension of the Twinscan algorithm, uses multiple pairwise informant genomes to predict human alternatively spliced transcripts de novo without expressed sequence information.
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Full-text index only
Genomes of Helicobacter pylori from native Peruvians suggest admixture of ancestral and modern lineages and reveal a western type cag-pathogenicity island.
PMID 16872520 · PMC1553449 · BMC genomics · 2006 · 8 claims · 7 setups
Native Peruvian H. pylori strains comprise two lineages: predominant hp-Europe and ~20% hsp-Amerind (Amerindian ancestry, closer to Alaska strains)
-
Full-text index only
Differences in the evolutionary history of disease genes affected by dominant or recessive mutations.
PMID 16817963 · PMC1534034 · BMC genomics · 2006 · 8 claims · 8 setups
Dominant disease genes are more conserved at the protein level (mouse orthologues) than recessive disease genes.
-
Full-text index only
Independent component analysis reveals new and biologically significant structures in micro array data.
PMID 16762055 · PMC1557674 · BMC bioinformatics · 2006 · 7 claims · 8 setups
ICA applied to three microarray datasets reveals many biologically significant components, including low-ranking ones not obvious by rank alone
-
Full-text index only
Adaptively inferring human transcriptional subnetworks.
PMID 16760900 · PMC1681499 · Molecular systems biology · 2006 · 8 claims · 7 setups
A multivariate linear spline (MARS-based) model correlating PWM binding scores with log expression ratios can identify active cis-motif combinations in mammalian promoters without requiring gene clustering.
-
Full-text index only
Epigenetics and phenotypic variation in mammals.
PMID 16688527 · PMC3906716 · Mammalian genome : official journal of the International Mammalian Genome Society · 2006 · 8 claims · 8 setups
Epigenetic modifications are mitotically heritable, but the fidelity of meiotic/transgenerational inheritance in mammals is poorly understood and evidence in mammals is scanty.
-
Full-text index only
Heterotachy in mammalian promoter evolution.
PMID 16683025 · PMC1449885 · PLoS genetics · 2006 · 8 claims · 5 setups
The rate of promoter evolution relative to control sequences is not consistent between or within mammalian lineages over time (heterotachy)
-
Full-text index only
A global definition of expression context is conserved between orthologs, but does not correlate with sequence conservation.
PMID 16423292 · PMC1382217 · BMC genomics · 2006 · 7 claims · 6 setups
Expression context is largely conserved between orthologs across four eukaryote species.
-
Full-text index only
Genome-wide identification of human functional DNA using a neutral indel model.
PMID 16410828 · PMC1326222 · PLoS computational biology · 2006 · 8 claims · 8 setups
A neutral indel model predicting a geometric distribution of intergap segment (IGS) lengths fits human-mouse ancestral repeat (AR) alignment data excellently
-
Has reproduction · 48
Rbfox2 controls autoregulation in RNA-binding protein networks.
PMID 24637117 · PMC3967051 · Genes & development · 2014 · 8 claims · 8 setups
Rbfox2 cross-regulates AS-NMD events within RNA-binding protein genes to alter their expression, tuning autoregulatory splicing networks and placing Rbfox2 at a critical node of a multilayer regulatory network.
-
Full-text index only
ARED 3.0: the large and diverse AU-rich transcriptome.
PMID 16381826 · PMC1347415 · Nucleic acids research · 2006 · 7 claims · 6 setups
ARED 3.0 computationally mapped more than 4000 ARE-mRNAs to the human genome, representing 5-8% of human genes.
-
Full-text index only
A Hidden Markov Model to estimate population mixture and allelic copy-numbers in cancers using Affymetrix SNP arrays.
PMID 17996079 · PMC2206057 · BMC bioinformatics · 2007 · 8 claims · 7 setups
An HMM using paired germline genotype calls and tumour allelic SNP intensities can estimate allele-specific copy-numbers, distinguishing events like uniparental disomy from allelic imbalance.
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.