Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 66
Prime editing in mice reveals the essentiality of a single base in driving tissue-specific gene expression.
PMID 33722289 · PMC7962346 · Genome biology · 2021 · 6 claims · 7 setups
A PE2-mediated single-base (C>G) substitution in the Tspan2 CArG box causes cell-specific loss of Tspan2 mRNA in aorta and bladder but not heart or brain, mirroring HDR-mediated 3-bp substitution.
-
Has reproduction · 79
Genome-wide prediction of DNase I hypersensitivity using gene expression.
PMID 29051481 · PMC5715040 · Nature communications · 2017 · 6 claims · 3 setups
Gene expression substantially predicts genome-wide DNase I hypersensitivity (DH), demonstrating transcriptome-based prediction as a feasible approach for regulome mapping
-
Full-text index only
Computational identification of transcriptional regulatory elements in DNA sequence.
PMID 16855295 · PMC1524905 · Nucleic acids research · 2006 · 8 claims · 3 setups
Weight matrix (PWM/PSSM) models of TF binding sites are grounded in biophysical theory of protein-DNA interactions, with position weights corresponding to log-odds contributions to binding free energy
-
Has reproduction · 76
The genome and development-dependent transcriptomes of Pyronema confluens: a window into fungal evolution.
PMID 24068976 · PMC3778014 · PLoS genetics · 2013 · 8 claims · 8 setups
The 50 Mb P. confluens genome with 13,369 predicted protein-coding genes is more characteristic of higher filamentous ascomycetes than of the large, repeat-rich Tuber melanosporum genome, showing that the truffle's expanded genome is not typical of the Pezizales.
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
The Bifidobacterium dentium Bd1 genome sequence reflects its genetic adaptation to the human oral cavity.
PMID 20041198 · PMC2788695 · PLoS genetics · 2009 · 8 claims · 8 setups
The B. dentium Bd1 genome was sequenced to completion, revealing a single circular 2,636,368 bp chromosome with 2,143 predicted ORFs
-
Full-text index only
High resolution analysis of the human transcriptome: detection of extensive alternative splicing independent of transcriptional activity.
PMID 19804644 · PMC2768739 · BMC genetics · 2009 · 8 claims · 6 setups
The human GWSA uses exon body and exon-exon junction probes to directly measure over 280,000 known and predicted splicing events genome-wide.
-
Full-text index only
Inference of transcriptional regulation using gene expression data from the bovine and human genomes.
PMID 17683551 · PMC1978505 · BMC genomics · 2007 · 7 claims · 8 setups
Using human reference promoter sequences is a useful approach for studying gene expression regulation in species with limited or non-existing genomic sequence, such as cattle.
-
Full-text index only
Dyneins across eukaryotes: a comparative genomic analysis.
PMID 17897317 · PMC2239267 · Traffic (Copenhagen, Denmark) · 2007 · 8 claims · 6 setups
Phylogenetic inference identified nine DHC families (two cytoplasmic, seven axonemal) and six IC families (one cytoplasmic)
-
Full-text index only
Twin peaks: the draft human genome sequence.
PMID 11276423 · PMC138909 · Genome biology · 2001 · 8 claims · 8 setups
The predicted number of human genes (~26,000-40,000) is far lower than the widely assumed ~100,000, though downstream RNA/protein complexity can still generate substantial biological complexity.
-
Full-text index only
Molecular phylogeny of the kelch-repeat superfamily reveals an expansion of BTB/kelch proteins in animals.
PMID 13678422 · PMC222960 · BMC bioinformatics · 2003 · 8 claims · 8 setups
The human genome encodes at least 71 kelch-repeat proteins
-
Full-text index only
F-SNP: computationally predicted functional SNPs for disease association studies.
PMID 17986460 · PMC2238878 · Nucleic acids research · 2008 · 6 claims · 8 setups
F-SNP is a database integrating functional effect predictions for SNPs from 16 bioinformatics tools/databases across four categories: splicing, transcription, translation, and post-translation
-
Full-text index only
A genomic pathway approach to a complex disease: axon guidance and Parkinson disease.
PMID 17571925 · PMC1904362 · PLoS genetics · 2007 · 8 claims · 5 setups
A genomic pathway approach using axon-guidance pathway SNPs strongly predicts PD susceptibility, survival free of PD, and age at onset of PD
-
Full-text index only
Rapid identification of PAX2/5/8 direct downstream targets in the otic vesicle by combinatorial use of bioinformatics tools.
PMID 18828907 · PMC2760872 · Genome biology · 2008 · 8 claims · 8 setups
A combinatorial bioinformatics pipeline (evolutionary double filtering comparative genomics, GXD/ZFIN database queries, MEDLINE text mining) can rapidly and specifically identify PAX2/5/8 direct downstream targets in the otic vesicle
-
Full-text index only
ProMiR II: a web server for the probabilistic prediction of clustered, nonclustered, conserved and nonconserved microRNAs.
PMID 16845048 · PMC1538778 · Nucleic acids research · 2006 · 6 claims · 4 setups
ProMiR II improves on the original ProMiR by integrating free energy, G/C ratio, conservation score and entropy for more controllable miRNA prediction
-
Full-text index only
Prediction of candidate primary immunodeficiency disease genes using a support vector machine learning approach.
PMID 19801557 · PMC2780952 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2009 · 6 claims · 3 setups
An SVM trained on 69 binary features of known PID genes can accurately classify PID vs non-PID genes and predict novel candidate PID genes
-
Has reproduction · 84
Improving recombinant protein production by yeast through genome-scale modeling using proteome constraints.
PMID 35624178 · PMC9142503 · Nature communications · 2022 · 7 claims · 5 setups
pcSecYeast, a proteome-constrained genome-scale model integrating metabolism, translation, and detailed secretory pathway processing (translocation, PTMs, folding, misfolding, degradation), was constructed for S. cerevisiae
-
Full-text index only
Non-EST based prediction of exon skipping and intron retention events using Pfam information.
PMID 16204458 · PMC1243800 · Nucleic acids research · 2005 · 7 claims · 5 setups
A novel ab initio method predicts exon skipping and intron retention events using only Pfam domain annotation, via a Viterbi-like dynamic programming algorithm applied to the Pfam alignment.
-
Full-text index only
EGASP: the human ENCODE Genome Annotation Assessment Project.
PMID 16925836 · PMC1810551 · Genome biology · 2006 · 8 claims · 6 setups
Best-performing computational gene prediction methods correctly predict at least one transcript for close to 70% of annotated genes in the ENCODE regions.
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.