Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
GeneTrail--advanced gene set enrichment analysis.
PMID 17526521 · PMC1933132 · Nucleic acids research · 2007 · 8 claims · 2 setups
GeneTrail is a comprehensive, easy-to-use web-based tool for gene set enrichment analysis supporting both Over-Representation Analysis (ORA) and Gene Set Enrichment Analysis (GSEA)
-
Full-text index only
DAVID Bioinformatics Resources: expanded annotation database and novel algorithms to better extract biology from large gene lists.
PMID 17576678 · PMC1933169 · Nucleic acids research · 2007 · 8 claims · 4 setups
The DAVID Gene Concept uses a single-linkage method to agglomerate tens of millions of gene/protein identifiers from NCBI, PIR, UniProt and other resources into unified DAVID genes.
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Full-text index only
A general definition and nomenclature for alternative splicing events.
PMID 18688268 · PMC2467475 · PLoS computational biology · 2008 · 6 claims · 4 setups
Existing AS nomenclatures (Malko et al.'s 5-letter strings, Nagasaki et al.'s bit matrices, and the ASD/ATD/AEdb system) are redundant, ambiguous, or incapable of representing complex or large splicing variations.
-
Full-text index only
BTW: a web server for Boltzmann time warping of gene expression time series.
PMID 16845055 · PMC1538860 · Nucleic acids research · 2006 · 5 claims · 4 setups
Symmetric time warping distance is more flexible than Euclidean distance or correlation coefficient for identifying genes with similar temporal expression profiles, especially across sequences of different length.
-
Full-text index only
Multiple whole genome alignments and novel biomedical applications at the VISTA portal.
PMID 17488840 · PMC1933192 · Nucleic acids research · 2007 · 8 claims · 4 setups
A novel multiple whole-genome alignment algorithm treats all genomes symmetrically, avoiding dependence on a single base/reference genome
-
Full-text index only
Gene loss rate: a probabilistic measure for the conservation of eukaryotic genes.
PMID 17158152 · PMC1802574 · Nucleic acids research · 2007 · 8 claims · 8 setups
GLR is a novel maximum-likelihood measure of gene loss rate that probabilistically weighs all possible ancestral phyletic patterns rather than relying on a single parsimonious reconstruction.
-
Full-text index only
Computational analysis of the synergy among multiple interacting genes.
PMID 17299419 · PMC1828751 · Molecular systems biology · 2007 · 8 claims · 3 setups
Multivariate synergy of a set of factors with respect to a phenotype can be defined via the maximum-information partition, i.e., comparing the mutual information of the full set to the best achievable sum of mutual information over any partition into disjoint subsets.
-
Full-text index only
Onto-Tools: new additions and improvements in 2006.
PMID 17584796 · PMC1933142 · Nucleic acids research · 2007 · 8 claims · 3 setups
OE2GO enables functional profiling for organisms lacking public-domain annotations by allowing users to supply custom GO-format annotation files and OBO-format ontology files
-
Full-text index only
Designing candidate gene and genome-wide case-control association studies.
PMID 17947991 · PMC4180089 · Nature protocols · 2007 · 8 claims · 4 setups
Common complex diseases are influenced by multiple genetic and environmental factors that individually are neither necessary nor sufficient to cause disease
-
Full-text index only
Testing whether genetic variation explains correlation of quantitative measures of gene expression, and application to genetic network analysis.
PMID 18444230 · PMC2729096 · Statistics in medicine · 2008 · 8 claims · 3 setups
A statistical test (delta method and Steiger-Browne optimal linear composites) is developed to test equality of the marginal correlation and the partial correlation of two gene expression traits conditional on a set of covariates.
-
Has reproduction · 67
Reducing language barriers, promoting information absorption, and communication using fanyi.
PMID 39039634 · PMC11332769 · Chinese medical journal · 2024 · 7 claims · 3 setups
The fanyi R package retrieves gene information from NCBI and translates it into multiple languages using AI-driven online translation services
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.
-
Full-text index only
BABELOMICS: a systems biology perspective in the functional annotation of genome-scale experiments.
PMID 16845052 · PMC1538844 · Nucleic acids research · 2006 · 8 claims · 8 setups
Babelomics is presented as an updated, complete suite of web tools for functional analysis of genome-scale experiments with new and improved modules
-
Full-text index only
Optimality driven nearest centroid classification from genomic data.
PMID 17912341 · PMC1991588 · PloS one · 2007 · 7 claims · 5 setups
A theoretical result determines the subset of features of a given size that minimizes the misclassification rate for a nearest-centroid (LDA) classifier, based on equation (4).
-
Full-text index only
Proteomic solutions for analytical challenges associated with alcohol research.
PMID 23584870 · PMC3860482 · Alcohol research & health : the journal of the National Institute on Alcohol Abuse and Alcoholism · 2008 · 7 claims · 4 setups
Protein-level meta-analyses analogous to the transcriptome meta-analysis by Mulligan et al. (2006) are not yet possible because proteins lack a uniform sample preparation/analysis method and span up to 8 orders of magnitude in abundance.
-
Full-text index only
SNPHunter: a bioinformatic software for single nucleotide polymorphism data acquisition and management.
PMID 15774022 · PMC1274256 · BMC bioinformatics · 2005 · 7 claims · 3 setups
SNPHunter allows ad hoc-mode and batch-mode SNP search, automatic SNP filtering, and retrieval of SNP data (physical position, function class, flanking sequences at user-defined lengths, heterozygosity) from NCBI dbSNP
-
Full-text index only
An integrated database-pipeline system for studying single nucleotide polymorphisms and diseases.
PMID 19091018 · PMC2638159 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Existing SNP/disease databases are fragmented; no combined resource widely supports gene-, SNP-, and disease-related information together
-
Full-text index only
YANA - a software tool for analyzing flux modes, gene-expression and enzyme activities.
PMID 15929789 · PMC1175843 · BMC bioinformatics · 2005 · 7 claims · 3 setups
YANA integrates METATOOL to provide a graphical, platform-independent front-end for editing, calculating, visualizing, and comparing elementary flux modes, with SBML support.
-
Full-text index only
Reconstructing transcriptional regulatory networks through genomics data.
PMID 20048387 · PMC3666560 · Statistical methods in medical research · 2009 · 7 claims · 5 setups
Location data (ChIP-chip/ChIP-seq) alone is insufficient for TRN inference because binding does not imply regulation, TF binding is dynamic across conditions/time, and TRNs involve combinatorial effects of multiple TFs not captured by single-TF ChIP experiments.