Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Prediction of candidate primary immunodeficiency disease genes using a support vector machine learning approach.
PMID 19801557 · PMC2780952 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2009 · 6 claims · 3 setups
An SVM trained on 69 binary features of known PID genes can accurately classify PID vs non-PID genes and predict novel candidate PID genes
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
Human PAML browser: a database of positive selection on human genes using phylogenetic methods.
PMID 17962310 · PMC2238824 · Nucleic acids research · 2008 · 8 claims · 5 setups
The Human PAML Browser is a web-accessible database of codeml-based positive selection test results for 13,721 human genes with orthologs in UCSC multispecies alignments.
-
Full-text index only
"Reverse ecology" and the power of population genomics.
PMID 18752601 · PMC2626434 · Evolution; international journal of organic evolution · 2008 · 8 claims · 7 setups
Population genomic data can be used to rapidly identify genes targeted by adaptive natural selection, an approach termed 'reverse ecology'.
-
Has reproduction · 92
A network-guided protocol to discover susceptibility genes in genome-wide association studies using stability selection.
PMID 36609152 · PMC9850185 · STAR protocols · 2023 · 5 claims · 5 setups
The protocol identifies genes that are both statistically associated with a phenotype and functionally interconnected in a biological network
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
SPRINT: a new parallel framework for R.
PMID 19114001 · PMC2628907 · BMC bioinformatics · 2008 · 8 claims · 1 setups
SPRINT is a prototype R framework that wraps parallelised functions, requiring minimal modification to existing sequential R scripts and no parallel programming expertise from the user
-
Full-text index only
WebGestalt: an integrated system for exploring gene sets in various biological contexts.
PMID 15980575 · PMC1160236 · Nucleic acids research · 2005 · 8 claims · 6 setups
WebGestalt is an integrated web-based system composed of four modules: gene set management, information retrieval, organization/visualization, and statistics.
-
Full-text index only
Evidence for positive selection in putative virulence factors within the Paracoccidioides brasiliensis species complex.
PMID 18820744 · PMC2553485 · PLoS neglected tropical diseases · 2008 · 8 claims · 8 setups
Positive selection has played an important role in the molecular evolution of putative virulence factors of P. brasiliensis
-
Full-text index only
Next-generation high-density self-assembling functional protein arrays.
PMID 18469824 · PMC3070491 · Nature methods · 2008 · 8 claims · 7 setups
A next-generation NAPPA method produces high-density protein microarrays displaying over 1500 unique proteins with >90% expression success
-
Has reproduction · 80
Colorectal Cancer Prediction Based on Weighted Gene Co-Expression Network Analysis and Variational Auto-Encoder.
PMID 32825264 · PMC7563725 · Biomolecules · 2020 · 6 claims · 7 setups
Combining WGCNA-derived hub genes with a VAE-derived 10-dimensional representation as features for an SVM classifier achieves high accuracy (0.9692) and AUC (0.9981) for colorectal cancer prediction.
-
Full-text index only
GoMiner: a resource for biological interpretation of genomic and proteomic data.
PMID 12702209 · PMC154579 · Genome biology · 2003 · 8 claims · 4 setups
GoMiner organizes 'interesting' gene lists (e.g., differentially expressed genes) into the Gene Ontology hierarchy for biological interpretation, displaying results as both a tree and a directed acyclic graph (DAG).
-
Full-text index only
AUGUSTUS at EGASP: using EST, protein and genomic alignments for improved gene prediction in the human genome.
PMID 16925833 · PMC1810548 · Genome biology · 2006 · 8 claims · 5 setups
AUGUSTUS predicted significantly more genes correctly than any other ab initio program in EGASP
-
Full-text index only
On the association between chromosomal rearrangements and genic evolution in humans and chimpanzees.
PMID 17971225 · PMC2246304 · Genome biology · 2007 · 8 claims · 4 setups
Genes located in rearranged chromosomes show lower non-coding (KI), synonymous (KS), and non-synonymous (KA) divergence than genes in colinear chromosomes.
-
Full-text index only
Genes implicated in multiple sclerosis pathogenesis from consilience of genotyping and expression profiles in relapse and remission.
PMID 18366677 · PMC2324081 · BMC medical genetics · 2008 · 8 claims · 7 setups
Distinct sets of dysregulated genes are found in peripheral blood during the relapse phase versus the remission phase of RRMS
-
Full-text index only
Indirect genomic effects on survival from gene expression data.
PMID 18358079 · PMC2397510 · Genome biology · 2008 · 7 claims · 6 setups
A novel methodology (dynamic path analysis combined with additive hazard survival regression) can detect and quantify indirect effects of gene expression on survival mediated through transcription factor target genes.
-
Full-text index only
Commonality of functional annotation: a method for prioritization of candidate genes from genome-wide linkage studies.
PMID 18263617 · PMC2275105 · Nucleic acids research · 2008 · 8 claims · 7 setups
Genes correlated with a common complex trait are more likely to share GO functional annotations than genes not correlated with that trait
-
Full-text index only
Gene- and evidence-based candidate gene selection for schizophrenia and gene feature analysis.
PMID 19944577 · PMC2826526 · Artificial intelligence in medicine · 2010 · 8 claims · 5 setups
The SCOR method outperforms the CCOR method for prioritizing schizophrenia candidate genes
-
Full-text index only
Examination of tetrahydrobiopterin pathway genes in autism.
PMID 19674121 · PMC2784255 · Genes, brain, and behavior · 2009 · 8 claims · 6 setups
PTS (6-pyruvoyl-tetrahydropterin synthase) shows significant nominal association with autism (p=0.009), not restricted to affected-male-only subset
-
Full-text index only
GeneTrail--advanced gene set enrichment analysis.
PMID 17526521 · PMC1933132 · Nucleic acids research · 2007 · 8 claims · 2 setups
GeneTrail is a comprehensive, easy-to-use web-based tool for gene set enrichment analysis supporting both Over-Representation Analysis (ORA) and Gene Set Enrichment Analysis (GSEA)