Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Prioritization of candidate cancer genes--an aid to oncogenomic studies.
PMID 18710882 · PMC2566894 · Nucleic acids research · 2008 · 8 claims · 8 setups
Computational classifiers using combinations of protein conservation, gene structure, protein domains, protein interactions, and regulatory data can distinguish known cancer genes (CD/CR) from unlabelled human genes
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Full-text index only
An oncogenomics-based in vivo RNAi screen identifies tumor suppressors in liver cancer.
PMID 19012953 · PMC2990916 · Cell · 2008 · 7 claims · 8 setups
shRNA pools targeting genes recurrently deleted in human HCC accelerate hepatocarcinogenesis in vivo, whereas randomly selected shRNA pools do not.
-
Full-text index only
A surrogate-based approach for post-genomic partner identification.
PMID 11602024 · PMC57814 · BMC biotechnology · 2001 · 8 claims · 5 setups
Peptide surrogates derived from random phage display libraries contain amino acid sequence information that identifies the natural biological partner of the panned target via database searching.
-
Full-text index only
Expansion of the Bactericidal/Permeability Increasing-like (BPI-like) protein locus in cattle.
PMID 17362520 · PMC1839098 · BMC genomics · 2007 · 8 claims · 8 setups
The bovine BPI-like locus spans 470 kbp and contains 14 contiguous genes (13 intact + 1 pseudogene); 9 are orthologous to human/mouse BPI-like genes and 4 (named BSP30A, BSP30B, BSP30C, BSP30D) arose through cattle-specific duplication of the PSP gene
-
Full-text index only
CanPredict: a computational tool for predicting cancer-associated missense mutations.
PMID 17537827 · PMC1933186 · Nucleic acids research · 2007 · 8 claims · 7 setups
CanPredict is a web application providing public access to a random forest classifier that combines SIFT, LogR.E-value, and GOSS scores to predict whether a missense mutation is cancer-associated
-
Full-text index only
"Reverse ecology" and the power of population genomics.
PMID 18752601 · PMC2626434 · Evolution; international journal of organic evolution · 2008 · 8 claims · 7 setups
Population genomic data can be used to rapidly identify genes targeted by adaptive natural selection, an approach termed 'reverse ecology'.
-
Full-text index only
Metagenomic analysis of human diarrhea: viral detection and discovery.
PMID 18398449 · PMC2290972 · PLoS pathogens · 2008 · 8 claims · 7 setups
Micro-mass sequencing (minimal stool input, minimal purification, ~384 reads/sample) can detect known enteric viruses in diarrhea specimens
-
Full-text index only
Ab initio identification of human microRNAs based on structure motifs.
PMID 18088431 · PMC2238772 · BMC bioinformatics · 2007 · 8 claims · 7 setups
MiRPred predicts miRNA precursors ab initio using only predicted secondary structure motifs, ignoring nucleotide sequence
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Full-text index only
Computational verification of protein-protein interactions by orthologous co-expression.
PMID 15740634 · PMC555590 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Co-expression of orthologous protein pairs across multiple species can verify/predict S. cerevisiae PPIs with better performance than S. cerevisiae co-expression alone.
-
Full-text index only
The use of edge-betweenness clustering to investigate biological function in protein interaction networks.
PMID 15740614 · PMC555937 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Edge-Betweenness clustering separates protein interaction graphs into subgraphs whose GO term distributions show significant correlations, revealing biologically meaningful functional modules.
-
Has reproduction
Unlocking the microbial studies through computational approaches: how far have we reached?
PMID 36920617 · PMC10016191 · Environmental science and pollution research international · 2023 · 8 claims · 8 setups
Metagenomics enables culture-independent study of microbial communities directly from their natural environments, bypassing the need for clonal isolation.
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
A clustering property of highly-degenerate transcription factor binding sites in the mammalian genome.
PMID 16670430 · PMC1456330 · Nucleic acids research · 2006 · 8 claims · 7 setups
Highly-degenerate RE1 sites are significantly enriched in promoters of validated and putative REST target genes compared to control promoters
-
Full-text index only
Pathogenic mitochondrial DNA mutations are common in the general population.
PMID 18674747 · PMC2495064 · American journal of human genetics · 2008 · 7 claims · 6 setups
At least 1 in 200 healthy humans harbors a pathogenic mtDNA mutation with potential to cause disease in offspring of female carriers
-
Full-text index only
Quadratic regression analysis for gene discovery and pattern recognition for non-cyclic short time-course microarray experiments.
PMID 15850479 · PMC1127068 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A step-down quadratic regression method (fitting quadratic, then linear, then null models per gene) identifies differentially expressed genes and classifies them into 9 temporal expression patterns using continuous time information.
-
Has reproduction · 30
Minimal metabolic pathway structure is consistent with associated biomolecular interactions.
PMID 24987116 · PMC4299494 · Molecular systems biology · 2014 · 8 claims · 8 setups
MinSpan, a mixed-integer linear optimization algorithm, computes the shortest, linearly independent pathways (sparsest basis of the null space of the stoichiometric matrix S) for genome-scale metabolic networks, which convex approaches (extreme pathways, elementary flux modes) cannot do at genome scale.
-
Has reproduction · 58
A comparative study of techniques for differential expression analysis on RNA-Seq data.
PMID 25119138 · PMC4132098 · PloS one · 2014 · 8 claims · 8 setups
edgeR performs slightly better than DESeq and Cuffdiff2 in terms of the ability to uncover true positives.