Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Has reproduction
Using random walks to identify cancer-associated modules in expression data.
PMID 24128261 · PMC4015830 · BioData mining · 2013 · 8 claims · 8 setups
Walktrap-GM, a random-walk community detection algorithm adapted with stopping criteria (maximum modularity, maximum size, maximum module score), identifies modules significantly enriched with cancer genes in expression-weighted interaction networks.
-
Full-text index only
Sample preparation for serum/plasma profiling and biomarker identification by mass spectrometry.
PMID 17166507 · PMC7094463 · Journal of chromatography. A · 2007 · 8 claims · 8 setups
Standardizing sample preparation procedures for serum/plasma profiling is critical for obtaining reliable biomarkers, since slight procedural changes can produce very different protein profiles.
-
Full-text index only
Pairagon+N-SCAN_EST: a model-based gene annotation pipeline.
PMID 16925839 · PMC1810554 · Genome biology · 2006 · 7 claims · 5 setups
Pairagon+N-SCAN_EST, using only native alignments, was as accurate as ENSEMBL and ExoGean in the EGASP mRNA/EST evidence assessment
-
Full-text index only
Comprehensive genome analysis of 203 genomes provides structural genomics with new insights into protein family space.
PMID 16481312 · PMC1373602 · Nucleic acids research · 2006 · 8 claims · 7 setups
The number of protein families continues to expand steadily as more genomes are sequenced, showing no sign of saturation.
-
Has reproduction · 48
Rbfox2 controls autoregulation in RNA-binding protein networks.
PMID 24637117 · PMC3967051 · Genes & development · 2014 · 8 claims · 8 setups
Rbfox2 cross-regulates AS-NMD events within RNA-binding protein genes to alter their expression, tuning autoregulatory splicing networks and placing Rbfox2 at a critical node of a multilayer regulatory network.
-
Full-text index only
Systems biology approach for mapping the response of human urothelial cells to infection by Enterococcus faecalis.
PMID 18047719 · PMC2099488 · BMC bioinformatics · 2007 · 8 claims · 5 setups
Deconvoluting gene expression variance into technical (Gaussian, ~6.5% relative SD) and biological components identifies hypervariable (HV) genes that reflect true biological response to infection without requiring replicates
-
Full-text index only
Optimality driven nearest centroid classification from genomic data.
PMID 17912341 · PMC1991588 · PloS one · 2007 · 7 claims · 5 setups
A theoretical result determines the subset of features of a given size that minimizes the misclassification rate for a nearest-centroid (LDA) classifier, based on equation (4).
-
Has reproduction · 59
Nucleosome regulatory dynamics in response to TGFβ.
PMID 24771338 · PMC4066760 · Nucleic acids research · 2014 · 8 claims · 7 setups
SuMMIt, a Bayesian strand-based mixture model requiring support from both ends of sequenced fragments, enables precise nucleosome mid-position calling, fuzziness scoring and between-condition change detection.
-
Full-text index only
Genome-wide analysis of KAP1 binding suggests autoregulation of KRAB-ZNFs.
PMID 17542650 · PMC1885280 · PLoS genetics · 2007 · 8 claims · 7 setups
H3me3K9 and H3me3K27 mark largely mutually exclusive, distinct classes of transcription factor genes: H3me3K9 at ZNF genes, H3me3K27 at homeobox genes
-
Full-text index only
Integrated analysis of genetic and proteomic data identifies biomarkers associated with adverse events following smallpox vaccination.
PMID 18923431 · PMC2692715 · Genes and immunity · 2009 · 7 claims · 6 setups
A two-stage strategy (Random Forest filtering followed by decision tree modeling) can integrate categorical genetic and continuous proteomic data to identify biomarkers of AE risk
-
Full-text index only
Exhaustive prediction of disease susceptibility to coding base changes in the human genome.
PMID 18793467 · PMC2537574 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Inter-species conservation is the strongest single predictor of disease-associated coding mutations among the factors tested.
-
Full-text index only
PRESTO: rapid calculation of order statistic distributions and multiple-testing adjusted P-values via permutation for one and two-stage genetic association studies.
PMID 18620604 · PMC2483288 · BMC bioinformatics · 2008 · 8 claims · 4 setups
PRESTO is an order of magnitude faster than other existing permutation testing software for genetic association studies.
-
Full-text index only
Seeded Bayesian Networks: constructing genetic networks from microarray data.
PMID 18601736 · PMC2474592 · BMC systems biology · 2008 · 8 claims · 4 setups
Seeding Bayesian Network analysis with prior networks derived from literature and/or PPI data improves recovery of known gene-gene interactions compared to BN analysis without a seed
-
Full-text index only
Managing incidental findings in human subjects research: analysis and recommendations.
PMID 18547191 · PMC2575242 · The Journal of law, medicine & ethics : a journal of the American Society of Law, Medicine & Ethics · 2008 · 8 claims · 5 setups
Little guidance currently exists on managing research IFs, and no consensus exists on the best approach across genetic/genomic and imaging research domains.
-
Full-text index only
Mapping proteins to disease terminologies: from UniProt to MeSH.
PMID 18460185 · PMC2367626 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Developed a three-step procedure (disease name extraction, exact matching, partial/similarity-based matching) to map UniProtKB/Swiss-Prot disease names to MeSH terms
-
Full-text index only
A hierarchical and modular approach to the discovery of robust associations in genome-wide association studies from pooled DNA samples.
PMID 18194558 · PMC2248205 · BMC genetics · 2008 · 8 claims · 5 setups
A hierarchical/modular approach integrating quality control, LD, physical distance, and gene ontology identifies authentic associations among those found by statistical tests in pooled DNA GWAS
-
Full-text index only
Network of Cancer Genes: a web resource to analyze duplicability, orthology and network properties of cancer genes.
PMID 19906700 · PMC2808873 · Nucleic acids research · 2010 · 7 claims · 4 setups
NCG is a web database integrating duplicability, orthology, evolutionary appearance, and network topology data for 736 human cancer genes
-
Full-text index only
A proteomic approach for plasma biomarker discovery with iTRAQ labelling and OFFGEL fractionation.
PMID 19888438 · PMC2771280 · Journal of biomedicine & biotechnology · 2010 · 6 claims · 6 setups
iTRAQ labelling combined with OFFGEL fractionation improves proteome coverage of human plasma compared to iTRAQ alone or no iTRAQ
-
Full-text index only
Metagenomic analysis of respiratory tract DNA viral communities in cystic fibrosis and non-cystic fibrosis individuals.
PMID 19816605 · PMC2756586 · PloS one · 2009 · 8 claims · 8 setups
CF phage communities are highly similar to each other, whereas Non-CF individuals have more distinct, variable phage communities reflecting transient environmental sampling