Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Full-text index only
Babelomics: advanced functional profiling of transcriptomics, proteomics and genomics experiments.
PMID 18515841 · PMC2447758 · Nucleic acids research · 2008 · 8 claims · 5 setups
Babelomics is a web suite offering both conventional functional enrichment methods and more advanced gene set analysis (GSA) methods, a combination offered by only one other tool (FuncAssociate) among competitors.
-
Has reproduction · 75
FEM: mining biological meaning from cell level in single-cell RNA sequencing data.
PMID 34909283 · PMC8641482 · PeerJ · 2021 · 7 claims · 5 setups
The FEM algorithm converts each cell's gene expression matrix (GEM) into a functional expression matrix by applying Fisher's exact test enrichment per cell and per gene set, then encoding adjusted p-values as information content.
-
Full-text index only
Motif discovery in promoters of genes co-localized and co-expressed during myeloid cells differentiation.
PMID 19059999 · PMC2632922 · Nucleic acids research · 2009 · 6 claims · 8 setups
A novel multi-step computational method (built on approximate pattern enumeration, binomial over-representation scoring with FDR correction, and k-medoids clustering) can identify over-represented motifs in a selected set of promoters relative to a background promoter set.
-
Full-text index only
Gene losses during human origins.
PMID 16464126 · PMC1361800 · PLoS biology · 2006 · 7 claims · 7 setups
A comparative genomic screen identified 67 new human-specific nonprocessed pseudogenes, bringing the total (with 13 from prior literature) to 80 human-specific pseudogenes.
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Full-text index only
Mining expressed sequence tags identifies cancer markers of clinical interest.
PMID 17078886 · PMC1635568 · BMC bioinformatics · 2006 · 8 claims · 6 setups
An EST-mining approach (Fisher Exact Test on tumor vs. non-tumor library hit counts) identifies differentially expressed transcripts with an estimated false discovery rate below 22% when human and mouse screens are combined.
-
Full-text index only
Exogean: a framework for annotating protein-coding genes in eukaryotic genomic DNA.
PMID 16925841 · PMC1810556 · Genome biology · 2006 · 8 claims · 5 setups
Exogean is a framework using directed acyclic coloured multigraphs (DACMs) to represent biological objects (mRNA, ESTs, protein alignments, exons) and iteratively combine them into complex protein-coding transcript models.
-
Full-text index only
Resequencing PNMT in European hypertensive and normotensive individuals: no common susceptibilily variants for hypertension and purifying selection on intron 1.
PMID 17645789 · PMC1947951 · BMC medical genetics · 2007 · 7 claims · 7 setups
Resequencing of PNMT found no common susceptibility variants that distinguish hypertensive from normotensive individuals
-
Full-text index only
FatiGO +: a functional profiling tool for genomic data. Integration of functional annotation, regulatory motifs and interaction data with microarray experiments.
PMID 17478504 · PMC1933151 · Nucleic acids research · 2007 · 8 claims · 8 setups
FatiGO+ is a web-based tool for functional profiling of genome-scale experiments that integrates functional annotation, regulatory motifs and interaction data
-
Full-text index only
Epidemiology of doublet/multiplet mutations in lung cancers: evidence that a subset arises by chronocoordinate events.
PMID 19005564 · PMC2579325 · PloS one · 2008 · 8 claims · 7 setups
Doublet mutations are significantly more frequent in EGFR (6.0%) and TP53 (2.3%) in human lung cancer than spontaneous doublets in mouse lacI (0.7%), about 8-fold and 3-fold higher respectively.
-
Full-text index only
Genetic diversity of vaccine candidate antigens in Plasmodium falciparum isolates from the Amazon basin of Peru.
PMID 18505558 · PMC2432069 · Malaria journal · 2008 · 8 claims · 7 setups
CSP (Th2R/Th3R region) is polymorphic in Peruvian isolates, with four distinct alleles identified, none identical to the 3D7 vaccine strain
-
Full-text index only
A non-parametric meta-analysis approach for combining independent microarray datasets: application using two microarray datasets pertaining to chronic allograft nephropathy.
PMID 18302764 · PMC2276496 · BMC genomics · 2008 · 8 claims · 6 setups
A novel non-parametric meta-analysis approach for combining independent microarray datasets is presented, requiring no distributional assumptions and being logically intuitive.
-
Full-text index only
A taxonomy of epithelial human cancer and their metastases.
PMID 20017941 · PMC2806369 · BMC medical genomics · 2009 · 8 claims · 6 setups
Unsupervised hierarchical clustering of 1566 primary epithelial tumors yields large tissue-enriched clusters (breast, colon/GI, lung, ovary, kidney) plus smaller prostate, thyroid-kidney, and mixed clusters
-
Full-text index only
Large-scale discovery of insertion hotspots and preferential integration sites of human transposed elements.
PMID 20008508 · PMC2836564 · Nucleic acids research · 2010 · 8 claims · 6 setups
Most TEs insert within specific 'hotspots' along the targeted TE rather than uniformly.
-
Full-text index only
Gene- and evidence-based candidate gene selection for schizophrenia and gene feature analysis.
PMID 19944577 · PMC2826526 · Artificial intelligence in medicine · 2010 · 8 claims · 5 setups
The SCOR method outperforms the CCOR method for prioritizing schizophrenia candidate genes
-
Has reproduction · 78
Requirements for Pseudomonas aeruginosa acute burn and chronic surgical wound infection.
PMID 25057820 · PMC4109851 · PLoS genetics · 2014 · 8 claims · 8 setups
In vivo gene expression is generally not correlated with a gene's importance for fitness, with the exception of metabolic genes, for which differential expression is more predictive of fitness.
-
Full-text index only
CLEAN: CLustering Enrichment ANalysis.
PMID 19640299 · PMC2734555 · BMC bioinformatics · 2009 · 8 claims · 4 setups
The gene-specific CLEAN score improves reproducibility of cluster analysis conclusions across independent datasets compared to the traditional cluster-wide score (cwCLEAN).
-
Has reproduction · 83
Gene Expression Atlas update--a value-added database of microarray and sequencing-based functional genomics experiments.
PMID 22064864 · PMC3245177 · Nucleic acids research · 2012 · 8 claims · 5 setups
Gene Expression Atlas is an added-value database providing curated, re-annotated and statistically analysed gene expression data across cell types, organism parts, developmental stages, disease states and other biological/experimental conditions, derived from ArrayExpress Archive and the European Nucleotide Archive.
-
Full-text index only
An online database for brain disease research.
PMID 16594998 · PMC1489945 · BMC genomics · 2006 · 7 claims · 5 setups
SMRIDB is a comprehensive web-based database integrating gene expression data and clinical metadata to aid understanding of the genetic effects of brain disease (bipolar disorder, schizophrenia, depression)