Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 86
Molecular Classification Models for Triple Negative Breast Cancer Subtype Using Machine Learning.
PMID 34575658 · PMC8472680 · Journal of personalized medicine · 2021 · 6 claims · 4 setups
A training gene set of 719 unique upregulated DEGs (subtype-specific) can be used to build ML models that classify TNBC into BLIA, BLIS, MES, and LAR subtypes.
-
Full-text index only
Two committees tackle toxicogenomics.
PMID 12501852 · PMC1241123 · Environmental health perspectives · 2002 · 8 claims · 8 setups
NIEHS funded a $37 million, five-year Toxicogenomics Research Consortium (TRC) linking the NIEHS Microarray Center with five academic institutions (UNC, Duke, Fred Hutchinson/UW, MIT, OHSU) to coordinate gene-expression research on environmental health effects.
-
Has reproduction · 90
Comparative Genomics Provides Insight into the Function of Broad-Host Range Sponge Symbionts.
PMID 34519538 · PMC8546597 · mBio · 2021 · 8 claims · 8 setups
Eleven new genomes were added to the Tethybacterales order and a novel family (Polydorabacteraceae) was identified
-
Has reproduction · 74
Discovery of a novel filamentous prophage in the genome of the Mimosa pudica microsymbiont Cupriavidus taiwanensis STM 6018.
PMID 36925474 · PMC10011098 · Frontiers in microbiology · 2023 · 8 claims · 8 setups
The STM 6018 genome contains two prophages: a complete Mu-like capsular phage and a filamentous phage that integrates into a putative dif site.
-
Full-text index only
G2Cdb: the Genes to Cognition database.
PMID 18984621 · PMC2686544 · Nucleic acids research · 2009 · 7 claims · 7 setups
G2Cdb integrates experimentally validated synapse proteome datasets with mouse/human genomic annotation, phenotype, and human disease data in a gene-centric database.
-
Full-text index only
Proteomic-based identification of maternal proteins in mature mouse oocytes.
PMID 19646285 · PMC2730056 · BMC genomics · 2009 · 8 claims · 6 setups
625 different proteins were identified from 2700 zona pellucida-free mature mouse MII oocytes, the largest oocyte proteome catalog to date
-
Has reproduction · 59
Metapangenomics of wild and cultivated banana microbiome reveals a plethora of host-associated protective functions.
PMID 37085932 · PMC10120106 · Environmental microbiome · 2023 · 8 claims · 8 setups
Root and corm endosphere communities are significantly richer and compositionally distinct from leaf endosphere communities across Musa genotypes
-
Full-text index only
Information extraction from full text scientific articles: where are the keywords?
PMID 12775220 · PMC166134 · BMC bioinformatics · 2003 · 8 claims · 5 setups
The keyword content of the five article sections (A, I, M, R, D) is heterogeneous, i.e., different sections carry different kinds of information.
-
Full-text index only
Outlook on Thailand's genomics and computational biology research and development.
PMID 18654621 · PMC2446437 · PLoS computational biology · 2008 · 8 claims · 8 setups
Thai government policy support, infrastructure investment, education programs, and human resource development have substantially advanced genomics and bioinformatics research capacity in Thailand
-
Full-text index only
DAVID Bioinformatics Resources: expanded annotation database and novel algorithms to better extract biology from large gene lists.
PMID 17576678 · PMC1933169 · Nucleic acids research · 2007 · 8 claims · 4 setups
The DAVID Gene Concept uses a single-linkage method to agglomerate tens of millions of gene/protein identifiers from NCBI, PIR, UniProt and other resources into unified DAVID genes.
-
Has reproduction · 75
Graph-Based Approaches Significantly Improve the Recovery of Antibiotic Resistance Genes From Complex Metagenomic Datasets.
PMID 34690959 · PMC8528159 · Frontiers in microbiology · 2021 · 8 claims · 6 setups
GraphAMR, a Nextflow pipeline that aligns AMR profile HMMs (or AA sequences) to metagenomic assembly graphs via PathRacer, then dereplicates and annotates hits, recovers more and more complete AMR genes than contig-based or read-based methods.
-
Full-text index only
ARED 3.0: the large and diverse AU-rich transcriptome.
PMID 16381826 · PMC1347415 · Nucleic acids research · 2006 · 7 claims · 6 setups
ARED 3.0 computationally mapped more than 4000 ARE-mRNAs to the human genome, representing 5-8% of human genes.
-
Full-text index only
Comparative genome and phenotypic analysis of Clostridium difficile 027 strains provides insight into the evolution of a hypervirulent bacterium.
PMID 19781061 · PMC2768977 · Genome biology · 2009 · 8 claims · 7 setups
The 027 genomes (CD196 and R20291) contain 234 additional genes compared to strain 630, spread across at least 50 regions of genetic difference, including a phage island, transposon genes, two-component response regulators, drug resistance genes and transporters
-
Full-text index only
3' tag digital gene expression profiling of human brain and universal reference RNA using Illumina Genome Analyzer.
PMID 19917133 · PMC2781828 · BMC genomics · 2009 · 8 claims · 4 setups
3' tag DGE transcript profiles are highly reproducible between technical and biological replicates, across libraries made at different labs, and across two generations of Illumina Genome Analyzers (GA I and GA II).
-
Full-text index only
Anopheles gambiae genome reannotation through synthesis of ab initio and comparative gene prediction algorithms.
PMID 16569258 · PMC1557760 · Genome biology · 2006 · 8 claims · 7 setups
An exon-gene-union (EGU) algorithm followed by an open-reading-frame-selection algorithm can synthesize ab initio (GENSCAN, GeneMark, SNAP) and comparative (Ensembl/Genewise) predictions into a single, more complete CDS set
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.
-
Has reproduction · 49
oPOSSUM-3: advanced analysis of regulatory motif over-representation across genes or ChIP-Seq datasets.
PMID 22973536 · PMC3429929 · G3 (Bethesda, Md.) · 2012 · 8 claims · 6 setups
oPOSSUM-3 is a web-accessible system that identifies over-represented TFBS and TFBS families in DNA sequences of co-expressed genes or in sequences from high-throughput methods such as ChIP-Seq.
-
Full-text index only
BioGPS: an extensible and customizable portal for querying and organizing gene annotation resources.
PMID 19919682 · PMC3091323 · Genome biology · 2009 · 8 claims · 4 setups
BioGPS aggregates distributed, third-party gene annotation resources into a single customizable portal for human, mouse, and rat genes.
-
Full-text index only
Recurrent and multiple bladder tumors show conserved expression profiles.
PMID 18590527 · PMC2483988 · BMC cancer · 2008 · 8 claims · 7 setups
Recurrent and multiple bladder tumors from the same patient display remarkably similar gene expression profiles despite genomic differences.
-
Full-text index only
Functional annotation and identification of candidate disease genes by computational analysis of normal tissue gene expression data.
PMID 18560577 · PMC2409962 · PloS one · 2008 · 7 claims · 5 setups
Ranked Coexpression Groups (RCG) built from k=6 nearest coexpressed genes, combined with a majority-rule functional characterization, integrate multiple datasets/coexpression measures to generate high-confidence functional annotation predictions