Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Of mice and men: comparative proteomics of bronchoalveolar fluid.
PMID 20032019 · PMC3049194 · The European respiratory journal · 2010 · 8 claims · 8 setups
Comparative shotgun proteomics of human and mouse BALF identifies conserved pathways (immunity, defence response, protease activity) alongside species-specific divergent pathways.
-
Full-text index only
Functional annotation and identification of candidate disease genes by computational analysis of normal tissue gene expression data.
PMID 18560577 · PMC2409962 · PloS one · 2008 · 7 claims · 5 setups
Ranked Coexpression Groups (RCG) built from k=6 nearest coexpressed genes, combined with a majority-rule functional characterization, integrate multiple datasets/coexpression measures to generate high-confidence functional annotation predictions
-
Full-text index only
DAVID Bioinformatics Resources: expanded annotation database and novel algorithms to better extract biology from large gene lists.
PMID 17576678 · PMC1933169 · Nucleic acids research · 2007 · 8 claims · 4 setups
The DAVID Gene Concept uses a single-linkage method to agglomerate tens of millions of gene/protein identifiers from NCBI, PIR, UniProt and other resources into unified DAVID genes.
-
Has reproduction · 100
Shared and unique phosphoproteomics responses in skeletal muscle from exercise models and in hyperammonemic myotubes.
PMID 36345342 · PMC9636548 · iScience · 2022 · 8 claims · 7 setups
Comparative phosphoproteomics of hyperammonemic myotubes and exercise-model muscle identifies shared enriched pathways: PKA, calcium signaling, MAPK signaling, and protein homeostasis.
-
Full-text index only
Comparative genomics and understanding of microbial biology.
PMID 10998382 · PMC2627966 · Emerging infectious diseases · 2000 · 8 claims · 7 setups
GC content varies widely among prokaryotic genomes (29% in B. burgdorferi to 68% in M. tuberculosis) and shapes codon usage and amino acid composition.
-
Full-text index only
Proteomics identifies multipotent and low oncogenic risk stem cells of the spleen.
PMID 20005973 · PMC2891339 · The international journal of biochemistry & cell biology · 2010 · 8 claims · 5 setups
CD45- splenic stem cell-specific proteins are identical to core iPS/ES markers OCT3/4, SOX2, KLF4, c-MYC and NANOG.
-
Has reproduction · 75
Identification of Key Differentially Expressed Genes in Arabidopsis thaliana Under Short- and Long-Term High Light Stress.
PMID 40869111 · PMC12386182 · International journal of molecular sciences · 2025 · 7 claims · 5 setups
Short- and long-term HL responses in Arabidopsis leaves are driven by distinct transcriptional programs, with duration of HL treatment as the primary factor separating transcriptomic clusters.
-
Full-text index only
Inferring combinatorial regulation of transcription in silico.
PMID 15647509 · PMC546154 · Nucleic acids research · 2005 · 8 claims · 5 setups
Combining Cluster-Buster (TFBS cluster prediction) with GOSSIP (rigorous GO enrichment statistics with multiple-testing/FDR correction) predicts biological functions controlled by combinatorial transcription factor action, without prior knowledge of factor targets
-
Has reproduction · 67
Essential Genes of Vibrio anguillarum and Other Vibrio spp. Guide the Development of New Drugs and Vaccines.
PMID 34745063 · PMC8564382 · Frontiers in microbiology · 2021 · 7 claims · 7 setups
Tn-seq using the TnSC189 mariner transposon identified 329 essential genes in V. anguillarum NB10Sm from a library of 52,662 insertion mutants.
-
Has reproduction · 61
Comprehensive transcriptome study to develop molecular resources of the copepod Calanus sinicus for their potential ecological applications.
PMID 24982883 · PMC4055022 · BioMed research international · 2014 · 8 claims · 8 setups
Illumina RNA-Seq with Trinity de novo assembly produced a C. sinicus transcriptome of 69,751 contigs (average 928.8 bp, N50 1,127 bp) from 58.9 million reads.
-
Full-text index only
Prediction-based approaches to characterize bidirectional promoters in the mammalian genome.
PMID 18366609 · PMC2386062 · BMC genomics · 2008 · 8 claims · 7 setups
The mapping algorithm identified 5,647 candidate bidirectional promoter regions in the mouse genome, similar in number to those previously found in human.
-
Has reproduction · 59
Metapangenomics of wild and cultivated banana microbiome reveals a plethora of host-associated protective functions.
PMID 37085932 · PMC10120106 · Environmental microbiome · 2023 · 8 claims · 8 setups
Root and corm endosphere communities are significantly richer and compositionally distinct from leaf endosphere communities across Musa genotypes
-
Full-text index only
Annotation and analysis of 10,000 expressed sequence tags from developing mouse eye and adult retina.
PMID 14519200 · PMC328454 · Genome biology · 2003 · 8 claims · 5 setups
Annotation of 8,633 high-quality non-mitochondrial/non-ribosomal ESTs shows 57% represent known genes and 43% are unknown or novel, with M15E having the highest proportion of novel ESTs
-
Has reproduction
Using random walks to identify cancer-associated modules in expression data.
PMID 24128261 · PMC4015830 · BioData mining · 2013 · 8 claims · 8 setups
Walktrap-GM, a random-walk community detection algorithm adapted with stopping criteria (maximum modularity, maximum size, maximum module score), identifies modules significantly enriched with cancer genes in expression-weighted interaction networks.
-
Full-text index only
ARED 3.0: the large and diverse AU-rich transcriptome.
PMID 16381826 · PMC1347415 · Nucleic acids research · 2006 · 7 claims · 6 setups
ARED 3.0 computationally mapped more than 4000 ARE-mRNAs to the human genome, representing 5-8% of human genes.
-
Full-text index only
An integrated database-pipeline system for studying single nucleotide polymorphisms and diseases.
PMID 19091018 · PMC2638159 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Existing SNP/disease databases are fragmented; no combined resource widely supports gene-, SNP-, and disease-related information together
-
Full-text index only
Metagenomic analysis of respiratory tract DNA viral communities in cystic fibrosis and non-cystic fibrosis individuals.
PMID 19816605 · PMC2756586 · PloS one · 2009 · 8 claims · 8 setups
CF phage communities are highly similar to each other, whereas Non-CF individuals have more distinct, variable phage communities reflecting transient environmental sampling
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
Two committees tackle toxicogenomics.
PMID 12501852 · PMC1241123 · Environmental health perspectives · 2002 · 8 claims · 8 setups
NIEHS funded a $37 million, five-year Toxicogenomics Research Consortium (TRC) linking the NIEHS Microarray Center with five academic institutions (UNC, Duke, Fred Hutchinson/UW, MIT, OHSU) to coordinate gene-expression research on environmental health effects.
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.