Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Proteomics as a method for early detection of cancer: a review of proteomics, exhaled breath condensate, and lung cancer screening.
PMID 18095050 · PMC2150625 · Journal of general internal medicine · 2008 · 8 claims · 7 setups
Protein expression is closely aligned with cellular activity, unlike genomic changes which may have no functional significance
-
Full-text index only
Comprehensive splice-site analysis using comparative genomics.
PMID 16914448 · PMC1557818 · Nucleic acids research · 2006 · 8 claims · 6 setups
Over half a million splice sites were collected from five species (H. sapiens, M. musculus, D. melanogaster, C. elegans, A. thaliana) and classified into four main subtypes: U2-type GT-AG and GC-AG, and U12-type GT-AG and AT-AC.
-
Full-text index only
Blood Pressure Sunday: introducing genomics to the community through family history.
PMID 15888234 · PMC1327717 · Preventing chronic disease · 2005 · 7 claims · 5 setups
Family history is a significant, longstanding risk factor for high blood pressure and can serve as a practical 'genomic tool' for public health given that routine DNA-based risk testing is not yet available.
-
Full-text index only
Genome Network and FANTOM3: assessing the complexity of the transcriptome.
PMID 16683037 · PMC1449904 · PLoS genetics · 2006 · 8 claims · 7 setups
63% of the genome is transcribed from at least one strand, versus the earlier belief that only 2% is transcribed into protein-coding mRNA
-
Full-text index only
Utah's Family High Risk Program: bridging the gap between genomics and public health.
PMID 15888235 · PMC1327718 · Preventing chronic disease · 2005 · 8 claims · 6 setups
Collection of family history through the Family High Risk Program (FHRP) is a cost-effective method for identifying and intervening with high-risk populations for chronic disease
-
Full-text index only
Sample preparation for serum/plasma profiling and biomarker identification by mass spectrometry.
PMID 17166507 · PMC7094463 · Journal of chromatography. A · 2007 · 8 claims · 8 setups
Standardizing sample preparation procedures for serum/plasma profiling is critical for obtaining reliable biomarkers, since slight procedural changes can produce very different protein profiles.
-
Full-text index only
MEROPS: the peptidase database.
PMID 19892822 · PMC2808883 · Nucleic acids research · 2010 · 8 claims · 5 setups
MEROPS is a manually curated hierarchical classification of peptidases and protein inhibitors organized into protein species, families, and clans based on sequence and structural homology.
-
Has reproduction · 68
Rfam 15: RNA families database in 2025.
PMID 39526405 · PMC11701678 · Nucleic acids research · 2025 · 8 claims · 6 setups
Rfamseq was expanded to 26 106 genomes, a 76% increase, by incorporating the latest UniProt reference proteomes and additional viral genomes
-
Has reproduction · 64
VIGET: A web portal for study of vaccine-induced host responses based on Reactome pathways and ImmPort data.
PMID 37180100 · PMC10172660 · Frontiers in immunology · 2023 · 7 claims · 7 setups
VIGET is a web portal that lets users select vaccines/ImmPort studies, run differential gene expression analysis, and perform Reactome-based pathway enrichment and functional interaction network construction
-
Has reproduction · 90
Sediment Resuspension as a System-Wide Driver of Legacy and Bioavailable Phosphorus Release in Lake Erie.
PMID 41960750 · PMC13130958 · Environmental science & technology · 2026 · 7 claims · 8 setups
Sediment resuspension is a major episodic internal phosphorus source releasing bioavailable P far exceeding previously reported aerobic diffusive fluxes.
-
Full-text index only
MSH6 missense mutations are often associated with no or low cancer susceptibility.
PMID 15354210 · PMC2409912 · British journal of cancer · 2004 · 7 claims · 8 setups
Most MSH6 missense changes found in MSI-positive tumours are likely clinically innocent or of low cancer-susceptibility significance
-
Full-text index only
Integrative annotation of 21,037 human genes validated by full-length cDNA clones.
PMID 15103394 · PMC393292 · PLoS biology · 2004 · 8 claims · 5 setups
41,118 full-length human cDNAs from six high-throughput sequencing projects were exhaustively integratively characterized
-
Full-text index only
Four genomic islands that mark post-1995 pandemic Vibrio parahaemolyticus isolates.
PMID 16672049 · PMC1464126 · BMC genomics · 2006 · 8 claims · 7 setups
Seven genomic islands (VPaI-1 to VPaI-7, 10-81 kb) were identified in V. parahaemolyticus RIMD2210633 by aberrant GC content, presence of integrases/transposases, flanking direct repeats, and absence from related Vibrionaceae genomes.
-
Full-text index only
Identification and characterization of insect-specific proteins by genome data analysis.
PMID 17407609 · PMC1852559 · BMC genomics · 2007 · 8 claims · 7 setups
Comparative genome analysis across five holometabolous insects and three non-insect eukaryotes (opisthokonts) identifies 154 insect-specific orthologous groups (refined to 51 proteins) and 466 eukaryote/opisthokont-core orthologous groups
-
Full-text index only
Antibody protein array analysis of the tear film cytokines.
PMID 18677223 · PMC3786218 · Optometry and vision science : official publication of the American Academy of Optometry · 2008 · 7 claims · 4 setups
Tear fluid contains factors with affinity for plastic, capture antibodies, and IgG that create matrix effects profoundly impacting dot ELISA/array reliability
-
Full-text index only
Complete genome sequence and comparative analysis of the wild-type commensal Escherichia coli strain SE11 isolated from a healthy adult.
PMID 18931093 · PMC2608844 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2008 · 8 claims · 6 setups
The SE11 genome comprises a 4.8 Mb chromosome encoding 4679 protein-coding genes and six plasmids encoding 323 protein-coding genes
-
Full-text index only
Adjustment of genomic waves in signal intensities from whole-genome SNP genotyping platforms.
PMID 18784189 · PMC2577347 · Nucleic acids research · 2008 · 8 claims · 6 setups
Genomic waves are present in both Illumina and Affymetrix SNP genotyping arrays, confirming they are not platform-specific
-
Full-text index only
Integrated proteomic analysis of human cancer cells and plasma from tumor bearing mice for ovarian cancer biomarker discovery.
PMID 19936259 · PMC2775948 · PloS one · 2009 · 8 claims · 8 setups
Integrated proteomic analysis of a cancer mouse model and human cancer cell populations provides an effective approach to identify potential circulating protein biomarkers.
-
Full-text index only
Bioinformatics methods for learning radiation-induced lung inflammation from heterogeneous retrospective and prospective data.
PMID 19704920 · PMC2688763 · Journal of biomedicine & biotechnology · 2009 · 8 claims · 3 setups
Kernel-based methods (e.g., SVM) can capture nonlinear dose-volume interactions relevant to predicting radiation pneumonitis
-
Full-text index only
Testing groups of genomic locations for enrichment in disease loci using linkage scan data: a method for hypothesis testing.
PMID 16848972 · PMC3525155 · Human genomics · 2006 · 8 claims · 2 setups
A method testing enrichment of a group of genomic locations for disease loci by comparing the average NPL score of the group to a null distribution from randomly drawn groups of equal size