Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Advancement of biomarker discovery and validation through the HUPO plasma proteome project.
PMID 15502245 · PMC3839274 · Disease markers · 2004 · 7 claims · 3 setups
Standardization of specimen collection, handling, storage, and choice of serum vs. plasma/anticoagulant is essential for comparable proteomic biomarker discovery.
-
Full-text index only
Antibody protein array analysis of the tear film cytokines.
PMID 18677223 · PMC3786218 · Optometry and vision science : official publication of the American Academy of Optometry · 2008 · 7 claims · 4 setups
Tear fluid contains factors with affinity for plastic, capture antibodies, and IgG that create matrix effects profoundly impacting dot ELISA/array reliability
-
Full-text index only
Utah's Family High Risk Program: bridging the gap between genomics and public health.
PMID 15888235 · PMC1327718 · Preventing chronic disease · 2005 · 8 claims · 6 setups
Collection of family history through the Family High Risk Program (FHRP) is a cost-effective method for identifying and intervening with high-risk populations for chronic disease
-
Full-text index only
Investigation of the human tear film proteome using multiple proteomic approaches.
PMID 18334958 · PMC2268847 · Molecular vision · 2008 · 8 claims · 6 setups
Tear collection method (capillary vs. Schirmer strip) significantly impacts which proteins are detected in the tear film proteome.
-
Full-text index only
Challenges and standards in integrating surveys of structural variation.
PMID 17597783 · PMC2698291 · Nature genetics · 2007 · 7 claims · 5 setups
There is no standard approach to collecting, assessing the quality of, or describing structural variants, risking the entire genome eventually being labeled 'structurally variant' based on uncurated nondisease-sample data.
-
Full-text index only
MEROPS: the peptidase database.
PMID 19892822 · PMC2808883 · Nucleic acids research · 2010 · 8 claims · 5 setups
MEROPS is a manually curated hierarchical classification of peptidases and protein inhibitors organized into protein species, families, and clans based on sequence and structural homology.
-
Full-text index only
NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins.
PMID 15608248 · PMC539979 · Nucleic acids research · 2005 · 7 claims · 5 setups
RefSeq provides a curated, non-redundant, explicitly linked collection of genomic, transcript and protein sequences spanning prokaryotes, eukaryotes and viruses.
-
Full-text index only
Genome annotation errors in pathway databases due to semantic ambiguity in partial EC numbers.
PMID 16034025 · PMC1179732 · Nucleic acids research · 2005 · 7 claims · 4 setups
Partial EC numbers are semantically ambiguous, and databases that assign a gene to all reactions sharing the same partial EC number make a faulty inference, causing systematic misannotation.
-
Full-text index only
Broad network-based predictability of Saccharomyces cerevisiae gene loss-of-function phenotypes.
PMID 18053250 · PMC2246260 · Genome biology · 2007 · 8 claims · 4 setups
Loss-of-function phenotypes in yeast are predictable from a gene's connections in a functional gene network via guilt-by-association.
-
Full-text index only
Extraction of human kinase mutations from literature, databases and genotyping studies.
PMID 19758464 · PMC2745582 · BMC bioinformatics · 2009 · 7 claims · 6 setups
A literature mining pipeline combining MutationFinder, false-positive filtering, and SVM-based classification can extract and disambiguate single-point mutation mentions from abstracts and full text
-
Full-text index only
Functional copy-number alterations in cancer.
PMID 18784837 · PMC2527508 · PloS one · 2008 · 8 claims · 3 setups
RAE is a comprehensive computational framework that robustly maps chromosomal alterations in tumor samples and statistically assesses their functional importance in cancer.
-
Full-text index only
The use of neuroproteomics in drug abuse research.
PMID 19926406 · PMC3947580 · Drug and alcohol dependence · 2010 · 8 claims · 8 setups
Neuroproteomic technologies (2D-DIGE, iTRAQ, ICAT, etc.) enable identification of protein-level changes underlying effects of drugs of abuse such as amphetamine, morphine, cocaine, and alcohol.
-
Full-text index only
Meeting highlights: beyond the genome 2000: the 18th International Congress of Biochemistry and Molecular Biology.
PMID 11119309 · PMC2448388 · Yeast (Chichester, England) · 2000 · 8 claims · 8 setups
Celera sequenced a human genome to ~45-fold coverage from one donor and used high-quality sequence stretches to define ~6 million SNPs
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
Identification and characterization of insect-specific proteins by genome data analysis.
PMID 17407609 · PMC1852559 · BMC genomics · 2007 · 8 claims · 7 setups
Comparative genome analysis across five holometabolous insects and three non-insect eukaryotes (opisthokonts) identifies 154 insect-specific orthologous groups (refined to 51 proteins) and 466 eukaryote/opisthokont-core orthologous groups
-
Full-text index only
Adjustment of genomic waves in signal intensities from whole-genome SNP genotyping platforms.
PMID 18784189 · PMC2577347 · Nucleic acids research · 2008 · 8 claims · 6 setups
Genomic waves are present in both Illumina and Affymetrix SNP genotyping arrays, confirming they are not platform-specific
-
Full-text index only
Genomic diversity and evolution of Mycobacterium ulcerans revealed by next-generation sequencing.
PMID 19806175 · PMC2736377 · PLoS pathogens · 2009 · 8 claims · 6 setups
Genome sequencing of three M. ulcerans strains (NM20/02, NM31/04, Jp8756) identified thousands of SNPs relative to reference strain Agy99
-
Full-text index only
DDBJ in collaboration with mass-sequencing teams on annotation.
PMID 15608189 · PMC539974 · Nucleic acids research · 2005 · 7 claims · 5 setups
DDBJ collected and released 1,066,084 entries (718,072,425 bases) in the past year, including the complete chimpanzee chromosome 22 sequence and silkworm whole-genome shotgun data
-
Has reproduction · 68
Cell-type annotation with accurate unseen cell-type identification using multiple references.
PMID 37379341 · PMC10335708 · PLoS computational biology · 2023 · 8 claims · 4 setups
mtANN integrates multiple reference datasets and eight gene selection methods via ensemble learning (multiple deep classification models + majority voting) to improve cell-type annotation accuracy
-
Has reproduction · 66
A global database for modeling tumor-immune cell communication.
PMID 37438390 · PMC10338499 · Scientific data · 2023 · 7 claims · 6 setups
TICCom integrates 739 experimentally-validated or manually-curated TIC interactions collected from more than 3,000 literatures