Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 78
Machine learning and free energy clustering reveal PAH protein binding linked to AD risk.
PMID 41953002 · PMC13053772 · iScience · 2026 · 7 claims · 8 setups
An integrated framework of bioinformatics, machine learning, and ΔG clustering can prioritize PAHs for AD-associated neurotoxicity.
-
Full-text index only
GLIDA: GPCR--ligand database for chemical genomics drug discovery--database and tools update.
PMID 17986454 · PMC2238933 · Nucleic acids research · 2008 · 7 claims · 5 setups
GLIDA is a public relational database integrating biological information on GPCRs with chemical information on their ligands and their binding interactions.
-
Full-text index only
Comparative genomics of vertebrate Fox cluster loci.
PMID 17062144 · PMC1634998 · BMC genomics · 2006 · 8 claims · 3 setups
Two additional human paralogous Fox cluster regions exist, on chromosomes 14 and 20, beyond the previously known chromosome 6 and 16 loci
-
Full-text index only
Comparative genomics of Lbx loci reveals conservation of identical Lbx ohnologs in bony vertebrates.
PMID 18541024 · PMC2446394 · BMC evolutionary biology · 2008 · 8 claims · 3 setups
Extant bony vertebrates (osteichthyans) retain only Lbx1- and Lbx2-type genes; no distinct Lbx3/Lbx4 proteins exist.
-
Full-text index only
PA-GOSUB: a searchable database of model organism protein sequences with their predicted Gene Ontology molecular function and subcellular localization.
PMID 15608166 · PMC540074 · Nucleic acids research · 2005 · 7 claims · 4 setups
PA-GOSUB significantly extends the coverage of GO molecular function and subcellular localization annotations for 10 model organism proteomes compared with existing databases (GOA, Swiss-Prot).
-
Full-text index only
Epidemiology of doublet/multiplet mutations in lung cancers: evidence that a subset arises by chronocoordinate events.
PMID 19005564 · PMC2579325 · PloS one · 2008 · 8 claims · 7 setups
Doublet mutations are significantly more frequent in EGFR (6.0%) and TP53 (2.3%) in human lung cancer than spontaneous doublets in mouse lacI (0.7%), about 8-fold and 3-fold higher respectively.
-
Full-text index only
L2L: a simple tool for discovering the hidden significance in microarray expression data.
PMID 16168088 · PMC1242216 · Genome biology · 2005 · 8 claims · 4 setups
L2L systematically compares a user's differentially expressed gene list against a database of published differentially expressed gene lists to find statistically significant overlaps and generate hypotheses about shared mechanisms
-
Full-text index only
The RCSB PDB information portal for structural genomics.
PMID 16381872 · PMC1347482 · Nucleic acids research · 2006 · 7 claims · 5 setups
The RCSB PDB Structural Genomics Information Portal integrates three resources: Structural Genomics Initiatives, Targets (TargetDB/PepcDB), and Structures (functional coverage analysis).
-
Full-text index only
The role of positive selection in determining the molecular cause of species differences in disease.
PMID 18837980 · PMC2576240 · BMC evolutionary biology · 2008 · 8 claims · 6 setups
Genes predicted to be under positive selection during human evolution are implicated in diseases (epithelial cancers, schizophrenia, autoimmune diseases, Alzheimer's disease) that differ in prevalence and symptomatology between humans and other mammals
-
Full-text index only
Investigation of the human tear film proteome using multiple proteomic approaches.
PMID 18334958 · PMC2268847 · Molecular vision · 2008 · 8 claims · 6 setups
Tear collection method (capillary vs. Schirmer strip) significantly impacts which proteins are detected in the tear film proteome.
-
Full-text index only
Proteomic analysis of human aqueous humor using multidimensional protein identification technology.
PMID 20019884 · PMC2793904 · Molecular vision · 2009 · 8 claims · 4 setups
Albumin/IgG depletion combined with MudPIT (2D-LC-MS/MS) enables high-confidence, extensive characterization of the human AH proteome
-
Full-text index only
iRefIndex: a consolidated protein interaction database with provenance.
PMID 18823568 · PMC2573892 · BMC bioinformatics · 2008 · 6 claims · 3 setups
A reproducible key (ROG) for each protein interactor and a corresponding key (RIG) for each interaction record can be generated by anyone using only primary sequence, taxonomy identifier, and the SHA-1 algorithm (SEGUID).
-
Full-text index only
Identification, characterization and comparative genomics of chimpanzee endogenous retroviruses.
PMID 16805923 · PMC1779541 · Genome biology · 2006 · 8 claims · 6 setups
The chimpanzee genome contains at least 42 separate families of endogenous retroviruses, 9 newly identified
-
Full-text index only
Gene losses during human origins.
PMID 16464126 · PMC1361800 · PLoS biology · 2006 · 7 claims · 7 setups
A comparative genomic screen identified 67 new human-specific nonprocessed pseudogenes, bringing the total (with 13 from prior literature) to 80 human-specific pseudogenes.
-
Full-text index only
The Proteomic Code: a molecular recognition code for proteins.
PMID 17999762 · PMC2206014 · Theoretical biology & medical modelling · 2007 · 8 claims · 8 setups
The Proteomic Code is a set of rules by which genetic information is transferred into the physico-chemical properties of amino acids, determining protein-protein interactions and folding; it is part of the redundant Genetic Code.
-
Full-text index only
Sequence occurrence and structural uniqueness of a G-quadruplex in the human c-kit promoter.
PMID 17720713 · PMC2034477 · Nucleic acids research · 2007 · 8 claims · 4 setups
The native 22-nt c-kit87 sequence occurs only once in the entire human genome.
-
Full-text index only
Comparative Toxicogenomics Database: a knowledgebase and discovery tool for chemical-gene-disease networks.
PMID 18782832 · PMC2686584 · Nucleic acids research · 2009 · 8 claims · 5 setups
CTD is a manually curated knowledgebase that integrates chemical-gene interactions, chemical-disease relationships, and gene-disease relationships into a chemical-gene-disease triad
-
Full-text index only
Diet and DNA.
PMID 15174456 · PMC1315985 · Environmental health perspectives · 2004 · 8 claims · 8 setups
The tombusvirus CIRV p19 protein selectively binds short (21-22 nt) silencing siRNAs, using tryptophan residues Trp39 and Trp42 as molecular 'calipers' that stack with the ends of the siRNA duplex
-
Full-text index only
PathogenMIPer: a tool for the design of molecular inversion probes to detect multiple pathogens.
PMID 17105657 · PMC1657037 · BMC bioinformatics · 2006 · 6 claims · 5 setups
PathogenMIPer designs unique, target-specific MIP probes, assembling all probe components (target-specific sequences, barcodes, universal primers, restriction sites) into ready-to-order probes for any genome.
-
Full-text index only
Sys-BodyFluid: a systematical database for human body fluid proteome research.
PMID 18978022 · PMC2686600 · Nucleic acids research · 2009 · 6 claims · 4 setups
Sys-BodyFluid is a web-based database integrating proteomic data from 11 human body fluids (plasma/serum, urine, cerebrospinal fluid, saliva, bronchoalveolar lavage fluid, synovial fluid, nipple aspirate fluid, tear fluid, seminal fluid, milk, amniotic fluid), containing over 10,000 proteins