Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 92
Systematic review of human post-mortem immunohistochemical studies and bioinformatics analyses unveil the complexity of astrocyte reaction in Alzheimer's disease.
PMID 34297416 · PMC8766893 · Neuropathology and applied neurobiology · 2022 · 8 claims · 5 setups
Systematic review of 306 eligible articles identified 196 distinct proteins constituting the ADRA (AD reactive astrocyte) protein set
-
Has reproduction · 50
Dissecting lncRNA-mRNA competitive regulatory network in human islet tissue exosomes of a type 1 diabetes model reveals exosome miRNA markers.
PMID 36440209 · PMC9682028 · Frontiers in endocrinology · 2022 · 6 claims · 7 setups
lncRNA-mRNA ceRNA networks differ substantially between control and cytokine-treated islet exosomes, sharing lncRNAs more than mRNAs, implying state-dependent regulatory functions
-
Full-text index only
Gene prediction in eukaryotes with a generalized hidden Markov model that uses hints from external sources.
PMID 16469098 · PMC1409804 · BMC bioinformatics · 2006 · 7 claims · 3 setups
AUGUSTUS+ extends the AUGUSTUS GHMM by combining intrinsic sequence information with extrinsic hints via an extended emission alphabet, so the GHMM jointly models the DNA sequence, gene structure, and hint collection.
-
Full-text index only
In silico discovery of gene-coding variants in murine quantitative trait loci using strain-specific genome sequence databases.
PMID 12537567 · PMC151180 · Genome biology · 2002 · 6 claims · 4 setups
Strain-specific mouse genome sequence databases can be used in a high-throughput in silico pipeline to discover gene-coding variants within murine QTLs, without de novo sequencing.
-
Full-text index only
Extraction of human kinase mutations from literature, databases and genotyping studies.
PMID 19758464 · PMC2745582 · BMC bioinformatics · 2009 · 7 claims · 6 setups
A literature mining pipeline combining MutationFinder, false-positive filtering, and SVM-based classification can extract and disambiguate single-point mutation mentions from abstracts and full text
-
Full-text index only
A systematic comparative and structural analysis of protein phosphorylation sites based on the mtcPTM database.
PMID 17521420 · PMC1929158 · Genome biology · 2007 · 7 claims · 6 setups
mtcPTM is a hierarchically organized database of human and mouse phosphosites that preserves experimental context, enabling comparison of phosphorylation patterns across conditions
-
Full-text index only
CRSD: a comprehensive web server for composite regulatory signature discovery.
PMID 16845073 · PMC1538777 · Nucleic acids research · 2006 · 7 claims · 5 setups
CRSD is a comprehensive web server integrating six large-scale databases (UniGene, mature microRNAs, putative promoter, TRANSFAC, pathway, GO) plus two newly constructed genome-wide databases (MRS and TRS) for composite regulatory signature discovery
-
Full-text index only
An analysis of human microRNA and disease associations.
PMID 18923704 · PMC2559869 · PloS one · 2008 · 8 claims · 8 setups
MicroRNAs tend to show similar dysfunctional evidence (both up- or both down-regulated) for diseases within the same disease cluster, and different dysfunctional evidence between different disease clusters.
-
Full-text index only
Aberrant 5' splice sites in human disease genes: mutation pattern, nucleotide structure and comparison of computational tools that predict their utilization.
PMID 17576681 · PMC1934990 · Nucleic acids research · 2007 · 8 claims · 4 setups
Cryptic 5'ss are best predicted by computational algorithms that accommodate nucleotide dependencies (e.g., Markov model, maximum entropy, maximum dependence decomposition) rather than by weight-matrix models
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Full-text index only
MACSIMS: multiple alignment of complete sequences information management system.
PMID 16792820 · PMC1539025 · BMC bioinformatics · 2006 · 8 claims · 5 setups
MACSIMS is a multiple alignment-based information management system combining knowledge-based database mining with ab initio sequence predictions
-
Full-text index only
Ontological visualization of protein-protein interactions.
PMID 15707487 · PMC550656 · BMC bioinformatics · 2005 · 8 claims · 8 setups
Aggregating independently made GO 'protein binding' (IPI) annotations reveals larger, previously undescribed mouse protein-protein interaction networks
-
Full-text index only
TPRpred: a tool for prediction of TPR-, PPR- and SEL1-like repeats from protein sequences.
PMID 17199898 · PMC1774580 · BMC bioinformatics · 2007 · 7 claims · 8 setups
TPRpred detects divergent/remote-homolog TPR repeat units that existing resources (Pfam, SMART, REP) fail to detect
-
Full-text index only
A searchable database of genetic evidence for psychiatric disorders.
PMID 18548508 · PMC2574546 · American journal of medical genetics. Part B, Neuropsychiatric genetics : the official publication of the International Society of Psychiatric Genetics · 2008 · 8 claims · 4 setups
SLEP (Sullivan Lab Evidence Project) is a freely available, searchable web database of findings from psychiatric genetics for non-commercial use.
-
Has reproduction · 84
Expanding the clinical spectrum of COL2A1 related disorders by a mass like phenotype.
PMID 35296718 · PMC8927422 · Scientific reports · 2022 · 8 claims · 8 setups
Four FBN1-negative patients from three families with a MASS-like phenotype carry likely pathogenic or uncertain-significance missense variants in the propeptide-coding regions of COL2A1
-
Full-text index only
Opportunities and challenges in synthetic oligosaccharide and glycoconjugate research.
PMID 20161474 · PMC2794050 · Nature chemistry · 2009 · 8 claims · 7 setups
A parallel combinatorial one-pot multi-step protecting-group procedure (Lewis acid catalyzed, up to seven steps) can transform tetra-O-TMS glucopyranosides into differentially protected monosaccharide building blocks without intermittent work-up/purification
-
Full-text index only
The Proteomic Code: a molecular recognition code for proteins.
PMID 17999762 · PMC2206014 · Theoretical biology & medical modelling · 2007 · 8 claims · 8 setups
The Proteomic Code is a set of rules by which genetic information is transferred into the physico-chemical properties of amino acids, determining protein-protein interactions and folding; it is part of the redundant Genetic Code.
-
Full-text index only
Comparison of multidimensional shotgun technologies targeting tissue proteomics.
PMID 19960471 · PMC3465977 · Electrophoresis · 2009 · 6 claims · 4 setups
CITP-based multidimensional separation achieves superior overall proteome performance (more total peptide, distinct peptide, and distinct protein identifications) than SCX/MuDPIT under matched conditions
-
Full-text index only
Cancer-specific high-throughput annotation of somatic mutations: computational prediction of driver missense mutations.
PMID 19654296 · PMC2763410 · Cancer research · 2009 · 7 claims · 7 setups
CHASM, a Random Forest-based computational method, was developed to identify and prioritize missense mutations likely to be functional drivers of tumor cell proliferation.
-
Has reproduction · 84
Expression Atlas update--a database of gene and transcript expression from microarray- and sequencing-based functional genomics experiments.
PMID 24304889 · PMC3964963 · Nucleic acids research · 2014 · 8 claims · 6 setups
Expression Atlas is a value-added database providing gene, protein and splice variant expression across cell types, organism parts, developmental stages, diseases and other biological/experimental conditions, built from manually curated high-quality microarray and RNA-sequencing experiments from ArrayExpress.