Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
GeneTide--Terra Incognita Discovery Endeavor: a new transcriptome focused member of the GeneCards/GeneNote suite of databases.
PMID 15608261 · PMC540076 · Nucleic acids research · 2005 · 8 claims · 7 setups
GeneTide integrates UniGene, DoTS, AceView, BLAT/GeneLoc genomic alignment, and GeneAnnot probe-set data into a unified Consensus/Uniqueness/Score scheme to associate ESTs with GeneCards genes
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
Elevated serum levels of interferon-regulated chemokines are biomarkers for active human systemic lupus erythematosus.
PMID 17177599 · PMC1702557 · PLoS medicine · 2006 · 8 claims · 4 setups
30 of 160 measured serum analytes (cytokines, chemokines, growth factors, soluble receptors) are dysregulated in SLE serum
-
Full-text index only
Genome assembly comparison identifies structural variants in the human genome.
PMID 17115057 · PMC2674632 · Nature genetics · 2006 · 7 claims · 7 setups
Genome assembly comparison is a robust approach for identifying all classes of genetic variation, with no lower size limit.
-
Full-text index only
The HIV positive selection mutation database.
PMID 17108357 · PMC1669717 · Nucleic acids research · 2007 · 8 claims · 5 setups
The database provides codon-level Ka/Ks selection pressure maps for HIV protease and the first 381 codons of RT, built from a novel ~50,000-sample clinical dataset.
-
Has reproduction · 43
StatsDB: platform-agnostic storage and understanding of next generation sequencing run metrics.
PMID 24627795 · PMC3938176 · F1000Research · 2013 · 8 claims · 6 setups
StatsDB is an open-source software package for storage and analysis of next generation sequencing run metrics, backed by an SQL (MySQL) database with Perl and Java APIs.
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
Identification of diagnostic markers for tuberculosis by proteomic fingerprinting of serum.
PMID 16980117 · PMC7159276 · Lancet (London, England) · 2006 · 8 claims · 5 setups
An SVM classifier trained on serum proteomic profiles discriminated patients with active tuberculosis from controls with clinically overlapping conditions
-
Has reproduction · 27
Transcriptome profiling of radish (Raphanus sativus L.) root and identification of genes involved in response to Lead (Pb) stress with next generation sequencing.
PMID 23840502 · PMC3688795 · PloS one · 2013 · 8 claims · 5 setups
A de novo radish root transcriptome of 68,940 assembled transcripts including 33,337 unigenes was generated, providing the first comprehensive molecular characterization of the radish root response to Pb stress.
-
Full-text index only
A global proteomics approach identifies novel phosphorylated signaling proteins in GPVI-activated platelets: involvement of G6f, a novel platelet Grb2-binding membrane adapter.
PMID 16941570 · PMC1869047 · Proteomics · 2006 · 8 claims · 7 setups
96 proteins undergo post-translational modification (phosphorylation) in response to CRP stimulation of human platelets, including 11 proteins not previously identified in platelets
-
Full-text index only
Large-scale identification and characterization of alternative splicing variants of human gene transcripts using 56,419 completely sequenced and manually annotated full-length cDNAs.
PMID 16914452 · PMC1557807 · Nucleic acids research · 2006 · 8 claims · 8 setups
Analysis of 56,419 full-length cDNAs identified 6877 alternative splicing genes encoding 18,297 alternative splicing variants made of 37,670 exons.
-
Full-text index only
An investigation of polymorphisms in the 17q11.2-12 CC chemokine gene cluster for association with multiple sclerosis in Australians.
PMID 16872505 · PMC1550395 · BMC medical genetics · 2006 · 7 claims · 7 setups
Marginally significant (uncorrected) transmission distortion was found for four SNPs after stratification by HLA-DRB1*1501 status, disease course, or gender.
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Full-text index only
Predicting survival within the lung cancer histopathological hierarchy using a multi-scale genomic model of development.
PMID 16800721 · PMC1483910 · PLoS medicine · 2006 · 8 claims · 8 setups
Multi-scale genomic similarities exist between four human lung cancer subtypes and the developing mouse lung, and these similarities are prognostically meaningful.
-
Full-text index only
Adaptively inferring human transcriptional subnetworks.
PMID 16760900 · PMC1681499 · Molecular systems biology · 2006 · 8 claims · 7 setups
A multivariate linear spline (MARS-based) model correlating PWM binding scores with log expression ratios can identify active cis-motif combinations in mammalian promoters without requiring gene clustering.
-
Full-text index only
Identification of serum biomarkers for colon cancer by proteomic analysis.
PMID 16755300 · PMC2361335 · British journal of cancer · 2006 · 8 claims · 8 setups
Complement C3a des-arg, α1-antitrypsin and transferrin were identified as serum proteins with diagnostic potential for CRC.
-
Has reproduction · 68
LaSSO, a strategy for genome-wide mapping of intronic lariats and branch points using RNA-seq.
PMID 24709818 · PMC4079972 · Genome research · 2014 · 8 claims · 8 setups
LaSSO (Lariat Sequence Site Origin) identifies intronic lariat reads and pinpoints branch points genome-wide from RNA-seq data by considering every intronic base as a potential branch point and including all possible exon-skipping lariats.
-
Full-text index only
Detection of atovaquone-proguanil resistance conferring mutations in Plasmodium falciparum cytochrome b gene in Luanda, Angola.
PMID 16597338 · PMC1513587 · Malaria journal · 2006 · 6 claims · 4 setups
No pfcytb mutations associated with atovaquone-proguanil treatment failure (codon 268 wild type, T802A, A803C) were found in the Luanda study population.
-
Full-text index only
Accurate splice site prediction using support vector machines.
PMID 18269701 · PMC2230508 · BMC bioinformatics · 2007 · 8 claims · 5 setups
Weighted degree (WD) kernel SVMs outperform Markov Chains, GeneSplicer and SpliceMachine for genome-wide splice site recognition
-
Full-text index only
Proteomics as a tool for biomarker discovery.
PMID 18057524 · PMC3851415 · Disease markers · 2007 · 8 claims · 7 setups
A useful clinical biomarker must be easily attainable, have adequate sensitivity, have adequate specificity, and lead to patient benefit through intervention