Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Shotgun haplotyping: a novel method for surveying allelic sequence variation.
PMID 16221968 · PMC1253838 · Nucleic acids research · 2005 · 8 claims · 7 setups
A novel shotgun haplotyping method generates haplotypic sequences from long PCR products by shotgun sequencing both alleles concurrently and using read-pair information to separate alleles during assembly
-
Full-text index only
The distribution of SNPs in human gene regulatory regions.
PMID 16209714 · PMC1260019 · BMC genomics · 2005 · 8 claims · 6 setups
SNPs occur with higher density closer to the transcriptional start site within gene promoter regions than in further upstream regions
-
Full-text index only
Functional analysis of human hematopoietic stem cell gene expression using zebrafish.
PMID 16089502 · PMC1166352 · PLoS biology · 2005 · 8 claims · 8 setups
277 unique transcripts are differentially expressed between Rho lo and Rho hi HSC-enriched/depleted populations, conserved across both umbilical cord blood and bone marrow
-
Full-text index only
In vitro identification and in silico utilization of interspecies sequence similarities using GeneChip technology.
PMID 15871745 · PMC1156887 · BMC genomics · 2005 · 7 claims · 6 setups
Only 14±2% of canine transcripts were detected by U133A probe sets versus 49±6% of human transcripts when hybridized to the same chip
-
Full-text index only
Large-scale analysis of human alternative protein isoforms: pattern classification and correlation with subcellular localization signals.
PMID 15860772 · PMC1087780 · Nucleic acids research · 2005 · 8 claims · 8 setups
Constructed a large-scale dataset of 6876 human alternative protein isoforms from 2624 genes by combining H-Invitational full-length cDNA data and SwissProt VARSPLIC entries
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Full-text index only
How to find soluble proteins: a comprehensive analysis of alpha/beta hydrolases for recombinant expression in E. coli.
PMID 15804363 · PMC1079826 · BMC genomics · 2005 · 7 claims · 7 setups
Predicted solubility in E. coli (via CV-CV') depends on hydrolase size, phylogenetic origin, homologous family, and superfamily
-
Full-text index only
Competitive enzymatic reaction to control allele-specific extensions.
PMID 15767273 · PMC1065263 · Nucleic acids research · 2005 · 6 claims · 7 setups
Protease-mediated allele-specific extension (PrASE) uses competition between polymerase activity and Proteinase K-mediated polymerase degradation to allow extension of perfectly matched primers while eliminating slower mismatched primer extension.
-
Has reproduction · 49
Aberration in DNA methylation in B-cell lymphomas has a complex origin and increases with disease severity.
PMID 23326238 · PMC3542081 · PLoS genetics · 2013 · 8 claims · 8 setups
B-cell non-Hodgkin lymphomas display striking intra-tumor (intra-sample) and inter-patient (inter-sample) cytosine methylation heterogeneity that increases progressively with disease aggressiveness (NBC<NGC<FL<GCB<ABC).
-
Full-text index only
Non-linear mapping for exploratory data analysis in functional genomics.
PMID 15661072 · PMC548129 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A relaxation method for non-linear mapping adapts one pair of points per step rather than all points at once, and was originally shown by Chang and Lee to outperform Sammon's mapping in cluster detection effectiveness and computational efficiency.
-
Full-text index only
GeneTide--Terra Incognita Discovery Endeavor: a new transcriptome focused member of the GeneCards/GeneNote suite of databases.
PMID 15608261 · PMC540076 · Nucleic acids research · 2005 · 8 claims · 7 setups
GeneTide integrates UniGene, DoTS, AceView, BLAT/GeneLoc genomic alignment, and GeneAnnot probe-set data into a unified Consensus/Uniqueness/Score scheme to associate ESTs with GeneCards genes
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
Elevated serum levels of interferon-regulated chemokines are biomarkers for active human systemic lupus erythematosus.
PMID 17177599 · PMC1702557 · PLoS medicine · 2006 · 8 claims · 4 setups
30 of 160 measured serum analytes (cytokines, chemokines, growth factors, soluble receptors) are dysregulated in SLE serum
-
Full-text index only
Genome assembly comparison identifies structural variants in the human genome.
PMID 17115057 · PMC2674632 · Nature genetics · 2006 · 7 claims · 7 setups
Genome assembly comparison is a robust approach for identifying all classes of genetic variation, with no lower size limit.
-
Full-text index only
The HIV positive selection mutation database.
PMID 17108357 · PMC1669717 · Nucleic acids research · 2007 · 8 claims · 5 setups
The database provides codon-level Ka/Ks selection pressure maps for HIV protease and the first 381 codons of RT, built from a novel ~50,000-sample clinical dataset.
-
Has reproduction · 43
StatsDB: platform-agnostic storage and understanding of next generation sequencing run metrics.
PMID 24627795 · PMC3938176 · F1000Research · 2013 · 8 claims · 6 setups
StatsDB is an open-source software package for storage and analysis of next generation sequencing run metrics, backed by an SQL (MySQL) database with Perl and Java APIs.
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
Identification of diagnostic markers for tuberculosis by proteomic fingerprinting of serum.
PMID 16980117 · PMC7159276 · Lancet (London, England) · 2006 · 8 claims · 5 setups
An SVM classifier trained on serum proteomic profiles discriminated patients with active tuberculosis from controls with clinically overlapping conditions
-
Has reproduction · 27
Transcriptome profiling of radish (Raphanus sativus L.) root and identification of genes involved in response to Lead (Pb) stress with next generation sequencing.
PMID 23840502 · PMC3688795 · PloS one · 2013 · 8 claims · 5 setups
A de novo radish root transcriptome of 68,940 assembled transcripts including 33,337 unigenes was generated, providing the first comprehensive molecular characterization of the radish root response to Pb stress.
-
Full-text index only
A global proteomics approach identifies novel phosphorylated signaling proteins in GPVI-activated platelets: involvement of G6f, a novel platelet Grb2-binding membrane adapter.
PMID 16941570 · PMC1869047 · Proteomics · 2006 · 8 claims · 7 setups
96 proteins undergo post-translational modification (phosphorylation) in response to CRP stimulation of human platelets, including 11 proteins not previously identified in platelets