Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Empirical codon substitution matrix.
PMID 15927081 · PMC1173088 · BMC bioinformatics · 2005 · 8 claims · 5 setups
The authors present the first empirical codon substitution matrix built entirely from alignments of vertebrate coding DNA sequences.
-
Full-text index only
Large-scale analysis of human alternative protein isoforms: pattern classification and correlation with subcellular localization signals.
PMID 15860772 · PMC1087780 · Nucleic acids research · 2005 · 8 claims · 8 setups
Constructed a large-scale dataset of 6876 human alternative protein isoforms from 2624 genes by combining H-Invitational full-length cDNA data and SwissProt VARSPLIC entries
-
Full-text index only
Protein ranking by semi-supervised network propagation.
PMID 16723003 · PMC1810311 · BMC bioinformatics · 2006 · 8 claims · 5 setups
RankProp, a diffusion-based network propagation algorithm on a PSI-BLAST-derived protein similarity network, significantly outperforms local search methods (BLAST/PSI-BLAST) at detecting remote homologs.
-
Full-text index only
Conservation, variability and the modeling of active protein kinases.
PMID 17912359 · PMC1989141 · PloS one · 2007 · 7 claims · 5 setups
A novel sequence-order independent (fold-independent) structural alignment algorithm was developed that maximizes side-chain similarity to produce a consensus kinase structure.
-
Full-text index only
Sources of variability and effect of experimental approach on expression profiling data interpretation.
PMID 11936955 · PMC65691 · BMC bioinformatics · 2002 · 8 claims · 7 setups
Intra-patient tissue heterogeneity (different regions of the same biopsy) is often the greatest source of variability in expression profiling
-
Full-text index only
Identification of gene interactions associated with disease from gene expression data using synergy networks.
PMID 18234101 · PMC2258206 · BMC systems biology · 2008 · 8 claims · 4 setups
Synergy of a gene pair with respect to disease, defined as I(G1,G2;C) - [I(G1;C)+I(G2;C)], identifies gene pairs that interact cooperatively with respect to a phenotype rather than independently.
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Full-text index only
The Princeton Protein Orthology Database (P-POD): a comparative genomics analysis tool for biologists.
PMID 17712414 · PMC1942082 · PloS one · 2007 · 8 claims · 5 setups
P-POD is the first comparative genomics database to combine results from multiple computational ortholog/homolog prediction methods with manually curated literature-derived experimental evidence of functional conservation.
-
Full-text index only
Characterisation of the genomic architecture of human chromosome 17q and evaluation of different methods for haplotype block definition.
PMID 15850495 · PMC1090572 · BMC genetics · 2005 · 8 claims · 6 setups
Haplotype block definitions based on LD measures (Definitions 1, 2, 3, 5) produce fewer, shorter blocks with limited sequence coverage compared to the haplotype diversity-based method (Definition 4)
-
Full-text index only
Haplotype frequencies at the DRD2 locus in populations of the East European Plain.
PMID 19793394 · PMC2765450 · BMC genetics · 2009 · 8 claims · 4 setups
The three-locus TaqI B-TaqI D-TaqI A haplotype at DRD2 constitutes a powerful genetic marker reflecting the most ancient dispersal of anatomically modern humans
-
Has reproduction · 57
Diapause vs. reproductive programs: transcriptional phenotypes in a keystone copepod.
PMID 33782539 · PMC8007741 · Communications biology · 2021 · 8 claims · 7 setups
t-SNE clustering of all-gene expression data groups field-collected (diapause program) samples into one cluster while early and late culture (reproductive program) samples separate into two distinct phenotypes
-
Full-text index only
Genetic analysis of the GLUT10 glucose transporter (SLC2A10) polymorphisms in Caucasian American type 2 diabetes.
PMID 16336637 · PMC1325051 · BMC medical genetics · 2005 · 7 claims · 5 setups
GLUT10 (SLC2A10) is a facilitative glucose transporter gene mapped within the T2DM-linked chromosome 20q12-13.1 region
-
Full-text index only
Comprehensive search for intra- and inter-specific sequence polymorphisms among coding envelope genes of retroviral origin found in the human genome: genes and pseudogenes.
PMID 16150157 · PMC1236922 · BMC genomics · 2005 · 8 claims · 5 setups
HERV-W (envW) and HERV-FRD (envFRD) envelope genes, both specifically expressed in placenta, show strong sequence conservation with only two nonsynonymous SNPs identified across 91 individuals
-
Full-text index only
Computer identification of snoRNA genes using a Mammalian Orthologous Intron Database.
PMID 16093549 · PMC1184218 · Nucleic acids research · 2005 · 8 claims · 5 setups
Created the Mammalian Orthologous Intron Database (MOID) containing orthologous introns of human, mouse and rat identified via conserved reading-frame position
-
Full-text index only
The estrogen hypothesis of schizophrenia implicates glucose metabolism: association study in three independent samples.
PMID 18460190 · PMC2391158 · BMC medical genetics · 2008 · 8 claims · 4 setups
A novel candidate-gene selection strategy combining unbiased schizophrenia expression/linkage data with the estrogen hypothesis as a biological filter identifies glycolysis as a candidate pathway network for schizophrenia
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.
-
Has reproduction · 89
Improved eukaryotic detection compatible with large-scale automated analysis of metagenomes.
PMID 37032329 · PMC10084625 · Microbiome · 2023 · 8 claims · 7 setups
MAPQ ≥30 filtering improves precision but substantially reduces recall, especially for unrepresented/divergent eukaryotic taxa
-
Has reproduction · 92
Evaluation of core genome and whole genome multilocus sequence typing schemes for Campylobacter jejuni and Campylobacter coli outbreak detection in the USA.
PMID 37133905 · PMC10272873 · Microbial genomics · 2023 · 8 claims · 8 setups
cgMLST, wgMLST and hqSNP WGS-based analysis methods clustered C. jejuni and C. coli isolates in concordance with epidemiological data.
-
Full-text index only
Tracing the origin of functional and conserved domains in the human proteome: implications for protein evolution at the modular level.
PMID 17090320 · PMC1654190 · BMC evolutionary biology · 2006 · 8 claims · 5 setups
HHpred (HMM-HMM comparison) detects remote homologs in the human proteome with higher sensitivity than hmmpfam (HMMER), giving 10% more functional domain coverage and 20% higher residue coverage against Pfam-A families.
-
Full-text index only
Eighth major clade for hepatitis delta virus.
PMID 17073101 · PMC3294742 · Emerging infectious diseases · 2006 · 7 claims · 7 setups
Three HDV isolates (dFr644, dFr2072, dFr2736) form a monophyletic group distinct from HDV-1 through HDV-7, constituting a new eighth major clade (HDV-8) of the Deltavirus genus.