Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Proteomic comparison of two-dimensional gel electrophoresis profiles from human lung squamous carcinoma and normal bronchial epithelial tissues.
PMID 15626334 · PMC5172349 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 5 setups
Well-resolved, reproducible 2-DE profiles of human lung squamous carcinoma and normal bronchial epithelial tissues were obtained under optimized conditions (0.75-mg protein load, TCA precipitation)
-
Full-text index only
InParanoid 7: new algorithms and tools for eukaryotic orthology analysis.
PMID 19892828 · PMC2808972 · Nucleic acids research · 2010 · 8 claims · 7 setups
InParanoid 7 expands the database by an order of magnitude to 100 species, 1.3 million proteins, and 42.7 million pairwise ortholog groups.
-
Full-text index only
Conservation, variability and the modeling of active protein kinases.
PMID 17912359 · PMC1989141 · PloS one · 2007 · 7 claims · 5 setups
A novel sequence-order independent (fold-independent) structural alignment algorithm was developed that maximizes side-chain similarity to produce a consensus kinase structure.
-
Full-text index only
InParanoid 6: eukaryotic ortholog clusters with inparalogs.
PMID 18055500 · PMC2238924 · Nucleic acids research · 2008 · 8 claims · 3 setups
InParanoid 6 is an updated eukaryotic ortholog database covering 35 species (34 eukaryotes plus E. coli as outgroup), providing pairwise ortholog clusters with inparalogs for all species pairs.
-
Full-text index only
Exhaustive prediction of disease susceptibility to coding base changes in the human genome.
PMID 18793467 · PMC2537574 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Inter-species conservation is the strongest single predictor of disease-associated coding mutations among the factors tested.
-
Full-text index only
CpG_MI: a novel approach for identifying functional CpG islands in mammalian genomes.
PMID 19854943 · PMC2800233 · Nucleic acids research · 2010 · 8 claims · 6 setups
Functional ('bona fide') CGIs show distinct average/cumulative mutual information (AMI/CMI) distributions of neighboring CpG distances compared to non-functional CGIs and random genome segments
-
Full-text index only
Genome-wide detection of segmental duplications and potential assembly errors in the human genome sequence.
PMID 12702206 · PMC154576 · Genome biology · 2003 · 8 claims · 6 setups
Segmental duplications comprise 3.53% (107.4/3,043.1 Mb) of the June 2002 human genome assembly
-
Full-text index only
Tracing the origin of functional and conserved domains in the human proteome: implications for protein evolution at the modular level.
PMID 17090320 · PMC1654190 · BMC evolutionary biology · 2006 · 8 claims · 5 setups
HHpred (HMM-HMM comparison) detects remote homologs in the human proteome with higher sensitivity than hmmpfam (HMMER), giving 10% more functional domain coverage and 20% higher residue coverage against Pfam-A families.
-
Full-text index only
Analysis of recent segmental duplications in the bovine genome.
PMID 19951423 · PMC2796684 · BMC genomics · 2009 · 8 claims · 6 setups
Recently duplicated sequence (≥1 kb, ≥90% identity) comprises 3.11% (94.4 Mb) of the bovine genome assembly (Btau_4.0)
-
Full-text index only
Network properties of complex human disease genes identified through genome-wide association studies.
PMID 19956617 · PMC2779513 · PloS one · 2009 · 7 claims · 6 setups
Complex disease genes are significantly less central (lower degree/closeness, higher eccentricity) in the human interactome than essential and monogenic disease genes, occupying an intermediate niche between monogenic disease genes and non-disease genes
-
Has reproduction · 67
Optimal scaling of digital transcriptomes.
PMID 24223126 · PMC3819321 · PloS one · 2013 · 8 claims · 8 setups
Fifteen existing and novel transcript-count normalization algorithms can be compared with two novel, mutually independent metrics: the number of "uniform" genes (sufficiently low coefficient of variation after normalization) and low average Spearman correlation between normalized expression profiles of gene pairs.
-
Full-text index only
Comparative genomic analysis of Mycobacterium avium subspecies obtained from multiple host species.
PMID 18366709 · PMC2323391 · BMC genomics · 2008 · 8 claims · 5 setups
Genome diversity among M. avium subspecies is mediated by large sequence polymorphisms (LSPs) commonly associated with mobile genetic elements.
-
Has reproduction · 91
Genomic Description of 'Candidatus Abyssubacteria,' a Novel Subsurface Lineage Within the Candidate Phylum Hydrogenedentes.
PMID 30210471 · PMC6121073 · Frontiers in microbiology · 2018 · 8 claims · 7 setups
SURF_5 and SURF_17 are the first full genomes of a novel bacterial lineage, 'Candidatus Abyssubacteria,' within the candidate phylum Hydrogenedentes
-
Has reproduction · 90
Comparative Genomics Provides Insight into the Function of Broad-Host Range Sponge Symbionts.
PMID 34519538 · PMC8546597 · mBio · 2021 · 8 claims · 8 setups
Eleven new genomes were added to the Tethybacterales order and a novel family (Polydorabacteraceae) was identified
-
Has reproduction · 67
binny: an automated binning algorithm to recover high-quality genomes from complex metagenomic datasets.
PMID 36239393 · PMC9677464 · Briefings in bioinformatics · 2022 · 8 claims · 8 setups
binny outperforms or is highly competitive with commonly used and state-of-the-art binning methods (MetaBAT2, MaxBin2, CONCOCT, VAMB, SemiBin, MetaDecoder)
-
Full-text index only
Protein length in eukaryotic and prokaryotic proteomes.
PMID 15951512 · PMC1150220 · Nucleic acids research · 2005 · 7 claims · 5 setups
Eukaryotic proteins are significantly longer than prokaryotic proteins across virtually all functional categories and the majority of protein families
-
Full-text index only
Increased DNA microarray hybridization specificity using sscDNA targets.
PMID 15847692 · PMC1090574 · BMC genomics · 2005 · 7 claims · 5 setups
A single round of ribo-SPIA amplification produces sufficient sscDNA for microarray hybridization from as little as 5 ng of starting total RNA
-
Full-text index only
Genotype differences in cognitive functioning in Noonan syndrome.
PMID 19077116 · PMC2760992 · Genes, brain, and behavior · 2009 · 8 claims · 6 setups
Genotype differences account for some of the variation in cognitive ability in Noonan syndrome
-
Full-text index only
Assessing the gene space in draft genomes.
PMID 19042974 · PMC2615622 · Nucleic acids research · 2009 · 6 claims · 7 setups
The proportion of mapped CEGs in a draft genome assembly is a useful metric for describing gene space completeness, complementing N50 and x-fold coverage.
-
Full-text index only
Information-based methods for predicting gene function from systematic gene knock-downs.
PMID 18959798 · PMC2596148 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Information-based metrics, which incorporate a phenotype's genomic frequency, outperform non-information-based metrics for detecting gene-gene functional similarity from phenotypic knock-down profiles.