Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Identifying repeat domains in large genomes.
PMID 16507140 · PMC1431705 · Genome biology · 2006 · 7 claims · 5 setups
A repeat domain graph, built using a modified A-Bruijn graph framework, decomposes a repeat library into shared repeat domains and reveals the mosaic structure of repeat families.
-
Full-text index only
Helminth genomics: The implications for human health.
PMID 19855829 · PMC2757907 · PLoS neglected tropical diseases · 2009 · 8 claims · 7 setups
More than two billion people (one-third of humanity) are infected with helminth parasites, causing major morbidity, mortality, and poverty maintenance
-
Full-text index only
From single cells to whole organisms.
PMID 16420683 · PMC1414103 · Genome biology · 2005 · 8 claims · 8 setups
The genetic-interaction map in S. cerevisiae is roughly four times as complex as the protein-protein interaction map, and genetic interactions do not overlap with physical interactions but instead predict functional neighborhoods
-
Full-text index only
Functional coverage of the human genome by existing structures, structural genomics targets, and homology models.
PMID 16118666 · PMC1188274 · PLoS computational biology · 2005 · 8 claims · 5 setups
Existing PDB structures provide single-domain coverage for 37% of functional classes in the human genome and complete (whole-protein) structure coverage for 25%.
-
Full-text index only
The dystrobrevin-binding protein 1 gene: features and networks.
PMID 18663367 · PMC2859304 · Molecular psychiatry · 2009 · 8 claims · 6 setups
DTNBP1 gene structure, protein-coding sequence, and dysbindin domain are conserved across 13 vertebrate species, while noncoding sequence is diverse.
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
Genomic view of the evolution of the complement system.
PMID 16896831 · PMC2480602 · Immunogenetics · 2006 · 8 claims · 6 setups
Bony fish and higher vertebrates share practically the same set of complement genes, indicating most complement gene duplications occurred by the teleost/mammalian divergence (~500 MYA)
-
Full-text index only
SUPERFAMILY--sophisticated comparative genomics, data mining, visualization and phylogeny.
PMID 19036790 · PMC2686452 · Nucleic acids research · 2009 · 7 claims · 6 setups
SUPERFAMILY provides structural, functional and evolutionary annotation for proteins from all completely sequenced genomes using SCOP-based hidden Markov models
-
Full-text index only
Structural evolution of the protein kinase-like superfamily.
PMID 16244704 · PMC1261164 · PLoS computational biology · 2005 · 8 claims · 5 setups
All kinases in the superfamily share a 'universal core' domain consisting only of the regions required for ATP binding and the phosphotransfer reaction.
-
Full-text index only
MEROPS: the peptidase database.
PMID 19892822 · PMC2808883 · Nucleic acids research · 2010 · 8 claims · 5 setups
MEROPS is a manually curated hierarchical classification of peptidases and protein inhibitors organized into protein species, families, and clans based on sequence and structural homology.
-
Full-text index only
DBD--taxonomically broad transcription factor predictions: new content and functionality.
PMID 18073188 · PMC2238844 · Nucleic acids research · 2008 · 8 claims · 3 setups
DBD is a database of predicted sequence-specific DNA-binding transcription factors covering over 700 publicly available proteomes, up from 150 in the initial version.
-
Has reproduction · 67
Satellitome Analysis and Transposable Elements Comparison in Geographically Distant Populations of Spodoptera frugiperda.
PMID 35455012 · PMC9026859 · Life (Basel, Switzerland) · 2022 · 8 claims · 5 setups
Most transposable elements are commonly shared across all eight geographically distant S. frugiperda samples, except Maverick and PIF/Harbinger elements which show divergent repeat copies
-
Has reproduction · 50
auts2 Features and Expression Are Highly Conserved during Evolution Despite Different Evolutionary Fates Following Whole Genome Duplication.
PMID 36078102 · PMC9454499 · Cells · 2022 · 8 claims · 7 setups
auts2a and auts2b originate from the teleost-specific whole genome duplication (TGD)
-
Full-text index only
Protein length in eukaryotic and prokaryotic proteomes.
PMID 15951512 · PMC1150220 · Nucleic acids research · 2005 · 7 claims · 5 setups
Eukaryotic proteins are significantly longer than prokaryotic proteins across virtually all functional categories and the majority of protein families
-
Full-text index only
Investigating hookworm genomes by comparative analysis of two Ancylostoma species.
PMID 15854223 · PMC1112591 · BMC genomics · 2005 · 8 claims · 8 setups
Nearly 20,000 ESTs from 7 cDNA libraries define nearly 7,000 hookworm genes across A. caninum and A. ceylanicum
-
Full-text index only
Genomic organization, annotation, and ligand-receptor inferences of chicken chemokines and chemokine receptor genes based on comparative genomics.
PMID 15790398 · PMC1082905 · BMC genomics · 2005 · 8 claims · 6 setups
Twenty-three chemokine genes and 14 chemokine receptor genes were identified in the chicken genome, including 12 new chemokines and 7 new receptors beyond prior reports.
-
Full-text index only
Evolutionary cores of domain co-occurrence networks.
PMID 15788102 · PMC1079808 · BMC evolutionary biology · 2005 · 8 claims · 4 setups
The innermost (globally central) cores of protein domain co-occurrence networks gradually grow in size with increasing evolutionary/developmental complexity of the organism.
-
Full-text index only
Comparative genomics of the syndecans defines an ancestral genomic context associated with matrilins in vertebrates.
PMID 16620374 · PMC1464127 · BMC genomics · 2006 · 8 claims · 6 setups
Syndecan-encoding sequences are present in Cnidaria and throughout the Bilateria, showing deep conservation of the family.
-
Full-text index only
Comparative genomic analysis of three Leishmania species that cause diverse human disease.
PMID 17572675 · PMC2592530 · Nature genetics · 2007 · 8 claims · 6 setups
L. infantum and L. braziliensis genomes were sequenced and show marked conservation of synteny with L. major, with only ~200 genes differentially distributed among the three species
-
Full-text index only
Comparative sequence analysis of leucine-rich repeats (LRRs) within vertebrate toll-like receptors.
PMID 17517123 · PMC1899181 · BMC genomics · 2007 · 8 claims · 4 setups
A new method combining known LRR structures, multiple sequence alignment, and secondary structure prediction identifies and aligns LRRs in TLRs more accurately than PFAM/InterPro/SMART