Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Evolutionary sequence analysis of complete eukaryote genomes.
PMID 15762985 · PMC1274250 · BMC bioinformatics · 2005 · 8 claims · 6 setups
A conservative genome-comparison method (MIA) identifies panorthologs — strict single-copy 1:1 orthologs containing only species divergences, no paralogy — to minimize errors from gene duplication in evolutionary sequence analysis.
-
Full-text index only
Coverage of whole proteome by structural genomics observed through protein homology modeling database.
PMID 17146617 · PMC1769342 · Journal of structural and functional genomics · 2006 · 8 claims · 7 setups
FAMSBASE, a homology-modeling database of whole-genome ORFs, currently covers about 50% of predicted ORFs (368,724 of 734,193) across 276 genomes with modeled 3D structures.
-
Full-text index only
CoMoDis: composite motif discovery in mammalian genomes.
PMID 17130158 · PMC1702496 · Nucleic acids research · 2007 · 7 claims · 4 setups
CoMoDis is a new bioinformatics tool that streamlines computational identification of novel regulatory modules starting from a single seed motif
-
Full-text index only
Insights into the coupling of duplication events and macroevolution from an age profile of animal transmembrane gene families.
PMID 16895434 · PMC1534073 · PLoS computational biology · 2006 · 8 claims · 7 setups
The density of transmembrane gene duplicates positively correlates with the estimated maximum number of cell types of common ancestors
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
Reference based annotation with GeneMapper.
PMID 16600017 · PMC1557983 · Genome biology · 2006 · 7 claims · 6 setups
GeneMapper transfers reference gene annotations to target genomes with higher accuracy than GeneWise and Projector
-
Full-text index only
TreeFam: a curated database of phylogenetic trees of animal gene families.
PMID 16381935 · PMC1347480 · Nucleic acids research · 2006 · 7 claims · 6 setups
Tree-based inference of orthologs and paralogs is more robust than BLAST-based methods because evolutionary rates (and thus pairwise BLAST scores) vary across gene family members
-
Full-text index only
EPGD: a comprehensive web resource for integrating and displaying eukaryotic paralog/paralogon information.
PMID 17984073 · PMC2238967 · Nucleic acids research · 2008 · 8 claims · 8 setups
EPGD is a gene-centered, internet-accessible database integrating paralog family and paralogon information for 26 eukaryotic genomes.
-
Full-text index only
The Princeton Protein Orthology Database (P-POD): a comparative genomics analysis tool for biologists.
PMID 17712414 · PMC1942082 · PloS one · 2007 · 8 claims · 5 setups
P-POD is the first comparative genomics database to combine results from multiple computational ortholog/homolog prediction methods with manually curated literature-derived experimental evidence of functional conservation.
-
Full-text index only
The human phylome.
PMID 17567924 · PMC2394744 · Genome biology · 2007 · 6 claims · 5 setups
Reconstruction of the human phylome: evolutionary trees for all human proteins and their homologs among 39 fully sequenced eukaryotic genomes, using a pipeline combining alignment trimming, NJ, ML (PhyML) and Bayesian (MrBayes) methods.
-
Full-text index only
Towards a comprehensive structural coverage of completed genomes: a structural genomics viewpoint.
PMID 17349043 · PMC1829165 · BMC bioinformatics · 2007 · 8 claims · 6 setups
A combined target-selection approach — pursuing both structurally uncharacterised domain families and additional targets from large structurally characterised superfamilies — is essential for comprehensive structural coverage of the genomes.
-
Full-text index only
Kinetoplastid genomics: the thin end of the wedge.
PMID 18675383 · PMC2676795 · Infection, genetics and evolution : journal of molecular epidemiology and evolutionary genetics in infectious diseases · 2008 · 8 claims · 8 setups
Completion of the T. brucei, T. cruzi, and L. major genome sequencing projects enabled numerous studies that would otherwise have been difficult or impossible.
-
Full-text index only
Exonic remnants of whole-genome duplication reveal cis-regulatory function of coding exons.
PMID 19969543 · PMC2831330 · Nucleic acids research · 2010 · 8 claims · 8 setups
38 candidate cis-regulatory coding exons (RCEs) with predicted target genes were identified genome-wide
-
Full-text index only
A pharmacogene database enhanced by the 1000 Genomes Project.
PMID 19745786 · PMC2935084 · Pharmacogenetics and genomics · 2009 · 7 claims · 4 setups
The database provides a convenient portal for immediate utilization of newly released 1000 Genomes Project (KGP) data in pharmacogenetic studies
-
Full-text index only
Fully haplotyped genome assemblies of healthy individuals reveal variability in 5'ss strength and support by splicing regulatory proteins.
PMID 40191587 · PMC11970367 · NAR genomics and bioinformatics · 2025 · 8 claims · 5 setups
44 individuals' fully haplotyped diploid genome assemblies (88 haplotypes) from the 1000 Genomes Project were used to comprehensively assess homozygous and heterozygous sequence variations around and within 5'ss
-
Full-text index only
Dynamics of gut bacteriophage in diversity outbred mice studied over lifespan and during extreme caloric restriction.
PMID 41772715 · PMC12983593 · Microbiome · 2026 · 8 claims · 8 setups
Quiescent prophages dominate gut viral metagenomes, consistent with 'piggyback-the-winner' dynamics
-
Has reproduction · 50
RNA-Seq alignment to individualized genomes improves transcript abundance estimates in multiparent populations.
PMID 25236449 · PMC4174954 · Genetics · 2014 · 8 claims · 7 setups
Genetic variants distinguishing an individual genome from the reference cause read misalignment and biased transcript abundance estimates, and fine-tuning of alignment algorithms does not correct this problem.
-
Full-text index only
Beyond blacklists: a critical assessment of exclusion set generation strategies and alternative approaches.
PMID 41826793 · PMC13020910 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 8 setups
Pre-generated Blacklist exclusion sets were difficult to reproduce due to sensitivity to input BAM data, aligner choice, and read length
-
Has reproduction · 83
Macrel: antimicrobial peptide screening in genomes and metagenomes.
PMID 33384902 · PMC7751412 · PeerJ · 2020 · 8 claims · 8 setups
Macrel introduces a novel set of 22 peptide features (6 local, 16 global), including a new Free Energy Transition (FET) feature group, for AMP and hemolytic activity classification
-
Full-text index only
From microarrays to genome duplications.
PMID 12914655 · PMC193639 · Genome biology · 2003 · 8 claims · 8 setups
Gene3D shows that most genes across sequenced genomes can be assigned to known structural domain families, many of which are shared across kingdoms of life