Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The cohesin complex: sequence homologies, interaction networks and shared motifs.
PMID 11276426 · PMC30708 · Genome biology · 2001 · 8 claims · 8 setups
Mouse Mmip1 and Smc3 (SMCD) share 99% sequence identity and are products of the same gene
-
Has reproduction · 75
Genomic regions and candidate genes selected during the breeding of rice in Vietnam.
PMID 35899250 · PMC9309459 · Evolutionary applications · 2022 · 8 claims · 7 setups
XP-CLR and FST scans identify genomic regions with distorted allele frequency/differentiation patterns resulting from differential selective pressures between Vietnamese rice subpopulations
-
Full-text index only
SysZNF: the C2H2 zinc finger gene database.
PMID 18974185 · PMC2686507 · Nucleic acids research · 2009 · 7 claims · 6 setups
SysZNF is a database that systematically catalogs C2H2-ZNF genes in human and mouse with physical location, gene models, expression probes, protein domains, homologs, and literature links
-
Full-text index only
The dystrobrevin-binding protein 1 gene: features and networks.
PMID 18663367 · PMC2859304 · Molecular psychiatry · 2009 · 8 claims · 6 setups
DTNBP1 gene structure, protein-coding sequence, and dysbindin domain are conserved across 13 vertebrate species, while noncoding sequence is diverse.
-
Full-text index only
Towards a comprehensive structural coverage of completed genomes: a structural genomics viewpoint.
PMID 17349043 · PMC1829165 · BMC bioinformatics · 2007 · 8 claims · 6 setups
A combined target-selection approach — pursuing both structurally uncharacterised domain families and additional targets from large structurally characterised superfamilies — is essential for comprehensive structural coverage of the genomes.
-
Full-text index only
Predicting positive p53 cancer rescue regions using Most Informative Positive (MIP) active learning.
PMID 19756158 · PMC2742196 · PLoS computational biology · 2009 · 8 claims · 4 setups
MIP active learning is a novel active learning method that preferentially seeks informative Positive (functionally active) examples rather than only maximizing classifier accuracy.
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
Bioinformatic mapping of AlkB homology domains in viruses.
PMID 15627404 · PMC544882 · BMC genomics · 2005 · 8 claims · 8 setups
AlkB-like domains are found in at least 22 different single-stranded RNA positive-strand plant viruses, mainly within a subgroup of the Flexiviridae family.
-
Full-text index only
Identification and characterisation of the angiotensin converting enzyme-3 (ACE3) gene: a novel mammalian homologue of ACE.
PMID 17597519 · PMC1925091 · BMC genomics · 2007 · 7 claims · 7 setups
A novel single-domain ACE-like gene, ACE3, exists in mouse, rat, cow, dog and human genomes, located on the same chromosome downstream of ACE.
-
Full-text index only
Rapid detection and curation of conserved DNA via enhanced-BLAT and EvoPrinterHD analysis.
PMID 18307801 · PMC2268679 · BMC genomics · 2008 · 8 claims · 8 setups
eBLAT detects up to 75% more conserved bases than original BLAT alignments, with the largest gains between evolutionarily distant orthologs
-
Full-text index only
Recent segmental and gene duplications in the mouse genome.
PMID 12914656 · PMC193640 · Genome biology · 2003 · 8 claims · 8 setups
33.6 Mb (1.2%) of the February 2003 mouse genome assembly (2,695 Mb) is involved in recent segmental duplications
-
Full-text index only
Systematic analysis of human kinase genes: a large number of genes and alternative splicing events result in functional and structural diversity.
PMID 16351747 · PMC1866387 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Systematic in silico search identified 5 novel human kinase genes (on chromosomes 1, 11, 13, 15, 16) and 1 pseudogene (chromosome X) absent from KinBase
-
Full-text index only
Computational comparison of two mouse draft genomes and the human golden path.
PMID 12537546 · PMC151282 · Genome biology · 2003 · 8 claims · 7 setups
The Celera and public mouse genome assemblies differ in about 10% of the mouse genome, with complementary strengths (Celera higher base-pair accuracy and overall coverage; public assembly higher quality in some finished BAC regions and freely accessible)
-
Full-text index only
Species-specific protein sequence and fold optimizations.
PMID 12487631 · PMC139977 · BMC bioinformatics · 2002 · 7 claims · 7 setups
Environmental niche is a significant factor explaining variability in amino acid composition across 100 complete genomes
-
Full-text index only
Novel gene and gene model detection using a whole genome open reading frame analysis in proteomics.
PMID 16646984 · PMC1557991 · Genome biology · 2006 · 8 claims · 4 setups
A six-frame genomic ORF translation used as an MS search database can detect novel peptides absent from standard protein databases, revealing incomplete genome annotation.
-
Full-text index only
Comprehensive genome analysis of 203 genomes provides structural genomics with new insights into protein family space.
PMID 16481312 · PMC1373602 · Nucleic acids research · 2006 · 8 claims · 7 setups
The number of protein families continues to expand steadily as more genomes are sequenced, showing no sign of saturation.
-
Full-text index only
Reconstructing the evolution of the mitochondrial ribosomal proteome.
PMID 17604309 · PMC1950548 · Nucleic acids research · 2007 · 8 claims · 6 setups
The ancestral mitoribosome was of alpha-proteobacterial descent and more than doubled its protein content in most eukaryotic lineages.
-
Full-text index only
The genome of the simian and human malaria parasite Plasmodium knowlesi.
PMID 18843368 · PMC2656934 · Nature · 2008 · 8 claims · 7 setups
The P. knowlesi (H strain) nuclear genome was sequenced and assembled: 23.5 Mb across 14 chromosomes with 5,188 predicted protein-encoding genes.
-
Full-text index only
NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins.
PMID 15608248 · PMC539979 · Nucleic acids research · 2005 · 7 claims · 5 setups
RefSeq provides a curated, non-redundant, explicitly linked collection of genomic, transcript and protein sequences spanning prokaryotes, eukaryotes and viruses.
-
Has reproduction · 53
iceDP: identifying inter-chromatin engagement via density peaks clustering algorithm.
PMID 41499218 · PMC12777978 · Briefings in bioinformatics · 2026 · 7 claims · 8 setups
iceDP uses the Density Peak clustering algorithm plus a distribution test and fold change filter to identify NHCCs from Hi-C interaction matrices while removing false positives