Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Comparative gene finding in chicken indicates that we are closing in on the set of multi-exonic widely expressed human genes.
PMID 15809229 · PMC1074396 · Nucleic acids research · 2005 · 8 claims · 6 setups
Comparative gene finding (SGP2) between human and chicken, followed by RT-PCR verification, adds at most ~0.2% new genes to the multi-exonic human gene catalog
-
Full-text index only
GeneAlign: a coding exon prediction tool based on phylogenetical comparisons.
PMID 16845010 · PMC1538901 · Nucleic acids research · 2006 · 8 claims · 5 setups
GeneAlign predicts coding exons by using signal detection (GeneSplicer/WMM) combined with CORAL, a heuristic linear-time alignment tool, to align candidate signal-flanked regions against annotated exons of a homologous organism's genes
-
Full-text index only
Automatic annotation of eukaryotic genes, pseudogenes and promoters.
PMID 16925832 · PMC1810547 · Genome biology · 2006 · 8 claims · 6 setups
Fgenesh++ gene prediction pipeline identifies 91% of coding nucleotides with 90% specificity
-
Full-text index only
How to find soluble proteins: a comprehensive analysis of alpha/beta hydrolases for recombinant expression in E. coli.
PMID 15804363 · PMC1079826 · BMC genomics · 2005 · 7 claims · 7 setups
Predicted solubility in E. coli (via CV-CV') depends on hydrolase size, phylogenetic origin, homologous family, and superfamily
-
Full-text index only
A survey of integral alpha-helical membrane proteins.
PMID 19760129 · PMC2780624 · Journal of structural and functional genomics · 2009 · 8 claims · 8 setups
An automated annotation pipeline defines the integral membrane genome and family associations for 21,379 proteins from 34 genomes, most belonging to 598 Pfam-derived membrane protein families.
-
Full-text index only
Mutation screening of HSF4 in 150 age-related cataract patients.
PMID 18941546 · PMC2569895 · Molecular vision · 2008 · 8 claims · 4 setups
Five new HSF4 sequence variants (c.1020-25G>A, c.1078A>G, c.1223C>T, c.1256+25C>T, c.1286C>T) were found in age-related cataract patients but not in 220 controls.
-
Full-text index only
Predicting the phenotypic effects of non-synonymous single nucleotide polymorphisms based on support vector machines.
PMID 18005451 · PMC2216041 · BMC bioinformatics · 2007 · 8 claims · 5 setups
Parepro, an SVM-based method integrating three attribute sets (RD, MI, IE) derived from evolutionary and residue-property information, predicts whether an nsSNP is deleterious or neutral.
-
Full-text index only
A space-efficient and accurate method for mapping and aligning cDNA sequences onto genomic sequence.
PMID 18344523 · PMC2377433 · Nucleic acids research · 2008 · 7 claims · 6 setups
Spaln maps and aligns large cDNA sequence sets onto whole mammalian genomes using substantially less memory than comparable existing tools
-
Full-text index only
Functional redundancy of exon 12 of BRCA2 revealed by a comprehensive analysis of the c.6853A>G (p.I2285V) variant.
PMID 19795481 · PMC3501199 · Human mutation · 2009 · 7 claims · 8 setups
BRCA2 c.6853A>G (p.I2285V) co-occurs in trans with the deleterious founder mutation c.5946delT, supporting classification as a neutral variant
-
Full-text index only
SelenoDB 1.0 : a database of selenoprotein genes, proteins and SECIS elements.
PMID 18174224 · PMC2238826 · Nucleic acids research · 2008 · 6 claims · 5 setups
Standard genome annotation pipelines misannotate selenoprotein genes because they rely on UGA as a universal stop codon, failing to recognize its dual role as the selenocysteine-recoding codon.
-
Full-text index only
PhylomeDB: a database for genome-wide collections of gene phylogenies.
PMID 17962297 · PMC2238872 · Nucleic acids research · 2008 · 7 claims · 6 setups
PhylomeDB is a publicly accessible database storing complete, genome-wide collections of gene phylogenies (phylomes).
-
Has reproduction · 95
Identification and Characterization of Small Noncoding RNAs in Genome Sequences of the Edible Fungus Pleurotus ostreatus.
PMID 27703969 · PMC5040776 · BioMed research international · 2016 · 7 claims · 8 setups
Genome-scale identification detected 254 small noncoding RNAs (snRNAs, snoRNAs, tRNAs, miRNAs, and other Rfam-classified sncRNAs) in the P. ostreatus CCEF00389 genome assembly
-
Full-text index only
Inventory and analysis of the protein subunits of the ribonucleases P and MRP provides further evidence of homology between the yeast and human enzymes.
PMID 16998185 · PMC1636426 · Nucleic acids research · 2006 · 8 claims · 6 setups
Fungal Pop8 is evolutionarily related to the Rpp14/Pop5 protein family, suggesting Pop8 is the fungal orthologue of Rpp14
-
Full-text index only
Evolutionary genomics reveals lineage-specific gene loss and rapid evolution of a sperm-specific ion channel complex: CatSpers and CatSperbeta.
PMID 18974790 · PMC2572835 · PloS one · 2008 · 8 claims · 6 setups
The CatSper channel complex (four CatSpers plus CatSperβ) originated as early as primitive metazoans such as the Cnidarian Nematostella vectensis
-
Has reproduction · 24
MiGPC: a comprehensive catalog of enzybiotics from environmental metagenomes.
PMID 41888223 · PMC13172421 · Scientific reports · 2026 · 8 claims · 8 setups
MiGPC is the first genome-resolved metagenomic gene and protein catalog specifically targeted to enzybiotics
-
Has reproduction · 92
Telomere-to-telomere reference genome for Panax ginseng highlights the evolution of saponin biosynthesis.
PMID 38883331 · PMC11179851 · Horticulture research · 2024 · 8 claims · 8 setups
A telomere-to-telomere reference genome of P. ginseng was assembled (3.45 Gb, 24 chromosomes, 77266 protein-coding genes)
-
Full-text index only
Evola: Ortholog database of all human genes in H-InvDB with manual curation of phylogenetic trees.
PMID 17982176 · PMC2238928 · Nucleic acids research · 2008 · 6 claims · 7 setups
Evola combines genome synteny-based computational ortholog detection with manual curation of phylogenetic trees by experts to yield more reliable orthologs than automated pairwise methods
-
Full-text index only
Mutation analysis in primary immunodeficiency diseases: case studies.
PMID 19841577 · PMC2774237 · Current opinion in allergy and clinical immunology · 2009 · 8 claims · 8 setups
Genomic DNA Sanger sequencing is the standard first-line approach for identifying PIDD-causing mutations but has limitations that can yield false-negative or false-positive results