Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
ECgene: genome annotation for alternative splicing.
PMID 15608289 · PMC540072 · Nucleic acids research · 2005 · 8 claims · 5 setups
ECgene combines genome-based EST clustering with a graph-theoretic transcript assembly procedure to predict gene models including alternative splicing events.
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome
-
Full-text index only
Analysis of expressed sequence tags from Actinidia: applications of a cross species EST database for gene discovery in the areas of flavor, health, color and ripening.
PMID 18655731 · PMC2515324 · BMC genomics · 2008 · 7 claims · 6 setups
A collection of 132,577 ESTs from four Actinidia species was generated and clustered into 41,858 non-redundant clusters (18,070 TCs and 23,788 singletons)
-
Has reproduction · 82
Reusable building blocks in biological systems.
PMID 30958230 · PMC6303794 · Journal of the Royal Society, Interface · 2018 · 8 claims · 4 setups
Biological systems can be decomposed into phenotypic building blocks (PBBs) via k-maximally reusable decompositions (k-MRD) that maximize average reusability across conditions.
-
Full-text index only
Genome wide identification of recessive cancer genes by combinatorial mutation analysis.
PMID 18846217 · PMC2557123 · PloS one · 2008 · 7 claims · 4 setups
A combinatorial mutation analysis identified 154 candidate recessive cancer genes (pRecessiveCancer<1.5x10-7, FDR=0.39)
-
Full-text index only
A rigorous method for multigenic families' functional annotation: the peptidyl arginine deiminase (PADs) proteins family example.
PMID 16271148 · PMC1310624 · BMC genomics · 2005 · 8 claims · 5 setups
Integrating EST-based expression data with phylogenetic analysis is a valid new method for functionally annotating multigenic protein families
-
Full-text index only
The ASAP II database: analysis and comparative genomics of alternative splicing in 15 animal species.
PMID 17108355 · PMC1669709 · Nucleic acids research · 2007 · 8 claims · 4 setups
ASAP II expands human alternative splicing data ~3-fold over the previous ASAP database, to ~89,078 distinct alternative splicing relationships in 11,717 genes
-
Full-text index only
NEIBank: genomics and bioinformatics resources for vision research.
PMID 18648525 · PMC2480482 · Molecular vision · 2008 · 8 claims · 7 setups
NEIBank is an integrated genomics and bioinformatics resource for vision research, combining EST/cDNA clone data, SAGE expression data, and eye disease gene databases.
-
Full-text index only
Comprehensive annotation of bidirectional promoters identifies co-regulation among breast and ovarian cancer genes.
PMID 17447839 · PMC1853124 · PLoS computational biology · 2007 · 8 claims · 8 setups
A new algorithm using spliced ESTs (cross-validated against Known Genes and GenBank mRNA) comprehensively maps bidirectional promoters in the human genome
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.
-
Full-text index only
MutScreener: primer design tool for PCR-direct sequencing.
PMID 16845093 · PMC1538803 · Nucleic acids research · 2006 · 8 claims · 4 setups
MutScreener is a web-based application that automates PCR-direct sequencing assay design by annotating gene structure and designing PCR and sequencing primers.
-
Has reproduction · 75
Roar: detecting alternative polyadenylation with standard mRNA sequencing libraries.
PMID 27756200 · PMC5069797 · BMC bioinformatics · 2016 · 8 claims · 5 setups
Roar, a method using PRE/POST read counts around annotated APA sites to compute an m/M ratio and a ratio-of-ratios (roar) statistic, detects differential 3'UTR shortening/lengthening from standard RNA-seq libraries.
-
Full-text index only
Phylogenetic profiling of the Arabidopsis thaliana proteome: what proteins distinguish plants from other organisms?
PMID 15287975 · PMC507878 · Genome biology · 2004 · 8 claims · 6 setups
3,848 Arabidopsis proteins were identified as likely plant-specific based on phylogenetic profiling and EST confirmation in multiple plant species
-
Full-text index only
Widespread ectopic expression of olfactory receptor genes.
PMID 16716209 · PMC1508154 · BMC genomics · 2006 · 8 claims · 6 setups
OR genes show widespread, locus-dependent, heterogeneous ectopic expression across dozens of non-olfactory human and mouse tissues
-
Full-text index only
Genome-wide survey of allele-specific splicing in humans.
PMID 18518984 · PMC2427040 · BMC genomics · 2008 · 8 claims · 5 setups
A genome-wide computational scan identified 30,977 SNPs located within predicted splicing regulatory sequences (donor sites, acceptor sites, branch points, and ESEs)
-
Full-text index only
Molecular evolution of Cide family proteins: novel domain formation in early vertebrates and the subsequent divergence.
PMID 18500987 · PMC2426694 · BMC evolutionary biology · 2008 · 8 claims · 5 setups
Sequences homologous to the CIDE-N domain/NCD show a wide phylogenetic distribution, from hydra and sea anemone to mammals, while true Cide proteins are restricted to vertebrates.
-
Full-text index only
The gene guessing game.
PMID 11025532 · PMC2448377 · Yeast (Chichester, England) · 2000 · 8 claims · 6 setups
Published methods for estimating human gene number diverge widely, from ~30,000 to over 140,000 genes.
-
Full-text index only
Polymorphix: a sequence polymorphism database.
PMID 15608242 · PMC540030 · Nucleic acids research · 2005 · 8 claims · 5 setups
Polymorphix is an ACNUC-structured database that organizes EMBL/GenBank sequences into within-species homologous sequence families using similarity and bibliographic criteria, with alignments, outgroups and phylogenetic trees provided.
-
Full-text index only
Predicting candidate genes for human deafness disorders: a bioinformatics approach.
PMID 16854223 · PMC1564145 · BMC genomics · 2006 · 8 claims · 4 setups
A bioinformatic approach combining expression databases and protein interaction data narrows ~2400 candidate genes across deafness loci to a manageable set of candidates.
-
Full-text index only
Does distance matter? Variations in alternative 3' splicing regulation.
PMID 17704130 · PMC2018619 · Nucleic acids research · 2007 · 8 claims · 7 setups
Alternative 3' splice sites can be distinguished from constitutive splice sites by a combination of sequence/conservation properties that vary depending on the distance between the splice sites.