Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
SNPmasker: automatic masking of SNPs and repeats across eukaryotic genomes.
PMID 16845091 · PMC1538889 · Nucleic acids research · 2006 · 8 claims · 4 setups
SNPmasker is a web service combining SNP masking and repeat masking, supporting both coordinate-defined and homology-search-defined input regions, a combination not offered by prior tools
-
Full-text index only
Wiggle-predicting functionally flexible regions from primary sequence.
PMID 16839194 · PMC1500818 · PLoS computational biology · 2006 · 7 claims · 6 setups
A GNM-derived, correlation-weighted 'FF score' can objectively define functionally flexible regions (FFRs) that match experimentally confirmed flexible/functional regions (hinges, recognition loops, catalytic loops).
-
Full-text index only
MACSIMS: multiple alignment of complete sequences information management system.
PMID 16792820 · PMC1539025 · BMC bioinformatics · 2006 · 8 claims · 5 setups
MACSIMS is a multiple alignment-based information management system combining knowledge-based database mining with ab initio sequence predictions
-
Full-text index only
Protein ranking by semi-supervised network propagation.
PMID 16723003 · PMC1810311 · BMC bioinformatics · 2006 · 8 claims · 5 setups
RankProp, a diffusion-based network propagation algorithm on a PSI-BLAST-derived protein similarity network, significantly outperforms local search methods (BLAST/PSI-BLAST) at detecting remote homologs.
-
Has reproduction · 68
LaSSO, a strategy for genome-wide mapping of intronic lariats and branch points using RNA-seq.
PMID 24709818 · PMC4079972 · Genome research · 2014 · 8 claims · 8 setups
LaSSO (Lariat Sequence Site Origin) identifies intronic lariat reads and pinpoints branch points genome-wide from RNA-seq data by considering every intronic base as a potential branch point and including all possible exon-skipping lariats.
-
Full-text index only
A clustering property of highly-degenerate transcription factor binding sites in the mammalian genome.
PMID 16670430 · PMC1456330 · Nucleic acids research · 2006 · 8 claims · 7 setups
Highly-degenerate RE1 sites are significantly enriched in promoters of validated and putative REST target genes compared to control promoters
-
Full-text index only
Comparative genomics of the syndecans defines an ancestral genomic context associated with matrilins in vertebrates.
PMID 16620374 · PMC1464127 · BMC genomics · 2006 · 8 claims · 6 setups
Syndecan-encoding sequences are present in Cnidaria and throughout the Bilateria, showing deep conservation of the family.
-
Full-text index only
Anopheles gambiae genome reannotation through synthesis of ab initio and comparative gene prediction algorithms.
PMID 16569258 · PMC1557760 · Genome biology · 2006 · 8 claims · 7 setups
An exon-gene-union (EGU) algorithm followed by an open-reading-frame-selection algorithm can synthesize ab initio (GENSCAN, GeneMark, SNAP) and comparative (Ensembl/Genewise) predictions into a single, more complete CDS set
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.
-
Full-text index only
Identification of polymorphisms and balancing selection in the male infertility candidate gene, ornithine decarboxylase antizyme 3.
PMID 16542438 · PMC1526716 · BMC medical genetics · 2006 · 8 claims · 6 setups
Mutations in the OAZ3 gene are not a common cause of male infertility
-
Full-text index only
Computational approaches for predicting the biological effect of p53 missense mutations: a comparison of three sequence analysis based methods.
PMID 16522644 · PMC1390679 · Nucleic acids research · 2006 · 7 claims · 6 setups
Align-GVGD predicts loss of transactivation activity with high specificity (~88%) but lower sensitivity (67.9-71.2%) for neutral mutants
-
Full-text index only
Comprehensive genome analysis of 203 genomes provides structural genomics with new insights into protein family space.
PMID 16481312 · PMC1373602 · Nucleic acids research · 2006 · 8 claims · 7 setups
The number of protein families continues to expand steadily as more genomes are sequenced, showing no sign of saturation.
-
Full-text index only
Gene losses during human origins.
PMID 16464126 · PMC1361800 · PLoS biology · 2006 · 7 claims · 7 setups
A comparative genomic screen identified 67 new human-specific nonprocessed pseudogenes, bringing the total (with 13 from prior literature) to 80 human-specific pseudogenes.
-
Full-text index only
Comparative analysis of cancer genes in the human and chimpanzee genomes.
PMID 16438707 · PMC1382208 · BMC genomics · 2006 · 7 claims · 6 setups
All 333 examined human cancer genes have intact, highly conserved orthologs in the chimpanzee genome (99.38% protein identity).
-
Full-text index only
Genome-wide identification of human functional DNA using a neutral indel model.
PMID 16410828 · PMC1326222 · PLoS computational biology · 2006 · 8 claims · 8 setups
A neutral indel model predicting a geometric distribution of intergap segment (IGS) lengths fits human-mouse ancestral repeat (AR) alignment data excellently
-
Full-text index only
Diversity of preferred nucleotide sequences around the translation initiation codon in eukaryote genomes.
PMID 18086709 · PMC2241899 · Nucleic acids research · 2008 · 8 claims · 5 setups
Preferred nucleotide sequences around the initiation codon are diverse among eukaryote species, but differences roughly reflect evolutionary relationships between species
-
Full-text index only
Inconsistencies in Neanderthal genomic DNA sequences.
PMID 17937503 · PMC2014787 · PLoS genetics · 2007 · 8 claims · 6 setups
The Noonan et al. and Green et al. Neanderthal nuclear DNA datasets yield mutually inconsistent estimates of population split time and Neanderthal admixture proportion when analyzed with the same method
-
Full-text index only
Molecular analysis of a leprosy immunotherapeutic bacillus provides insights into Mycobacterium evolution.
PMID 17912347 · PMC1989137 · PloS one · 2007 · 8 claims · 8 setups
MIP is the evolutionary predecessor/ancestor of the pathogenic Mycobacterium avium intracellulare complex (MAIC), having retained a free-living lifestyle rather than undergoing parasitic reductive genome evolution
-
Full-text index only
Dyneins across eukaryotes: a comparative genomic analysis.
PMID 17897317 · PMC2239267 · Traffic (Copenhagen, Denmark) · 2007 · 8 claims · 6 setups
Phylogenetic inference identified nine DHC families (two cytoplasmic, seven axonemal) and six IC families (one cytoplasmic)