Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Human-zebrafish non-coding conserved elements act in vivo to regulate transcription.
PMID 16179648 · PMC1236720 · Nucleic acids research · 2005 · 8 claims · 4 setups
Deeply conserved human-zebrafish non-coding elements are enriched for in vivo cis-acting transcriptional regulatory activity.
-
Full-text index only
CONTRAST: a discriminative, phylogeny-free approach to multiple informant de novo gene prediction.
PMID 18096039 · PMC2246271 · Genome biology · 2007 · 8 claims · 5 setups
CONTRAST predicts exact coding region structures for 65% more human genes than the previous state-of-the-art de novo predictor (N-SCAN)
-
Full-text index only
Retropseudogenes derived from the human Ro/SS-A autoantigen-associated hY RNAs.
PMID 15817567 · PMC1074747 · Nucleic acids research · 2005 · 8 claims · 8 setups
966 pseudogenes derived from the four human Y (hY) RNAs were characterized in the human genome
-
Full-text index only
Genome-wide analyses of retrogenes derived from the human box H/ACA snoRNAs.
PMID 17175533 · PMC1802619 · Nucleic acids research · 2007 · 8 claims · 6 setups
202 novel box H/ACA RNA-related sequences were identified in the human genome
-
Full-text index only
Using several pair-wise informant sequences for de novo prediction of alternatively spliced transcripts.
PMID 16925842 · PMC1810557 · Genome biology · 2006 · 8 claims · 4 setups
MARS, an extension of the Twinscan algorithm, uses multiple pairwise informant genomes to predict human alternatively spliced transcripts de novo without expressed sequence information.
-
Full-text index only
Steps toward broad-spectrum therapeutics: discovering virulence-associated genes present in diverse human pathogens.
PMID 19874620 · PMC2774872 · BMC genomics · 2009 · 8 claims · 8 setups
Phylogenetic profiling of protein clusters across pathogen and non-pathogen genomes can identify candidate generic virulence factors
-
Full-text index only
Functional coverage of the human genome by existing structures, structural genomics targets, and homology models.
PMID 16118666 · PMC1188274 · PLoS computational biology · 2005 · 8 claims · 5 setups
Existing PDB structures provide single-domain coverage for 37% of functional classes in the human genome and complete (whole-protein) structure coverage for 25%.
-
Full-text index only
Structure SNP (StSNP): a web server for mapping and modeling nsSNPs on protein structures with linkage to metabolic pathways.
PMID 17537826 · PMC1933130 · Nucleic acids research · 2007 · 7 claims · 5 setups
StSNP integrates dbSNP, PDB, KEGG, and NCBI Entrez data into a single web server for nsSNP analysis
-
Full-text index only
Evidence for a novel gene associated with human influenza A viruses.
PMID 19917120 · PMC2780412 · Virology journal · 2009 · 8 claims · 8 setups
A 167-codon ORF (NEG8) on the negative-sense genomic strand of segment 8 is associated with early-20th-century human influenza A isolates
-
Full-text index only
Comparative genomics of Lbx loci reveals conservation of identical Lbx ohnologs in bony vertebrates.
PMID 18541024 · PMC2446394 · BMC evolutionary biology · 2008 · 8 claims · 3 setups
Extant bony vertebrates (osteichthyans) retain only Lbx1- and Lbx2-type genes; no distinct Lbx3/Lbx4 proteins exist.
-
Full-text index only
Identification and characterization of insect-specific proteins by genome data analysis.
PMID 17407609 · PMC1852559 · BMC genomics · 2007 · 8 claims · 7 setups
Comparative genome analysis across five holometabolous insects and three non-insect eukaryotes (opisthokonts) identifies 154 insect-specific orthologous groups (refined to 51 proteins) and 466 eukaryote/opisthokont-core orthologous groups
-
Full-text index only
Comparative analysis of genome tiling array data reveals many novel primate-specific functional RNAs in human.
PMID 17288572 · PMC1796608 · BMC evolutionary biology · 2007 · 8 claims · 6 setups
Widespread transcription occurs across the human genome outside known gene annotations, and the bulk of TARs represent genuine transcripts rather than experimental artifacts
-
Has reproduction · 78
A case study for large-scale human microbiome analysis using JCVI's metagenomics reports (METAREP).
PMID 22719821 · PMC3374610 · PloS one · 2012 · 8 claims · 7 setups
METAREP version 1.3.1 is an open-source, scalable tool for querying, browsing and comparing extremely large volumes of metagenomic annotations, with an extended data model, dynamic weighting, distributed searches and advanced clustering.
-
Full-text index only
Genome-wide detection of segmental duplications and potential assembly errors in the human genome sequence.
PMID 12702206 · PMC154576 · Genome biology · 2003 · 8 claims · 6 setups
Segmental duplications comprise 3.53% (107.4/3,043.1 Mb) of the June 2002 human genome assembly
-
Full-text index only
Genome assembly comparison identifies structural variants in the human genome.
PMID 17115057 · PMC2674632 · Nature genetics · 2006 · 7 claims · 7 setups
Genome assembly comparison is a robust approach for identifying all classes of genetic variation, with no lower size limit.
-
Full-text index only
Function2Gene: a gene selection tool to increase the power of genetic association studies by utilizing public databases and expert knowledge.
PMID 18631403 · PMC2500032 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Function2Gene is a set of Perl programs that queries public databases (NCBI, GeneCards, Harvester, with Uniprot/Ensembl also supported) using expert-selected keywords to rank genes by prior probability of disease association.
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs
-
Full-text index only
Identification of the REST regulon reveals extensive transposable element-mediated binding site duplication.
PMID 16899447 · PMC1557810 · Nucleic acids research · 2006 · 8 claims · 8 setups
The RE1 PSSM identifies functional RE1 binding sites with greater sensitivity and selectivity than the previously used RE1 consensus sequence
-
Full-text index only
Molecular evolution of Cide family proteins: novel domain formation in early vertebrates and the subsequent divergence.
PMID 18500987 · PMC2426694 · BMC evolutionary biology · 2008 · 8 claims · 5 setups
Sequences homologous to the CIDE-N domain/NCD show a wide phylogenetic distribution, from hydra and sea anemone to mammals, while true Cide proteins are restricted to vertebrates.
-
Has reproduction · 94
Large-Scale Phylogenomics of the Lactobacillus casei Group Highlights Taxonomic Inconsistencies and Reveals Novel Clade-Associated Features.
PMID 28845461 · PMC5566788 · mSystems · 2017 · 8 claims · 8 setups
The L. casei group resolves into three distinct clades (A, B, C) supported by phylogeny, GC content, ANI, and TETRA, and many strains are misclassified relative to their nearest type strain.