Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome-wide analysis of the human Alu Yb-lineage.
PMID 15588477 · PMC3525081 · Human genomics · 2004 · 8 claims · 6 setups
1,733 Alu Yb-lineage elements are present on human autosomal chromosomes
-
Full-text index only
Genomic rearrangements by LINE-1 insertion-mediated deletion in the human and chimpanzee lineages.
PMID 16034026 · PMC1179734 · Nucleic acids research · 2005 · 8 claims · 6 setups
L1 insertions are directly responsible for genomic deletions (L1IMDs) confirmed in both human and chimpanzee genomes
-
Full-text index only
The European Bioinformatics Institute's data resources: towards systems biology.
PMID 15608238 · PMC539980 · Nucleic acids research · 2005 · 8 claims · 5 setups
Since 2003 the EBI has launched new databases covering protein-protein interactions (IntAct), pathways (Reactome) and small molecules (ChEBI)
-
Full-text index only
Steps toward broad-spectrum therapeutics: discovering virulence-associated genes present in diverse human pathogens.
PMID 19874620 · PMC2774872 · BMC genomics · 2009 · 8 claims · 8 setups
Phylogenetic profiling of protein clusters across pathogen and non-pathogen genomes can identify candidate generic virulence factors
-
Full-text index only
Target SNP selection in complex disease association studies.
PMID 15248903 · PMC487897 · BMC bioinformatics · 2004 · 7 claims · 3 setups
A computational pipeline can retrieve gene sequence, collect SNP variation data, and annotate SNPs falling in functional motifs (promoter, exon-intron structure, AU-rich elements, TF binding sites, splice sites) with expression in target tissue
-
Full-text index only
Non-EST based prediction of exon skipping and intron retention events using Pfam information.
PMID 16204458 · PMC1243800 · Nucleic acids research · 2005 · 7 claims · 5 setups
A novel ab initio method predicts exon skipping and intron retention events using only Pfam domain annotation, via a Viterbi-like dynamic programming algorithm applied to the Pfam alignment.
-
Full-text index only
The Princeton Protein Orthology Database (P-POD): a comparative genomics analysis tool for biologists.
PMID 17712414 · PMC1942082 · PloS one · 2007 · 8 claims · 5 setups
P-POD is the first comparative genomics database to combine results from multiple computational ortholog/homolog prediction methods with manually curated literature-derived experimental evidence of functional conservation.
-
Full-text index only
GeneSeer: a sage for gene names and genomic resources.
PMID 16176584 · PMC1266031 · BMC genomics · 2005 · 7 claims · 4 setups
GeneSeer aggregates gene name synonyms from GenBank, FlyBase, ExPASy, HUGO, ENSEMBL, UCSC and Gene Ontology into a name-translation database that maps any familiar name to a reference (SOFAR) identifier.
-
Full-text index only
TassDB: a database of alternative tandem splice sites.
PMID 17142241 · PMC1669710 · Nucleic acids research · 2007 · 7 claims · 3 setups
TassDB is a relational database storing GYNGYN donor and NAGNAG acceptor tandem splice sites across eight species
-
Full-text index only
The functional importance of disease-associated mutation.
PMID 12220483 · PMC128831 · BMC bioinformatics · 2002 · 6 claims · 1 setups
Disease-associated mutations occur in conserved regions of genes and can be used to identify likely disease-causing mutations
-
Has reproduction · 95
MetaMap: an atlas of metatranscriptomic reads in human disease-related RNA-seq data.
PMID 29901703 · PMC6025204 · GigaScience · 2018 · 6 claims · 7 setups
A two-step 'omni' RNA-seq pipeline (MetaMap) combining STAR human alignment with CLARK-S metagenomic classification can quantify archaeal, bacterial, and viral reads from the non-human read fraction of human RNA-seq data
-
Full-text index only
An integrative genomic and epigenomic approach for the study of transcriptional regulation.
PMID 18365023 · PMC2266992 · PloS one · 2008 · 8 claims · 7 setups
Integrative analysis combining gene expression, DNA methylation, and H3K9 acetylation data reveals hundreds of additional differentially expressed genes missed by gene expression arrays alone
-
Full-text index only
Pegasys: software for executing and integrating analyses of biological sequences.
PMID 15096276 · PMC406494 · BMC bioinformatics · 2004 · 8 claims · 7 setups
Pegasys is a flexible, modular, customizable software system for executing and integrating heterogeneous biological sequence analysis tools
-
Full-text index only
Computer identification of snoRNA genes using a Mammalian Orthologous Intron Database.
PMID 16093549 · PMC1184218 · Nucleic acids research · 2005 · 8 claims · 5 setups
Created the Mammalian Orthologous Intron Database (MOID) containing orthologous introns of human, mouse and rat identified via conserved reading-frame position
-
Full-text index only
Similarities and differences in genome-wide expression data of six organisms.
PMID 14737187 · PMC300882 · PLoS biology · 2004 · 8 claims · 8 setups
Coexpression of functionally related genes is frequently conserved across evolutionarily distant organisms
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Has reproduction · 38
RNA-Seq transcriptome profiling of upland cotton (Gossypium hirsutum L.) root tissue under water-deficit stress.
PMID 24324815 · PMC3855774 · PloS one · 2013 · 8 claims · 8 setups
A total of 1,530 transcripts were differentially expressed between well-watered and water-deficit stressed field-grown upland cotton root tissues (913 up-regulated, 617 down-regulated).