Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
SARS-CoV genome polymorphism: a bioinformatics study.
PMID 16144519 · PMC5172477 · Genomics, proteomics & bioinformatics · 2005 · 8 claims · 6 setups
SARS-CoV isolates can be classified into groups/subgroups based on the number and distribution of SNVs and INDELs relative to a 'profile' sequence, and this classification aligns with phylogenetic tree relationships and epidemiological spread.
-
Full-text index only
Molecular evolution and multilocus sequence typing of 145 strains of SARS-CoV.
PMID 16112670 · PMC7118731 · FEBS letters · 2005 · 8 claims · 7 setups
145 SARS-CoV genomes can be divided into three groups: animal-origin viruses, first-epidemic clinical viruses, and GD03T0013
-
Full-text index only
SVC: structured visualization of evolutionary sequence conservation.
PMID 15991338 · PMC1160265 · Nucleic acids research · 2005 · 7 claims · 5 setups
SVC aligns protein-coding sequences of orthologous gene pairs and maps them back onto their encoding exons/introns to generate a scaffold of conserved gene structure.
-
Full-text index only
Structural organization and interactions of transmembrane domains in tetraspanin proteins.
PMID 15985154 · PMC1190194 · BMC structural biology · 2005 · 8 claims · 5 setups
TM1, TM2 and TM3 of human tetraspanins display a distinct heptad repeat motif (abcdefg)n, while TM4 lacks this motif.
-
Full-text index only
Dcode.org anthology of comparative genomic tools.
PMID 15980535 · PMC1160116 · Nucleic acids research · 2005 · 8 claims · 7 setups
The dcode.org suite (zPicture, Mulan, eShadow, rVista 2.0, multiTF, Creme 2.0, ECR Browser) provides integrated tools for comparative genomic analysis and non-coding regulatory element discovery.
-
Full-text index only
Empirical codon substitution matrix.
PMID 15927081 · PMC1173088 · BMC bioinformatics · 2005 · 8 claims · 5 setups
The authors present the first empirical codon substitution matrix built entirely from alignments of vertebrate coding DNA sequences.
-
Full-text index only
Ab initio identification of putative human transcription factor binding sites by comparative genomics.
PMID 15865625 · PMC1097714 · BMC bioinformatics · 2005 · 8 claims · 5 setups
An integrated algorithm combining human-mouse genomic comparison, motif overrepresentation, and coregulation filters (GO annotation and microarray coexpression) can identify candidate transcription factor binding sites genome-wide
-
Full-text index only
Large-scale analysis of human alternative protein isoforms: pattern classification and correlation with subcellular localization signals.
PMID 15860772 · PMC1087780 · Nucleic acids research · 2005 · 8 claims · 8 setups
Constructed a large-scale dataset of 6876 human alternative protein isoforms from 2624 genes by combining H-Invitational full-length cDNA data and SwissProt VARSPLIC entries
-
Full-text index only
Phylogenomic approaches to common problems encountered in the analysis of low copy repeats: the sulfotransferase 1A gene family example.
PMID 15752422 · PMC555591 · BMC evolutionary biology · 2005 · 8 claims · 8 setups
A previously unidentified fourth human SULT1A gene (SULT1A4) exists on chromosome 16 and is transcriptionally active
-
Full-text index only
Evolutionarily conserved human targets of adenosine to inosine RNA editing.
PMID 15731336 · PMC549564 · Nucleic acids research · 2005 · 8 claims · 6 setups
Identified four novel human ADAR editing substrates causing amino acid changes: FLNA, BLCAP, CYFIP2 and IGFBP7
-
Full-text index only
Integrating alternative splicing detection into gene prediction.
PMID 15705189 · PMC550657 · BMC bioinformatics · 2005 · 8 claims · 4 setups
An integrative intrinsic/extrinsic method was implemented in the gene finder EuGÈNE (as EuGÈNE-M) to detect AS evidence from aligned transcripts and generate alternative optimal gene predictions consistent with each detected AS event.
-
Full-text index only
Linking disease-associated genes to regulatory networks via promoter organization.
PMID 15701758 · PMC549397 · Nucleic acids research · 2005 · 8 claims · 7 setups
Pairs of TFBSs conserved both vertically (orthologous genes) and horizontally (co-regulated genes) can serve as seeds to build promoter models representing potential co-regulation networks
-
Full-text index only
ECgene: genome annotation for alternative splicing.
PMID 15608289 · PMC540072 · Nucleic acids research · 2005 · 8 claims · 5 setups
ECgene combines genome-based EST clustering with a graph-theoretic transcript assembly procedure to predict gene models including alternative splicing events.
-
Full-text index only
DG-CST (Disease Gene Conserved Sequence Tags), a database of human-mouse conserved elements associated to disease genes.
PMID 15608249 · PMC539965 · Nucleic acids research · 2005 · 5 claims · 8 setups
Comparative human-mouse genome analysis identifies conserved sequence tags (CSTs, >=70% identity over >=100bp) that frequently correspond to non-coding elements with putative regulatory or structural roles
-
Full-text index only
Improvements to GALA and dbERGE II: databases featuring genomic sequence alignment, annotation and experimental results.
PMID 15608239 · PMC539999 · Nucleic acids research · 2005 · 8 claims · 8 setups
GALA is now a set of interlinked relational databases covering five vertebrate species: human, chimpanzee, mouse, rat and chicken.
-
Full-text index only
GenomeTrafac: a whole genome resource for the detection of transcription factor binding site clusters associated with conventional and microRNA encoding genes conserved between mouse and human gene orthologs.
PMID 17178752 · PMC1781107 · Nucleic acids research · 2007 · 8 claims · 5 setups
GenomeTrafac is a web-accessible database enabling genome-wide detection of conserved cis-element clusters in human-mouse gene orthologs, covering both conventional and microRNA genes
-
Full-text index only
Use of modified U1 snRNAs to inhibit HIV-1 replication.
PMID 17158512 · PMC1802557 · Nucleic acids research · 2007 · 7 claims · 6 setups
U1 snRNAs complementary to 5 of 15 targeted conserved regions in the HIV-1 terminal exon significantly suppress HIV-1 protein expression and viral replication, coincident with loss of viral RNA
-
Full-text index only
Phylogenetic analysis of RhoGAP domain-containing proteins.
PMID 17127216 · PMC5054073 · Genomics, proteomics & bioinformatics · 2006 · 7 claims · 6 setups
RhoGAP domain-containing proteins, sharing the conserved arginine residue, form a monophyletic group with a common ancestor.
-
Full-text index only
Genome assembly comparison identifies structural variants in the human genome.
PMID 17115057 · PMC2674632 · Nature genetics · 2006 · 7 claims · 7 setups
Genome assembly comparison is a robust approach for identifying all classes of genetic variation, with no lower size limit.
-
Full-text index only
Systematic identification of pseudogenes through whole genome expression evidence profiling.
PMID 16945953 · PMC1636364 · Nucleic acids research · 2006 · 8 claims · 8 setups
Developed a novel bioinformatics method that identifies pseudogenes by profiling whole-genome transcript and protein expression evidence