Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
DNA sequence of human chromosome 17 and analysis of rearrangement in the human lineage.
PMID 16625196 · PMC2610434 · Nature · 2006 · 8 claims · 7 setups
A finished sequence of human chromosome 17 (78,839,971 bases, ~2.8% of the euchromatic genome) was generated.
-
Full-text index only
Large-scale identification and characterization of alternative splicing variants of human gene transcripts using 56,419 completely sequenced and manually annotated full-length cDNAs.
PMID 16914452 · PMC1557807 · Nucleic acids research · 2006 · 8 claims · 8 setups
Analysis of 56,419 full-length cDNAs identified 6877 alternative splicing genes encoding 18,297 alternative splicing variants made of 37,670 exons.
-
Full-text index only
Mitochondrial diversity within modern human populations.
PMID 17439969 · PMC1888801 · Nucleic acids research · 2007 · 8 claims · 5 setups
Modern humans show extremely low divergence from the mitochondrial consensus sequence, differing on average by only 21.6 nucleotide sites
-
Full-text index only
A genome annotation-driven approach to cloning the human ORFeome.
PMID 15461802 · PMC545604 · Genome biology · 2004 · 8 claims · 8 setups
Existing human cDNA clone collections together provide only 60% coverage of full-length chromosome 22 ORFs, with the best single collection (MGC) providing 48%
-
Full-text index only
Large-scale analysis of human alternative protein isoforms: pattern classification and correlation with subcellular localization signals.
PMID 15860772 · PMC1087780 · Nucleic acids research · 2005 · 8 claims · 8 setups
Constructed a large-scale dataset of 6876 human alternative protein isoforms from 2624 genes by combining H-Invitational full-length cDNA data and SwissProt VARSPLIC entries
-
Full-text index only
Integrative annotation of 21,037 human genes validated by full-length cDNA clones.
PMID 15103394 · PMC393292 · PLoS biology · 2004 · 8 claims · 5 setups
41,118 full-length human cDNAs from six high-throughput sequencing projects were exhaustively integratively characterized
-
Full-text index only
MACSIMS: multiple alignment of complete sequences information management system.
PMID 16792820 · PMC1539025 · BMC bioinformatics · 2006 · 8 claims · 5 setups
MACSIMS is a multiple alignment-based information management system combining knowledge-based database mining with ab initio sequence predictions
-
Full-text index only
Low conservation and species-specific evolution of alternative splicing in humans and mice: comparative genomics analysis using well-annotated full-length cDNAs.
PMID 18838389 · PMC2582632 · Nucleic acids research · 2008 · 7 claims · 8 setups
Although 86% of individual human exons are conserved in the mouse genome, only a small fraction (431/20392, ~2%) of human AS variants are perfectly conserved AS variants in mice.
-
Full-text index only
Eighth major clade for hepatitis delta virus.
PMID 17073101 · PMC3294742 · Emerging infectious diseases · 2006 · 7 claims · 7 setups
Three HDV isolates (dFr644, dFr2072, dFr2736) form a monophyletic group distinct from HDV-1 through HDV-7, constituting a new eighth major clade (HDV-8) of the Deltavirus genus.
-
Full-text index only
Identification, characterization and comparative genomics of chimpanzee endogenous retroviruses.
PMID 16805923 · PMC1779541 · Genome biology · 2006 · 8 claims · 6 setups
The chimpanzee genome contains at least 42 separate families of endogenous retroviruses, 9 newly identified
-
Full-text index only
Novel dengue virus type 1 from travelers to Yap State, Micronesia.
PMID 16494770 · PMC3373118 · Emerging infectious diseases · 2006 · 8 claims · 5 setups
DENV-1 responsible for the 2004 Yap State dengue outbreak was isolated from serum of 4 Japanese travelers returning from Yap
-
Full-text index only
Large-scale analysis of Macaca fascicularis transcripts and inference of genetic divergence between M. fascicularis and M. mulatta.
PMID 18294402 · PMC2287170 · BMC genomics · 2008 · 8 claims · 6 setups
Constructed full-length-enriched cDNA libraries and determined 85,721 EST sequences and 9407 full-insert sequences from cynomolgus macaque brain (7 regions), testis, and liver
-
Full-text index only
Ensembl 2007.
PMID 17148474 · PMC1761443 · Nucleic acids research · 2007 · 8 claims · 7 setups
Ensembl added 18 new chordate genomes this year, increasing total genomes available from 15 to 33, the largest yearly increase to date.
-
Full-text index only
ASPIC: a web resource for alternative splicing prediction and transcript isoforms characterization.
PMID 16845044 · PMC1538898 · Nucleic acids research · 2006 · 8 claims · 2 setups
The ASPIC algorithm, using an optimization procedure that minimizes splice site predictions and transcript isoforms from multiple EST-genome alignments, outperforms other similar AS-prediction tools in sensitivity and selectivity
-
Has reproduction · 80
Progressive transformation of the HIV-1 reservoir cell profile over two decades of antiviral therapy.
PMID 36596305 · PMC9839361 · Cell host & microbe · 2023 · 8 claims · 7 setups
After long-term ART, intact HIV-1 proviruses are predominantly integrated in heterochromatin locations, most prominently centromeric satellite/micro-satellite DNA.
-
Has reproduction · 78
QuasiFlow: a Nextflow pipeline for analysis of NGS-based HIV-1 drug resistance data.
PMID 36699347 · PMC9722223 · Bioinformatics advances · 2022 · 6 claims · 8 setups
QuasiFlow is a Nextflow pipeline that runs entirely locally via command-line tools and a local HIVdb database copy to analyze NGS-based HIV-1 drug resistance testing data.
-
Full-text index only
Genetic variation of SARS coronavirus in Beijing Hospital.
PMID 15200810 · PMC3323231 · Emerging infectious diseases · 2004 · 8 claims · 4 setups
113 sequence variations at 9 recurrent variant sites were identified in 29 full-length S-gene sequences compared to the BJ01 reference strain
-
Has reproduction · 86
RNASEQR--a streamlined and accurate RNA-seq sequence analysis program.
PMID 22199257 · PMC3315322 · Nucleic acids research · 2012 · 8 claims · 7 setups
RNASEQR is a new RNA-seq mapper/aligner that combines a BWT-based (Bowtie) transcriptomic/genomic alignment with hash-based BLAT local alignment in three sequential steps: transcriptome mapping, novel exon detection, and anchor-and-align novel splice junction identification.
-
Full-text index only
Comparative genomics of the syndecans defines an ancestral genomic context associated with matrilins in vertebrates.
PMID 16620374 · PMC1464127 · BMC genomics · 2006 · 8 claims · 6 setups
Syndecan-encoding sequences are present in Cnidaria and throughout the Bilateria, showing deep conservation of the family.
-
Full-text index only
L1Base: from functional annotation to prediction of active LINE-1 elements.
PMID 15608246 · PMC539998 · Nucleic acids research · 2005 · 7 claims · 6 setups
L1Base is a database of putatively active LINE-1 insertions in human, mouse and rat genomes, containing FLI-L1s (intact in both ORFs), ORF2-L1s (intact ORF2, disrupted ORF1), and FLnI-L1s (full-length, >6000 bp, non-intact)