Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The most frequent short sequences in non-coding DNA.
PMID 19966278 · PMC2831315 · Nucleic acids research · 2010 · 8 claims · 2 setups
Short frequent sequences (9-14 bases) in non-coding DNA may play a role in maintaining chromosome structure and function
-
Full-text index only
Compensatory mutations cause excess of antagonistic epistasis in RNA secondary structure folding.
PMID 12590655 · PMC149451 · BMC evolutionary biology · 2003 · 6 claims · 1 setups
RNA secondary structure folding shows a clear prevalence of antagonistic epistasis (β < 1) among reference sequences
-
Has reproduction · 92
Draft Genome Sequences of Antimicrobial-Resistant Shigella Clinical Isolates from Pakistan.
PMID 31346012 · PMC6658682 · Microbiology resource announcements · 2019 · 6 claims · 6 setups
Draft genome sequences are reported for three multidrug-resistant Shigella clinical isolates from Pakistan (two S. flexneri and one S. sonnei).
-
Full-text index only
Gene function in the mammalian genome, courtesy of the mouse.
PMID 12537544 · PMC151280 · Genome biology · 2003 · 8 claims · 8 setups
Mosaicism of Mus musculus domesticus and Mus musculus musculus haplotypes exists across the inbred laboratory mouse genome, and genome-wide haplotype mapping can enhance positional cloning
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
MODBASE: a database of annotated comparative protein structure models and associated resources.
PMID 16381869 · PMC1347422 · Nucleic acids research · 2006 · 8 claims · 7 setups
MODBASE is a database of automatically calculated comparative protein structure models covering all UniProt sequences matchable to a known structure
-
Full-text index only
The Functional RNA Database 3.0: databases to support mining and annotation of functional RNAs.
PMID 18948287 · PMC2686472 · Nucleic acids research · 2009 · 8 claims · 5 setups
fRNAdb 3.0 is a completely rebuilt sequence database hosting a much larger collection of known/predicted non-coding RNA sequences with improved search functionality
-
Full-text index only
GenBank.
PMID 18940867 · PMC2686462 · Nucleic acids research · 2009 · 8 claims · 4 setups
GenBank is a comprehensive public database of nucleotide sequences with bibliographic and biological annotation, growing exponentially with a current doubling time of ~30 months.
-
Full-text index only
MitoVariome: a variome database of human mitochondrial DNA.
PMID 19958475 · PMC2788364 · BMC genomics · 2009 · 8 claims · 5 setups
MitoVariome is a web-based, integrated variome database for human mitochondrial DNA that unifies sequence variation, haplogroup, and disease annotation information not jointly available in prior databases (MITOMAP, mtDB, Mitome, MitoRes).
-
Full-text index only
G-compass: a web-based comparative genome browser between human and other vertebrate genomes.
PMID 19846439 · PMC2788932 · Bioinformatics (Oxford, England) · 2009 · 7 claims · 2 setups
G-compass is a web-based tool that displays two corresponding genomic regions from human and another vertebrate species simultaneously in parallel, without requiring client installation.
-
Full-text index only
mtDB: Human Mitochondrial Genome Database, a resource for population genetics and medical sciences.
PMID 16381973 · PMC1347373 · Nucleic acids research · 2006 · 8 claims · 3 setups
mtDB is a comprehensive, actively maintained database of published human mitochondrial genome sequences, providing a common resource for population genetics and medical research
-
Has reproduction · 69
Discovery and characterization of Alu repeat sequences via precise local read assembly.
PMID 26503250 · PMC4666360 · Nucleic acids research · 2015 · 7 claims · 8 setups
Combining Alu-supporting read detection (RetroSeq) with local de novo assembly (CAP3) reconstructs the full sequence of non-reference Alu insertions from Illumina paired-end WGS reads
-
Full-text index only
Epidemiologic and evolutionary relationships between Romanian and Brazilian HIV-subtype F strains.
PMID 8903171 · PMC2626880 · Emerging infectious diseases · 1995 · 7 claims · 4 setups
Romanian and Brazilian HIV-1 subtype F envelope C2-V3 sequences cluster into two related but distinct phylogenetic groups.
-
Full-text index only
Computational comparison of two mouse draft genomes and the human golden path.
PMID 12537546 · PMC151282 · Genome biology · 2003 · 8 claims · 7 setups
The Celera and public mouse genome assemblies differ in about 10% of the mouse genome, with complementary strengths (Celera higher base-pair accuracy and overall coverage; public assembly higher quality in some finished BAC regions and freely accessible)
-
Full-text index only
A statistical approach designed for finding mathematically defined repeats in shotgun data and determining the length distribution of clone-inserts.
PMID 15626332 · PMC5172250 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 6 setups
Repeats of different copy number have distinct probabilities of appearance in shotgun data, which can be modeled statistically to define recognition thresholds (MDRs) at different shotgun coverages.
-
Full-text index only
Benchmarking tools for the alignment of functional noncoding DNA.
PMID 14736341 · PMC344529 · BMC bioinformatics · 2004 · 8 claims · 4 setups
Global alignment tools (Avid, ClustalW, Lagan, Needle, DiAlign-G) typically have higher sensitivity over entire noncoding sequences and within constrained blocks than local tools
-
Full-text index only
New approaches to the analysis of palindromic sequences from the human genome: evolution and polymorphism of an intronic site at the NF1 locus.
PMID 16340004 · PMC1310899 · Nucleic acids research · 2005 · 7 claims · 8 setups
Long pure palindromes (>~200 bp) cannot be stably cloned in E.coli due to cruciform-driven instability, and no E.coli mutant fully overcomes this cloning block.
-
Full-text index only
Characterization of 954 bovine full-CDS cDNA sequences.
PMID 16305752 · PMC1314900 · BMC genomics · 2005 · 7 claims · 8 setups
954 bovine full-length insert cDNA (bFLIC) clones representing 762 distinct loci were sequenced and characterized
-
Full-text index only
Coiled-coil protein composition of 22 proteomes--differences and common themes in subcellular infrastructure and traffic control.
PMID 16288662 · PMC1322226 · BMC evolutionary biology · 2005 · 7 claims · 5 setups
Proteins with extended coiled-coil domains (>250 amino acids) are largely absent from bacterial genomes but present in archaea and eukaryotes.
-
Full-text index only
The distribution of SNPs in human gene regulatory regions.
PMID 16209714 · PMC1260019 · BMC genomics · 2005 · 8 claims · 6 setups
SNPs occur with higher density closer to the transcriptional start site within gene promoter regions than in further upstream regions