Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genomic organization and recombinational unit duplication-driven evolution of ovine and bovine T cell receptor gamma loci.
PMID 18282289 · PMC2270265 · BMC genomics · 2008 · 8 claims · 6 setups
The sheep TRG1 and TRG2 loci evolved through a series of duplication events involving either entire V-J-J-C recombinational cassettes or single V genes
-
Full-text index only
Large-scale discovery of insertion hotspots and preferential integration sites of human transposed elements.
PMID 20008508 · PMC2836564 · Nucleic acids research · 2010 · 8 claims · 6 setups
Most TEs insert within specific 'hotspots' along the targeted TE rather than uniformly.
-
Full-text index only
Empirical codon substitution matrix.
PMID 15927081 · PMC1173088 · BMC bioinformatics · 2005 · 8 claims · 5 setups
The authors present the first empirical codon substitution matrix built entirely from alignments of vertebrate coding DNA sequences.
-
Full-text index only
Developments in CORG: a gene-centric comparative genomics resource.
PMID 17135197 · PMC1751536 · Nucleic acids research · 2007 · 7 claims · 4 setups
CORG provides pairwise and multiple sequence alignments of upstream promoter regions and whole gene loci across 10 vertebrate species.
-
Full-text index only
Flanking p10 contribution and sequence bias in matrix based epitope prediction: revisiting the assumption of independent binding pockets.
PMID 18925947 · PMC2600787 · BMC structural biology · 2008 · 8 claims · 3 setups
The extended matrix PP10 (built from a proline-containing peptide library) shows significant improvement in binding prediction over the original nine-residue matrix P9
-
Full-text index only
GeneSeer: a sage for gene names and genomic resources.
PMID 16176584 · PMC1266031 · BMC genomics · 2005 · 7 claims · 4 setups
GeneSeer aggregates gene name synonyms from GenBank, FlyBase, ExPASy, HUGO, ENSEMBL, UCSC and Gene Ontology into a name-translation database that maps any familiar name to a reference (SOFAR) identifier.
-
Full-text index only
Reverse polarization in amino acid and nucleotide substitution patterns between human-mouse orthologs of two compositional extrema.
PMID 17895298 · PMC2533592 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2007 · 8 claims · 7 setups
Nucleotide and amino acid substitution trends between human-mouse orthologs are highly asymmetric and polarized in opposite directions for high-GC versus low-GC gene groups.
-
Full-text index only
Characterisation of the genetic diversity of Brucella by multilocus sequencing.
PMID 17448232 · PMC1877810 · BMC microbiology · 2007 · 8 claims · 5 setups
Brucella isolates show low overall genetic diversity (1.5%) across nine sequenced loci, confirming the genus is genetically conserved.
-
Full-text index only
Grammar-based distance in progressive multiple sequence alignment.
PMID 18616828 · PMC2478692 · BMC bioinformatics · 2008 · 7 claims · 3 setups
A grammar-based (LZ complexity) distance metric can be used to determine the order in which sequences are progressively pairwise aligned
-
Full-text index only
Molecular epidemiology of O139 Vibrio cholerae: mutation, lateral gene transfer, and founder flush.
PMID 12890320 · PMC3023423 · Emerging infectious diseases · 2003 · 8 claims · 4 setups
Lateral gene transfer (LGT) produced roughly three times as many nucleotide changes as point mutation among the 96 O139 isolates.
-
Full-text index only
Biased exon/intron distribution of cryptic and de novo 3' splice sites.
PMID 16141195 · PMC1197134 · Nucleic acids research · 2005 · 7 claims · 5 setups
Cryptic 3'ss (from 3'YAG consensus mutations) are significantly more frequent in exons than in introns
-
Full-text index only
Coxiella burnetii genotyping.
PMID 16102309 · PMC3320512 · Emerging infectious diseases · 2005 · 8 claims · 5 setups
Multispacer sequence typing (MST) is the first reliable method for typing Coxiella burnetii isolates
-
Full-text index only
Ab initio identification of putative human transcription factor binding sites by comparative genomics.
PMID 15865625 · PMC1097714 · BMC bioinformatics · 2005 · 8 claims · 5 setups
An integrated algorithm combining human-mouse genomic comparison, motif overrepresentation, and coregulation filters (GO annotation and microarray coexpression) can identify candidate transcription factor binding sites genome-wide
-
Full-text index only
FeatureScan: revealing property-dependent similarity of nucleotide sequences.
PMID 16845077 · PMC1538849 · Nucleic acids research · 2006 · 6 claims · 5 setups
FeatureScan transforms nucleotide sequences into numerical signals of physico-chemical/conformational properties and compares them via a convolution/correlation (Fourier transform) method rather than comparing letters
-
Full-text index only
Comparative genomics of the syndecans defines an ancestral genomic context associated with matrilins in vertebrates.
PMID 16620374 · PMC1464127 · BMC genomics · 2006 · 8 claims · 6 setups
Syndecan-encoding sequences are present in Cnidaria and throughout the Bilateria, showing deep conservation of the family.
-
Full-text index only
Computational approaches for predicting the biological effect of p53 missense mutations: a comparison of three sequence analysis based methods.
PMID 16522644 · PMC1390679 · Nucleic acids research · 2006 · 7 claims · 6 setups
Align-GVGD predicts loss of transactivation activity with high specificity (~88%) but lower sensitivity (67.9-71.2%) for neutral mutants
-
Full-text index only
EPGD: a comprehensive web resource for integrating and displaying eukaryotic paralog/paralogon information.
PMID 17984073 · PMC2238967 · Nucleic acids research · 2008 · 8 claims · 8 setups
EPGD is a gene-centered, internet-accessible database integrating paralog family and paralogon information for 26 eukaryotic genomes.
-
Has reproduction · 98
Sequence-based pangenomic core detection.
PMID 35663029 · PMC9160775 · iScience · 2022 · 7 claims · 3 setups
Sequence-based pangenomic core detection can be performed directly on unannotated genome sequences using a colored de Bruijn graph, avoiding bias from error-prone gene annotations
-
Full-text index only
Having a BLAST with bioinformatics (and avoiding BLASTphemy).
PMID 11597340 · PMC138974 · Genome biology · 2001 · 8 claims · 4 setups
BLAST is the most widely used tool for searching biological sequences for regions of local similarity
-
Full-text index only
Improvements to GALA and dbERGE II: databases featuring genomic sequence alignment, annotation and experimental results.
PMID 15608239 · PMC539999 · Nucleic acids research · 2005 · 8 claims · 8 setups
GALA is now a set of interlinked relational databases covering five vertebrate species: human, chimpanzee, mouse, rat and chicken.