Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
CoMoDis: composite motif discovery in mammalian genomes.
PMID 17130158 · PMC1702496 · Nucleic acids research · 2007 · 7 claims · 4 setups
CoMoDis is a new bioinformatics tool that streamlines computational identification of novel regulatory modules starting from a single seed motif
-
Full-text index only
Sequence determinants of human microsatellite variability.
PMID 20015383 · PMC2806349 · BMC genomics · 2009 · 6 claims · 4 setups
Mean and maximum number of repeats across individuals are positively correlated with heterozygosity
-
Full-text index only
Local combinational variables: an approach used in DNA-binding helix-turn-helix motif prediction with sequence information.
PMID 19651875 · PMC2761287 · Nucleic acids research · 2009 · 8 claims · 7 setups
The LCV approach predicts HTH motifs with 93.29% accuracy, 93.93% sensitivity and 92.66% specificity using only primary sequence information
-
Full-text index only
Classification of real and pseudo microRNA precursors using local structure-sequence features and support vector machine.
PMID 16381612 · PMC1360673 · BMC bioinformatics · 2005 · 7 claims · 7 setups
A 32-dimensional triplet structure-sequence feature vector combined with SVM (triplet-SVM) can distinguish real human pre-miRNAs from pseudo pre-miRNA hairpins with ~90% accuracy.
-
Full-text index only
ORFer--retrieval of protein sequences and open reading frames from GenBank and storage into relational databases or text files.
PMID 12493080 · PMC139979 · BMC bioinformatics · 2002 · 6 claims · 6 setups
ORFer retrieves protein and nucleic acid sequences and annotations from NCBI GenBank using the XML sequence format
-
Full-text index only
CompMoby: comparative MobyDick for detection of cis-regulatory motifs.
PMID 18950538 · PMC2605473 · BMC bioinformatics · 2008 · 7 claims · 4 setups
CompMoby identifies cis-regulatory binding sites at both transcriptional and post-transcriptional levels in metazoans without prior knowledge of the trans-acting factor
-
Full-text index only
Genomic sequencing of the severe acute respiratory syndrome-coronavirus.
PMID 16916263 · PMC7121524 · Methods in molecular biology (Clifton, N.J.) · 2006 · 7 claims · 7 setups
PCR-based amplification and direct sequencing of SARS-CoV genome fragments is feasible from uncultured clinical specimens (serum, nasopharyngeal aspirate, stool), avoiding culture-derived artifacts and biohazard risk.
-
Has reproduction · 78
Chromosome-Scale Assembly of the Complete Genome Sequence of Leishmania (Mundinia) orientalis, Isolate LSCM4, Strain LV768.
PMID 34498920 · PMC8428255 · Microbiology resource announcements · 2021 · 6 claims · 8 setups
The complete genome sequence of Leishmania (Mundinia) orientalis, isolate LSCM4, strain LV768, was determined using combined short-read and long-read sequencing.
-
Full-text index only
Genome assembly comparison identifies structural variants in the human genome.
PMID 17115057 · PMC2674632 · Nature genetics · 2006 · 7 claims · 7 setups
Genome assembly comparison is a robust approach for identifying all classes of genetic variation, with no lower size limit.
-
Full-text index only
Improved mutation tagging with gene identifiers applied to membrane protein stability prediction.
PMID 19758467 · PMC2745585 · BMC bioinformatics · 2009 · 8 claims · 4 setups
MutationTagger achieves 87% F-measure for the mutation retrieval task on a benchmark dataset
-
Full-text index only
Sequence analysis of MYOC and CYP1B1 in a Chinese pedigree of juvenile glaucoma with goniodysgenesis.
PMID 19668597 · PMC2722712 · Molecular vision · 2009 · 7 claims · 4 setups
A heterozygous MYOC mutation c.1109C>T (P370L) in exon 3 cosegregates with disease, present in all 6 affected members and absent in asymptomatic members.
-
Full-text index only
Identification of "pathologs" (disease-related genes) from the RIKEN mouse cDNA dataset using human curation plus FACTS, a new biological information extraction system.
PMID 15115540 · PMC420239 · BMC genomics · 2004 · 6 claims · 3 setups
Bioinformatic sequence comparison of 60,770 RIKEN FANTOM2 mouse cDNA clones identified 2,578 sequences with 70-85% identity to known human disease genes/proteins
-
Full-text index only
HLA-A gene polymorphism defined by high-resolution sequence-based typing in 161 Northern Chinese Han people.
PMID 15629059 · PMC5172246 · Genomics, proteomics & bioinformatics · 2003 · 7 claims · 5 setups
HLA-A gene shows high polymorphism in the Northern Chinese Han population, with 74 gene types and 36 alleles detected in 161 individuals
-
Full-text index only
SNP-specific extraction of haplotype-resolved targeted genomic regions.
PMID 18611953 · PMC2528194 · Nucleic acids research · 2008 · 6 claims · 4 setups
SNP-specific extraction isolates haploid genomic DNA flanking targeted SNPs by hybridizing an allele-specific oligo, enzymatically extending it with biotinylated nucleotides, and capturing the tagged allele with streptavidin-coated magnetic particles.
-
Has reproduction · 87
A target enrichment method for gathering phylogenetic information from hundreds of loci: An example from the Compositae.
PMID 25202605 · PMC4103609 · Applications in plant sciences · 2014 · 8 claims · 8 setups
A custom sequence capture probe set (9678 baits targeting 1061 orthologous genes) was designed to enrich COS loci across the Compositae.
-
Has reproduction · 100
Intra-Host Co-Existing Strains of SARS-CoV-2 Reference Genome Uncovered by Exhaustive Computational Search.
PMID 37243151 · PMC10224212 · Viruses · 2023 · 8 claims · 7 setups
An exhaustive-search workflow can recover intra-host co-existing SARS-CoV-2 strains from the reference-genome read set (SRR11092062) that de Bruijn-graph assemblers discard.
-
Full-text index only
DNA sequence and analysis of human chromosome 9.
PMID 15164053 · PMC2734081 · Nature · 2004 · 8 claims · 8 setups
The finished euchromatic sequence of chromosome 9 comprises 109,044,351 base pairs, representing >99.6% of the region.
-
Full-text index only
Clinical and genetic analysis of Korean patients with Miyoshi myopathy: identification of three novel mutations in the DYSF gene.
PMID 16891820 · PMC2729898 · Journal of Korean medical science · 2006 · 7 claims · 7 setups
All three unrelated Korean MM patients carried compound heterozygous mutations in the DYSF gene.
-
Full-text index only
Assignment of Streptococcus agalactiae isolates to clonal complexes using a small set of single nucleotide polymorphisms.
PMID 18710585 · PMC2533671 · BMC microbiology · 2008 · 7 claims · 6 setups
A four-SNP set (glnA36, glnA429, glcK180, adhP111) identified via the Not-N algorithm plus empirical testing divides GBS into 10 groups concordant with eBURST-defined population structure.
-
Full-text index only
EGenBio: a data management system for evolutionary genomics and biodiversity.
PMID 17118150 · PMC1683573 · BMC bioinformatics · 2006 · 7 claims · 7 setups
EGenBio is a web-based system for integrated management, filtering, curation, and visualization of large-scale genomic sequences, alignments, and phylogenetic trees for evolutionary genomics and biodiversity research.