Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 67
Reducing language barriers, promoting information absorption, and communication using fanyi.
PMID 39039634 · PMC11332769 · Chinese medical journal · 2024 · 7 claims · 3 setups
The fanyi R package retrieves gene information from NCBI and translates it into multiple languages using AI-driven online translation services
-
Full-text index only
The distribution of SNPs in human gene regulatory regions.
PMID 16209714 · PMC1260019 · BMC genomics · 2005 · 8 claims · 6 setups
SNPs occur with higher density closer to the transcriptional start site within gene promoter regions than in further upstream regions
-
Full-text index only
DAVID Bioinformatics Resources: expanded annotation database and novel algorithms to better extract biology from large gene lists.
PMID 17576678 · PMC1933169 · Nucleic acids research · 2007 · 8 claims · 4 setups
The DAVID Gene Concept uses a single-linkage method to agglomerate tens of millions of gene/protein identifiers from NCBI, PIR, UniProt and other resources into unified DAVID genes.
-
Full-text index only
BLASTO: a tool for searching orthologous groups.
PMID 17483516 · PMC1933156 · Nucleic acids research · 2007 · 7 claims · 2 setups
BLASTO treats each orthologous group as a unit and outputs a ranked list of orthologous groups instead of single sequences
-
Full-text index only
Functional coverage of the human genome by existing structures, structural genomics targets, and homology models.
PMID 16118666 · PMC1188274 · PLoS computational biology · 2005 · 8 claims · 5 setups
Existing PDB structures provide single-domain coverage for 37% of functional classes in the human genome and complete (whole-protein) structure coverage for 25%.
-
Full-text index only
Single-molecule sequencing of an individual human genome.
PMID 19668243 · PMC4117198 · Nature biotechnology · 2009 · 8 claims · 7 setups
Single-molecule sequencing without cloning, amplification or ligation can sequence an individual human genome on one instrument by a single operator in four runs
-
Full-text index only
Cryptic loxP sites in mammalian genomes: genome-wide distribution and relevance for the efficiency of BAC/PAC recombineering techniques.
PMID 17284462 · PMC1865043 · Nucleic acids research · 2007 · 6 claims · 6 setups
Cryptic lox P sites occur frequently and are homogeneously distributed across the mouse genome (1.2 primary sites per megabase).
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.
-
Has reproduction · 68
Complete Genome Sequencing of Lactobacillus plantarum ZLP001, a Potential Probiotic That Enhances Intestinal Epithelial Barrier Function and Defense Against Pathogens in Pigs.
PMID 30542296 · PMC6277807 · Frontiers in physiology · 2018 · 8 claims · 8 setups
The complete genome of L. plantarum ZLP001 comprises a single 3,164,369 bp circular chromosome (GC 44.65%) plus seven plasmids (A–G), encoding 3,264 protein-coding sequences.
-
Has reproduction · 94
Large-Scale Phylogenomics of the Lactobacillus casei Group Highlights Taxonomic Inconsistencies and Reveals Novel Clade-Associated Features.
PMID 28845461 · PMC5566788 · mSystems · 2017 · 8 claims · 8 setups
The L. casei group resolves into three distinct clades (A, B, C) supported by phylogeny, GC content, ANI, and TETRA, and many strains are misclassified relative to their nearest type strain.
-
Has reproduction · 75
Graph-Based Approaches Significantly Improve the Recovery of Antibiotic Resistance Genes From Complex Metagenomic Datasets.
PMID 34690959 · PMC8528159 · Frontiers in microbiology · 2021 · 8 claims · 6 setups
GraphAMR, a Nextflow pipeline that aligns AMR profile HMMs (or AA sequences) to metagenomic assembly graphs via PathRacer, then dereplicates and annotates hits, recovers more and more complete AMR genes than contig-based or read-based methods.
-
Has reproduction · 86
RNASEQR--a streamlined and accurate RNA-seq sequence analysis program.
PMID 22199257 · PMC3315322 · Nucleic acids research · 2012 · 8 claims · 7 setups
RNASEQR is a new RNA-seq mapper/aligner that combines a BWT-based (Bowtie) transcriptomic/genomic alignment with hash-based BLAT local alignment in three sequential steps: transcriptome mapping, novel exon detection, and anchor-and-align novel splice junction identification.
-
Full-text index only
PathogenMIPer: a tool for the design of molecular inversion probes to detect multiple pathogens.
PMID 17105657 · PMC1657037 · BMC bioinformatics · 2006 · 6 claims · 5 setups
PathogenMIPer designs unique, target-specific MIP probes, assembling all probe components (target-specific sequences, barcodes, universal primers, restriction sites) into ready-to-order probes for any genome.
-
Full-text index only
Automatic annotation of eukaryotic genes, pseudogenes and promoters.
PMID 16925832 · PMC1810547 · Genome biology · 2006 · 8 claims · 6 setups
Fgenesh++ gene prediction pipeline identifies 91% of coding nucleotides with 90% specificity
-
Full-text index only
Strategies for folding of affinity tagged proteins using GroEL and osmolytes.
PMID 19082872 · PMC3693453 · Journal of structural and functional genomics · 2009 · 8 claims · 8 setups
GroEL/osmolyte mixtures can be used to refold difficult-to-fold chimeric affinity-tagged proteins by exploiting intrinsic chaperonin binding.
-
Has reproduction · 30
taxize: taxonomic search and retrieval in R.
PMID 24555091 · PMC3901538 · F1000Research · 2013 · 8 claims · 8 setups
taxize is an open-source R package (on CRAN) giving simple programmatic access to taxonomic data from 13 web data sources.
-
Full-text index only
In vitro and in silico analysis reveals an efficient algorithm to predict the splicing consequences of mutations at the 5' splice sites.
PMID 17726045 · PMC2094079 · Nucleic acids research · 2007 · 8 claims · 6 setups
Two exonic mutations, PINK1 E417G and PARK7 E64D, disrupt binding to U1 snRNA and cause skipping of the mutation-harboring exon