Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Mutation analysis of the MSMB gene in familial prostate cancer.
PMID 19997100 · PMC2816656 · British journal of cancer · 2010 · 8 claims · 5 setups
No deleterious mutations were found in the MSMB coding region in 192 familial prostate cancer cases
-
Full-text index only
A comprehensive resequence analysis of the KLK15-KLK3-KLK2 locus on chromosome 19q13.33.
PMID 19823874 · PMC2793378 · Human genetics · 2010 · 7 claims · 7 setups
Deep resequencing of a 56 kb region on chr19q13.33 identified 555 polymorphic loci, including 116 novel SNPs and 182 novel indels.
-
Full-text index only
NeMeSys: a biological resource for narrowing the gap between sequence and function in the human pathogen Neisseria meningitidis.
PMID 19818133 · PMC2784325 · Genome biology · 2009 · 8 claims · 5 setups
Determined and manually annotated the complete genome sequence of N. meningitidis clinical isolate strain 8013
-
Full-text index only
Prevalence and functional analysis of sequence variants in the ATR checkpoint mediator Claspin.
PMID 19737971 · PMC2994259 · Molecular cancer research : MCR · 2009 · 8 claims · 8 setups
CLSPN is a mediator protein essential for the ATR- and CHK1-dependent checkpoint response to replicative stress or single-stranded DNA
-
Full-text index only
The Neandertal genome and ancient DNA authenticity.
PMID 19661919 · PMC2725275 · The EMBO journal · 2009 · 8 claims · 6 setups
Only direct assays of DNA sequence positions where Neandertals differ from all contemporary humans can reliably estimate human contamination.
-
Has reproduction · 77
Ecotype diversity and conversion in Photobacterium profundum strains.
PMID 24824441 · PMC4019646 · PloS one · 2014 · 8 claims · 8 setups
No single gene restricts the environmental niche of each bathytype; instead a set of strain-specific genetic features confers depth-specific stress tolerance (temperature, pressure, nutrients).
-
Has reproduction · 58
MZPAQ: a FASTQ data compression tool.
PMID 31171931 · PMC6547476 · Source code for biology and medicine · 2019 · 7 claims · 3 setups
MZPAQ, a hybrid of MFCompress and ZPAQ, outperforms state-of-the-art and general-purpose compression tools on all benchmark datasets in terms of compression ratio
-
Full-text index only
A SNP-centric database for the investigation of the human genome.
PMID 15046636 · PMC395999 · BMC bioinformatics · 2004 · 8 claims · 3 setups
SNPper is a web-based, integrated SNP database combining dbSNP, the Human Genome sequence (Goldenpath), LocusLink, GeneOntology, and SWISS-PROT data with querying, visualization, and export tools.
-
Has reproduction · 61
lncEvo: automated identification and conservation study of long noncoding RNAs.
PMID 33563213 · PMC7871587 · BMC bioinformatics · 2021 · 8 claims · 5 setups
lncEvo is an integrated Nextflow/Docker pipeline combining transcriptome assembly, lncRNA identification, and cross-species conservation analysis into a single workflow.
-
Full-text index only
VISTA Enhancer Browser--a database of tissue-specific human enhancers.
PMID 17130149 · PMC1716724 · Nucleic acids research · 2007 · 8 claims · 2 setups
Comparative genome analysis can identify candidate human enhancer elements whose tissue-specific in vivo activity can then be experimentally validated in transgenic mice.
-
Full-text index only
The many uses of a genome sequence.
PMID 11423005 · PMC138940 · Genome biology · 2001 · 8 claims · 8 setups
Solved protein structures from structural genomics efforts can be used to model many other proteins by homology, aiding function prediction
-
Full-text index only
Twin peaks: the draft human genome sequence.
PMID 11276423 · PMC138909 · Genome biology · 2001 · 8 claims · 8 setups
The predicted number of human genes (~26,000-40,000) is far lower than the widely assumed ~100,000, though downstream RNA/protein complexity can still generate substantial biological complexity.
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.
-
Has reproduction · 84
Genome of the Asian longhorned beetle (Anoplophora glabripennis), a globally significant invasive species, reveals key functional and evolutionary innovations at the beetle-plant interface.
PMID 27832824 · PMC5105290 · Genome biology · 2016 · 8 claims · 7 setups
The A. glabripennis genome encodes a uniquely diverse arsenal of enzymes that degrade plant cell wall polysaccharide networks (cellulose, hemicellulose, pectin) and detoxify plant allelochemicals.
-
Has reproduction · 89
DFAST and DAGA: web-based integrated genome annotation tools and resources.
PMID 27867804 · PMC5107635 · Bioscience of microbiota, food and health · 2016 · 8 claims · 7 setups
DFAST is a web-based genome annotation pipeline with integrated quality assessment (CheckM) and taxonomic assessment (ANI) that produces DDBJ submission-ready files
-
Has reproduction · 83
Current status of use of high throughput nucleotide sequencing in rheumatology.
PMID 33408124 · PMC7789458 · RMD open · 2021 · 8 claims · 8 setups
RNA-Seq is the most represented HTS assay used in rheumatology research, primarily for biomarker identification in blood or synovial tissue.
-
Has reproduction · 71
Prospects of telomere-to-telomere assembly in barley: Analysis of sequence gaps in the MorexV3 reference genome.
PMID 35338551 · PMC9241371 · Plant biotechnology journal · 2022 · 7 claims · 8 setups
Almost all centromeric sequences and 45S ribosomal DNA repeat arrays are absent from the MorexV3 pseudomolecules
-
Has reproduction · 57
KARAJ: An Efficient Adaptive Multi-Processor Tool to Streamline Genomic and Transcriptomic Sequence Data Acquisition.
PMID 36430895 · PMC9694301 · International journal of molecular sciences · 2022 · 8 claims · 6 setups
KARAJ automates end-to-end querying and downloading of genomic/transcriptomic sequence data from a list of PMCIDs, URLs, or accession numbers
-
Has reproduction · 99
getSequenceInfo: a suite of tools allowing to get genome sequence information from public repositories.
PMID 35804320 · PMC9264741 · BMC bioinformatics · 2022 · 8 claims · 8 setups
getSequenceInfo (gSeqI) allows programmatic (CLI) or GUI-based retrieval of sequence data and metadata from GenBank, RefSeq, and ENA across Linux, MacOS, and Windows.
-
Has reproduction · 92
Evaluation of core genome and whole genome multilocus sequence typing schemes for Campylobacter jejuni and Campylobacter coli outbreak detection in the USA.
PMID 37133905 · PMC10272873 · Microbial genomics · 2023 · 8 claims · 8 setups
cgMLST, wgMLST and hqSNP WGS-based analysis methods clustered C. jejuni and C. coli isolates in concordance with epidemiological data.