Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
CONTRAST: a discriminative, phylogeny-free approach to multiple informant de novo gene prediction.
PMID 18096039 · PMC2246271 · Genome biology · 2007 · 8 claims · 5 setups
CONTRAST predicts exact coding region structures for 65% more human genes than the previous state-of-the-art de novo predictor (N-SCAN)
-
Full-text index only
Simple models of genomic variation in human SNP density.
PMID 17553150 · PMC1919371 · BMC genomics · 2007 · 6 claims · 4 setups
Hierarchical Poisson model B, which allows both the mutation-rate proxy (Beta-distributed Λ) and the ARG-size proxy (Gamma-distributed T) to vary, fits the observed SNP density distribution significantly better than models with only one or neither varying.
-
Full-text index only
SilkDB v2.0: a platform for silkworm (Bombyx mori ) genome biology.
PMID 19793867 · PMC2808975 · Nucleic acids research · 2010 · 8 claims · 8 setups
A new 8.5x-coverage silkworm genome assembly with N50 scaffold size of ~3.7 Mb over a 432 Mb genome represents a significant quality improvement over the prior draft.
-
Has reproduction · 72
Analysis of the genome of the New Zealand giant collembolan (Holacanthella duospinosa) sheds light on hexapod evolution.
PMID 29041914 · PMC5644144 · BMC genomics · 2017 · 8 claims · 8 setups
A high-quality ~375 Mbp draft genome and transcriptome of Holacanthella duospinosa was assembled and annotated, providing a genomic resource for hexapod evolution.
-
Full-text index only
Cryptic loxP sites in mammalian genomes: genome-wide distribution and relevance for the efficiency of BAC/PAC recombineering techniques.
PMID 17284462 · PMC1865043 · Nucleic acids research · 2007 · 6 claims · 6 setups
Cryptic lox P sites occur frequently and are homogeneously distributed across the mouse genome (1.2 primary sites per megabase).
-
Full-text index only
Genomic analysis of the chromosome 15q11-q13 Prader-Willi syndrome region and characterization of transcripts for GOLGA8E and WHCD1L1 from the proximal breakpoint region.
PMID 18226259 · PMC2268926 · BMC genomics · 2008 · 8 claims · 7 setups
GOLGA8E and WHDC1L1 are characterized for the first time as protein-coding transcripts from the PWS proximal breakpoint region.
-
Full-text index only
Genome-wide detection of segmental duplications and potential assembly errors in the human genome sequence.
PMID 12702206 · PMC154576 · Genome biology · 2003 · 8 claims · 6 setups
Segmental duplications comprise 3.53% (107.4/3,043.1 Mb) of the June 2002 human genome assembly
-
Full-text index only
DBD--taxonomically broad transcription factor predictions: new content and functionality.
PMID 18073188 · PMC2238844 · Nucleic acids research · 2008 · 8 claims · 3 setups
DBD is a database of predicted sequence-specific DNA-binding transcription factors covering over 700 publicly available proteomes, up from 150 in the initial version.
-
Full-text index only
Molecular evolution of Cide family proteins: novel domain formation in early vertebrates and the subsequent divergence.
PMID 18500987 · PMC2426694 · BMC evolutionary biology · 2008 · 8 claims · 5 setups
Sequences homologous to the CIDE-N domain/NCD show a wide phylogenetic distribution, from hydra and sea anemone to mammals, while true Cide proteins are restricted to vertebrates.
-
Has reproduction · 69
A comparison across non-model animals suggests an optimal sequencing depth for de novo transcriptome assembly.
PMID 23496952 · PMC3655071 · BMC genomics · 2013 · 8 claims · 8 setups
Representative de novo transcriptome assemblies are generated with as few as ~20 million reads for single-tissue samples and ~30 million reads for whole animals at the mRNA-coverage level.
-
Has reproduction · 90
PrimerSeq: Design and visualization of RT-PCR primers for alternative splicing using RNA-seq data.
PMID 24747190 · PMC4411361 · Genomics, proteomics & bioinformatics · 2014 · 8 claims · 3 setups
PrimerSeq is a user-friendly stand-alone software with a GUI for systematic design and visualization of RT-PCR primers for alternative splicing analysis using user-provided RNA-seq data.
-
Has reproduction · 57
Regulatory Noncoding Small RNAs Are Diverse and Abundant in an Extremophilic Microbial Community.
PMID 32019831 · PMC7002113 · mSystems · 2020 · 8 claims · 7 setups
Hundreds of intergenic (itsRNAs) and antisense (asRNAs) sRNAs are diverse and abundant in the halite endolithic microbial community, with 1,538 total ncRNAs discovered across Archaea and Bacteria.
-
Full-text index only
Identifying related L1 retrotransposons by analyzing 3' transduced sequences.
PMID 12734010 · PMC156586 · Genome biology · 2003 · 8 claims · 6 setups
L1 elements with transduction-derived 3' sequence (L1-TDs) can be computationally identified using RepeatMasker/TSDfinder and grouped into families sharing a common progenitor via BLAST comparison of downstream sequences.
-
Full-text index only
Capturing genomic signatures of DNA sequence variation using a standard anonymous microarray platform.
PMID 17000641 · PMC1636412 · Nucleic acids research · 2006 · 8 claims · 6 setups
An anonymous SHyP oligonucleotide microarray can capture genomic signatures of DNA sequence variation from any organism, including a previously unsequenced species
-
Full-text index only
Methylation of class II transactivator gene promoter IV is not associated with susceptibility to multiple sclerosis.
PMID 18606010 · PMC2464579 · BMC medical genetics · 2008 · 6 claims · 4 setups
Methylation of the MHC2TA promoter pIV is not associated with MS susceptibility; no methylation was detected in any twin sample regardless of disease status
-
Full-text index only
Predicting failure rate of PCR in large genomes.
PMID 18492719 · PMC2441781 · Nucleic acids research · 2008 · 7 claims · 8 setups
The number of predicted primer-binding sites in genomic DNA is the most important factor determining PCR failure.
-
Has reproduction · 79
TSUNAMI: Translational Bioinformatics Tool Suite for Network Analysis and Mining.
PMID 33705981 · PMC9403021 · Genomics, proteomics & bioinformatics · 2021 · 8 claims · 6 setups
TSUNAMI is a freely accessible web-based tool suite that mines gene co-expression network (GCN) modules from public (GEO, TCGA) or user-uploaded numerical omics data and performs downstream gene set enrichment analysis.
-
Has reproduction · 89
A near complete genome for goat genetic and genomic research.
PMID 34507524 · PMC8434745 · Genetics, selection, evolution : GSE · 2021 · 8 claims · 8 setups
Saanen_v1 is a high-quality de novo goat genome assembly from a male Saanen buck, including the first goat Y chromosome scaffold.
-
Has reproduction · 86
RNASEQR--a streamlined and accurate RNA-seq sequence analysis program.
PMID 22199257 · PMC3315322 · Nucleic acids research · 2012 · 8 claims · 7 setups
RNASEQR is a new RNA-seq mapper/aligner that combines a BWT-based (Bowtie) transcriptomic/genomic alignment with hash-based BLAT local alignment in three sequential steps: transcriptome mapping, novel exon detection, and anchor-and-align novel splice junction identification.
-
Has reproduction · 68
Complete Genome Sequencing of Lactobacillus plantarum ZLP001, a Potential Probiotic That Enhances Intestinal Epithelial Barrier Function and Defense Against Pathogens in Pigs.
PMID 30542296 · PMC6277807 · Frontiers in physiology · 2018 · 8 claims · 8 setups
The complete genome of L. plantarum ZLP001 comprises a single 3,164,369 bp circular chromosome (GC 44.65%) plus seven plasmids (A–G), encoding 3,264 protein-coding sequences.