Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 90
Gap-free telomere-to-telomere haplotype assembly of the tomato hind (Cephalopholis sonnerati).
PMID 39578472 · PMC11584678 · Scientific data · 2024 · 8 claims · 8 setups
Two T2T gap-free haplotype assemblies of C. sonnerati (YSFRI_Csonn_HA_1.0 and YSFRI_Csonn_HB_1.0) were successfully generated, each spanning 24 chromosomes with no gaps.
-
Has reproduction · 90
An improved assembly of the pearl millet reference genome using Oxford Nanopore long reads and optical mapping.
PMID 36891809 · PMC10151396 · G3 (Bethesda, Md.) · 2023 · 8 claims · 8 setups
Combining ONT long reads with Bionano optical maps produced a substantially more complete and contiguous pearl millet Tift 23D2B1-P1-P5 assembly than the prior short-read assembly.
-
Full-text index only
Versatile and open software for comparing large genomes.
PMID 14759262 · PMC395750 · Genome biology · 2004 · 8 claims · 8 setups
MUMmer 3.0 efficiently handles comparisons of large eukaryotic genomes at varying evolutionary distances
-
Full-text index only
Genomic analysis of the chromosome 15q11-q13 Prader-Willi syndrome region and characterization of transcripts for GOLGA8E and WHCD1L1 from the proximal breakpoint region.
PMID 18226259 · PMC2268926 · BMC genomics · 2008 · 8 claims · 7 setups
GOLGA8E and WHDC1L1 are characterized for the first time as protein-coding transcripts from the PWS proximal breakpoint region.
-
Has reproduction · 58
HGA: de novo genome assembly method for bacterial genomes using high coverage short sequencing reads.
PMID 26945881 · PMC4779561 · BMC genomics · 2016 · 8 claims · 7 setups
HGA leads to significant improvement in assembly quality (N50 and corrected N50) for all 7 evaluated GAGE-B bacterial datasets using most of the 8 evaluated assemblers
-
Has reproduction · 75
A step forward for Shiga toxin-producing Escherichia coli identification and characterization in raw milk using long-read metagenomics.
PMID 36748417 · PMC9836091 · Microbial genomics · 2022 · 8 claims · 6 setups
Long-read metagenomics enables isolation-independent identification and characterization of eae-positive STEC directly from raw milk.
-
Has reproduction · 76
What the Phage: a scalable workflow for the identification and analysis of phage sequences.
PMID 36399058 · PMC9673492 · GigaScience · 2022 · 8 claims · 7 setups
WtP combines 11 tools (14 approaches) for phage prediction in a parallel, containerized Nextflow workflow
-
Has reproduction · 95
transXpress: a Snakemake pipeline for streamlined de novo transcriptome assembly and annotation.
PMID 37016291 · PMC10074830 · BMC bioinformatics · 2023 · 6 claims · 7 setups
transXpress is a Snakemake pipeline that streamlines de novo transcriptome assembly, quantification, and annotation for non-model organisms
-
Full-text index only
Investigating hookworm genomes by comparative analysis of two Ancylostoma species.
PMID 15854223 · PMC1112591 · BMC genomics · 2005 · 8 claims · 8 setups
Nearly 20,000 ESTs from 7 cDNA libraries define nearly 7,000 hookworm genes across A. caninum and A. ceylanicum
-
Full-text index only
Inference of transcriptional regulation using gene expression data from the bovine and human genomes.
PMID 17683551 · PMC1978505 · BMC genomics · 2007 · 7 claims · 8 setups
Using human reference promoter sequences is a useful approach for studying gene expression regulation in species with limited or non-existing genomic sequence, such as cattle.
-
Has reproduction · 86
LMAS: evaluating metagenomic short de novo assembly methods through defined communities.
PMID 36576131 · PMC9795473 · GigaScience · 2022 · 8 claims · 5 setups
LMAS (Last Metagenomic Assembler Standing) is a flexible, Nextflow-based, Docker-containerized automated workflow for benchmarking de novo metagenomic assemblers against defined mock communities, producing an interactive HTML report.
-
Full-text index only
Analysis of human sarcospan as a candidate gene for CFEOM1.
PMID 11180757 · PMC29083 · BMC genetics · 2001 · 7 claims · 5 setups
Sarcospan sequence is unmutated in all six CFEOM1 families studied
-
Full-text index only
Assessing the gene space in draft genomes.
PMID 19042974 · PMC2615622 · Nucleic acids research · 2009 · 6 claims · 7 setups
The proportion of mapped CEGs in a draft genome assembly is a useful metric for describing gene space completeness, complementing N50 and x-fold coverage.
-
Has reproduction · 99
A platinum standard pan-genome resource that represents the population structure of Asian rice.
PMID 32265447 · PMC7138821 · Scientific data · 2020 · 6 claims · 6 setups
The 3,000 Rice Genomes (3K-RG) dataset can be subdivided into 15 subpopulations (K=15), refining the previous K=9 population structure.
-
Full-text index only
Sequence variation and linkage disequilibrium in the GABA transporter-1 gene (SLC6A1) in five populations: implications for pharmacogenetic research.
PMID 17941974 · PMC2175509 · BMC genetics · 2007 · 8 claims · 7 setups
SLC6A1 shows low levels of LD and an absence of major LD blocks across all five populations studied
-
Full-text index only
Comparative genomic analysis of three Leishmania species that cause diverse human disease.
PMID 17572675 · PMC2592530 · Nature genetics · 2007 · 8 claims · 6 setups
L. infantum and L. braziliensis genomes were sequenced and show marked conservation of synteny with L. major, with only ~200 genes differentially distributed among the three species
-
Full-text index only
Processing and population genetic analysis of multigenic datasets with ProSeq3 software.
PMID 19797407 · PMC2778335 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 7 setups
ProSeq3 is a program with a graphic user interface that simplifies preparation and basic population genetic analysis of multigenic DNA polymorphism datasets
-
Has reproduction · 91
Genome-wide identification of conserved and novel microRNAs in one bud and two tender leaves of tea plant (Camellia sinensis) by small RNA sequencing, microarray-based hybridization and genome survey scaffold sequences.
PMID 29157210 · PMC5697157 · BMC plant biology · 2017 · 7 claims · 8 setups
175 conserved and 83 novel miRNAs were identified mainly in one bud and two tender leaves of tea plant via small RNA sequencing combined with genome survey data
-
Has reproduction · 92
Acquisition and loss of CTX-M plasmids in Shigella species associated with MSM transmission in the UK.
PMID 34427554 · PMC8549364 · Microbial genomics · 2021 · 8 claims · 8 setups
bla_CTX-M-27 is located on IncFII pKSR100-like plasmids, flanked by IS26 and IS903B
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.