Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 69
Discovery and characterization of Alu repeat sequences via precise local read assembly.
PMID 26503250 · PMC4666360 · Nucleic acids research · 2015 · 7 claims · 8 setups
Combining Alu-supporting read detection (RetroSeq) with local de novo assembly (CAP3) reconstructs the full sequence of non-reference Alu insertions from Illumina paired-end WGS reads
-
Has reproduction
Predicting favorable landing pads for targeted integrations in Chinese hamster ovary cell lines by learning stability characteristics from random transgene integrations.
PMID 33304461 · PMC7710658 · Computational and structural biotechnology journal · 2020 · 7 claims · 6 setups
Expression stability in CHO cell lines is controlled at three levels: choice of integration site, integrity/concatemerization pattern of the transgene, and stress-related cellular processes.
-
Full-text index only
Design and analysis issues in genome-wide somatic mutation studies of cancer.
PMID 18692126 · PMC2820387 · Genomics · 2009 · 6 claims · 4 setups
Two-stage (discovery + validation) sequencing designs efficiently allocate resources and can produce highly informative candidate driver gene lists even with relatively small sample sizes.
-
Full-text index only
6th annual meeting of the Complex Trait Consortium.
PMID 17906895 · PMC2042027 · Mammalian genome : official journal of the International Mammalian Genome Society · 2007 · 8 claims · 7 setups
The NIEHS Perlegen/resequencing project has generated over 8.5 million SNPs from 15 inbred mouse strains but shows a high false-negative discovery rate, with an estimated 45 million SNPs actually present.
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.
-
Full-text index only
Discovering the phylodynamics of RNA viruses.
PMID 19855824 · PMC2756585 · PLoS computational biology · 2009 · 8 claims · 8 setups
A comprehensive 'discovery phase' integrating ecological and evolutionary dynamics of RNA viruses across scales is needed to achieve a full phylodynamic understanding
-
Has reproduction · 63
A methyl-sensitive element induces bidirectional transcription in TATA-less CpG island-associated promoters.
PMID 30332484 · PMC6192621 · PloS one · 2018 · 7 claims · 8 setups
The CGCG element (consensus TCTCGCGAGA) is a conserved 10-bp motif enriched in TATA-less, CpG island-associated promoters of ribosomal protein and housekeeping genes.
-
Full-text index only
Hit selection with false discovery rate control in genome-scale RNAi screens.
PMID 18628291 · PMC2504311 · Nucleic acids research · 2008 · 8 claims · 3 setups
A Bayesian FDR-controlling methodology for hit selection in genome-scale RNAi HTS is proposed, using a direct posterior probability approach analogous to Newton et al.
-
Full-text index only
Science star over Asia.
PMID 16149850 · PMC1201306 · PLoS biology · 2005 · 8 claims · 7 setups
Ariff Bongso and associates at Singapore's National University Hospital were the first to derive human embryonic stem cells, from a five-day-old discarded human embryo in 1994, and showed the cells were pluripotent with therapeutic transplant potential.
-
Full-text index only
Human and mouse introns are linked to the same processes and functions through each genome's most frequent non-conserved motifs.
PMID 18450818 · PMC2425492 · Nucleic acids research · 2008 · 8 claims · 5 setups
Pyknons (recurrent, genome-specific, ≥16nt motifs with ≥30 intact intergenic/intronic copies and ≥1 exonic copy) span a substantial fraction of previously uncharacterized intronic space (7.4% human, 4.4% mouse)
-
Has reproduction · 80
DMN-seq enriches DNA hypomethylated regions for biomarker discovery using 5-methylcytosine glycosylase.
PMID 41673887 · PMC13097799 · Genome biology · 2026 · 8 claims · 9 setups
DMN-seq (DMN+) uses DME to nick DNA specifically at 5mC sites, enabling 5mC detection at single-base resolution via selective adaptor ligation
-
Full-text index only
CRSD: a comprehensive web server for composite regulatory signature discovery.
PMID 16845073 · PMC1538777 · Nucleic acids research · 2006 · 7 claims · 5 setups
CRSD is a comprehensive web server integrating six large-scale databases (UniGene, mature microRNAs, putative promoter, TRANSFAC, pathway, GO) plus two newly constructed genome-wide databases (MRS and TRS) for composite regulatory signature discovery
-
Full-text index only
Motif discovery in promoters of genes co-localized and co-expressed during myeloid cells differentiation.
PMID 19059999 · PMC2632922 · Nucleic acids research · 2009 · 6 claims · 8 setups
A novel multi-step computational method (built on approximate pattern enumeration, binomial over-representation scoring with FDR correction, and k-medoids clustering) can identify over-represented motifs in a selected set of promoters relative to a background promoter set.
-
Has reproduction · 45
Accurate sequence variant genotyping in cattle using variation-aware genome graphs.
PMID 31092189 · PMC6521551 · Genetics, selection, evolution : GSE · 2019 · 8 claims · 7 setups
Graphtyper outperformed GATK and SAMtools in genotype concordance, non-reference sensitivity, and non-reference discrepancy compared to microarray genotypes
-
Full-text index only
CoMoDis: composite motif discovery in mammalian genomes.
PMID 17130158 · PMC1702496 · Nucleic acids research · 2007 · 7 claims · 4 setups
CoMoDis is a new bioinformatics tool that streamlines computational identification of novel regulatory modules starting from a single seed motif
-
Full-text index only
Genome-wide identification of in vivo protein-DNA binding sites from ChIP-Seq data.
PMID 18684996 · PMC2532738 · Nucleic acids research · 2008 · 8 claims · 7 setups
SISSRs identifies binding sites from ChIP-Seq short reads with much higher resolution than the standard region-clustering approach
-
Full-text index only
SNP selection for genes of iron metabolism in a study of genetic modifiers of hemochromatosis.
PMID 18366708 · PMC2289803 · BMC medical genetics · 2008 · 7 claims · 6 setups
Illumina validation/design scores above 0.6 are not strongly correlated with actual SNP genotyping performance (Gentrain score)
-
Full-text index only
Outlook on Thailand's genomics and computational biology research and development.
PMID 18654621 · PMC2446437 · PLoS computational biology · 2008 · 8 claims · 8 setups
Thai government policy support, infrastructure investment, education programs, and human resource development have substantially advanced genomics and bioinformatics research capacity in Thailand
-
Has reproduction
Trans-ethnic association study of blood pressure determinants in over 750,000 individuals.
PMID 30578418 · PMC6365102 · Nature genetics · 2019 · 8 claims · 8 setups
Discovery and replication GWAS of SBP, DBP and pulse pressure in up to 776,078 individuals identified 208 novel common blood pressure SNPs and 53 rare variants.
-
Has reproduction · 75
Sequencing of human genomes with nanopore technology.
PMID 31015479 · PMC6478738 · Nature communications · 2019 · 8 claims · 7 setups
A novel single-sample, reference panel-free, read-based phasing algorithm built on the STITCH model improves nanopore SNV calling from modest baseline levels.