Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The relationship of potential G-quadruplex sequences in cis-upstream regions of the human genome to SP1-binding elements.
PMID 18353860 · PMC2377421 · Nucleic acids research · 2008 · 7 claims · 1 setups
A large number of upstream PQSSs incorporate the SP1-binding element, establishing a clear link between PQSS occurrence and SP1 elements
-
Full-text index only
Fast and systematic genome-wide discovery of conserved regulatory elements using a non-alignment based approach.
PMID 15693947 · PMC551538 · Genome biology · 2005 · 7 claims · 8 setups
FastCompare, a non-alignment-based, linear-time algorithm, computes a genome-wide conservation score for all k-mers (7-9 nt) between two genomes to identify conserved regulatory elements
-
Full-text index only
Ab initio identification of human microRNAs based on structure motifs.
PMID 18088431 · PMC2238772 · BMC bioinformatics · 2007 · 8 claims · 7 setups
MiRPred predicts miRNA precursors ab initio using only predicted secondary structure motifs, ignoring nucleotide sequence
-
Full-text index only
Prediction-based approaches to characterize bidirectional promoters in the mammalian genome.
PMID 18366609 · PMC2386062 · BMC genomics · 2008 · 8 claims · 7 setups
The mapping algorithm identified 5,647 candidate bidirectional promoter regions in the mouse genome, similar in number to those previously found in human.
-
Full-text index only
Annotation and analysis of 10,000 expressed sequence tags from developing mouse eye and adult retina.
PMID 14519200 · PMC328454 · Genome biology · 2003 · 8 claims · 5 setups
Annotation of 8,633 high-quality non-mitochondrial/non-ribosomal ESTs shows 57% represent known genes and 43% are unknown or novel, with M15E having the highest proportion of novel ESTs
-
Full-text index only
Consolidating the set of known human protein-protein interactions in preparation for large-scale mapping of the human interactome.
PMID 15892868 · PMC1175952 · Genome biology · 2005 · 8 claims · 6 setups
Two quantitative benchmarks (functional-annotation-based and physical-interaction-based log likelihood ratio scores) can measure relative accuracy of human PPI datasets
-
Full-text index only
Duplication count distributions in DNA sequences.
PMID 19256873 · PMC3121164 · Physical review. E, Statistical, nonlinear, and soft matter physics · 2008 · 8 claims · 8 setups
Duplication count distributions N(c) for complex 40-mers show power-law-like decay for c roughly 3 to 50 (or higher) across human, C. elegans, A. thaliana, and D. melanogaster genomes.
-
Full-text index only
ATBF1 and NQO1 as candidate targets for allelic loss at chromosome arm 16q in breast cancer: absence of somatic ATBF1 mutations and no role for the C609T NQO1 polymorphism.
PMID 18416817 · PMC2377272 · BMC cancer · 2008 · 8 claims · 7 setups
Five genes (NQO1, ATBF1, DBNDD1, HSBP1, CGI-38) at 16q show significantly lower mRNA expression in breast tumors with LOH at 16q compared to tumors without LOH
-
Full-text index only
Structure of protein interaction networks and their implications on drug design.
PMID 19876376 · PMC2760708 · PLoS computational biology · 2009 · 8 claims · 6 setups
Budding yeast and human PINs are scale-rich and configured as highly optimized tolerance (HOT) networks similar to Internet router-level topology, rather than scale-free networks formed by preferential attachment.
-
Full-text index only
BTW: a web server for Boltzmann time warping of gene expression time series.
PMID 16845055 · PMC1538860 · Nucleic acids research · 2006 · 5 claims · 4 setups
Symmetric time warping distance is more flexible than Euclidean distance or correlation coefficient for identifying genes with similar temporal expression profiles, especially across sequences of different length.
-
Full-text index only
Human SNPs resulting in premature stop codons and protein truncation.
PMID 16595072 · PMC3500177 · Human genomics · 2006 · 8 claims · 6 setups
Genome-wide screening of dbSNP identified 28 validated X-SNPs from 28 genes with known minor allele frequencies.
-
Full-text index only
Tumor mapping in 2 large multigenerational families with CYLD mutations: implications for disease management and tumor induction.
PMID 19917957 · PMC2935681 · Archives of dermatology · 2009 · 8 claims · 4 setups
The clinical distinction between FC, BSS, and MFT has little prognostic or clinical utility, even within the same family, warranting a unifying diagnosis of 'CYLD cutaneous syndrome'.
-
Full-text index only
Assaying chromosomal inversions by single-molecule haplotyping.
PMID 16721377 · PMC2690135 · Nature methods · 2006 · 8 claims · 4 setups
Haplotype Fusion PCR (HF-PCR) juxtaposes sequences flanking an inversion breakpoint on single DNA molecules via emulsion PCR, generating orientation-specific fusion products diagnostic of inversion genotype
-
Has reproduction · 68
Bayesian transcriptome assembly.
PMID 25367074 · PMC4397945 · Genome biology · 2014 · 8 claims · 8 setups
Bayesembler, a probabilistic transcriptome assembler built on a Bayesian model of the RNA sequencing process with Gibbs sampling over expressed candidates, abundances and read assignments, is introduced.
-
Full-text index only
Whole genome amplification and de novo assembly of single bacterial cells.
PMID 19724646 · PMC2731171 · PloS one · 2009 · 8 claims · 6 setups
FACS-based single-cell isolation combined with strict handling procedures virtually eliminates contaminating DNA from single-cell MDA reactions
-
Has reproduction · 67
Leveraging RNA-seq deconvolution to improve complex in vitro model characterization.
PMID 40701251 · PMC12391696 · The Journal of biological chemistry · 2025 · 8 claims · 6 setups
RNA-seq deconvolution can predict cell type proportions from bulk RNA-seq using scRNA-seq references, offering a useful characterization tool for CIVMs where single-cell methods are impractical
-
Has reproduction · 83
Accurate prediction of metagenome-assembled genome completeness by MAGISTA, a random forest model built on alignment-free intra-bin statistics.
PMID 35248155 · PMC8898458 · Environmental microbiome · 2022 · 7 claims · 7 setups
MAGISTA, a random forest model built on alignment-free intra-bin distance-distribution statistics, can estimate MAG completeness and purity without relying on reference marker genes.
-
Has reproduction · 65
FusionQ: a novel approach for gene fusion detection and quantification from paired-end RNA-Seq.
PMID 23768108 · PMC3691734 · BMC bioinformatics · 2013 · 8 claims · 8 setups
FusionQ is a novel tool that detects gene fusions, constructs chimerical transcript structures, and estimates their abundances from paired-end RNA-Seq data.
-
Has reproduction · 44
Detecting DNA modifications from SMRT sequencing data by modeling sequence context dependence of polymerase kinetic.
PMID 23516341 · PMC3597545 · PLoS computational biology · 2013 · 8 claims · 7 setups
Local sequence context strongly determines position-specific polymerase kinetic rate: roughly 80% of IPD variation is explained by a 10 bp context (7 bases upstream, 2 bases downstream of the incorporation site), saturating at 7 bases upstream.
-
Full-text index only
How to find soluble proteins: a comprehensive analysis of alpha/beta hydrolases for recombinant expression in E. coli.
PMID 15804363 · PMC1079826 · BMC genomics · 2005 · 7 claims · 7 setups
Predicted solubility in E. coli (via CV-CV') depends on hydrolase size, phylogenetic origin, homologous family, and superfamily