Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A genome-wide screen for copy number alterations in Aicardi syndrome.
PMID 19760649 · PMC3640635 · American journal of medical genetics. Part A · 2009 · 7 claims · 4 setups
Aicardi syndrome is thought to result from heterozygous defects in an essential X-linked gene, or from a sex-limited autosomal gene defect, due to its occurrence almost exclusively in females and in 47,XXY males.
-
Full-text index only
How to find soluble proteins: a comprehensive analysis of alpha/beta hydrolases for recombinant expression in E. coli.
PMID 15804363 · PMC1079826 · BMC genomics · 2005 · 7 claims · 7 setups
Predicted solubility in E. coli (via CV-CV') depends on hydrolase size, phylogenetic origin, homologous family, and superfamily
-
Has reproduction · 91
Genome-wide identification of conserved and novel microRNAs in one bud and two tender leaves of tea plant (Camellia sinensis) by small RNA sequencing, microarray-based hybridization and genome survey scaffold sequences.
PMID 29157210 · PMC5697157 · BMC plant biology · 2017 · 7 claims · 8 setups
175 conserved and 83 novel miRNAs were identified mainly in one bud and two tender leaves of tea plant via small RNA sequencing combined with genome survey data
-
Has reproduction · 100
A comprehensive framework for analysis of microRNA sequencing data in metastatic colorectal cancer.
PMID 35047825 · PMC8759566 · NAR cancer · 2022 · 7 claims · 7 setups
Five miRNAs (Mir-210_3p, Mir-191_5p, Mir-8-P1b_3p [miR-141-3p], Mir-1307_5p, Mir-155_5p) are up-regulated at multiple metastatic sites in colorectal cancer.
-
Has reproduction · 42
KAGE: fast alignment-free graph-based genotyping of SNPs and short indels.
PMID 36195962 · PMC9531401 · Genome biology · 2022 · 7 claims · 7 setups
KAGE combines population-based kmer count modeling with single-variant prior adjustment into an alignment-free genotyper that matches the accuracy of the best existing alignment-free genotypers while being an order of magnitude faster.
-
Full-text index only
A small-cell lung cancer genome with complex signatures of tobacco exposure.
PMID 20016488 · PMC2880489 · Nature · 2010 · 8 claims · 7 setups
22,910 somatic substitutions, including 132 in coding exons, were identified genome-wide in the NCI-H209 SCLC genome.
-
Full-text index only
Genome-wide analysis of small RNA and novel MicroRNA discovery in human acute lymphoblastic leukemia based on extensive sequencing approach.
PMID 19724645 · PMC2731166 · PloS one · 2009 · 7 claims · 5 setups
159 novel miRNAs and 116 novel miRNA*s were identified from ALL patient and normal donor small RNA libraries
-
Has reproduction · 37
A Bayesian approach to accurate and robust signature detection on LINCS L1000 data.
PMID 32003771 · PMC7203754 · Bioinformatics (Oxford, England) · 2020 · 6 claims · 4 setups
A novel Bayesian-based peak deconvolution algorithm gives unbiased likelihood estimations for peak locations and characterizes peaks with probability-based z-scores.
-
Full-text index only
Prediction of specificity-determining residues for small-molecule kinase inhibitors.
PMID 19032760 · PMC2655090 · BMC bioinformatics · 2008 · 8 claims · 5 setups
S-Filter is a novel method combining sequence and structural information (within PFAAT) to predict specificity-determining residues and selectivity profiles for small-molecule kinase inhibitors
-
Full-text index only
Ratiocinative screen of eukaryotic integral membrane protein expression and solubilization for structure determination.
PMID 19031011 · PMC2756966 · Journal of structural and functional genomics · 2009 · 8 claims · 6 setups
A discovery-oriented pipeline using standardized single-condition methods (one expression system, one detergent, one SEC buffer) can efficiently triage large numbers of eukaryotic IMP targets to identify well-behaved candidates for crystallization
-
Full-text index only
SeqBuster, a bioinformatic tool for the processing and analysis of small RNAs datasets, reveals ubiquitous miRNA modifications in human embryonic cells.
PMID 20008100 · PMC2836562 · Nucleic acids research · 2010 · 8 claims · 6 setups
SeqBuster is a versatile web-based and stand-alone bioinformatic toolkit for processing and analyzing large-scale small RNA deep sequencing datasets.
-
Full-text index only
BioDrugScreen: a computational drug design resource for ranking molecules docked to the human proteome.
PMID 19923229 · PMC2808957 · Nucleic acids research · 2010 · 6 claims · 5 setups
BioDrugScreen is a web resource providing pre-docked and pre-scored receptor-ligand complexes for ranking molecules against human proteome targets
-
Full-text index only
Using multiple alignments to improve seeded local alignment algorithms.
PMID 16100379 · PMC1185574 · Nucleic acids research · 2005 · 8 claims · 2 setups
Using information implicit in a multiple alignment to dynamically build a spaced-seed index weighted toward promising regions increases sensitivity of local alignment search compared to indexing a sequence alone
-
Full-text index only
The human L-threonine 3-dehydrogenase gene is an expressed pseudogene.
PMID 12361482 · PMC131051 · BMC genetics · 2002 · 8 claims · 7 setups
The human TDH gene is located at chromosome 8p23-22, spans 10 kb, and has 8 exons that would be expected to encode a 369-residue ORF.
-
Full-text index only
Empirical Bayes analysis of quantitative proteomics experiments.
PMID 19829701 · PMC2759080 · PloS one · 2009 · 8 claims · 4 setups
Developed a new empirical Bayes framework that models log2 SILAC protein ratios and is robust to non-Gaussian tails and data sparsity, unlike Gaussian mixture models or Efron's original spline-based approach
-
Has reproduction · 88
Transcriptome-wide analyses of piRNA binding sites suggest distinct mechanisms regulate piRNA binding and silencing in C. elegans.
PMID 36737102 · PMC10158993 · RNA (New York, N.Y.) · 2023 · 8 claims · 7 setups
C. elegans piRNAs preferentially bind the coding regions (CDS) of target mRNAs in vivo, rather than 3' UTRs.
-
Full-text index only
The Functional RNA Database 3.0: databases to support mining and annotation of functional RNAs.
PMID 18948287 · PMC2686472 · Nucleic acids research · 2009 · 8 claims · 5 setups
fRNAdb 3.0 is a completely rebuilt sequence database hosting a much larger collection of known/predicted non-coding RNA sequences with improved search functionality
-
Has reproduction · 100
Conservation and losses of non-coding RNAs in avian genomes.
PMID 25822729 · PMC4378963 · PloS one · 2015 · 8 claims · 6 setups
34 lncRNA-associated loci are conserved between birds and mammals, and 12 of these were validated in chicken by RNA-seq.
-
Has reproduction · 74
An open RNA-Seq data analysis pipeline tutorial with an example of reprocessing data from a recent Zika virus study.
PMID 27583132 · PMC4972086 · F1000Research · 2016 · 6 claims · 6 setups
An open-source, reproducible RNA-seq pipeline delivered as an IPython notebook and Docker image can process raw RNA-seq data into interactive PCA/HC plots, enrichment results, and small-molecule predictions with minimal setup overhead
-
Full-text index only
GENCODE: producing a reference annotation for ENCODE.
PMID 16925838 · PMC1810553 · Genome biology · 2006 · 8 claims · 8 setups
GENCODE annotation combines initial manual annotation by HAVANA, experimental validation, and refinement based on results to identify protein-coding genes in ENCODE regions