Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome-wide detection of segmental duplications and potential assembly errors in the human genome sequence.
PMID 12702206 · PMC154576 · Genome biology · 2003 · 8 claims · 6 setups
Segmental duplications comprise 3.53% (107.4/3,043.1 Mb) of the June 2002 human genome assembly
-
Has reproduction · 95
Comparative Genome Analysis of 16SrXII-A 'Candidatus Phytoplasma solani' POT Transmitted by Hyalesthes obsoletus.
PMID 41597744 · PMC12843639 · Microorganisms · 2026 · 7 claims · 8 setups
The complete 832,614 bp circular chromosome of the H. obsoletus-transmissible 'Ca. P. solani' 16SrXII-A strain POT was assembled and functionally reconstructed.
-
Full-text index only
Genetic affinities among the lower castes and tribal groups of India: inference from Y chromosome and mitochondrial DNA.
PMID 16893451 · PMC1569435 · BMC genetics · 2006 · 8 claims · 4 setups
Mitochondrial DNA shows no significant difference between Indian tribal and caste populations except higher frequency of west Eurasian-specific haplogroups in upper castes, especially in northwest India
-
Full-text index only
Sex-biased evolutionary forces shape genomic patterns of human diversity.
PMID 18818765 · PMC2538571 · PLoS genetics · 2008 · 7 claims · 5 setups
X-linked diversity is higher than the neutral expectation (0.75) relative to autosomal diversity in all six sampled human populations
-
Full-text index only
Versatile and open software for comparing large genomes.
PMID 14759262 · PMC395750 · Genome biology · 2004 · 8 claims · 8 setups
MUMmer 3.0 efficiently handles comparisons of large eukaryotic genomes at varying evolutionary distances
-
Has reproduction · 66
HTSstation: a web application and open-access libraries for high-throughput sequencing data analysis.
PMID 24475057 · PMC3903476 · PloS one · 2014 · 8 claims · 5 setups
HTSstation is a web application suite coupling simple web forms to modular analysis pipelines for ChIP-seq, RNA-seq, 4C-seq and re-sequencing HTS applications, accessible at http://htsstation.epfl.ch.
-
Full-text index only
Selecting additional tag SNPs for tolerating missing data in genotyping.
PMID 16259642 · PMC1316880 · BMC bioinformatics · 2005 · 7 claims · 6 setups
There exists a subset of SNPs (robust tag SNPs) that can distinguish all distinct haplotypes even when up to m SNPs are missing
-
Full-text index only
A Hidden Markov Model to estimate population mixture and allelic copy-numbers in cancers using Affymetrix SNP arrays.
PMID 17996079 · PMC2206057 · BMC bioinformatics · 2007 · 8 claims · 7 setups
An HMM using paired germline genotype calls and tumour allelic SNP intensities can estimate allele-specific copy-numbers, distinguishing events like uniparental disomy from allelic imbalance.
-
Full-text index only
A global view of protein expression in human cells, tissues, and organs.
PMID 20029370 · PMC2824494 · Molecular systems biology · 2009 · 7 claims · 6 setups
A high fraction (>65%) of proteins is expressed in most human cells and tissues, while very few proteins (<2%) are detected in any single cell type.
-
Full-text index only
Identifying cis-regulatory sequences by word profile similarity.
PMID 19730735 · PMC2731932 · PloS one · 2009 · 8 claims · 8 setups
WPH-finder identifies putative co-regulated CRMs by scanning the genome for sequences with word profiles similar to a known CRM, without explicitly defining binding sites
-
Full-text index only
The many uses of a genome sequence.
PMID 11423005 · PMC138940 · Genome biology · 2001 · 8 claims · 8 setups
Solved protein structures from structural genomics efforts can be used to model many other proteins by homology, aiding function prediction
-
Full-text index only
Direct inference of SNP heterozygosity rates and resolution of LOH detection.
PMID 18052545 · PMC2098867 · PLoS computational biology · 2007 · 6 claims · 7 setups
A large proportion of SNPs in dbSNP have high-variance HET rate estimates, limiting their reliability for LOH study design.
-
Has reproduction · 50
Polymorphism identification and improved genome annotation of Brassica rapa through Deep RNA sequencing.
PMID 25122667 · PMC4232532 · G3 (Bethesda, Md.) · 2014 · 8 claims · 8 setups
330,995 SNPs were identified in transcribed regions between B. rapa genotypes R500 and IMB211, at an average frequency of one SNP per 200 bases.
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
Using ESTs to improve the accuracy of de novo gene prediction.
PMID 16817966 · PMC1534067 · BMC bioinformatics · 2006 · 8 claims · 8 setups
TWINSCAN_EST combines EST alignments with TWINSCAN via a trainable 'ESTseq' representation and improves exact gene structure prediction accuracy on the whole C. elegans genome