Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
FeatureScan: revealing property-dependent similarity of nucleotide sequences.
PMID 16845077 · PMC1538849 · Nucleic acids research · 2006 · 6 claims · 5 setups
FeatureScan transforms nucleotide sequences into numerical signals of physico-chemical/conformational properties and compares them via a convolution/correlation (Fourier transform) method rather than comparing letters
-
Has reproduction · 76
What the Phage: a scalable workflow for the identification and analysis of phage sequences.
PMID 36399058 · PMC9673492 · GigaScience · 2022 · 8 claims · 7 setups
WtP combines 11 tools (14 approaches) for phage prediction in a parallel, containerized Nextflow workflow
-
Full-text index only
Sequence analysis of p53 response-elements suggests multiple binding modes of the p53 tetramer to DNA targets.
PMID 17439973 · PMC1888811 · Nucleic acids research · 2007 · 8 claims · 5 setups
p53REs are not simple direct repeats of half-sites; the two half-sites couple to form a higher-order 20-bp full-site palindrome
-
Full-text index only
'Chumanzee' evolution: the urge to diverge and merge.
PMID 17129363 · PMC1794591 · Genome biology · 2006 · 8 claims · 3 setups
Human-chimpanzee divergence was not a simple clean split; evidence suggests hybridization continued after an initial split.
-
Has reproduction · 44
Detecting DNA modifications from SMRT sequencing data by modeling sequence context dependence of polymerase kinetic.
PMID 23516341 · PMC3597545 · PLoS computational biology · 2013 · 8 claims · 7 setups
Local sequence context strongly determines position-specific polymerase kinetic rate: roughly 80% of IPD variation is explained by a 10 bp context (7 bases upstream, 2 bases downstream of the incorporation site), saturating at 7 bases upstream.
-
Full-text index only
The sequence and de novo assembly of the giant panda genome.
PMID 20010809 · PMC3951497 · Nature · 2010 · 8 claims · 8 setups
A draft giant panda genome was successfully generated and assembled de novo using only Illumina Genome Analyser short-read sequencing
-
Full-text index only
Improved tagging strategy for protein identification in mammalian cells.
PMID 16138932 · PMC1250225 · BMC genomics · 2005 · 7 claims · 7 setups
EGFP tagging via the artificial exon did not affect the subcellular localization of the tagged endogenous proteins
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
INDELSCAN: a web server for comparative identification of species-specific and non-species-specific insertion/deletion events.
PMID 17517762 · PMC1933116 · Nucleic acids research · 2007 · 8 claims · 3 setups
Pair-wise sequence alignment-based indel identification lacks discrimination of species specificity and cannot distinguish insertions from deletions.
-
Full-text index only
In silico segmentations of lentivirus envelope sequences.
PMID 17376229 · PMC1847453 · BMC bioinformatics · 2007 · 8 claims · 8 setups
C and V regions of lentivirus SU sequences have distinct statistical (oligonucleotide/amino-acid) compositions that HMMs can learn and use to delimit them.
-
Full-text index only
Characterization of 954 bovine full-CDS cDNA sequences.
PMID 16305752 · PMC1314900 · BMC genomics · 2005 · 7 claims · 8 setups
954 bovine full-length insert cDNA (bFLIC) clones representing 762 distinct loci were sequenced and characterized
-
Full-text index only
In vitro and in silico analysis reveals an efficient algorithm to predict the splicing consequences of mutations at the 5' splice sites.
PMID 17726045 · PMC2094079 · Nucleic acids research · 2007 · 8 claims · 6 setups
Two exonic mutations, PINK1 E417G and PARK7 E64D, disrupt binding to U1 snRNA and cause skipping of the mutation-harboring exon
-
Full-text index only
Decoding of superimposed traces produced by direct sequencing of heterozygous indels.
PMID 18654614 · PMC2429969 · PLoS computational biology · 2008 · 7 claims · 3 setups
A dynamic programming method (implemented as web app Indelligent) can decode superimposed allelic sequences from a single mixed trace, using only the observed string of ambiguous peak calls, without a reference sequence or reverse trace.
-
Full-text index only
Repeating patterns of mimicry.
PMID 17048984 · PMC1617347 · PLoS biology · 2006 · 7 claims · 4 setups
The Yb locus controls presence of a yellow wing band in H. melpomene
-
Full-text index only
Identification of serum biomarkers for colon cancer by proteomic analysis.
PMID 16755300 · PMC2361335 · British journal of cancer · 2006 · 8 claims · 8 setups
Complement C3a des-arg, α1-antitrypsin and transferrin were identified as serum proteins with diagnostic potential for CRC.
-
Full-text index only
A high-throughput method for quantifying alleles and haplotypes of the malaria vaccine candidate Plasmodium falciparum merozoite surface protein-1 19 kDa.
PMID 16626494 · PMC1459863 · Malaria journal · 2006 · 8 claims · 5 setups
Pyrosequencing, after adjustment to a standard curve, provides accurate and precise estimates of allele frequencies in mixed MSP-1_19 infections
-
Full-text index only
Towards alignment independent quantitative assessment of homology detection.
PMID 17205117 · PMC1762415 · PloS one · 2006 · 8 claims · 6 setups
The Fhom Estimator uses the prevalence of a conserved protein feature (X) in two protein sets to estimate the fraction of true homologs among paired proteins, independent of alignment quality.
-
Full-text index only
Comparative genomic study reveals a transition from TA richness in invertebrates to GC richness in vertebrates at CpG flanking sites: an indication for context-dependent mutagenicity of methylated CpG sites.
PMID 19329065 · PMC5054122 · Genomics, proteomics & bioinformatics · 2008 · 8 claims · 8 setups
Nucleotide preference at CpG flanking sites transitions from 5' T (invertebrates) to 5' A (vertebrates) at the invertebrate-vertebrate boundary