Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Optimal step length EM algorithm (OSLEM) for the estimation of haplotype frequency and its application in lipoprotein lipase genotyping.
PMID 12529185 · PMC149347 · BMC bioinformatics · 2003 · 5 claims · 4 setups
OSLEM (Optimal Step Length EM), which approximates an optimal step length via a fixed-point search (D_N = D_{N-1} + λ(D_preN - D_{N-1})), runs about twice as fast as standard EM while producing the same haplotype frequency estimates.
-
Full-text index only
Genome-wide identification of specific oligonucleotides using artificial neural network and computational genomic analysis.
PMID 17518996 · PMC1892811 · BMC bioinformatics · 2007 · 7 claims · 4 setups
The IAB algorithm (integration of ANN and BLAST) identifies genome-wide specific oligos much faster than pure BLAST search while maintaining comparable success rate and cross homology
-
Full-text index only
SNP haplotype tagging from DNA pools of two individuals.
PMID 12709267 · PMC156884 · BMC bioinformatics · 2003 · 8 claims · 3 setups
An algorithm can reconstruct haplotypes from pools of two individuals' DNA under very general conditions, without requiring Hardy-Weinberg equilibrium.
-
Has reproduction · 70
GAUGE-Annotated Microbial Transcriptomic Data Facilitate Parallel Mining and High-Throughput Reanalysis To Form Data-Driven Hypotheses.
PMID 33758032 · PMC8547006 · mSystems · 2021 · 7 claims · 6 setups
GAUGE automatically annotates GEO microbial microarray and RNA-seq data sets, increasing the percentage amenable to analysis from 4% to 33%.
-
Has reproduction · 67
binny: an automated binning algorithm to recover high-quality genomes from complex metagenomic datasets.
PMID 36239393 · PMC9677464 · Briefings in bioinformatics · 2022 · 8 claims · 8 setups
binny outperforms or is highly competitive with commonly used and state-of-the-art binning methods (MetaBAT2, MaxBin2, CONCOCT, VAMB, SemiBin, MetaDecoder)
-
Has reproduction · 85
Digital sorting of complex tissues for cell type-specific gene expression profiles.
PMID 23497278 · PMC3626856 · BMC bioinformatics · 2013 · 8 claims · 8 setups
The Digital Sorting Algorithm (DSA) deconvolves mixed tissue expression into cell type-specific profiles using only marker genes, without requiring prior knowledge of cell type frequencies or in vitro pure-cell profiles.
-
Full-text index only
Reconstructing the genomic architecture of mammalian ancestors using multispecies comparative maps.
PMID 15601531 · PMC3525001 · Human genomics · 2003 · 8 claims · 4 setups
The MGR algorithm applied to human, mouse, cat and cattle comparative maps can impute an ancestral mammalian genome composed of conserved segments.
-
Full-text index only
PPC: an algorithm for accurate estimation of SNP allele frequencies in small equimolar pools of DNA using data from high density microarrays.
PMID 16199750 · PMC1240117 · Nucleic acids research · 2005 · 7 claims · 6 setups
The PPC algorithm, which applies a probe-pair-specific second-degree polynomial correction, increases the accuracy of allele frequency estimates from pooled DNA compared with previously described algorithms
-
Full-text index only
Transcriptome annotation using tandem SAGE tags.
PMID 17709346 · PMC2034470 · Nucleic acids research · 2007 · 8 claims · 7 setups
A novel algorithm pairs tandem SAGE tags anchored on two different restriction sites (CATG and GATC) to define tag-delimited genomic sequences (TDGS)
-
Full-text index only
Computer-aided identification of polymorphism sets diagnostic for groups of bacterial and viral genetic variants.
PMID 17672919 · PMC1973086 · BMC bioinformatics · 2007 · 6 claims · 8 setups
The Not-N algorithm, incorporated into the Minimum SNPs program, identifies small marker sets diagnostic for user-defined subgroups of genetic variants with 0% false negatives
-
Full-text index only
Decision tree-driven tandem mass spectrometry for shotgun proteomics.
PMID 18931669 · PMC2597439 · Nature methods · 2008 · 8 claims · 5 setups
A decision tree (DT) algorithm that selects CAD or ETD per precursor based on z and m/z yields more peptide identifications than either CAD or ETD alone
-
Full-text index only
Assessing batch effects of genotype calling algorithm BRLMM for the Affymetrix GeneChip Human Mapping 500 K array set using 270 HapMap samples.
PMID 18793462 · PMC2537568 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Batch size affects genotype calling results (call rate and concordance) and the resulting lists of significantly associated SNPs.
-
Full-text index only
Evaluation of two methods for computational HLA haplotypes inference using a real dataset.
PMID 18230173 · PMC2268655 · BMC bioinformatics · 2008 · 8 claims · 5 setups
PHASE v2.1.1 had the best overall performance in both haplotype construction and frequency calculation compared to Arlequin V3.0
-
Full-text index only
Local combinational variables: an approach used in DNA-binding helix-turn-helix motif prediction with sequence information.
PMID 19651875 · PMC2761287 · Nucleic acids research · 2009 · 8 claims · 7 setups
The LCV approach predicts HTH motifs with 93.29% accuracy, 93.93% sensitivity and 92.66% specificity using only primary sequence information
-
Full-text index only
ASPIC: a web resource for alternative splicing prediction and transcript isoforms characterization.
PMID 16845044 · PMC1538898 · Nucleic acids research · 2006 · 8 claims · 2 setups
The ASPIC algorithm, using an optimization procedure that minimizes splice site predictions and transcript isoforms from multiple EST-genome alignments, outperforms other similar AS-prediction tools in sensitivity and selectivity
-
Full-text index only
htSNPer1.0: software for haplotype block partition and htSNPs selection.
PMID 15740612 · PMC1274247 · BMC bioinformatics · 2005 · 6 claims · 1 setups
The GBB algorithm finds the globally optimal minimal htSNP set with far less computing time than exhaustive/enumeration search.
-
Full-text index only
Evolutionary algorithms for the selection of single nucleotide polymorphisms.
PMID 12875658 · PMC183839 · BMC bioinformatics · 2003 · 8 claims · 3 setups
Evolutionary algorithms are well suited to multiobjective optimization problems with large, intractable search spaces such as SNP selection, unlike exact methods (exhaustive enumeration) or single-objective search techniques (tabu search, simulated annealing).
-
Full-text index only
SNP-RFLPing: restriction enzyme mining for SNPs in genomes.
PMID 16503968 · PMC1386656 · BMC genomics · 2006 · 8 claims · 2 setups
SNP-RFLPing accepts three flexible input types (dbSNP rs#/ss# IDs, HUGO gene name/Entrez gene ID, or free-form SNP-in-sequence including IUPAC or [dNTP1/dNTP2] formats) for human, rat, and mouse genomes
-
Has reproduction · 51
A platelet-related signature for predicting the prognosis and immunotherapy benefit in bladder cancer based on machine learning combinations.
PMID 39280688 · PMC11399026 · Translational andrology and urology · 2024 · 8 claims · 8 setups
An Enet (alpha=0.4) machine-learning algorithm built from 10 platelet-related genes yields the optimal platelet-related signature (PRS) for bladder cancer prognosis, with average C-index 0.73
-
Full-text index only
Proteomics: characterizing the cogs in the machinery of life.
PMID 14630521 · PMC1241753 · Environmental health perspectives · 2003 · 8 claims · 5 setups
Protein expression patterns in blood serum, detected via SELDI-TOF mass spectrometry and analyzed with a genetic algorithm, can distinguish ovarian cancer patients from healthy individuals with very high sensitivity and specificity.