Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Sequence similarity network reveals common ancestry of multidomain proteins.
PMID 18475320 · PMC2377100 · PLoS computational biology · 2008 · 8 claims · 6 setups
Traditional homology definitions do not capture multidomain evolution; the authors extend the definition to include domain insertion via a common ancestral locus model.
-
Full-text index only
iMapper: a web application for the automated analysis and mapping of insertional mutagenesis sequence data against Ensembl genomes.
PMID 18974167 · PMC2639305 · Bioinformatics (Oxford, England) · 2008 · 6 claims · 3 setups
iMapper is a web application for automated analysis and mapping of insertional mutagenesis sequence data against vertebrate and invertebrate Ensembl genomes (human, mouse, rat, zebrafish, Drosophila, S. cerevisiae).
-
Full-text index only
Prediction of catalytic residues using Support Vector Machine with selected protein sequence and structural properties.
PMID 16790052 · PMC1534064 · BMC bioinformatics · 2006 · 8 claims · 7 setups
The Sequential Minimal Optimization (SMO) SVM algorithm was the best-performing classifier among 26 WEKA classifiers for predicting catalytic residues
-
Full-text index only
An efficient method for the prediction of deleterious multiple-point mutations in the secondary structure of RNAs using suboptimal folding solutions.
PMID 18445289 · PMC2386494 · BMC bioinformatics · 2008 · 8 claims · 6 setups
Using RNAsubopt suboptimal solutions computed once for the wild-type sequence, specific multiple-point mutations likely to cause conformational rearrangement can be selected without brute-force enumeration.
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
Having a BLAST with bioinformatics (and avoiding BLASTphemy).
PMID 11597340 · PMC138974 · Genome biology · 2001 · 8 claims · 4 setups
BLAST is the most widely used tool for searching biological sequences for regions of local similarity
-
Full-text index only
A statistical approach designed for finding mathematically defined repeats in shotgun data and determining the length distribution of clone-inserts.
PMID 15626332 · PMC5172250 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 6 setups
Repeats of different copy number have distinct probabilities of appearance in shotgun data, which can be modeled statistically to define recognition thresholds (MDRs) at different shotgun coverages.
-
Full-text index only
Clustering of phosphorylation site recognition motifs can be exploited to predict the targets of cyclin-dependent kinase.
PMID 17316440 · PMC1852407 · Genome biology · 2007 · 8 claims · 6 setups
CDK consensus motifs are frequently clustered (closely spaced) in known CDK substrate proteins rather than uniformly distributed
-
Has reproduction · 83
Hobbes: optimized gram-based methods for efficient read alignment.
PMID 22199254 · PMC3315303 · Nucleic acids research · 2012 · 8 claims · 4 setups
Hobbes, a gram-based short-read mapper supporting Hamming and edit distance, is faster than all other read-mapping programs tested while maintaining high mapping quality.
-
Full-text index only
targetTB: a target identification pipeline for Mycobacterium tuberculosis through an interactome, reactome and genome-scale structural analysis.
PMID 19099550 · PMC2651862 · BMC systems biology · 2008 · 8 claims · 8 setups
A comprehensive in silico target identification pipeline (targetTB) integrating interactome, reactome, essentiality, sequence and structural analyses can identify high-confidence drug targets for Mtb
-
Has reproduction · 84
Strong population differentiation in lingcod (Ophiodon elongatus) is driven by a small portion of the genome.
PMID 33294007 · PMC7691466 · Evolutionary applications · 2020 · 7 claims · 8 setups
Lingcod comprise two distinct genetic clusters separated latitudinally at a break near Point Reyes off Northern California, with a high frequency of admixed individuals near the break.
-
Full-text index only
Comparative analysis of genome tiling array data reveals many novel primate-specific functional RNAs in human.
PMID 17288572 · PMC1796608 · BMC evolutionary biology · 2007 · 8 claims · 6 setups
Widespread transcription occurs across the human genome outside known gene annotations, and the bulk of TARs represent genuine transcripts rather than experimental artifacts
-
Full-text index only
Bushes in the tree of life.
PMID 17105342 · PMC1637082 · PLoS biology · 2006 · 8 claims · 8 setups
Bush-shaped clades, produced by short internal stems and long external branches, resist phylogenetic resolution regardless of the amount of conventional data collected.
-
Full-text index only
Inter-individual variation of DNA methylation and its implications for large-scale epigenome mapping.
PMID 18413340 · PMC2425484 · Nucleic acids research · 2008 · 8 claims · 8 setups
CpG-rich regions (CpG islands) show low and similar methylation levels across individuals, but the sequential order of the few methylated CpGs among the many unmethylated ones varies randomly between individuals.