Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 84
An accurate method for identifying recent recombinants from unaligned sequences.
PMID 35025988 · PMC8963311 · Bioinformatics (Oxford, England) · 2022 · 8 claims · 4 setups
A novel algorithm combining the JHMM (Zilversmit et al. 2013) mosaic representation with a distance-based triple comparison can identify recombinant sequences and their parents from unaligned, gene-length sequences without a reference panel.
-
Full-text index only
Evaluation of six methods for estimating synonymous and nonsynonymous substitution rates.
PMID 17127215 · PMC5054070 · Genomics, proteomics & bioinformatics · 2006 · 8 claims · 4 setups
Incorporating more sequence evolution features (transition/transversion bias, nucleotide/codon frequency bias) into Ka/Ks estimation methods yields more accurate and reliable estimates.
-
Full-text index only
Computing Ka and Ks with a consideration of unequal transitional substitutions.
PMID 16740169 · PMC1552089 · BMC evolutionary biology · 2006 · 7 claims · 7 setups
MYN, a modified version of the Yang-Nielsen (YN) algorithm based on the Tamura-Nei Model, allows unequal transitional substitution rates between purines (κR) and pyrimidines (κY) plus codon frequency bias
-
Full-text index only
A statistical approach designed for finding mathematically defined repeats in shotgun data and determining the length distribution of clone-inserts.
PMID 15626332 · PMC5172250 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 6 setups
Repeats of different copy number have distinct probabilities of appearance in shotgun data, which can be modeled statistically to define recognition thresholds (MDRs) at different shotgun coverages.
-
Full-text index only
Benchmarking tools for the alignment of functional noncoding DNA.
PMID 14736341 · PMC344529 · BMC bioinformatics · 2004 · 8 claims · 4 setups
Global alignment tools (Avid, ClustalW, Lagan, Needle, DiAlign-G) typically have higher sensitivity over entire noncoding sequences and within constrained blocks than local tools
-
Full-text index only
An empirical study of choosing efficient discriminative seeds for oligonucleotide design.
PMID 19958494 · PMC2788383 · BMC genomics · 2009 · 8 claims · 3 setups
The spaced seed is the most efficient discriminative seed for oligonucleotide design among the five algorithms tested.
-
Has reproduction · 83
MetaGT: A pipeline for de novo assembly of metatranscriptomes with the aid of metagenomic data.
PMID 36386613 · PMC9651917 · Frontiers in microbiology · 2022 · 7 claims · 4 setups
MetaGT is a pipeline that combines metatranscriptomic and metagenomic data from the same sample to assemble complete transcript sequences
-
Full-text index only
Direct maximum parsimony phylogeny reconstruction from genotype data.
PMID 18053244 · PMC2222657 · BMC bioinformatics · 2007 · 6 claims · 4 setups
The paper presents the first practical method for computing maximum parsimony phylogenies directly from genotype data, using integer linear programming.
-
Has reproduction · 78
IsomiR_Window: a system for analyzing small-RNA-seq data in an integrative and user-friendly manner.
PMID 33522913 · PMC7852101 · BMC bioinformatics · 2021 · 8 claims · 2 setups
IsomiR Window is an integrated, user-friendly platform that systematically identifies, quantifies, and functionally explores isomiR expression in small-RNA-seq datasets without requiring computational skills
-
Full-text index only
Evolutionary distance estimation and fidelity of pair wise sequence alignment.
PMID 15840174 · PMC1087827 · BMC bioinformatics · 2005 · 8 claims · 8 setups
Evolutionary distance estimation is relatively unaffected by alignment error as long as 50% or more of homologous sites remain identical between sequences
-
Has reproduction · 97
CRISPRbuilder-TB: "CRISPR-builder for tuberculosis". Exhaustive reconstruction of the CRISPR locus in mycobacterium tuberculosis complex using SRA.
PMID 33667225 · PMC7968741 · PLoS computational biology · 2021 · 8 claims · 7 setups
CRISPRbuilder-TB is a new pipeline that reconstructs MTC CRISPR-Cas loci directly from short SRA reads without requiring genome assembly
-
Full-text index only
A space-efficient and accurate method for mapping and aligning cDNA sequences onto genomic sequence.
PMID 18344523 · PMC2377433 · Nucleic acids research · 2008 · 7 claims · 6 setups
Spaln maps and aligns large cDNA sequence sets onto whole mammalian genomes using substantially less memory than comparable existing tools
-
Has reproduction · 94
Deep learning from phylogenies to uncover the epidemiological dynamics of outbreaks.
PMID 35794110 · PMC9258765 · Nature communications · 2022 · 8 claims · 5 setups
Deep learning (FFNN-SS and CNN-CBLV) enables accurate and fast likelihood-free estimation of epidemiological parameters and model selection from phylogenies
-
Full-text index only
Modeling the amplification dynamics of human Alu retrotransposons.
PMID 16201008 · PMC1239904 · PLoS computational biology · 2005 · 8 claims · 4 setups
Combining sequence diversity (π) and insertion polymorphism level (IPL) statistics can statistically exclude implausible Alu amplification scenarios and narrow the range of plausible ones for individual subfamilies.
-
Full-text index only
A computational study of off-target effects of RNA interference.
PMID 15800213 · PMC1072799 · Nucleic acids research · 2005 · 8 claims · 5 setups
The chance of RNAi off-target effects is considerable, ranging from 5% to 80% depending on organism and parameters, when using exact sequence identity between siRNA and transcripts.
-
Full-text index only
Multiplexed discovery of sequence polymorphisms using base-specific cleavage and MALDI-TOF MS.
PMID 15731331 · PMC549577 · Nucleic acids research · 2005 · 8 claims · 7 setups
Multiplexed base-specific cleavage/MALDI-TOF MS (Multiplexed Comparative Sequence Analysis) enables simultaneous discovery of sequence polymorphisms across multiple target regions
-
Full-text index only
Sequence analysis of p53 response-elements suggests multiple binding modes of the p53 tetramer to DNA targets.
PMID 17439973 · PMC1888811 · Nucleic acids research · 2007 · 8 claims · 5 setups
p53REs are not simple direct repeats of half-sites; the two half-sites couple to form a higher-order 20-bp full-site palindrome
-
Full-text index only
Effect of the assignment of ancestral CpG state on the estimation of nucleotide substitution rates in mammals.
PMID 18826599 · PMC2576242 · BMC evolutionary biology · 2008 · 7 claims · 4 setups
CpG/non-CpG assignment based on presence/absence of a CpG dinucleotide seriously biases substitution rate estimates, overestimating CpG changes and underestimating non-CpG changes.
-
Has reproduction · 87
Forseti: a mechanistic and predictive model of the splicing status of scRNA-seq reads.
PMID 38940130 · PMC11256924 · Bioinformatics (Oxford, England) · 2024 · 7 claims · 5 setups
Forseti is the first probabilistic model for resolving the splicing status of exonic scRNA-seq reads by scoring putative fragments linking read alignments to proximate priming sites
-
Full-text index only
Structural organization and interactions of transmembrane domains in tetraspanin proteins.
PMID 15985154 · PMC1190194 · BMC structural biology · 2005 · 8 claims · 5 setups
TM1, TM2 and TM3 of human tetraspanins display a distinct heptad repeat motif (abcdefg)n, while TM4 lacks this motif.