Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Using multiple alignments to improve seeded local alignment algorithms.
PMID 16100379 · PMC1185574 · Nucleic acids research · 2005 · 8 claims · 2 setups
Using information implicit in a multiple alignment to dynamically build a spaced-seed index weighted toward promising regions increases sensitivity of local alignment search compared to indexing a sequence alone
-
Full-text index only
Serum diagnosis of diffuse large B-cell lymphomas and further identification of response to therapy using SELDI-TOF-MS and tree analysis patterning.
PMID 18163913 · PMC2242801 · BMC cancer · 2007 · 8 claims · 8 setups
SELDI-TOF-MS serum proteomic patterns analyzed by decision tree classification (Biomarker Pattern Software) can discriminate DLBCL patients from healthy controls with high sensitivity and specificity.
-
Full-text index only
Comparative genomic analysis of Campylobacter jejuni associated with Guillain-Barré and Miller Fisher syndromes: neuropathogenic and enteritis-associated isolates can share high levels of genomic similarity.
PMID 17919333 · PMC2174954 · BMC genomics · 2007 · 8 claims · 4 setups
GBS/MFS strains are genomically heterogeneous, falling into about six major lineages rather than a single clonal group
-
Full-text index only
Intrinsic structural disorder confers cellular viability on oncogenic fusion proteins.
PMID 19888473 · PMC2768585 · PLoS computational biology · 2009 · 8 claims · 5 setups
Translocation-related human proteins are significantly enriched in intrinsic structural disorder compared to all human proteins
-
Full-text index only
PedGenie: an analysis approach for genetic association testing in extended pedigrees and genealogies of arbitrary size.
PMID 16620382 · PMC1459209 · BMC bioinformatics · 2006 · 7 claims · 3 setups
PedGenie is a valid, flexible statistical tool for genetic association analysis in pedigrees of arbitrary size and structure using Monte Carlo significance testing
-
Full-text index only
Species-specific protein sequence and fold optimizations.
PMID 12487631 · PMC139977 · BMC bioinformatics · 2002 · 7 claims · 7 setups
Environmental niche is a significant factor explaining variability in amino acid composition across 100 complete genomes
-
Full-text index only
A model-based approach to selection of tag SNPs.
PMID 16776821 · PMC1525207 · BMC bioinformatics · 2006 · 7 claims · 5 setups
The Li and Stephens hidden Markov model outperforms other tested models (simple Markov, two-state HMM, HMM-4D, greedy GR-1/GR-2) in description code-length, tag set information content, and prediction of tagged SNPs.
-
Full-text index only
Network properties of complex human disease genes identified through genome-wide association studies.
PMID 19956617 · PMC2779513 · PloS one · 2009 · 7 claims · 6 setups
Complex disease genes are significantly less central (lower degree/closeness, higher eccentricity) in the human interactome than essential and monogenic disease genes, occupying an intermediate niche between monogenic disease genes and non-disease genes
-
Full-text index only
A high-temporal resolution technology for dynamic proteomic analysis based on 35S labeling.
PMID 18714357 · PMC2500177 · PloS one · 2008 · 6 claims · 6 setups
35S pulse-labeling (phosphor-imaging) detects more protein spots with higher sensitivity than CBB staining on the same 2-DE gels
-
Full-text index only
The genomic distribution of intraspecific and interspecific sequence divergence of human segmental duplications relative to human/chimpanzee chromosomal rearrangements.
PMID 18699995 · PMC2542386 · BMC genomics · 2008 · 8 claims · 5 setups
Some relatively recent (young) SDs accumulate in regions homologous to chromosomal inversions that occurred in the sister lineage
-
Full-text index only
Increased frequency of the k-ras G12C mutation in MYH polyposis colorectal adenomas.
PMID 15083190 · PMC2410274 · British journal of cancer · 2004 · 8 claims · 2 setups
Somatic k-ras mutations occur in 9 of 54 (16.7%) MYH polyposis colorectal tumours, with frequency increasing with degree of dysplasia.
-
Full-text index only
MiPred: classification of real and pseudo microRNA precursors using random forest prediction model with combined features.
PMID 17553836 · PMC1933124 · Nucleic acids research · 2007 · 8 claims · 8 setups
A hybrid feature combining local contiguous triplet structure-sequence composition, MFE of the secondary structure, and P-value of a randomization test improves classification of real vs pseudo pre-miRNAs
-
Has reproduction · 71
Sustainable data analysis with Snakemake.
PMID 34035898 · PMC8114187 · F1000Research · 2021 · 8 claims · 4 setups
Reproducibility alone is insufficient for sustainable data analysis; transparency and adaptability are equally important additional properties.
-
Full-text index only
SeqDoC: rapid SNP and mutation detection by direct comparison of DNA sequence chromatograms.
PMID 15927052 · PMC1156871 · BMC bioinformatics · 2005 · 8 claims · 6 setups
SeqDoC generates a subtracted difference trace between a reference and test chromatogram that highlights single base changes
-
Full-text index only
Computational comparison of two mouse draft genomes and the human golden path.
PMID 12537546 · PMC151282 · Genome biology · 2003 · 8 claims · 7 setups
The Celera and public mouse genome assemblies differ in about 10% of the mouse genome, with complementary strengths (Celera higher base-pair accuracy and overall coverage; public assembly higher quality in some finished BAC regions and freely accessible)
-
Full-text index only
Unusual linkage patterns of ligands and their cognate receptors indicate a novel reason for non-random gene order in the human genome.
PMID 16277660 · PMC1309615 · BMC evolutionary biology · 2005 · 8 claims · 5 setups
Ligands are not more closely linked (shorter physical distance) to their cognate receptors than expected by chance
-
Full-text index only
A comparison of random sequence reads versus 16S rDNA sequences for estimating the biodiversity of a metagenomic library.
PMID 18682527 · PMC2532719 · Nucleic acids research · 2008 · 8 claims · 7 setups
Biodiversity observed by RSR analysis is consistent with that obtained by 16S rDNA analysis
-
Full-text index only
Next-generation high-density self-assembling functional protein arrays.
PMID 18469824 · PMC3070491 · Nature methods · 2008 · 8 claims · 7 setups
A next-generation NAPPA method produces high-density protein microarrays displaying over 1500 unique proteins with >90% expression success
-
Full-text index only
Genomic mutation rates: what high-throughput methods can tell us.
PMID 19644920 · PMC2952423 · BioEssays : news and reviews in molecular, cellular and developmental biology · 2009 · 8 claims · 8 setups
High-throughput DNA analyses yield genome mutation rate estimates markedly higher than those obtained with pre-genomic strategies
-
Has reproduction · 87
Enhanced Generalizability of RNA Secondary Structure Prediction via Convolutional Block Attention Network and Ensemble Learning.
PMID 40871599 · PMC12388828 · Molecules (Basel, Switzerland) · 2025 · 8 claims · 8 setups
TrioFold integrates base-pairing clues from thermodynamic- and DL-based methods via ensemble learning and a convolutional block attention mechanism to enhance RSS prediction generalizability.