Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genome comparison without alignment using shortest unique substrings.
PMID 15910684 · PMC1166540 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A number of sequence comparison tasks, including detection of unique genomic regions, can be accomplished efficiently without an alignment step using shortest unique substrings.
-
Full-text index only
SW-ARRAY: a dynamic programming solution for the identification of copy-number changes in genomic DNA using array comparative genome hybridization data.
PMID 15961730 · PMC1151590 · Nucleic acids research · 2005 · 7 claims · 5 setups
SW-ARRAY, an adaptation of the Smith-Waterman dynamic programming algorithm, provides a sensitive and robust method for identifying copy-number changes in array CGH data
-
Full-text index only
Large-scale analysis of human alternative protein isoforms: pattern classification and correlation with subcellular localization signals.
PMID 15860772 · PMC1087780 · Nucleic acids research · 2005 · 8 claims · 8 setups
Constructed a large-scale dataset of 6876 human alternative protein isoforms from 2624 genes by combining H-Invitational full-length cDNA data and SwissProt VARSPLIC entries
-
Full-text index only
Insights into the coupling of duplication events and macroevolution from an age profile of animal transmembrane gene families.
PMID 16895434 · PMC1534073 · PLoS computational biology · 2006 · 8 claims · 7 setups
The density of transmembrane gene duplicates positively correlates with the estimated maximum number of cell types of common ancestors
-
Full-text index only
Predicting failure rate of PCR in large genomes.
PMID 18492719 · PMC2441781 · Nucleic acids research · 2008 · 7 claims · 8 setups
The number of predicted primer-binding sites in genomic DNA is the most important factor determining PCR failure.
-
Full-text index only
Genotyping of genetically monomorphic bacteria: DNA sequencing in Mycobacterium tuberculosis highlights the limitations of current methodologies.
PMID 19915672 · PMC2772813 · PloS one · 2009 · 8 claims · 8 setups
MLSA of 89 genes across 108 global MTBC strains yields a single, highly robust phylogeny with virtually no homoplasy, congruent across parsimony, NJ, ML, and Bayesian methods.
-
Has reproduction · 58
Mucospheres produced by a mixotrophic protist impact ocean carbon cycling.
PMID 35288549 · PMC8921327 · Nature communications · 2022 · 8 claims · 8 setups
P. cf. balticum constructs carbon-rich mucospheres that attract, capture and immobilise microbial prey to facilitate phago-heterotrophic consumption
-
Full-text index only
Implementation of a data repository-driven approach for targeted proteomics experiments by multiple reaction monitoring.
PMID 19121650 · PMC2706936 · Journal of proteomics · 2009 · 7 claims · 5 setups
A new MRM worksheet was implemented in The Global Proteome Machine database (GPMDB) that provides all information needed to design MRM transitions based solely on archived observations from previous experiments by other researchers.
-
Full-text index only
Probing the cancer genome.
PMID 18492227 · PMC2441462 · Genome biology · 2008 · 8 claims · 8 setups
Combined Sanger and 454 pyrosequencing of MCF-7 BAC clones identified 157 PCR-confirmed translocation breakpoint junctions, including 10 in-frame junctions confirmed at the transcript level
-
Full-text index only
Optimal step length EM algorithm (OSLEM) for the estimation of haplotype frequency and its application in lipoprotein lipase genotyping.
PMID 12529185 · PMC149347 · BMC bioinformatics · 2003 · 5 claims · 4 setups
OSLEM (Optimal Step Length EM), which approximates an optimal step length via a fixed-point search (D_N = D_{N-1} + λ(D_preN - D_{N-1})), runs about twice as fast as standard EM while producing the same haplotype frequency estimates.
-
Full-text index only
A novel wavelet-based thresholding method for the pre-processing of mass spectrometry data that accounts for heterogeneous noise.
PMID 18615428 · PMC2855839 · Proteomics · 2008 · 6 claims · 4 setups
Noise in SELDI-TOF/MALDI-TOF mass spectrometry data is heteroscedastic across the m/z range, with larger variance at lower m/z values, contrary to the homogeneous noise assumption of existing wavelet denoising methods.
-
Full-text index only
Genomic and bioinformatics analysis of human adenovirus type 37: new insights into corneal tropism.
PMID 18471294 · PMC2397415 · BMC genomics · 2008 · 7 claims · 7 setups
The complete genome of HAdV-37 was sequenced and annotated (35,213 bp, 56.6% GC content, 35 predicted coding sequences plus 8 hypothetical ORFs)
-
Full-text index only
Detecting purely epistatic multi-locus interactions by an omnibus permutation test on ensembles of two-locus analyses.
PMID 19761607 · PMC2759961 · BMC bioinformatics · 2009 · 8 claims · 5 setups
2LOmb performs an omnibus permutation test on ensembles of two-locus analyses via a four-step algorithm (two-locus analysis, permutation test, global p-value determination, progressive ensemble search)
-
Full-text index only
Nucleotide-resolution analysis of structural variants using BreakSeq and a breakpoint library.
PMID 20037582 · PMC2951730 · Nature biotechnology · 2010 · 8 claims · 7 setups
A standardized, non-redundant library of 1,889 breakpoint-resolved SVs was assembled from eight published surveys
-
Full-text index only
Assessing the genomic evidence for conserved transcribed pseudogenes under selection.
PMID 19754956 · PMC2753554 · BMC genomics · 2009 · 8 claims · 8 setups
1750 transcribed pseudogene annotations (TPAs) were identified in the human genome, ~11.5% of all human pseudogene annotations.
-
Has reproduction · 86
LMAS: evaluating metagenomic short de novo assembly methods through defined communities.
PMID 36576131 · PMC9795473 · GigaScience · 2022 · 8 claims · 5 setups
LMAS (Last Metagenomic Assembler Standing) is a flexible, Nextflow-based, Docker-containerized automated workflow for benchmarking de novo metagenomic assemblers against defined mock communities, producing an interactive HTML report.
-
Has reproduction · 30
Minimal metabolic pathway structure is consistent with associated biomolecular interactions.
PMID 24987116 · PMC4299494 · Molecular systems biology · 2014 · 8 claims · 8 setups
MinSpan, a mixed-integer linear optimization algorithm, computes the shortest, linearly independent pathways (sparsest basis of the null space of the stoichiometric matrix S) for genome-scale metabolic networks, which convex approaches (extreme pathways, elementary flux modes) cannot do at genome scale.
-
Full-text index only
The Bifidobacterium dentium Bd1 genome sequence reflects its genetic adaptation to the human oral cavity.
PMID 20041198 · PMC2788695 · PLoS genetics · 2009 · 8 claims · 8 setups
The B. dentium Bd1 genome was sequenced to completion, revealing a single circular 2,636,368 bp chromosome with 2,143 predicted ORFs
-
Full-text index only
Comprehensive genomic analysis reveals clinically relevant molecular distinctions between thymic carcinomas and thymomas.
PMID 19861435 · PMC2783876 · Clinical cancer research : an official journal of the American Association for Cancer Research · 2009 · 7 claims · 7 setups
Comprehensive genomic analysis shows thymic carcinomas are molecularly distinct from thymomas
-
Has reproduction · 61
A comprehensive resource of genomic, epigenomic and transcriptomic sequencing data for the black truffle Tuber melanosporum.
PMID 25392735 · PMC4228822 · GigaScience · 2014 · 8 claims · 8 setups
T. melanosporum shows a high rate of cytosine methylation (>44%) that selectively targets transposable elements rather than genes, with a strong preference for CpG sites.