Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 93
aCLImatise: automated generation of tool definitions for bioinformatics workflows.
PMID 33325479 · PMC8016486 · Bioinformatics (Oxford, England) · 2021 · 6 claims · 3 setups
aCLImatise automatically generates workflow-language tool definitions by parsing a command-line tool's help output
-
Has reproduction · 100
Integrative transcriptome sequencing identifies trans-splicing events with important roles in human embryonic stem cell pluripotency.
PMID 24131564 · PMC3875859 · Genome research · 2014 · 8 claims · 8 setups
TSscan, a computational pipeline integrating long- and short-read transcriptome sequencing from multiple hESC lines, can detect trans-splicing while minimizing false positives from experimental artifacts and genetic rearrangements.
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
Grammar-based distance in progressive multiple sequence alignment.
PMID 18616828 · PMC2478692 · BMC bioinformatics · 2008 · 7 claims · 3 setups
A grammar-based (LZ complexity) distance metric can be used to determine the order in which sequences are progressively pairwise aligned
-
Full-text index only
Genome annotation errors in pathway databases due to semantic ambiguity in partial EC numbers.
PMID 16034025 · PMC1179732 · Nucleic acids research · 2005 · 7 claims · 4 setups
Partial EC numbers are semantically ambiguous, and databases that assign a gene to all reactions sharing the same partial EC number make a faulty inference, causing systematic misannotation.
-
Full-text index only
FeatureScan: revealing property-dependent similarity of nucleotide sequences.
PMID 16845077 · PMC1538849 · Nucleic acids research · 2006 · 6 claims · 5 setups
FeatureScan transforms nucleotide sequences into numerical signals of physico-chemical/conformational properties and compares them via a convolution/correlation (Fourier transform) method rather than comparing letters
-
Full-text index only
The association of Alu repeats with the generation of potential AU-rich elements (ARE) at 3' untranslated regions.
PMID 15610565 · PMC544599 · BMC genomics · 2004 · 6 claims · 4 setups
Alu repeats are a source of AREs at 3' UTRs of human mRNA, via poly-A regions of Alu generating complementary poly-T/poly-U regions that acquire regular adenine insertions to form ARE motifs.
-
Full-text index only
Adaptive discriminant function analysis and reranking of MS/MS database search results for improved peptide identification in shotgun proteomics.
PMID 18788775 · PMC3744223 · Journal of proteome research · 2008 · 7 claims · 4 setups
PeptideProphet's fixed LDA coefficients for combining search scores (Xcorr', ΔCn, SpRank) may not be optimal under all search/instrument conditions.
-
Full-text index only
Identifying the important HIV-1 recombination breakpoints.
PMID 18787691 · PMC2522274 · PLoS computational biology · 2008 · 8 claims · 3 setups
Local sequence identity between co-packaged parental RNAs strongly influences the probability of strand-transfer/breakpoint location, with fewer breakpoints occurring near mismatches
-
Has reproduction · 86
RNASEQR--a streamlined and accurate RNA-seq sequence analysis program.
PMID 22199257 · PMC3315322 · Nucleic acids research · 2012 · 8 claims · 7 setups
RNASEQR is a new RNA-seq mapper/aligner that combines a BWT-based (Bowtie) transcriptomic/genomic alignment with hash-based BLAT local alignment in three sequential steps: transcriptome mapping, novel exon detection, and anchor-and-align novel splice junction identification.
-
Has reproduction · 42
The electrostatic profile of consecutive Cβ atoms applied to protein structure quality assessment.
PMID 25506420 · PMC4257144 · F1000Research · 2013 · 8 claims · 8 setups
The EPD between Cβ atoms of consecutive residues provides unique signatures of amino acid pair types and can discriminate native from decoy protein structures.
-
Has reproduction · 67
Cyrface: An interface from Cytoscape to R that provides a user interface to R packages.
PMID 24715956 · PMC3962008 · F1000Research · 2013 · 8 claims · 6 setups
Cyrface is a Cytoscape app/Java library providing a general interface from Cytoscape (Java) to any R function or package.
-
Full-text index only
A computational screen for type I polyketide synthases in metagenomics shotgun data.
PMID 18953415 · PMC2568958 · PloS one · 2008 · 8 claims · 6 setups
Combining HMM domain searches with maximum-likelihood phylogenetic trees can discriminate true PKS I sequences from evolutionarily related but functionally different enzymes (e.g., FAS I) in metagenomic data.
-
Full-text index only
An XML standard for the dissemination of annotated 2D gel electrophoresis data complemented with mass spectrometry results.
PMID 15005801 · PMC341449 · BMC bioinformatics · 2004 · 7 claims · 3 setups
An XML schema called Annotated Gel Markup Language (AGML) is proposed to manage, analyze, and disseminate annotated 2D gel electrophoresis and MS results.