Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
GeneKeyDB: a lightweight, gene-centric, relational database to support data mining environments.
PMID 15790402 · PMC1274265 · BMC bioinformatics · 2005 · 8 claims · 6 setups
GeneKeyDB is a lightweight, gene-centric relational database that supports data mining and integration with computational analysis tools.
-
Full-text index only
Web services and workflow management for biological resources.
PMID 16351751 · PMC1866383 · BMC bioinformatics · 2005 · 8 claims · 4 setups
Workflow management systems combined with Web Services are a promising ICT approach for automating access to and integration of biomedical data.
-
Full-text index only
Gene Prospector: an evidence gateway for evaluating potential susceptibility genes and interacting risk factors for human diseases.
PMID 19063745 · PMC2613935 · BMC bioinformatics · 2008 · 8 claims · 5 setups
Gene Prospector is a Web-based application that selects and prioritizes potential disease-related genes using a curated, updated literature database of genetic association studies
-
Full-text index only
BIPASS: BioInformatics Pipeline Alternative Splicing Services.
PMID 17584795 · PMC1933140 · Nucleic acids research · 2007 · 8 claims · 4 setups
BIPASS offers two complementary services for alternative splicing (AS) research: BIPAS-SpliceDB, a queryable pre-computed AS data warehouse, and BIPAS-Align&Splice, an online pipeline for user-submitted sequences.
-
Has reproduction · 74
Wide-Open: Accelerating public data release by automating detection of overdue datasets.
PMID 28594819 · PMC5464523 · PLoS biology · 2017 · 6 claims · 5 setups
A general text-mining + API-query approach (Wide-Open) can automatically identify datasets that are overdue for public release in a repository
-
Full-text index only
Natural history of S-adenosylmethionine-binding proteins.
PMID 16225687 · PMC1282579 · BMC structural biology · 2005 · 8 claims · 6 setups
The last universal common ancestor (LUCA) of cellular life had between 10 and 20 SAM-binding proteins from at least 5 fold classes
-
Full-text index only
RotaC: a web-based tool for the complete genome classification of group A rotaviruses.
PMID 19930627 · PMC2785824 · BMC microbiology · 2009 · 7 claims · 4 setups
RotaC is a freely available web-based tool for complete genome classification of group A rotaviruses across all 11 gene segments.
-
Has reproduction · 87
R2DT is a framework for predicting and visualising RNA secondary structure using templates.
PMID 34108470 · PMC8190129 · Nature communications · 2021 · 8 claims · 6 setups
R2DT is a template-based computational framework/pipeline that predicts and visualises RNA 2D structure in standardised, community-accepted layouts
-
Has reproduction · 60
TRAPID 2.0: a web application for taxonomic and functional analysis of de novo transcriptomes.
PMID 34197621 · PMC8464036 · Nucleic acids research · 2021 · 8 claims · 8 setups
TRAPID 2.0 is a web application performing global characterization of de novo transcriptomes via structural, functional, and taxonomic annotation in an initial processing phase, followed by an exploratory phase of downstream analyses.
-
Has reproduction · 30
taxize: taxonomic search and retrieval in R.
PMID 24555091 · PMC3901538 · F1000Research · 2013 · 8 claims · 8 setups
taxize is an open-source R package (on CRAN) giving simple programmatic access to taxonomic data from 13 web data sources.
-
Full-text index only
Finding disease candidate genes by liquid association.
PMID 17915034 · PMC2246280 · Genome biology · 2007 · 7 claims · 6 setups
LA can detect functionally associated genes that are not directly co-expressed by identifying a mediating gene Z whose expression level changes the correlation between X and Y.
-
Full-text index only
Cataloging coding sequence variations in human genome databases.
PMID 18974781 · PMC2570488 · PloS one · 2008 · 8 claims · 7 setups
A significant proportion of CVs overlap between HGMD and dbSNP (4.36% of HGMD CVs registered in dbSNP; 8.11% of dbSNP CVs registered in HGMD), warranting caution when interpreting phenotypic relevance of concurrent CVs.
-
Full-text index only
PPDPF is not a key regulator of human pancreas development.
PMID 40193385 · PMC12037078 · PLoS genetics · 2025 · 8 claims · 8 setups
PPDPF is not a key regulator of human pancreas development, in contrast to its zebrafish orthologue exdpf
-
Has reproduction · 68
Cell-type annotation with accurate unseen cell-type identification using multiple references.
PMID 37379341 · PMC10335708 · PLoS computational biology · 2023 · 8 claims · 4 setups
mtANN integrates multiple reference datasets and eight gene selection methods via ensemble learning (multiple deep classification models + majority voting) to improve cell-type annotation accuracy
-
Has reproduction · 71
RNAmountAlign: Efficient software for local, global, semiglobal pairwise and multiple RNA sequence/structure alignment.
PMID 31978147 · PMC6980424 · PloS one · 2020 · 7 claims · 6 setups
RNAmountAlign performs pairwise local, global, and semiglobal (query search) alignment and progressive multiple alignment (global and local) using incremental ensemble mountain height, running in O(n^3) time and O(n^2) space for two sequences of length n
-
Full-text index only
Genome annotation errors in pathway databases due to semantic ambiguity in partial EC numbers.
PMID 16034025 · PMC1179732 · Nucleic acids research · 2005 · 7 claims · 4 setups
Partial EC numbers are semantically ambiguous, and databases that assign a gene to all reactions sharing the same partial EC number make a faulty inference, causing systematic misannotation.
-
Full-text index only
Peptide bioinformatics: peptide classification using peptide machines.
PMID 19065810 · PMC7122642 · Methods in molecular biology (Clifton, N.J.) · 2008 · 8 claims · 4 setups
The bio-basis function, which converts peptides into numerical vectors using nongapped pairwise homology alignment scores against indicator peptides, can statistically quantify peptide similarity for classification.
-
Full-text index only
Identification and recovery of minor HIV-1 variants using the heteroduplex tracking assay and biotinylated probes.
PMID 18948297 · PMC2602764 · Nucleic acids research · 2008 · 6 claims · 8 setups
Incorporating a biotin tag into the HTA probe enables purification of labeled heteroduplexes and direct sequencing of the separated query strand, allowing recovery of minor variant sequences
-
Full-text index only
SPSmart: adapting population based SNP genotype databases for fast and comprehensive web access.
PMID 18847484 · PMC2576268 · BMC bioinformatics · 2008 · 7 claims · 8 setups
SPSmart is a novel tool for accessing and combining large-scale SNP genotype databases with population information
-
Full-text index only
WebScipio: an online tool for the determination of gene structures using protein sequences.
PMID 18801164 · PMC2644328 · BMC genomics · 2008 · 7 claims · 4 setups
WebScipio, a web interface to Scipio, determines gene structure from a query protein sequence against an assembled eukaryotic genome with quality approaching manual annotation.