Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Reconstruction of pathways associated with amino acid metabolism in human mitochondria.
PMID 18267298 · PMC5054205 · Genomics, proteomics & bioinformatics · 2007 · 8 claims · 5 setups
Out of 20 amino acids, the metabolic pathways of 17 utilize mitochondrial enzymes, and dysfunction of these enzymes causes over 40 known human mitochondrial diseases/disorders
-
Full-text index only
Ensembl 2005.
PMID 15608235 · PMC540092 · Nucleic acids research · 2005 · 8 claims · 4 setups
Ensembl's automatic gene build system can flexibly and reliably annotate a wide variety of genomes with limited species-specific evidence.
-
Has reproduction · 95
Comparative Genome Analysis of 16SrXII-A 'Candidatus Phytoplasma solani' POT Transmitted by Hyalesthes obsoletus.
PMID 41597744 · PMC12843639 · Microorganisms · 2026 · 7 claims · 8 setups
The complete 832,614 bp circular chromosome of the H. obsoletus-transmissible 'Ca. P. solani' 16SrXII-A strain POT was assembled and functionally reconstructed.
-
Full-text index only
A genome-wide survey of Major Histocompatibility Complex (MHC) genes and their paralogues in zebrafish.
PMID 16271140 · PMC1309616 · BMC genomics · 2005 · 8 claims · 4 setups
149 putative MHC gene loci and their paralogues were identified in the zebrafish genome using sequence similarity searches against the Zv4 draft assembly.
-
Full-text index only
The RCSB PDB information portal for structural genomics.
PMID 16381872 · PMC1347482 · Nucleic acids research · 2006 · 7 claims · 5 setups
The RCSB PDB Structural Genomics Information Portal integrates three resources: Structural Genomics Initiatives, Targets (TargetDB/PepcDB), and Structures (functional coverage analysis).
-
Full-text index only
In silico segmentations of lentivirus envelope sequences.
PMID 17376229 · PMC1847453 · BMC bioinformatics · 2007 · 8 claims · 8 setups
C and V regions of lentivirus SU sequences have distinct statistical (oligonucleotide/amino-acid) compositions that HMMs can learn and use to delimit them.
-
Has reproduction · 63
Genetic analysis of Leishmania donovani tropism using a naturally attenuated cutaneous strain.
PMID 24992200 · PMC4081786 · PLoS pathogens · 2014 · 8 claims · 8 setups
The CL-SL L. donovani isolate is severely attenuated for survival in visceral organs (liver, spleen) of BALB/c mice compared with the VL-SL isolate, while inducing transient footpad swelling that VL-SL does not.
-
Full-text index only
Using multiple alignments to improve seeded local alignment algorithms.
PMID 16100379 · PMC1185574 · Nucleic acids research · 2005 · 8 claims · 2 setups
Using information implicit in a multiple alignment to dynamically build a spaced-seed index weighted toward promising regions increases sensitivity of local alignment search compared to indexing a sequence alone
-
Has reproduction · 77
Representing and querying disease networks using graph databases.
PMID 27462371 · PMC4960687 · BioData mining · 2016 · 7 claims · 8 setups
Graph databases are well suited for representing biological information that is highly connected, semi-structured, and unpredictable.
-
Full-text index only
Candida albicans genome sequence: a platform for genomics in the absence of genetics.
PMID 15239821 · PMC463275 · Genome biology · 2004 · 8 claims · 5 setups
Publication of the complete diploid C. albicans genome sequence will accelerate research into the pathogenesis of Candida infections
-
Full-text index only
Extending Asia Pacific bioinformatics into new realms in the "-omics" era.
PMID 19958472 · PMC2788361 · BMC genomics · 2009 · 8 claims · 6 setups
88 full paper submissions were peer-reviewed for InCoB2009, with 49 shortlisted for oral presentation and 34 accepted into this BMC Genomics supplement, reflecting an overall acceptance rate of 50% across venues.
-
Has reproduction · 86
RNASEQR--a streamlined and accurate RNA-seq sequence analysis program.
PMID 22199257 · PMC3315322 · Nucleic acids research · 2012 · 8 claims · 7 setups
RNASEQR is a new RNA-seq mapper/aligner that combines a BWT-based (Bowtie) transcriptomic/genomic alignment with hash-based BLAT local alignment in three sequential steps: transcriptome mapping, novel exon detection, and anchor-and-align novel splice junction identification.
-
Full-text index only
Twin peaks: the draft human genome sequence.
PMID 11276423 · PMC138909 · Genome biology · 2001 · 8 claims · 8 setups
The predicted number of human genes (~26,000-40,000) is far lower than the widely assumed ~100,000, though downstream RNA/protein complexity can still generate substantial biological complexity.
-
Has reproduction · 57
KARAJ: An Efficient Adaptive Multi-Processor Tool to Streamline Genomic and Transcriptomic Sequence Data Acquisition.
PMID 36430895 · PMC9694301 · International journal of molecular sciences · 2022 · 8 claims · 6 setups
KARAJ automates end-to-end querying and downloading of genomic/transcriptomic sequence data from a list of PMCIDs, URLs, or accession numbers
-
Full-text index only
Functional coverage of the human genome by existing structures, structural genomics targets, and homology models.
PMID 16118666 · PMC1188274 · PLoS computational biology · 2005 · 8 claims · 5 setups
Existing PDB structures provide single-domain coverage for 37% of functional classes in the human genome and complete (whole-protein) structure coverage for 25%.
-
Full-text index only
TB database: an integrated platform for tuberculosis research.
PMID 18835847 · PMC2686437 · Nucleic acids research · 2009 · 8 claims · 8 setups
TBDB is an integrated database providing access to TB genomic data and resources relevant to discovery/development of TB drugs, vaccines and biomarkers.
-
Full-text index only
Columba: an integrated database of proteins, structures, and annotations.
PMID 15801979 · PMC1087474 · BMC bioinformatics · 2005 · 8 claims · 6 setups
COLUMBA physically integrates data from twelve protein structure-related databases (PDB, KEGG, Swiss-Prot, CATH, SCOP, Gene Ontology, ENZYME, etc.) into a single PostgreSQL data warehouse.
-
Full-text index only
Exogean: a framework for annotating protein-coding genes in eukaryotic genomic DNA.
PMID 16925841 · PMC1810556 · Genome biology · 2006 · 8 claims · 5 setups
Exogean is a framework using directed acyclic coloured multigraphs (DACMs) to represent biological objects (mRNA, ESTs, protein alignments, exons) and iteratively combine them into complex protein-coding transcript models.
-
Full-text index only
Comparing whole genomes using DNA microarrays.
PMID 18347592 · PMC7097741 · Nature reviews. Genetics · 2008 · 8 claims · 6 setups
DNA microarrays offer a relatively inexpensive and efficient alternative to genome sequencing for comparing all known classes of genomic diversity between closely related genomes.
-
Full-text index only
Systematic analysis of human kinase genes: a large number of genes and alternative splicing events result in functional and structural diversity.
PMID 16351747 · PMC1866387 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Systematic in silico search identified 5 novel human kinase genes (on chromosomes 1, 11, 13, 15, 16) and 1 pseudogene (chromosome X) absent from KinBase