Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 100
Viral Diagnostics in Plants Using Next Generation Sequencing: Computational Analysis in Practice.
PMID 29123534 · PMC5662881 · Frontiers in plant science · 2017 · 8 claims · 8 setups
NGS/RNA-seq enables unbiased, hypothesis-free detection of multiple known and emergent plant viruses, unlike RT-PCR which only detects one or a few known viruses per test.
-
Full-text index only
Inparanoid: a comprehensive database of eukaryotic orthologs.
PMID 15608241 · PMC540061 · Nucleic acids research · 2005 · 8 claims · 4 setups
The Inparanoid algorithm identifies true ortholog clusters by seeding on reciprocal best-matching pairs, gathering inparalogs (post-speciation duplicates) while excluding outparalogs (pre-speciation duplicates)
-
Full-text index only
Evolutionary sequence analysis of complete eukaryote genomes.
PMID 15762985 · PMC1274250 · BMC bioinformatics · 2005 · 8 claims · 6 setups
A conservative genome-comparison method (MIA) identifies panorthologs — strict single-copy 1:1 orthologs containing only species divergences, no paralogy — to minimize errors from gene duplication in evolutionary sequence analysis.
-
Full-text index only
Development of an integrated genome informatics, data management and workflow infrastructure: a toolbox for the study of complex disease genetics.
PMID 15601538 · PMC3525068 · Human genomics · 2004 · 8 claims · 8 setups
An integrated system combining Ensembl, ACeDB, Gbrowse and custom relational databases provides a scalable genome informatics and workflow infrastructure for complex disease gene discovery.
-
Full-text index only
Protein ranking by semi-supervised network propagation.
PMID 16723003 · PMC1810311 · BMC bioinformatics · 2006 · 8 claims · 5 setups
RankProp, a diffusion-based network propagation algorithm on a PSI-BLAST-derived protein similarity network, significantly outperforms local search methods (BLAST/PSI-BLAST) at detecting remote homologs.
-
Full-text index only
Nature of protein family signatures: insights from singular value analysis of position-specific scoring matrices.
PMID 18398479 · PMC2276316 · PloS one · 2008 · 8 claims · 6 setups
The first singular component of a PSSM acts to disfavor substitutions, penalizing potentially functionally important residues at conserved sites more severely.
-
Full-text index only
Molecular phylogeny of the kelch-repeat superfamily reveals an expansion of BTB/kelch proteins in animals.
PMID 13678422 · PMC222960 · BMC bioinformatics · 2003 · 8 claims · 8 setups
The human genome encodes at least 71 kelch-repeat proteins
-
Full-text index only
Genome-wide survey for biologically functional pseudogenes.
PMID 16680195 · PMC1456316 · PLoS computational biology · 2006 · 8 claims · 6 setups
A subset of ancient, cross-species-conserved pseudogenes (30 of 1,453 candidate quartets) show evidence consistent with retained biological function
-
Full-text index only
The global landscape of sequence diversity.
PMID 17996061 · PMC2258180 · Genome biology · 2007 · 7 claims · 5 setups
Eukaryotic sequence datasets show substantially greater genetic diversity (higher sequence/gene family discovery rates) than bacterial datasets, likely related to differences in modes of genetic inheritance.
-
Full-text index only
Pseudofam: the pseudogene families database.
PMID 18957444 · PMC2686518 · Nucleic acids research · 2009 · 8 claims · 7 setups
Pseudofam is an online database of pseudogene families built by mapping pseudogenes to Pfam protein families, providing query tools, statistics, and sequence alignments
-
Full-text index only
Genome wide identification of recessive cancer genes by combinatorial mutation analysis.
PMID 18846217 · PMC2557123 · PloS one · 2008 · 7 claims · 4 setups
A combinatorial mutation analysis identified 154 candidate recessive cancer genes (pRecessiveCancer<1.5x10-7, FDR=0.39)
-
Full-text index only
Generation of a restriction minus enteropathogenic Escherichia coli E2348/69 strain that is efficiently transformed with large, low copy plasmids.
PMID 18681975 · PMC2518929 · BMC microbiology · 2008 · 8 claims · 7 setups
E2348/69 possesses a type I restriction-modification system encoded by an hsdMSR-like operon identified by homology to known Hsd proteins.
-
Full-text index only
Comparative genomics of Lbx loci reveals conservation of identical Lbx ohnologs in bony vertebrates.
PMID 18541024 · PMC2446394 · BMC evolutionary biology · 2008 · 8 claims · 3 setups
Extant bony vertebrates (osteichthyans) retain only Lbx1- and Lbx2-type genes; no distinct Lbx3/Lbx4 proteins exist.
-
Full-text index only
The functional importance of disease-associated mutation.
PMID 12220483 · PMC128831 · BMC bioinformatics · 2002 · 6 claims · 1 setups
Disease-associated mutations occur in conserved regions of genes and can be used to identify likely disease-causing mutations
-
Has reproduction · 59
eDNAmap: A Metabarcoding Web Tool for Comparing Marine Biodiversity, With Special Reference to Teleost Fish.
PMID 41189540 · PMC12627913 · Molecular ecology resources · 2026 · 7 claims · 5 setups
eDNAmap is a web-based platform that maps sampling locations, generates heatmaps to evaluate batch effects, and performs nMDS and cluster analyses using similarity indices on uploaded eDNA composition data
-
Full-text index only
Genome-wide detection of segmental duplications and potential assembly errors in the human genome sequence.
PMID 12702206 · PMC154576 · Genome biology · 2003 · 8 claims · 6 setups
Segmental duplications comprise 3.53% (107.4/3,043.1 Mb) of the June 2002 human genome assembly
-
Full-text index only
Genome-wide analysis of the human Alu Yb-lineage.
PMID 15588477 · PMC3525081 · Human genomics · 2004 · 8 claims · 6 setups
1,733 Alu Yb-lineage elements are present on human autosomal chromosomes
-
Full-text index only
EPGD: a comprehensive web resource for integrating and displaying eukaryotic paralog/paralogon information.
PMID 17984073 · PMC2238967 · Nucleic acids research · 2008 · 8 claims · 8 setups
EPGD is a gene-centered, internet-accessible database integrating paralog family and paralogon information for 26 eukaryotic genomes.
-
Full-text index only
Identification and characterization of insect-specific proteins by genome data analysis.
PMID 17407609 · PMC1852559 · BMC genomics · 2007 · 8 claims · 7 setups
Comparative genome analysis across five holometabolous insects and three non-insect eukaryotes (opisthokonts) identifies 154 insect-specific orthologous groups (refined to 51 proteins) and 466 eukaryote/opisthokont-core orthologous groups
-
Full-text index only
Origin and diversification of the basic helix-loop-helix gene family in metazoans: insights from comparative genomics.
PMID 17335570 · PMC1828162 · BMC evolutionary biology · 2007 · 8 claims · 4 setups
An initial diversification of bHLHs occurred in the pre-Cambrian, prior to metazoan cladogenesis