Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Molecular evolution of Cide family proteins: novel domain formation in early vertebrates and the subsequent divergence.
PMID 18500987 · PMC2426694 · BMC evolutionary biology · 2008 · 8 claims · 5 setups
Sequences homologous to the CIDE-N domain/NCD show a wide phylogenetic distribution, from hydra and sea anemone to mammals, while true Cide proteins are restricted to vertebrates.
-
Has reproduction · 71
RNAmountAlign: Efficient software for local, global, semiglobal pairwise and multiple RNA sequence/structure alignment.
PMID 31978147 · PMC6980424 · PloS one · 2020 · 8 claims · 6 setups
RNAmountAlign is the first RNA sequence/structure pairwise alignment algorithm based on incremental ensemble mountain distance, running in O(n^3) time and O(n^2) space for two sequences of length n.
-
Has reproduction · 79
Detection of Virus-Related Sequences Associated With Potential Etiologies of Hepatitis in Liver Tissue Samples From Rats, Mice, Shrews, and Bats.
PMID 34177835 · PMC8221242 · Frontiers in microbiology · 2021 · 8 claims · 6 setups
Viral metagenomics of liver tissue from rats, mice, shrews, and bats revealed a diverse set of sequences related to herpesviruses, orthomyxoviruses, anelloviruses, hepeviruses, hepadnaviruses, flaviviruses, parvoviruses, and picornaviruses
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences
-
Full-text index only
Shotgun haplotyping: a novel method for surveying allelic sequence variation.
PMID 16221968 · PMC1253838 · Nucleic acids research · 2005 · 8 claims · 7 setups
A novel shotgun haplotyping method generates haplotypic sequences from long PCR products by shotgun sequencing both alleles concurrently and using read-pair information to separate alleles during assembly
-
Full-text index only
Evaluation of NTHL1, NEIL1, NEIL2, MPG, TDG, UNG and SMUG1 genes in familial colorectal cancer predisposition.
PMID 17029639 · PMC1624846 · BMC cancer · 2006 · 6 claims · 4 setups
Coding sequences and intron-exon boundaries of NTHL1, NEIL1, NEIL2, MPG, TDG, UNG and SMUG1 were screened in 94 familial CRC cases with known genes excluded
-
Full-text index only
Dog Y chromosomal DNA sequence: identification, sequencing and SNP discovery.
PMID 17026745 · PMC1630699 · BMC genetics · 2006 · 8 claims · 6 setups
Identified 32 male-specific Y-chromosome sequences totaling 24159 bp via combined Blast (human Y chromosome match, absence from female dog genome) and PCR male-specificity screening of a male poodle shotgun genome.
-
Full-text index only
Inventory and analysis of the protein subunits of the ribonucleases P and MRP provides further evidence of homology between the yeast and human enzymes.
PMID 16998185 · PMC1636426 · Nucleic acids research · 2006 · 8 claims · 6 setups
Fungal Pop8 is evolutionarily related to the Rpp14/Pop5 protein family, suggesting Pop8 is the fungal orthologue of Rpp14
-
Full-text index only
Comparison of characteristics and function of translation termination signals between and within prokaryotic and eukaryotic organisms.
PMID 16614446 · PMC1435984 · Nucleic acids research · 2006 · 8 claims · 5 setups
A core termination signal of 4 nt (stop codon plus the following nucleotide) is preferred across most prokaryotic and eukaryotic genomes
-
Full-text index only
Identifying repeat domains in large genomes.
PMID 16507140 · PMC1431705 · Genome biology · 2006 · 7 claims · 5 setups
A repeat domain graph, built using a modified A-Bruijn graph framework, decomposes a repeat library into shared repeat domains and reveals the mosaic structure of repeat families.
-
Full-text index only
TreeFam: a curated database of phylogenetic trees of animal gene families.
PMID 16381935 · PMC1347480 · Nucleic acids research · 2006 · 7 claims · 6 setups
Tree-based inference of orthologs and paralogs is more robust than BLAST-based methods because evolutionary rates (and thus pairwise BLAST scores) vary across gene family members
-
Full-text index only
EPGD: a comprehensive web resource for integrating and displaying eukaryotic paralog/paralogon information.
PMID 17984073 · PMC2238967 · Nucleic acids research · 2008 · 8 claims · 8 setups
EPGD is a gene-centered, internet-accessible database integrating paralog family and paralogon information for 26 eukaryotic genomes.
-
Full-text index only
piRNABank: a web resource on classified and clustered Piwi-interacting RNAs.
PMID 17881367 · PMC2238943 · Nucleic acids research · 2008 · 6 claims · 4 setups
piRNABank is a web-accessible database storing empirically known piRNA sequences and annotations for human, mouse and rat.
-
Full-text index only
The Princeton Protein Orthology Database (P-POD): a comparative genomics analysis tool for biologists.
PMID 17712414 · PMC1942082 · PloS one · 2007 · 8 claims · 5 setups
P-POD is the first comparative genomics database to combine results from multiple computational ortholog/homolog prediction methods with manually curated literature-derived experimental evidence of functional conservation.
-
Full-text index only
Retroposition and evolution of the DNA-binding motifs of YY1, YY2 and REX1.
PMID 17478514 · PMC1904287 · Nucleic acids research · 2007 · 8 claims · 5 setups
62 YY1-related sequences were identified across genomes ranging from flying insects to humans, with high zinc finger domain conservation
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
Efficient algorithms for probing the RNA mutation landscape.
PMID 18688270 · PMC2475669 · PLoS computational biology · 2008 · 8 claims · 4 setups
RNAmutants generalizes McCaskill's partition function algorithm to sum over the grand canonical ensemble of all secondary structures of all k-point mutants, simultaneously computing MFE(k) and Z(k) for each k
-
Full-text index only
Genotypic analysis of two hypervariable human cytomegalovirus genes.
PMID 18649324 · PMC2658010 · Journal of medical virology · 2008 · 8 claims · 8 setups
UL146 sequences from 184 samples fall into the same 14 genotypes (G1-G14) previously defined, with no new genotypes found.
-
Full-text index only
Diversity of the parB and repA genes of the Burkholderia cepacia complex and their utility for rapid identification of Burkholderia cenocepacia.
PMID 18328098 · PMC2324101 · BMC microbiology · 2008 · 8 claims · 7 setups
repA sequences show distinct clustering of B. cenocepacia relative to other Bcc species, enabling design of a species-specific multiplex PCR
-
Full-text index only
MEROPS: the peptidase database.
PMID 19892822 · PMC2808883 · Nucleic acids research · 2010 · 8 claims · 5 setups
MEROPS is a manually curated hierarchical classification of peptidases and protein inhibitors organized into protein species, families, and clans based on sequence and structural homology.