Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction
miRge3.0: a comprehensive microRNA and tRF sequencing analysis pipeline.
PMID 34308351 · PMC8294687 · NAR genomics and bioinformatics · 2021 · 8 claims · 6 setups
miRge3.0 is a Python 3-based small RNA-seq and tRF analysis pipeline that improves on miRge2.0 (which was Python 2.7-based)
-
Full-text index only
Highly individual methylation patterns of alternative glucocorticoid receptor promoters suggest individualized epigenetic regulatory mechanisms.
PMID 19004867 · PMC2602793 · Nucleic acids research · 2008 · 7 claims · 4 setups
Methylation patterns of the five GR promoters activated in PBMCs are highly variable between individuals.
-
Full-text index only
A screen for proteins that interact with PAX6: C-terminal mutations disrupt interaction with HOMER3, DNCL1 and TRIM11.
PMID 16098226 · PMC1208879 · BMC genetics · 2005 · 8 claims · 7 setups
PAX6 interacts with three novel proteins: HOMER3, DNCL1 and TRIM11
-
Full-text index only
Dcode.org anthology of comparative genomic tools.
PMID 15980535 · PMC1160116 · Nucleic acids research · 2005 · 8 claims · 7 setups
The dcode.org suite (zPicture, Mulan, eShadow, rVista 2.0, multiTF, Creme 2.0, ECR Browser) provides integrated tools for comparative genomic analysis and non-coding regulatory element discovery.
-
Full-text index only
Simple models of genomic variation in human SNP density.
PMID 17553150 · PMC1919371 · BMC genomics · 2007 · 6 claims · 4 setups
Hierarchical Poisson model B, which allows both the mutation-rate proxy (Beta-distributed Λ) and the ARG-size proxy (Gamma-distributed T) to vary, fits the observed SNP density distribution significantly better than models with only one or neither varying.
-
Full-text index only
Extreme conservation of noncoding DNA near HoxD complex of vertebrates.
PMID 15462684 · PMC524357 · BMC genomics · 2004 · 7 claims · 7 setups
Three blocks of extremely conserved non-coding DNA (CR1, CR2, CR3) exist within 7 kb upstream of the HoxD complex, 3' of Evx-2, conserved from fish to human.
-
Full-text index only
Improving the specificity of exon prediction using comparative genomics.
PMID 18831778 · PMC2559877 · BMC genomics · 2008 · 8 claims · 6 setups
A log-odds ratio scoring method based on codon conservation across human-mouse/human-dog alignments and adjacent-codon dependency can classify putative exons as coding vs non-coding.
-
Full-text index only
SARS-CoV genome polymorphism: a bioinformatics study.
PMID 16144519 · PMC5172477 · Genomics, proteomics & bioinformatics · 2005 · 8 claims · 6 setups
SARS-CoV isolates can be classified into groups/subgroups based on the number and distribution of SNVs and INDELs relative to a 'profile' sequence, and this classification aligns with phylogenetic tree relationships and epidemiological spread.
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
How accurately is ncRNA aligned within whole-genome multiple alignments?
PMID 17963514 · PMC2206062 · BMC bioinformatics · 2007 · 7 claims · 4 setups
MULTIZ does a fairly accurate job of aligning ncRNA regions across 17 vertebrate genomes, but better alignments exist in some regions.
-
Full-text index only
MultiPhyl: a high-throughput phylogenomics webserver using distributed computing.
PMID 17553837 · PMC1933173 · Nucleic acids research · 2007 · 8 claims · 8 setups
MultiPhyl is the first high-throughput distributed phylogenetics platform capable of using idle computational resources of many heterogeneous non-dedicated machines to form a phylogenetics supercomputer
-
Full-text index only
Developments in CORG: a gene-centric comparative genomics resource.
PMID 17135197 · PMC1751536 · Nucleic acids research · 2007 · 7 claims · 4 setups
CORG provides pairwise and multiple sequence alignments of upstream promoter regions and whole gene loci across 10 vertebrate species.
-
Full-text index only
Molecular evolution and multilocus sequence typing of 145 strains of SARS-CoV.
PMID 16112670 · PMC7118731 · FEBS letters · 2005 · 8 claims · 7 setups
145 SARS-CoV genomes can be divided into three groups: animal-origin viruses, first-epidemic clinical viruses, and GD03T0013
-
Full-text index only
A macaque's-eye view of human insertions and deletions: differences in mechanisms.
PMID 17941704 · PMC1976337 · PLoS computational biology · 2007 · 7 claims · 4 setups
Insertion and deletion rates are differentially associated with replication- versus recombination-related genomic features, indicating the two mutation types are driven in part by distinct mechanisms
-
Full-text index only
Retroposition and evolution of the DNA-binding motifs of YY1, YY2 and REX1.
PMID 17478514 · PMC1904287 · Nucleic acids research · 2007 · 8 claims · 5 setups
62 YY1-related sequences were identified across genomes ranging from flying insects to humans, with high zinc finger domain conservation
-
Full-text index only
Rapid detection and curation of conserved DNA via enhanced-BLAT and EvoPrinterHD analysis.
PMID 18307801 · PMC2268679 · BMC genomics · 2008 · 8 claims · 8 setups
eBLAT detects up to 75% more conserved bases than original BLAT alignments, with the largest gains between evolutionarily distant orthologs
-
Full-text index only
Network of Cancer Genes: a web resource to analyze duplicability, orthology and network properties of cancer genes.
PMID 19906700 · PMC2808873 · Nucleic acids research · 2010 · 7 claims · 4 setups
NCG is a web database integrating duplicability, orthology, evolutionary appearance, and network topology data for 736 human cancer genes
-
Has reproduction · 79
pyrpipe: a Python package for RNA-Seq workflows.
PMID 34085037 · PMC8168212 · NAR genomics and bioinformatics · 2021 · 8 claims · 3 setups
pyrpipe enables development of flexible, reproducible, and easy-to-debug RNA-Seq computational pipelines purely in Python, in an object-oriented manner
-
Full-text index only
Grammar-based distance in progressive multiple sequence alignment.
PMID 18616828 · PMC2478692 · BMC bioinformatics · 2008 · 7 claims · 3 setups
A grammar-based (LZ complexity) distance metric can be used to determine the order in which sequences are progressively pairwise aligned
-
Full-text index only
NCBI Reference Sequences: current status, policy and new initiatives.
PMID 18927115 · PMC2686572 · Nucleic acids research · 2009 · 7 claims · 5 setups
RefSeq is a curated, non-redundant, explicitly linked database of nucleotide and protein sequences spanning genomes, transcripts and proteins across prokaryotes, eukaryotes and viruses