Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 88
AuPairWise: A Method to Estimate RNA-Seq Replicability through Co-expression.
PMID 27082953 · PMC4833304 · PLoS computational biology · 2016 · 7 claims · 4 setups
Sample-sample correlation of transcript abundances is trivially high regardless of condition and gives misleading estimates of the replicability of conditional (differential) variation in expression.
-
Has reproduction · 71
Parsimonious Gene Correlation Network Analysis (PGCNA): a tool to define modular gene co-expression for refined molecular stratification in cancer.
PMID 30993001 · PMC6459838 · NPJ systems biology and applications · 2019 · 8 claims · 7 setups
Retaining only the top ~3 most correlated edges per gene (EPG3) combined with FastUnfold clustering (termed PGCNA) produces gene co-expression modules with significantly better separation and enrichment of known biology than using all edges or other clustering methods.
-
Has reproduction · 96
GC-biased gene conversion conceals the prediction of the nearly neutral theory in avian genomes.
PMID 30616647 · PMC6322265 · Genome biology · 2019 · 8 claims · 6 setups
gBGC conceals the correlation between life-history traits and dN/dS in birds; accounting for it reveals correlations consistent with nearly neutral theory
-
Full-text index only
Use of genomic data in risk assessment.
PMID 11983054 · PMC139345 · Genome biology · 2002 · 8 claims · 4 setups
New genomic technologies (CGH, SNP analysis, restriction landmark genome scanning, spectral karyotyping, transcript profiling) can be applied to improve risk assessment accuracy
-
Has reproduction · 44
Detecting DNA modifications from SMRT sequencing data by modeling sequence context dependence of polymerase kinetic.
PMID 23516341 · PMC3597545 · PLoS computational biology · 2013 · 8 claims · 7 setups
Local sequence context strongly determines position-specific polymerase kinetic rate: roughly 80% of IPD variation is explained by a 10 bp context (7 bases upstream, 2 bases downstream of the incorporation site), saturating at 7 bases upstream.
-
Full-text index only
Computational verification of protein-protein interactions by orthologous co-expression.
PMID 15740634 · PMC555590 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Co-expression of orthologous protein pairs across multiple species can verify/predict S. cerevisiae PPIs with better performance than S. cerevisiae co-expression alone.
-
Full-text index only
A high throughput method for genome-wide analysis of retroviral integration.
PMID 17028098 · PMC1636494 · Nucleic acids research · 2006 · 8 claims · 8 setups
VITA uses MmeI to cleave DNA at a fixed distance from its recognition site, generating 21-22 bp genomic tags that serve as signatures of lentiviral integration sites.
-
Full-text index only
FeatureScan: revealing property-dependent similarity of nucleotide sequences.
PMID 16845077 · PMC1538849 · Nucleic acids research · 2006 · 6 claims · 5 setups
FeatureScan transforms nucleotide sequences into numerical signals of physico-chemical/conformational properties and compares them via a convolution/correlation (Fourier transform) method rather than comparing letters
-
Full-text index only
Benchmarking ortholog identification methods using functional genomics data.
PMID 16613613 · PMC1557999 · Genome biology · 2006 · 8 claims · 7 setups
InParanoid is the best overall ortholog identification method for identifying functionally equivalent proteins when sensitivity and selectivity are combined into an overall score.
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Full-text index only
miRNAMap 2.0: genomic maps of microRNAs in metazoan genomes.
PMID 18029362 · PMC2238982 · Nucleic acids research · 2008 · 8 claims · 6 setups
miRNAMap 2.0 is a resource collecting experimentally verified miRNAs and experimentally verified miRNA target genes in human, mouse, rat and other metazoan genomes
-
Full-text index only
fREDUCE: detection of degenerate regulatory elements using correlation with expression.
PMID 17941998 · PMC2174516 · BMC bioinformatics · 2007 · 6 claims · 5 setups
fREDUCE is a computational method that detects weak or degenerate binding motifs from gene expression or ChIP-chip data by exhaustive search of degenerate IUPAC oligonucleotides
-
Full-text index only
Applicability of DNA pools on 500 K SNP microarrays for cost-effective initial screens in genomewide association studies.
PMID 17610740 · PMC1925094 · BMC genomics · 2007 · 8 claims · 5 setups
SNP-MaP can be effectively applied to the Affymetrix 500K GeneChip, providing a cost-effective, reliable and valid initial genomewide screen
-
Full-text index only
Analysis of proteomic profiles and functional properties of human peripheral blood myeloid dendritic cells, monocyte-derived dendritic cells and the dendritic cell-like KG-1 cells reveals distinct characteristics.
PMID 17331236 · PMC1868942 · Genome biology · 2007 · 8 claims · 8 setups
moDCs and KG-1 cells show significant proteomic differences from primary mDCs, particularly in proteins involved in cell growth/maintenance and cell-cell interaction/integrity.
-
Full-text index only
Long-term trends in evolution of indels in protein sequences.
PMID 17298668 · PMC1805498 · BMC evolutionary biology · 2007 · 8 claims · 5 setups
More than one third of protein domains show a statistically significant tendency to increase or decrease in size over evolutionary distance.
-
Full-text index only
Functional importance of different patterns of correlation between adjacent cassette exons in human and mouse.
PMID 18439302 · PMC2432081 · BMC genomics · 2008 · 8 claims · 7 setups
Adjacent cassette exon pairs can be categorized by EST-derived correlation coefficient into three groups: mutually exclusive (ME, r<=-0.7), independent (IND, -0.2<=r<=0.2), and linked (LNK, r>=0.7)
-
Full-text index only
A map of human protein interactions derived from co-expression of human mRNAs and their orthologs.
PMID 18414481 · PMC2387231 · Molecular systems biology · 2008 · 8 claims · 6 setups
Comparing human mRNA co-expression with co-expression of orthologous gene pairs in five other organisms identifies proteins that physically associate
-
Full-text index only
Commonality of functional annotation: a method for prioritization of candidate genes from genome-wide linkage studies.
PMID 18263617 · PMC2275105 · Nucleic acids research · 2008 · 8 claims · 7 setups
Genes correlated with a common complex trait are more likely to share GO functional annotations than genes not correlated with that trait
-
Has reproduction · 78
Requirements for Pseudomonas aeruginosa acute burn and chronic surgical wound infection.
PMID 25057820 · PMC4109851 · PLoS genetics · 2014 · 8 claims · 8 setups
In vivo gene expression is generally not correlated with a gene's importance for fitness, with the exception of metabolic genes, for which differential expression is more predictive of fitness.
-
Full-text index only
High resolution analysis of the human transcriptome: detection of extensive alternative splicing independent of transcriptional activity.
PMID 19804644 · PMC2768739 · BMC genetics · 2009 · 8 claims · 6 setups
The human GWSA uses exon body and exon-exon junction probes to directly measure over 280,000 known and predicted splicing events genome-wide.