Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
GenBank.
PMID 17202161 · PMC1781245 · Nucleic acids research · 2007 · 8 claims · 1 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI
-
Has reproduction · 97
Metatranscriptomics From a Small Aquatic System: Microeukaryotic Community Functions Through the Diurnal Cycle.
PMID 32523568 · PMC7261829 · Frontiers in microbiology · 2020 · 6 claims · 6 setups
Photosynthesis-related and translational transcripts are upregulated at midday (high light) compared to night/darkness in the pond microeukaryotic community
-
Has reproduction · 75
Genomic regions and candidate genes selected during the breeding of rice in Vietnam.
PMID 35899250 · PMC9309459 · Evolutionary applications · 2022 · 8 claims · 7 setups
XP-CLR and FST scans identify genomic regions with distorted allele frequency/differentiation patterns resulting from differential selective pressures between Vietnamese rice subpopulations
-
Has reproduction · 74
Transcriptome profiling of Giardia intestinalis using strand-specific RNA-seq.
PMID 23555231 · PMC3610916 · PLoS computational biology · 2013 · 8 claims · 8 setups
Most of the G. intestinalis genome is transcribed in in vitro-grown trophozoites, but at vastly different expression levels.
-
Full-text index only
In silico and in vitro comparative analysis to select, validate and test SNPs for human identification.
PMID 18076761 · PMC2222643 · BMC genomics · 2007 · 8 claims · 7 setups
A panel of 24 SNPs was selected and validated for human identification using 1,040 unrelated samples from three populations (Italian, Benin Gulf, Mongolian)
-
Full-text index only
Reconstructing the evolution of the mitochondrial ribosomal proteome.
PMID 17604309 · PMC1950548 · Nucleic acids research · 2007 · 8 claims · 6 setups
The ancestral mitoribosome was of alpha-proteobacterial descent and more than doubled its protein content in most eukaryotic lineages.
-
Full-text index only
Simultaneous analysis of all SNPs in genome-wide and re-sequencing association studies.
PMID 18654633 · PMC2464715 · PLoS genetics · 2008 · 8 claims · 5 setups
A Bayesian-inspired penalised maximum likelihood stochastic search method can simultaneously analyse all SNPs (up to 500K) from a GWA study in a few hours on a desktop workstation
-
Full-text index only
Function2Gene: a gene selection tool to increase the power of genetic association studies by utilizing public databases and expert knowledge.
PMID 18631403 · PMC2500032 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Function2Gene is a set of Perl programs that queries public databases (NCBI, GeneCards, Harvester, with Uniprot/Ensembl also supported) using expert-selected keywords to rank genes by prior probability of disease association.
-
Has reproduction · 67
A consensus approach to vertebrate de novo transcriptome assembly from RNA-seq data: assembly of the duck (Anas platyrhynchos) transcriptome.
PMID 25009556 · PMC4070175 · Frontiers in genetics · 2014 · 8 claims · 8 setups
Multiple k-mer (MK) assemblies are more complete than single k-mer (SK) assemblies, showing higher reads-mapped-back-to-transcripts (RMBT) and higher CEGMA complete-gene percentages for all three tools.
-
Has reproduction · 100
Integrative transcriptome sequencing identifies trans-splicing events with important roles in human embryonic stem cell pluripotency.
PMID 24131564 · PMC3875859 · Genome research · 2014 · 8 claims · 8 setups
TSscan, a computational pipeline integrating long- and short-read transcriptome sequencing from multiple hESC lines, can detect trans-splicing while minimizing false positives from experimental artifacts and genetic rearrangements.
-
Full-text index only
Ensembl 2005.
PMID 15608235 · PMC540092 · Nucleic acids research · 2005 · 8 claims · 4 setups
Ensembl's automatic gene build system can flexibly and reliably annotate a wide variety of genomes with limited species-specific evidence.
-
Full-text index only
The promoter and the enhancer region of the KLK 3 (prostate specific antigen) gene is frequently mutated in breast tumours and in breast carcinoma cell lines.
PMID 10188912 · PMC2362704 · British journal of cancer · 1999 · 7 claims · 5 setups
No mutations were found in the protein-coding exons of the PSA gene in breast tumours or cell lines
-
Full-text index only
Functional genomics of early cortex patterning.
PMID 16515721 · PMC1431711 · Genome biology · 2006 · 7 claims · 6 setups
The neocortical protomap is established as continuous rostro-caudal gradients of gene expression in progenitor cells, not discrete compartments
-
Full-text index only
High quality catalog of proteotypic peptides from human heart.
PMID 18803417 · PMC2765113 · Journal of proteome research · 2008 · 8 claims · 4 setups
A catalog of 4476 proteotypic peptides representing 2558 human heart proteins was generated from MudPIT analyses of nonfailing and failing left-ventricular explants
-
Full-text index only
A simple and efficient algorithm for genome-wide homozygosity analysis in disease.
PMID 19756043 · PMC2758715 · Molecular systems biology · 2009 · 8 claims · 4 setups
A genome-wide AH analysis (GAHA) algorithm can identify disease-associated loci by comparing frequencies of homozygous segments between cases and controls using a z-statistic proportion test
-
Full-text index only
Software for tag single nucleotide polymorphism selection.
PMID 16004730 · PMC3525260 · Human genomics · 2005 · 8 claims · 3 setups
Pairwise R2 methods tend to pick more tagging SNPs than strictly needed because they miss redundancy where two or more tag SNPs jointly predict an untagged SNP with no single direct surrogate.
-
Has reproduction · 95
The archives are half-empty: an assessment of the availability of microbial community sequencing data.
PMID 32859925 · PMC7455719 · Communications biology · 2020 · 8 claims · 6 setups
A large proportion of 16S rRNA amplicon sequencing studies contain data that is not available or not reusable despite being reported as deposited.
-
Has reproduction · 84
Deep transcriptomics reveals cell-specific isoforms of pan-neuronal genes.
PMID 40379625 · PMC12084633 · Nature communications · 2025 · 8 claims · 5 setups
Pan-neuronal genes (expressed in many/all neurons) harbor highly cell-specific splice variants/isoforms restricted to single or few neuron types.
-
Full-text index only
A cell biological perspective on genome research.
PMID 8522596 · PMC2120688 · The Journal of cell biology · 1995 · 7 claims · 7 setups
Genome sequencing represents a sixth stage in the historical progression of structural biology (comparative anatomy through crystallography), and will be similarly valuable once related to function.
-
Full-text index only
Cell clusters overlying focally disrupted mammary myoepithelial cell layers and adjacent cells within the same duct display different immunohistochemical and genetic features: implications for tumor progression and invasion.
PMID 14580259 · PMC314413 · Breast cancer research : BCR · 2003 · 7 claims · 4 setups
ER-negative cell clusters are far more likely than ER-positive clusters to overlie disrupted myoepithelial cell layers, both at the case level and the individual-disruption level