Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Diversity of preferred nucleotide sequences around the translation initiation codon in eukaryote genomes.
PMID 18086709 · PMC2241899 · Nucleic acids research · 2008 · 8 claims · 5 setups
Preferred nucleotide sequences around the initiation codon are diverse among eukaryote species, but differences roughly reflect evolutionary relationships between species
-
Full-text index only
Evidence for a novel gene associated with human influenza A viruses.
PMID 19917120 · PMC2780412 · Virology journal · 2009 · 8 claims · 8 setups
A 167-codon ORF (NEG8) on the negative-sense genomic strand of segment 8 is associated with early-20th-century human influenza A isolates
-
Full-text index only
Genomic variability associated with the presence of occult hepatitis B virus in HIV co-infected individuals.
PMID 19889143 · PMC3032083 · Journal of viral hepatitis · 2010 · 7 claims · 8 setups
O-HBV-associated mutations in PreS/S/polymerase regions likely contribute to undetectable HBsAg by interfering with serologic detection, altering antigen secretion, and/or decreasing replicative fitness
-
Full-text index only
Genome-wide prioritization of disease genes and identification of disease-disease associations from an integrated human functional linkage network.
PMID 19728866 · PMC2768980 · Genome biology · 2009 · 6 claims · 6 setups
Integrating 16 genomic features (32 sub-features) via a naïve Bayes classifier produces a genome-scale FLN of 21,657 human genes and 22,388,609 weighted links that outperforms any individual data source for inferring functional linkages.
-
Full-text index only
Multi-omics feature engineering driven by biomedical foundation models improves drug response prediction for inflammatory bowel disease patients.
PMID 41844950 · PMC13129071 · Scientific reports · 2026 · 8 claims · 7 setups
FM (MAMMAL)-derived drug-target binding affinity (BA) inference can be used to rank/select biologically relevant protein targets and their associated genes/SNPs for a drug of interest without knowledge of protein structure or active sites
-
Full-text index only
Update of the G2D tool for prioritization of gene candidates to inherited diseases.
PMID 17478516 · PMC1933178 · Nucleic acids research · 2007 · 8 claims · 4 setups
G2D is a web server that prioritizes candidate genes for inherited diseases using three distinct algorithms based on different input information.
-
Full-text index only
Impact of short-read sequencing on the misassembly of a plant genome.
PMID 33530937 · PMC7852129 · BMC genomics · 2021 · 7 claims · 6 setups
Short-read tomato assembly has substantial high-coverage (0.6%, 5.1 Mb) and low-coverage (9.7%, 79.6 Mb) regions relative to background coverage
-
Full-text index only
Geometry-aware graph attention networks to explain single-cell chromatin states and gene expression with SEAGALL.
PMID 42026624 · PMC13238118 · Genome biology · 2026 · 8 claims · 6 setups
SEAGALL combines a geometry-regularised autoencoder (GRAE) to embed cells and build a cell-cell graph with a graph attention network (GAT) classifier and GNNExplainer-based XAI to identify features driving cell type/phenotype.
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
Optimizing Single-Cell Long-Read Sequencing for Enhanced Isoform Detection in Pancreatic Islets.
PMID 41563441 · PMC13007207 · Diabetes · 2026 · 8 claims · 7 setups
5′ single-cell library preparation protocols outperform 3′ protocols for transcript identification and read length