Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Evaluation of multiple displacement amplification in a 5 cM STR genome-wide scan.
PMID 16055919 · PMC1182175 · Nucleic acids research · 2005 · 7 claims · 5 setups
MDA genotyping call rates and accuracy are only marginally lower than for genomic DNA
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Has reproduction · 96
Deep learning based protocol to construct an immune-related gene network of host-pathogen interactions in plants.
PMID 36525344 · PMC9791427 · STAR protocols · 2023 · 7 claims · 6 setups
DLNet algorithm ranks genes based on their contribution to classifying treatment vs. control expression data
-
Has reproduction · 71
Newborn sex-specific transcriptome signatures and gestational exposure to fine particles: findings from the ENVIRONAGE birth cohort.
PMID 28583124 · PMC5458481 · Environmental health : a global access science source · 2017 · 8 claims · 6 setups
Gestational PM2.5 exposure is associated with sex-specific transcriptome signatures in cord blood, with a significant sex-by-exposure interaction for many genes.
-
Full-text index only
The DAVID Gene Functional Classification Tool: a novel biological module-centric algorithm to functionally analyze large gene lists.
PMID 17784955 · PMC2375021 · Genome biology · 2007 · 8 claims · 6 setups
Gene-gene functional similarity can be measured using kappa statistics applied to a binary gene-annotation-term matrix built from 14 annotation categories.
-
Full-text index only
metaFun: An analysis pipeline for metagenomic big data with fast and unified functional searches.
PMID 41530917 · PMC12818822 · Gut microbes · 2026 · 8 claims · 8 setups
metaFun is an open-source, end-to-end Nextflow/Apptainer pipeline integrating quality control, taxonomic profiling, functional profiling, de novo assembly, binning, genome assessment, comparative genomics, network analysis, and strain-level microdiversity analysis into a unified framework
-
Full-text index only
ChromAcS: an automated and flexible GUI for end-to-end reproducible ATAC-seq analysis across multiple species.
PMID 41639613 · PMC12973882 · BMC bioinformatics · 2026 · 8 claims · 8 setups
ChromAcS is a comprehensive open-source, GUI-based ATAC-seq analysis pipeline supporting multi-species genomes with real-time progress monitoring and modular re-execution.
-
Has reproduction · 86
Improving the annotation of the cattle genome by annotating transcription start sites in a diverse set of tissues and populations using Cap Analysis Gene Expression sequencing.
PMID 37216666 · PMC10411599 · G3 (Bethesda, Md.) · 2023 · 7 claims · 8 setups
CAGE sequencing of 24 tissues from 3 cattle populations (dairy, beef-dairy cross, Kinsella composite) defines TSS and coexpressed short-range enhancers in the ARS-UCD1.2 reference genome
-
Has reproduction · 67
A consensus approach to vertebrate de novo transcriptome assembly from RNA-seq data: assembly of the duck (Anas platyrhynchos) transcriptome.
PMID 25009556 · PMC4070175 · Frontiers in genetics · 2014 · 8 claims · 8 setups
Multiple k-mer (MK) assemblies are more complete than single k-mer (SK) assemblies, showing higher reads-mapped-back-to-transcripts (RMBT) and higher CEGMA complete-gene percentages for all three tools.
-
Full-text index only
SeqDoC: rapid SNP and mutation detection by direct comparison of DNA sequence chromatograms.
PMID 15927052 · PMC1156871 · BMC bioinformatics · 2005 · 8 claims · 6 setups
SeqDoC generates a subtracted difference trace between a reference and test chromatogram that highlights single base changes
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
Retentive Network promotes efficient RNA language modeling of long sequences.
PMID 41814064 · PMC13111708 · Communications biology · 2026 · 8 claims · 6 setups
RNAret, a RetNet-based RNA language model with O(n) complexity, achieves training parallelism and low computational overhead while processing long RNA sequences
-
Full-text index only
PHScaffolding: a hypergraph clustering and dual-weight integration strategy for scaffolding with Pore-C reads.
PMID 41569288 · PMC12825295 · Briefings in bioinformatics · 2026 · 7 claims · 6 setups
PHScaffolding builds a weighted hypergraph from Pore-C read-to-contig alignments, with hyperedges representing multi-way contig interactions and weighted by the geometric mean of per-contig alignment lengths
-
Full-text index only
Evaluating the performance of ancient DNA genetic relatedness estimation methods using high-fidelity pedigree simulations.
PMID 41796349 · PMC13081257 · Genome biology · 2026 · 8 claims · 5 setups
BADGER, an automated snakemake pipeline, was developed to simulate high-fidelity pedigrees and raw ancient DNA sequence data for benchmarking genetic relatedness methods
-
Has reproduction · 45
Accurate sequence variant genotyping in cattle using variation-aware genome graphs.
PMID 31092189 · PMC6521551 · Genetics, selection, evolution : GSE · 2019 · 8 claims · 7 setups
Graphtyper outperformed GATK and SAMtools in genotype concordance, non-reference sensitivity, and non-reference discrepancy compared to microarray genotypes
-
Full-text index only
Recent segmental and gene duplications in the mouse genome.
PMID 12914656 · PMC193640 · Genome biology · 2003 · 8 claims · 8 setups
33.6 Mb (1.2%) of the February 2003 mouse genome assembly (2,695 Mb) is involved in recent segmental duplications
-
Full-text index only
Developing a set of ancestry-sensitive DNA markers reflecting continental origins of humans.
PMID 19860882 · PMC2775748 · BMC genetics · 2009 · 8 claims · 8 setups
A set of 47 SNPs selected via the 4gen pairwise F_ST approach serves as an ASM panel distinguishing four continental groups (African, Eurasian, Asian/Oceanian, Native American)
-
Full-text index only
SNPdetector: a software tool for sensitive and accurate SNP detection.
PMID 16261194 · PMC1274293 · PLoS computational biology · 2005 · 7 claims · 7 setups
SNPdetector, which models human visual inspection of sequencing traces, achieves low false positive and false negative rates in automated SNP and mutation detection
-
Full-text index only
Comparative analysis reveals signatures of differentiation amid genomic polymorphism in Lake Malawi cichlids.
PMID 18616806 · PMC2530870 · Genome biology · 2008 · 8 claims · 8 setups
Lake Malawi cichlids are phenotypically and behaviorally diverse but appear genetically like a single subdivided population rather than distinct species
-
Full-text index only
Genome- and Transcriptome-Wide Characterization of AP2/ERF Transcription Factor Superfamily Reveals Their Relevance in Stylosanthes scabra Vogel Under Water Deficit Stress.
PMID 41515103 · PMC12787715 · Plants (Basel, Switzerland) · 2026 · 8 claims · 8 setups
295 AP2/ERF transcription factor genes were identified and classified in the S. scabra genome