Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Microarray analysis: genome-scale hypothesis scanning.
PMID 14551912 · PMC212694 · PLoS biology · 2003 · 8 claims · 5 setups
Microarrays can be used to both test and generate hypotheses, not merely to fish for candidate genes.
-
Full-text index only
A genome-wide survey demonstrates widespread non-linear mRNA in expressed sequences from multiple species.
PMID 16237125 · PMC1258171 · Nucleic acids research · 2005 · 8 claims · 6 setups
A genome-wide computational survey identifies 245 genes in mammals (264 across six species) that produce RREO events in expressed sequences
-
Full-text index only
Upgrades to StellaBase facilitate medical and genetic studies on the starlet sea anemone, Nematostella vectensis.
PMID 17982171 · PMC2238866 · Nucleic acids research · 2008 · 6 claims · 5 setups
StellaBase Disease houses homology data for 155,904 invertebrate isoforms of human disease genes across four model systems, including 14,874 predicted Nematostella genes
-
Full-text index only
GenBank.
PMID 16381837 · PMC1347519 · Nucleic acids research · 2006 · 8 claims · 8 setups
GenBank is a comprehensive public database of nucleotide sequences with supporting bibliographic and biological annotation, built and distributed by NCBI.
-
Full-text index only
The revolution of the biology of the genome.
PMID 15040884 · PMC7091781 · Cell research · 2004 · 8 claims · 6 setups
Polyploidization and gene duplication are the major mechanisms increasing eukaryotic genome size.
-
Full-text index only
Functional dissection of siRNA sequence by systematic DNA substitution: modified siRNA with a DNA seed arm is a powerful tool for mammalian gene silencing with significantly reduced off-target effect.
PMID 18267968 · PMC2367719 · Nucleic acids research · 2008 · 6 claims · 8 setups
The seed arm (guide strand positions 2-8), its complementary passenger-strand sequence, the 5' end of the guide strand, and the 3' overhang of the passenger strand can be simultaneously replaced with DNA without substantial loss of gene-silencing activity.
-
Has reproduction · 65
Lineage-specific, fast-evolving GATA-like gene regulates zygotic gene activation to promote endoderm specification and pattern formation in the Theridiidae spider.
PMID 36203191 · PMC9535882 · BMC biology · 2022 · 8 claims · 8 setups
Comparative RNA-seq of cells isolated from central, intermediate, and peripheral regions of stage-3 embryos identifies genes with locally restricted expression genome-wide
-
Full-text index only
The MAPPER database: a multi-genome catalog of putative transcription factor binding sites.
PMID 15608292 · PMC540057 · Nucleic acids research · 2005 · 8 claims · 6 setups
Built a library of 1134 HMM models (359 matrix-derived, 718 factor-derived, 57 JASPAR-derived), corresponding to 863 distinct TF names, from TRANSFAC and JASPAR binding site data
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Full-text index only
ECgene: an alternative splicing database update.
PMID 17132829 · PMC1716719 · Nucleic acids research · 2007 · 8 claims · 5 setups
ECgene provides functional annotation (domain, GO, expression pattern) for alternatively spliced genes
-
Full-text index only
DDBJ dealing with mass data produced by the second generation sequencer.
PMID 18927114 · PMC2686496 · Nucleic acids research · 2009 · 8 claims · 7 setups
DDBJ collected and released 2,368,110 entries (1,415,106,598 bases) of original DNA sequence data from July 2007 to June 2008.