SKILL·44DB21

primekg

Name: primekg
Author: K-Dense-AI

K-Dense-AI

Updated 1 month ago

2 views

31,025

3,113

31,025

View on GitHub

Otherdata

About

This skill enables programmatic querying of the PrimeKG knowledge graph to retrieve interconnected biomedical data on genes, drugs, and diseases. Developers can use it to search for biological entities, analyze their associations, and explore paths for insights like drug repurposing. It's ideal for integrating structured, multiscale medical relationships into bioinformatics applications.

Quick Install

Claude Code

Recommended

Primary

npx skills add K-Dense-AI/claude-scientific-skills -a claude-code

Plugin CommandAlternative

/plugin add https://github.com/K-Dense-AI/claude-scientific-skills

Git CloneAlternative

git clone https://github.com/K-Dense-AI/claude-scientific-skills.git ~/.claude/skills/primekg

Copy and paste this command in Claude Code to install this skill

Documentation

PrimeKG Knowledge Graph Skill

Overview

PrimeKG is a precision medicine knowledge graph that integrates over 20 primary databases and high-quality scientific literature into a single resource. It contains over 100,000 nodes and 4 million edges across 29 relationship types, including drug-target, disease-gene, and phenotype-disease associations.

Key capabilities:

Search for nodes (genes, proteins, drugs, diseases, phenotypes)
Retrieve direct neighbors (associated entities and clinical evidence)
Analyze local disease context (related genes, drugs, phenotypes)
Identify drug-disease paths (potential repurposing opportunities)

Data access: Programmatic access via query_primekg.py. Data is stored at C:\Users\eamon\Documents\Data\PrimeKG\kg.csv.

When to Use This Skill

This skill should be used when:

Knowledge-based drug discovery: Identifying targets and mechanisms for diseases.
Drug repurposing: Finding existing drugs that might have evidence for new indications.
Phenotype analysis: Understanding how symptoms/phenotypes relate to diseases and genes.
Multiscale biology: Bridging the gap between molecular targets (genes) and clinical outcomes (diseases).
Network pharmacology: Investigating the broader network effects of drug-target interactions.

Core Workflow

1. Search for Entities

Find identifiers for genes, drugs, or diseases.

from scripts.query_primekg import search_nodes

# Search for Alzheimer's disease nodes
results = search_nodes("Alzheimer", node_type="disease")
# Returns: [{"id": "EFO_0000249", "type": "disease", "name": "Alzheimer's disease", ...}]

2. Get Neighbors (Direct Associations)

Retrieve all connected nodes and relationship types.

from scripts.query_primekg import get_neighbors

# Get all neighbors of a specific disease ID
neighbors = get_neighbors("EFO_0000249")
# Returns: List of neighbors like {"neighbor_name": "APOE", "relation": "disease_gene", ...}

3. Analyze Disease Context

A high-level function to summarize associations for a disease.

from scripts.query_primekg import get_disease_context

# Comprehensive summary for a disease
context = get_disease_context("Alzheimer's disease")
# Access: context['associated_genes'], context['associated_drugs'], context['phenotypes']

Relationship Types in PrimeKG

The graph contains several key relationship types including:

protein_protein: Physical PPIs
drug_protein: Drug target/mechanism associations
disease_gene: Genetic associations
drug_disease: Indications and contraindications
disease_phenotype: Clinical signs and symptoms
gwas: Genome-wide association studies evidence

Best Practices

Use specific IDs: When using get_neighbors, ensure you have the correct ID from search_nodes.
Context first: Use get_disease_context for a broad overview before diving into specific genes or drugs.
Filter relationships: Use the relation_type filter in get_neighbors to focus on specific evidence (e.g., only drug_protein).
Multiscale integration: Combine with OpenTargets for deeper genetic evidence or Semantic Scholar for the latest literature context.

Resources

Scripts

scripts/query_primekg.py: Core functions for searching and querying the knowledge graph.

Data Path

Data: /mnt/c/Users/eamon/Documents/Data/PrimeKG/kg.csv
Total nodes: ~129,000
Total edges: ~4,000,000
Database: CSV-based, optimized for pandas querying.

GitHub Repository

K-Dense-AI/claude-scientific-skills

Path: skills/primekg

agent-skillsai-scientistbioinformaticschemoinformaticsclaudeclaude-skills

FAQ

Frequently asked questions

What is the primekg skill?

primekg is a Claude Skill by K-Dense-AI. Skills package instructions and resources that Claude loads on demand, so Claude can perform primekg-related tasks without extra prompting.

How do I install primekg?

Use the install commands on this page: add primekg to Claude Code as a plugin, or clone its repository into your skills directory, then restart Claude so it picks up the skill.

What category does primekg belong to?

primekg is in the Other category, tagged data.

Is primekg free to use?

Yes. primekg is listed on AIMCP and free to install. It runs inside Claude, so no separate service account is required to use the skill itself.

Related Skills

llamaguard

Other

LlamaGuard is Meta's 7-8B parameter model for moderating LLM inputs and outputs across six safety categories like violence and hate speech. It offers 94-95% accuracy and can be deployed using vLLM, Hugging Face, or Amazon SageMaker. Use this skill to easily integrate content filtering and safety guardrails into your AI applications.

View skill

cost-optimization

Other

This Claude Skill helps developers optimize cloud costs through resource rightsizing, tagging strategies, and spending analysis. It provides a framework for reducing cloud expenses and implementing cost governance across AWS, Azure, and GCP. Use it when you need to analyze infrastructure costs, right-size resources, or meet budget constraints.

View skill

sports-betting-analyzer

Other

This Claude Skill analyzes sports betting markets including spreads, over/unders, and prop bets by examining historical trends and situational statistics to identify value bets. It provides structured markdown output with actionable recommendations for educational purposes. Developers should use this for sports betting analysis tools while noting it's designed for entertainment/education only.

View skill

quantizing-models-bitsandbytes

Other

This skill quantizes LLMs to 8-bit or 4-bit precision using bitsandbytes, achieving 50-75% memory reduction with minimal accuracy loss. It's ideal for running larger models on limited GPU memory or accelerating inference, supporting formats like INT8, NF4, and FP4. The skill integrates with HuggingFace Transformers and enables QLoRA training and 8-bit optimizers.

View skill