st-k8s MCP Server
Exposes Kubernetes cluster management tools to LLMs, enabling querying pods, deployments, logs, metrics, and managing port forwards via natural language.
README
ST-K8s
View and chat to your Kubernetes cluster and container log files.
brew tap bhf/st-k8s
brew install st-k8s
st-k8s
Features a dashboard (with a K9s inspired dark theme and keyboard navigation), REST API, port forwarding management, resource monitoring, and MCP server. In browser AI chat powered by the Copilot SDK, any OpenAI API compatible provider, or local WebLLM models (requires WebGPU support).

Uses Github Projects for planning and tracking.
Keyboard Navigation
The dashboard supports K9s-style keyboard navigation. Press : to open the command palette and navigate between resources using commands or aliases:
:podsor:po:deploymentsor:deploy:servicesor:svc- ...and many more standard K8s shortcuts.

Log Viewer
View, copy and download streaming logs.

Port Forwarding
Manage Kubernetes port forwarding sessions directly from the dashboard or through AI chat. Supports both Pods and Services.
- Dynamic Config: Specify target ports and local interface bindings.
- Service Mapping: Automatically resolves Service targets to active Pods.
- Agentic Control: Start or stop forwards using natural language through the Copilot integration or MCP server.

Resource Monitoring
Monitor CPU and memory usage for Nodes and Pods directly in the dashboard using interactive charts. Requires that your cluster has metrics server installed.
- Real-time Data: Fetches live metrics from the Kubernetes Metrics Server.
- Node Metrics: View cluster-wide resource utilization across all nodes.
- Pod Metrics: Inspect resource consumption for individual pods in any namespace.
- Visual Charts: Interactive Recharts-based visualizations for easier performance analysis.
Hardware Acceleration & WebGPU
ST-K8s supports local AI models running directly in your browser using WebLLM. This requires WebGPU and hardware acceleration to be enabled.
Google Chrome / Chromium
- Ensure you are on a recent version of Chrome.
- Enable WebGPU: Paste
chrome://flags/#enable-unsafe-webgpuinto your address bar and set it to Enabled. - Enable Vulkan (Linux/Windows): Paste
chrome://flags/#enable-vulkanand set it to Enabled. - Relaunch Chrome.
Mozilla Firefox
- Type
about:configin the address bar. - Search for
dom.webgpu.enabledand set it to true. - Search for
gfx.webgpu.force-enabledand set it to true if WebGPU doesn't work by default. - MacOS users may also need to ensure
gfx.webrender.allis true.
Verification
You can verify WebGPU support by visiting webgpu.github.io/webgpu-samples. If the samples run, ST-K8s will be able to load local models.
Table of Contents
- ST-K8s
How to Run
Using Homebrew (macOS/Linux)
The easiest way to install and run st-k8s is via Homebrew:
brew tap bhf/st-k8s
brew install st-k8s
st-k8s
From Source
To use the browser based chat feature make sure you install the Copilot CLI.
git clone https://github.com/bhf/st-k8s
cd st-k8s
npm run build
npm run start
Using the st-k8s CLI
You can install the project as a global CLI to run the app using the st-k8s command.
# From the repo root — install globally (or publish and install from a registry)
npm install -g .
# During development, link the local package to make `st-k8s` available globally
npm link
# Then launch the app with the CLI (it will build if no build exists)
st-k8s
Notes:
npm install -g .requires appropriate permissions (usesudoon some systems).npm linkis useful when iterating locally — run it once from the repo root.- The
st-k8scommand will attempt to use a Next.js standalone server if present (fromnext build), otherwise it runsnpm run start.
Running Tests
This project uses Vitest for testing.
# Run all tests
npm test
# Run tests in watch mode
npm test -- --watch
# Run tests with coverage
npm run test:coverage
End-to-End Tests
This project uses Playwright for End-to-End testing.
# Run E2E tests
npm run test:e2e
API
Swagger spec available at http://localhost:3000/openapi.json after starting the server or from the public folder.
Model Context Protocol (MCP) Server
This project includes an MCP server that exposes Kubernetes tools to LLMs over stdio. Here are some example uses:
- List of pods
- Rank containers by their memory requests and limits
- Summary of the last events in the namespace
- Get the last 100 lines of logs for a specific pod


Features
Exposes read-only Kubernetes operations as tools:
list_namespaceslist_podslist_deploymentslist_serviceslist_daemonsetslist_replicasetslist_statefulsetslist_ingresseslist_endpointslist_eventslist_pvcslist_nodeslist_configmapslist_jobslist_cronjobslist_serviceaccountslist_roleslist_rolebindingsget_pod_logslist_port_forwardsstart_port_forwardstop_port_forwardget_node_metricsget_pod_metrics
Running the MCP Server
Make sure to auth your kubectl context in your preferred way before running the MCP server.
You can run the MCP server directly using:
npm run mcp
You can also run it from VSCode or any MCP-compatible client by configuring it as shown below.
Configuring for VSCode
Add the following to your mcp.json
{
"servers": {
"k8s-tools": {
"command": "npm",
"args": ["run", "mcp"],
"cwd": "/absolute/path/to/st-k8s",
"disabled": false,
"autoApprove": []
}
}
}
Make sure to replace /absolute/path/to/st-k8s with the actual path to this repository on your machine.
LLM Integration Techniques
This project uses several LLM-based techniques to enhance the development lifecycle and user experience. These artifacts are located in the .github directory:
- Agents: Domain-specific personas which embody specialized knowledge for consistent code generation.
- Instructions: Contextual guidelines that enforce coding standards and architectural patterns.
- Skills: Reusable capabilities that allow the model to perform complex tasks.
- Prompts: Curated prompt templates ensuring high-quality, reproducible outputs for specific tasks.
High Level Architecture

Accessibility
We are committed to making the dashboard accessible to all users. Please refer to our Accessibility Statement and Guidelines for details on current status, findings, and remediation plans.
Security
We take security seriously. Please refer to our Security Review for details on our security posture, findings, and recommendations.
Recommended Servers
playwright-mcp
A Model Context Protocol server that enables LLMs to interact with web pages through structured accessibility snapshots without requiring vision models or screenshots.
Magic Component Platform (MCP)
An AI-powered tool that generates modern UI components from natural language descriptions, integrating with popular IDEs to streamline UI development workflow.
Audiense Insights MCP Server
Enables interaction with Audiense Insights accounts via the Model Context Protocol, facilitating the extraction and analysis of marketing insights and audience data including demographics, behavior, and influencer engagement.
VeyraX MCP
Single MCP tool to connect all your favorite tools: Gmail, Calendar and 40 more.
graphlit-mcp-server
The Model Context Protocol (MCP) Server enables integration between MCP clients and the Graphlit service. Ingest anything from Slack to Gmail to podcast feeds, in addition to web crawling, into a Graphlit project - and then retrieve relevant contents from the MCP client.
Kagi MCP Server
An MCP server that integrates Kagi search capabilities with Claude AI, enabling Claude to perform real-time web searches when answering questions that require up-to-date information.
E2B
Using MCP to run code via e2b.
Neon Database
MCP server for interacting with Neon Management API and databases
Exa Search
A Model Context Protocol (MCP) server lets AI assistants like Claude use the Exa AI Search API for web searches. This setup allows AI models to get real-time web information in a safe and controlled way.
Qdrant Server
This repository is an example of how to create a MCP server for Qdrant, a vector search engine.