# Shaide

> Distributed, multi-model LLM inference on Kubernetes you own - axem-solutions/shaide

- **Website:** https://github.com/axem-solutions/shaide
- **Pricing:** open source
- **Categories:** Infrastructure, Developer Tools
- **Tags:** infrastructure, developer-tools, ai-agents
- **Platforms:** CLI
- **Last verified:** 2026-09-10
- **Canonical page:** https://linkrena.com/tools/shaide

## About

Distributed, multi-model LLM inference on Kubernetes you own - installed by a single command, all the way down to air-gapped clusters.

Standing up enterprise AI infrastructure today means assembling a long list of moving parts - an inference engine, a serving orchestrator, a gateway, a model registry, storage, observability, and that is only the beginning. Each one has to be chosen, configured and glued to the next, component by component, then again for every environment. shaide ships that whole stack as one installable platform.

The entire platform is managed as infrastructure as code . Every layer - the internal registry, the gateway, model serving and the application layer - is defined as a Pulumi project in this repository, so your AI infrastructure is versioned and reproducible.

And it stays inside your perimeter. shaide is built for organisations whose data cannot leave their infrastructure: regulated industries, defence, public sector, or anyone who simply will not send prompts to a third-party API. There is nothing phoning home and no dependency on a vendor's cloud - including fully air-gapped installations with no internet access at all.

Sovereign by design. Everything runs in your infrastructure. An internal OCI registry mirrors every container image and model weight, so a cluster can operate with no egress whatsoever.

## Related tools

- [Supabase](https://linkrena.com/tools/supabase): Build production-grade applications with a Postgres database, Authentication, instant APIs, Realtime, Functions, Storage and Vector embeddings.
- [Hillock](https://linkrena.com/tools/hillock): Local, gradient-free neuro-symbolic memory engine combining Hyperdimensional Computing (HDC/VSA), Hebbian plasticity, and graph triples for offline AI.
- [Web LLM](https://linkrena.com/tools/web-llm): High-performance In-browser LLM Inference Engine .
- [Agentic Data Kernel](https://linkrena.com/tools/agentic-data-kernel): Temporal knowledge and durable workflow infrastructure for software agents - Jason-Doyle/agentic-data-kernel
- [Picolm](https://linkrena.com/tools/picolm): SIMD/GPU pure C inference for Llama 2-family, GPT-2, Gemma-3n, and Qwen 3.x GGUF files.
- [Pushin.eu](https://linkrena.com/tools/pushin-eu): Sovereign, GDPR-native git hosting — code, issues, and CI, all on European soil.
