LLMKube – Kubernetes Operator for Local LLM Inference
LLMKube is an open-source Kubernetes operator that manages local LLM inference across NVIDIA, Apple Silicon Metal, and AMD GPUs from a single YAML spec, with an OpenAI-compatible API.
Tag
5 posts tagged #apple-silicon
Browse 5 posts tagged Apple Silicon, including practical setup notes, reviews, comparisons, and workflow patterns for engineers working with AI tools.
LLMKube is an open-source Kubernetes operator that manages local LLM inference across NVIDIA, Apple Silicon Metal, and AMD GPUs from a single YAML spec, with an OpenAI-compatible API.
Phi-3-MLX brings Microsoft's Phi-3.5-vision and Phi-3.5-mini models to Apple Silicon Macs via MLX. Install with pip and run locally in minutes.
RCLI is a local voice AI for macOS running a complete STT + LLM + TTS + VLM pipeline on Apple Silicon. Sub-200ms latency, 40 voice-controlled macOS actions, local RAG over your documents.
RCLI runs a full STT + LLM + TTS pipeline on Apple Silicon with MetalRT acceleration. No cloud, no API keys — just 40 macOS voice actions and sub-200ms latency.
Lume is a free open-source CLI to create macOS and Linux VMs on Apple Silicon. Zero-touch setup, MCP integration, isolated sandboxes for AI agents. Native Virtualization Framework — no emulation.