Peridot
ShippedA sovereign, local-first AI inference kernel. No cloud dependency, no telemetry, and you decide how the model runs.
Peridot runs large language models on hardware you own, fully offline, on Windows and Linux. Version 1.6.0 moved to a native CUDA engine: a 27B model now fits entirely on an 8 GB laptop GPU and decodes at 23.6 tokens a second. It can call skills and sandboxed plugins you approve, and Claude Code, Codex or Gemini CLI can hand it private jobs over MCP. You review what goes back. Behavior is set by a versioned constitution file you can read and edit.