Kayon
An honest local-LLM workstation.
Every quant of every model gets a verdict computed for your GPU at your context length, and you can expand any one of them to see the arithmetic it came from.
- Windows 10/11 x64 · NVIDIA optional
- Adopts Ollama models by hard link, zero bytes moved
- llama.cpp ships inside the installer
- No account, no cloud, MIT licensed
Not yet code-signed, so SmartScreen will warn. Code signing is on the roadmap.