Website · Models & datasets · Build log · X · Get in touch
I'm SSH. I build open-weight models, fast inference stacks, and tools that run on your machine. I like measured results, source you can inspect, and projects you can actually use.
The workstation above is a detailed Blender reconstruction of my real machine, rendered with blue lighting and spinning fans. Explore the interactive version on ssh.codes to rotate it and change the lighting.
The Grug Series — Open-weight models trained to spend fewer tokens thinking, with GGUF builds for local use.
Auto — Local permission intelligence for coding agents. Checks proposed tool calls against your request and agent history, with support for Pi, OpenCode, and Hermes.
Black Hole Benchmark — One prompt, different models, actual compiled C++ renders. Inspect the submissions →
| Project | What I’m building |
|---|---|
| GLM on Blackwell | Single-GPU GLM-5.3-Flash serving, RAM expert caching, DFlash2, and reproducible benchmarks. |
| Qwen on Blackwell | Qwen3.8-Flash-Next inference tuning, full-context serving, and correctness checks on an RTX PRO 6000. |
| Second Opinion | Let Claude bring other AI models into the work through MCP. |
| Videopaper | Native macOS live wallpapers, including a real-time black hole ray tracer in Metal. |
| Codex RPC | Discord Rich Presence for the Codex CLI, with animated activity states. |
Models & inference PyTorch · LoRA · GGUF / llama.cpp · CUDA · MLX
Interfaces & systems Python · C++ · TypeScript · Swift · Metal · WebGPU · Three.js · MCP
Hardware & hosting RTX PRO 6000 Blackwell · Apple Silicon · Raspberry Pi · Docker
Contribution trail
Have something interesting to build? me@ssh.codes


