Blog

Working notes.

Technical writing from the team about routing internals, benchmark data, EU compliance, and the rare opinion piece. Quiet by design, because we'd rather publish three good things a year than thirty mediocre ones.

May 2026
Benchmark

What 240 local-model benchmarks say

We ran the same 16-category eval across qwen2.5-coder, deepseek-r1, gemma3, phi4-mini and the qwen3 family on three Apple Silicon configs. The results say less about the models and more about hardware sensitivity.

Read more →