Testing Qwen3.6-27B, llama.cpp, MTP speculative decoding, 131K context, and whether cheap server GPUs are actually worth it for local AI agents.
I’ve been experimenting with local AI agents for a while now.
I previously wrote a few articles around OpenClaw and running AI locally, but the