Guide
Layering a CLI agent's models under a budget cap
Under a budget-capped subscription like opencode Go, picking a model isn't about token price but requests per 5 hours, the one quality source still alive (vals.ai), and tool-calling reliability — the decisive criterion no benchmark measures. The result: one model per preset, a frontier delegated outside the agent, and tools constrained rather than instructed.
model-routing opencode-go inference-cost