Files
ai-agent/runner/cli.mjs
Gabriel Vidal 62a60ba5fb feat(models): run Qwen 3.8 27B on the pi harness, locally or via OpenRouter
Adds both halves of the new release as composer chips: `qwen/qwen3.8-27b`
on OpenRouter and ggml-org's Q4_K_M build served by the EVOX2 LM Studio.

The local one exposed a wrong assumption. Provider was inferred from the
id's shape — a `/` meant OpenRouter — but LM Studio serves this model as
`ggml-org/qwen3.8-27b`, prefix and all, while OpenRouter has a
`qwen/qwen3.8-27b` of its own. The two differ by provider, not by shape,
so the runner now resolves the provider from `~/.pi/agent/models.json`
(a model must be listed there for pi to run it locally anyway) and only
falls back to the shape for ids it doesn't know — which keeps the whole
OpenRouter catalogue runnable without registering ids by hand. That also
fixes `laguna-s-2.1`, which the bare-id rule aimed at LM Studio's port
instead of its own llama.cpp one.

Session pricing carried the same assumption, so a local run with a
prefixed id would have been billed at OpenRouter's estimate; it now
prices from the declared provider and stays $0.

Qwen 3.6 stays pickable, demoted from `latest`; the pi default model
follows to 3.8.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
2026-08-15 12:42:17 +02:00

19 KiB
Executable File