SophyAI Super Local Intelligence

2026-10-06

Qwen3.8-Flash-Next, a preview of Qwen4

Qwen3.8-Flash-Next FP8 runs on 4× RTX PRO 6000 at about $1.3/h measured, or in Q4 GGUF on an RTX 5090.

QwenOpenWeightsvLLMGGUF

Qwen3.8-Flash-Next is the preview of the Qwen4 family. The catalog has three routes, from cheapest to fastest.

The recipes

Prices are the ones measured in the catalog and follow the Vast market: the app shows the real price before every start.

For security work

It's a good base for code review and triage on customer repositories: light enough to keep running for a long time, without sending code to an external API.

← All posts