Qwen/Qwen3.8-Flash-Next-FP8Qwen3.8 Flash Next is an LLM listed in RunInfra Model APIs. RunInfra serves it as Qwen/Qwen3.8-Flash-Next-FP8 at $0.12 per 1M input tokens and $0.40 per 1M output tokens. Its context window is 262,144 tokens. The API provides OpenAI-compatible chat completions and Anthropic-compatible Messages, POST /v1/messages.
USD, pay per token
Confirm how your client reaches this model.
Check the limits your workload must fit.
See which request modes the API supports.
Set RUNINFRA_GATEWAY_KEY to your workspace API key before using an example.
Verify the company and operating credentials behind this API.