Qwen3.8 Flash
inferway/qwen3.8-flashQwen's efficiency-focused Flash model.
- Model input
- TextImage
- Model output
- Text
Your prompts stay yours. Your bill stays small.
Already have an account?Sign in
Prompts and completions stay in memory and never touch disk. Request metadata is kept for up to 90 days.
TLS-encrypted, straight to a dedicated endpoint.
No shared queue in the path.
Paged KV, never written to disk.
Content never hits disk; only metadata remains.
Per token, metered from metadata. No seats, no minimums.
Per second of video ordered. No seats, no minimums.
Coming soon
inferway/qwen3.8-flashQwen's efficiency-focused Flash model.
inferway/deepseek-v4.1-flashDeepSeek's efficiency-focused V4.1 Flash release.
inferway/glm-5.3-flashZ.ai's Flash model for coding and agentic workloads.
Free trial — requests are not counted in your account.
Checking live model availability…
curl https://api.inferway.ai/v1/chat/completions \
-H "Authorization: Bearer $INFERWAY_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "inferway/mimo-v2.6-flash",
"stream": true,
"messages": [{"role": "user", "content": "Say hello in one short sentence."}]
}'Docs · Privacy · Transparency