Qwen3.8 speed on AxForge
How fast is Qwen3.8 on AxForge?
qwen3.8-27b-nvfp4
How fast is Qwen3.8 on AxForge?
AxForge currently cites a community GB10 result of about 25.1 tokens/s for Qwen3.8 27B NVFP4 + MTP at 4k context.
That is not the same thing as a guaranteed serverless API speed, and AxForge explicitly says it has not published its own tuned number yet.
Real speed changes with prompt length, concurrency, serving engine and batching, so I would not turn 25.1 t/s into an SLA.
Does Qwen3.8 understand images too?
The Qwen3.8 27B model family is vision + text capable.
Whether a specific hosted endpoint exposes every upstream multimodal feature is something I would verify against the current API docs before promising a particular image workflow.
For a FAQ, I would say Qwen3.8 is a vision-language model, but point users to the current endpoint documentation for supported request formats and limits.
Sign in to reply.
Posting guidelines
Accounts that break these can lose forum access — paid plans included. Read the full guidelines →
Related topics