Two new multimodal models are now available on Dell Enterprise Hub. Qwen3.8-Flash-Next-FP8, a MoE model with 125B parameters and just 6B active parameters, is the preview version of Qwen4, pairing efficient sparse attention with FP8 weights for long-context workloads. Meanwhile, GLM-5.3-Flash, a 320B MoE model with 18B active parameters, is built for coding and agentic tasks. Both frontier open models achieve great performance and you already can deploy them on-premise on Dell platforms.



