Radar / AI models / Qwen3.8-Max open weights

Qwen3.8-Max open weights (Qwen3.8-2.4T-A95B)

Alibaba's frontier Qwen-Max tier as downloadable open weights — a 2.4T-param sparse MoE (95B active across 512 experts, 262K native context). Released August, newly trending this week: Fixstars benchmarked day-1 serving at ~2x Kimi-K3 throughput on 8x B300, and AWS published a full SageMaker HyperPod deployment path.

trendingmodels · open-weightsJEV traction 0.67added 2026-10-08
Open Qwen3.8-Max open weights →View on the radar

Why it matters

The open-weight frontier model everyone is benchmarking against this week — a renewed community moment with fresh serving benchmarks, NVFP4 quants, and full vLLM/SGLang support; datacenter-class (~4.8TB bf16), custom licence.

What you could build with it

A research lab or GPU-cloud startup could host Qwen3.8-Max weights on a managed inference cluster, offering frontier-class agentic coding endpoints at self-hosted cost to teams that need data sovereignty.

Does it hold up?

Grounded evidence is unusually strong for a datacenter-class release: Fixstars benchmarked it live on 8x B300 at ~2x Kimi-K3 throughput, AWS documented a full HyperPod path, and community quants (Unsloth GGUF, RadixArk NVFP4, Inferact NVFP4) exist. But the bar is a multi-GPU cluster — individuals cannot run it, so for most developers the 27B sibling is the usable artifact.

Built with Qwen3.8-Max open weights

Learn more

technical deep dive →

First spotted on modelscope: source.

More AI models

OpenAI publishes 722 AI-generated math manuscripts from an unreleased modelOpenAI released 722 mathematical manuscripts in 372 result families to GitHub under Apache-2.0, produced by an…models · JEV 0.74Mistral Large 4 (Le Chonk): 1.05T-param open-weight MoEMistral's new flagship: a 1.05T-parameter mixture-of-experts multimodal model with 49B active params and 1M-token…models · JEV 0.72ElevenLabs Eleven v4 + v4 Turbo: new emotive TTS architecture with a 100ms real-time variantElevenLabs released Eleven v4 and Eleven v4 Turbo on September 28: a new text-to-speech architecture with inline…models · JEV 0.72VoiceStudio: open-source, fully-local ElevenLabs alternative with voice cloning, dubbing and dictation in 646 languagesVoiceStudio is an open-source, fully-local voice AI studio: voice cloning from a 3-second sample, voice design, video…models · JEV 0.71Mistral Forge: enterprise platform for training and continuously improving proprietary modelsMistral AI launched Forge (Sep 28) — a full-lifecycle model training platform (pre-training, SFT, DPO/ODPO, RL…models · JEV 0.69DeepSeek V4-Flash official API: public beta with upgraded agent capabilities, Responses API and Codex supportDeepSeek's official V4-Flash is now a public beta API with massively upgraded agent capabilities - benchmark scores…models · JEV 0.69

Get the week's best AI launches, plus 3 ideas worth building

One email every Saturday. Ranked by traction, not hype. Free.