Qwen3.8-2.4T
Qwen releases FP8-quantized weights for its 2.4T-parameter Qwen3.8 model, with 95B parameters activated and support for vLLM, SGLang, and other inference stacks. HN discusses its agent and coding benchmarks, upcoming smaller variants, and the enormous hardware required to run it.