China’s Alibaba released a 27-billion-parameter open model.

FP8 quantization halves memory, letting the model fit on one commodity data-centre GPU.

Qwen3.8-27B natively handles a 262,144-token context, extensible to 1,000,000, and reads images and video, its team said.

A hosted Qwen Cloud version with 1 million tokens by default is coming soon, the company said.

Sources: Hacker News