Serve Qwen3.8-2.4T-A95B, a 2.4T-Parameter Model, with Configurable Reasoning on NVIDIA GB300 NVL72

Alibaba released the open weights for Qwen3.8-2.4T-A95B (Qwen3.8-Max), its largest open-weight model, bringing near-frontier capabilities to the open…

Alibaba released the open weights for Qwen3.8-2.4T-A95B (Qwen3.8-Max), its largest open-weight model, bringing near-frontier capabilities to the open ecosystem. It has 2.4T total parameters with 95B activated per token. It has 2.4T total parameters with 95B activated per token. It’s a fine-grained mixture of experts (MoE) architecture with a hybrid of full and linear attention, a context window of…

Source

Leave a Reply

Your email address will not be published.

Previous post Dawn of War 4 has been delayed: ‘We saw an opportunity to further enhance Dawn of War 4 and deliver the quality of experience we’ve always envisioned’
Next post CD Projekt cuts more than 20% of The Witcher spinoff development team ‘to reflect the project’s needs at this stage of development’