Tag: qwen
All the articles with the tag "qwen".
-
Deploying Qwen3.6-27B-W8A8 on Huawei Ascend 910B with vLLM Ascend
Updated:A field-tested deployment guide for Qwen3.6-27B-W8A8 on Huawei Ascend 910B/A2 with vLLM Ascend v0.19.1rc1 and vLLM v0.19.1, covering CANN 8.5.2, Python 3.11, torch-npu 2.9, Triton Ascend, native kernels, ATB, graph mode, and the failure fixes needed to make deployment repeatable.
-
From Zero to a Working Coding Assistant: Deploying Qwen3.5-9B on a Huawei Cloud 910B Notebook
A practical guide to deploying Qwen3.5-9B on Huawei Cloud ModelArts with Ascend 910B, including model selection, runtime setup, and opencode integration for coding assistance.
-
Qwen3-8B Deployment on Huawei Cloud ModelArts Notebook (Ascend 910B + PyTorch 2.6.0)
Comprehensive operational runbook: deploy Qwen3-8B on Huawei Cloud ModelArts Notebook with Ascend 910B and PyTorch 2.6.0 (Ascend build), including custom image, model download, inference, and optional API serving.