动态

介绍使用llama serve运行Qwen模型的命令

Victor M
简单:

llama serve -hf ggml-org/Qwen3.8-27B-GGUF --spec-type draft-mtp
动态Victor M2026-08-14原文

相关内容