中文简介
这是 Qwythos-9B-Claude-Mythos-5-1M 的 GGUF 量化仓库,可用于 llama.cpp、Ollama、LM Studio 等 GGUF 运行环境。上游称模型基于 Qwen3.5-9B 继续训练,具备推理、函数调用、多模态和长上下文能力;训练数据来源、1M 上下文效果与基准提升均需参考基础模型卡并自行验证。
上游模型卡 / 数据集卡
这是 Qwythos-9B-Claude-Mythos-5-1M 的 GGUF 量化仓库,可用于 llama.cpp、Ollama、LM Studio 等 GGUF 运行环境。上游称模型基于 Qwen3.5-9B 继续训练,具备推理、函数调用、多模态和长上下文能力;训练数据来源、1M 上下文效果与基准提升均需参考基础模型卡并自行验证。
Qwythos-9B-Claude-Mythos-5-1M-GGUF Developed by Empero** GGUF quantizations of **empero-ai/Qwythos-9B-Claude-Mythos-5-1M** for llama.cpp, Ollama, LM Studio, jan, KoboldCpp, and other GGUF runtimes. Qwythos-9B is a full-parameter reasoning model post-trained on over 500 million tokens of high-quality Claude Mythos / Claude Fable traces with chain-of-thought generated in-house by Empero AI's internal `rethink` tool. It dominates the base Qwen3.5-9B under matched evaluation (**+34 pts MMLU, +30 pts gsm8k-strict, +19 pts gsm8k-flex**), supports **native function calling** per the Qwen3.5 spec, and ships with a **1,048,576-token (1M) context window** via YaRN rope-scaling enabled by default. For full training details, evaluation numbers, and capability writeup, see the **base model card**. Files Normal text weights — fixed v3 replacements | File | Quant | Size | Notes | |---|---|---|---| | `Qwythos-9B-Claude-Mythos-5-1M-Q4_K_M.gguf` | Q4_K_M | 5.24 GiB / 5.63 GB | **recommended default** — fixed v3, best compatibility | | `Qwythos-9B-Claude-Mythos-5-1M-Q5_K_M.gguf` | Q5_K_M | 6.02 GiB / 6.47 GB | fixed v3, balanced quality / size | | `Qwythos-9B-Claude-Mythos-5-1M-Q6_K.gguf` | Q6_K | 6.85 GiB / 7.36 GB | fixed v3, high quality | | `Qwythos-9B-Claude-Mythos-5-1M-Q8_0.gguf` | Q8_0 | 8.87 GiB / 9.53 GB | fixed v3, near-lossless | | `Qwythos-9B-Claude-Mythos-5-1M-BF16.gguf` | BF16 | 16.69 GiB / 17.92 GB | fixed v3, full precision conversion base | If you don't know which to pick, **Q4_K_M is the right starting point** — it's the smallest practical quant with good quality preservation. MTP-enabled text weights — fixed v3 variants These include the restored Qwen3.5-compatible MTP head inside the GGUF. Use them with llama.cpp builds that support MTP draft speculation, for example `--spec-type draft-mtp`. | File | Quant | Size | Notes | |---|---|---|---| | `Qwythos-9B-Claude-Mythos-5-1M-MTP-Q4_K_M.gguf` | Q4_K_M + MTP | 5.48 GiB / 5.89 GB | **recommended MTP default** | | `Qwyth
上游文件元数据
.gitattributes2.41 KBmmproj-Qwythos-9B-Claude-Mythos-5-1M-F16.gguf875.63 MBQwythos-9B-Claude-Mythos-5-1M-BF16.gguf16.69 GBQwythos-9B-Claude-Mythos-5-1M-MTP-BF16.gguf17.14 GBQwythos-9B-Claude-Mythos-5-1M-MTP-Q4_K_M.gguf5.48 GBQwythos-9B-Claude-Mythos-5-1M-MTP-Q5_K_M.gguf6.26 GBQwythos-9B-Claude-Mythos-5-1M-MTP-Q6_K.gguf7.09 GBQwythos-9B-Claude-Mythos-5-1M-MTP-Q8_0.gguf9.11 GBQwythos-9B-Claude-Mythos-5-1M-Q4_K_M.gguf5.24 GBQwythos-9B-Claude-Mythos-5-1M-Q5_K_M.gguf6.02 GBQwythos-9B-Claude-Mythos-5-1M-Q6_K.gguf6.85 GBQwythos-9B-Claude-Mythos-5-1M-Q8_0.gguf8.87 GB
本页面为橙子AI科技的中文整理与服务说明,不代表资源作者或平台官方页面。实际许可、访问和使用条件以上游原文为准。