model

Qwythos-9B-Claude-Mythos-5-1M-GGUF

huggingface.co/empero-ai/Qwythos-9B-Claude-Mythos-5-1M-GGUF ↗

486810 downloads·587 likes·text-generation·gguf

from the model card

🚨 v2 released — please redownload the GGUFs The v2 GGUFs replace the original normal filenames and add explicit -MTP- variants. If you downloaded this repo before v2, please redownload your GGUF. Fixes in v2: tokenizer metadata normalized for Qwen3.5 GGUF runtimes; embedded chat template updated for reliable tool/function calling and OpenCode-style agent loops; Qwythos/Empero identity prompt embedded in the template; MTP-enabled variants added as Qwythos-9B-Claude-Mythos-5-1M-MTP-*.gguf; Q4/Q8 tool-calling, MTP draft speculation, 1M-context allocation, and vision projector smoke-tested with current llama.cpp. Use the normal files for maximum runtime compatibility. Use the -MTP- files when you want llama.cpp MTP draft speculation. Qwythos-9B-Claude-Mythos-5-1M-GGUF Developed by Empero GGUF quantizations of empero-ai/Qwythos-9B-Claude-Mythos-5-1M for llama.cpp, Ollama, LM Studio, jan, KoboldCpp, and other GGUF runtimes. Qwythos-9B is a full-parameter reasoning model post-trained on over 500 million tokens of high-quality Claude Mythos / Claude Fable traces with chain-of-thought generated in-house by Empero AI's internal rethink tool. It dominates the base Qwen3.5-9B under matched evaluation (+34 pts MMLU, +30 pts gsm8k-strict, +19 pts gsm8k-flex), supports native function calling per the Qwen3.5 spec, and ships with a 1,048,576-token (1M) context window via YaRN rope-scaling enab…

discussions

recent items

← all models