Qwythos-9B (Claude-Mythos-5 1M)
9B parameter open-weight model. Community model from Empero AI built on Qwen3.5-9B. No official Ollama library tag (registry checked 2026-07-06) — run via vLLM or SGLang, or community GGUF builds in llama.cpp / LM Studio. Full 1M-token context requires tensor-parallel multi-GPU or KV-cache offloading; on a single 12 GB GPU target Q4_K_M at reduced context.
Empero AI · Qwythos
Editorial review
VRAM figures are empirical estimates. Actual usage varies by runtime, context length, and system configuration. Verify on your specific hardware before production use.
Will Qwythos-9B (Claude-Mythos-5 1M) run on your machine?
Qwythos-9B (Claude-Mythos-5 1M) is 9B parameters and needs 8 GB of VRAM at Q4_K_M — 6.5 GB of weights plus 1.5 GB of runtime overhead for the inference server itself.
VRAM by quantization
| Quantization | Weights | Needs (with overhead) | Quality |
|---|---|---|---|
| Q4_K_M | 6.5 GB | 8 GB | good |
| Q8_0 | 11 GB | 12.5 GB | high |
Fit on common hardware at Q4_K_M
| Hardware | Memory the model can use | System RAM | Verdict |
|---|---|---|---|
| CPU Only | None (CPU only) | 16 GB | CPU offload |
| RTX 4060 Laptop | 8 GB | 16 GB | Tight |
| RTX 3060 (12GB) | 12 GB | 32 GB | Comfortable |
| RTX 4060 Ti (16GB) | 16 GB | 32 GB | Comfortable |
| RTX 3090 | 24 GB | 64 GB | Comfortable |
Comfortable means VRAM clears the requirement by 2 GB or more. Tight means it covers the requirement with no margin. CPU offload means the model does not fit in VRAM but system RAM is at least 1.6× the weights, so it will run at reduced speed — expect roughly 1–5 tokens per second. Figures are weights plus a fixed runtime overhead and exclude KV-cache growth, which scales with context length.
Need more hardware for Qwythos-9B (Claude-Mythos-5 1M)? Open the PC Builder for the 7B / 8B tier →
This catalog has no verified local Ollama tag for this checkpoint, so fit grades do not include a local run command.
Ready to run this model locally?
Find a compatible interface in our Local AI Tools directory →