Q4_K_M · 4K context
K2 Horizon 7B 7B dense model · estimated Q4_K_M memory about 6GB
A 7B dense model that sits between quality and local execution burden. We support the official 512K contexts, but on desktops it is better for responsiveness and memory management to limit this to the length required for the task.
- Estimated Q4_K_M memory
- About 6GB
- total parameters
- 7B
- active parameter
- 7B
- native context
- 512K
Start with a model
Find hardware for this model
Also running document search?
Reserve this for embedding and reranking models sharing the same GPU or unified memory. Leave it at none if they run separately on the CPU.
Matching devices 44· 5 shown
Comfortable fit · Q4_K_M · Headroom 8.0GB
₩1,180,000
19.8 ~ 22.1 tok/s
Comfortable fit · Q4_K_M · Headroom 8.0GB
₩1,525,000
19.2 ~ 21.5 tok/s
Comfortable fit · Q4_K_M · Headroom 8.0GB
₩1,669,000
25.9 ~ 28.9 tok/s
Comfortable fit · Q4_K_M · Headroom 15.0GB
₩1,925,000
19.8 ~ 22.1 tok/s
Comfortable fit · Q4_K_M · Headroom 15.0GB
₩1,950,000
19.2 ~ 21.5 tok/s
Price is the range midpoint, with a basic host added for GPU-only prices. Speed is an estimate for the same model and 4K input.
If you want to choose equipment again
You can first narrow down the model configurations based on budget, noise level, operating system and whether it's acceptable used.