Q4_K_M · 4K context

K2 Horizon 7B 7B dense model · estimated Q4_K_M memory about 6GB

A 7B dense model that sits between quality and local execution burden. We support the official 512K contexts, but on desktops it is better for responsiveness and memory management to limit this to the length required for the task.

Estimated Q4_K_M memory
About 6GB
total parameters
7B
active parameter
7B
native context
512K

Start with a model

Find hardware for this model

Also running document search?

Reserve this for embedding and reranking models sharing the same GPU or unified memory. Leave it at none if they run separately on the CPU.

Matching devices 44· 5 shown

Mac mini M4 16GB

Comfortable fit · Q4_K_M · Headroom 8.0GB

₩1,180,000

19.8 ~ 22.1 tok/s

MBA M4 16GB

Comfortable fit · Q4_K_M · Headroom 8.0GB

₩1,525,000

19.2 ~ 21.5 tok/s

Mac mini M6 16GB

Comfortable fit · Q4_K_M · Headroom 8.0GB

₩1,669,000

25.9 ~ 28.9 tok/s

Mac mini M4 24GB

Comfortable fit · Q4_K_M · Headroom 15.0GB

₩1,925,000

19.8 ~ 22.1 tok/s

MBA M4 24GB

Comfortable fit · Q4_K_M · Headroom 15.0GB

₩1,950,000

19.2 ~ 21.5 tok/s

Price is the range midpoint, with a basic host added for GPU-only prices. Speed is an estimate for the same model and 4K input.

If you want to choose equipment again

You can first narrow down the model configurations based on budget, noise level, operating system and whether it's acceptable used.