Q4_K_M · 4K context

Holo4 27B 27B dense model · estimated Q4_K_M memory about 18GB

It is a 27B dense computer usage model that handles web, desktop, and mobile screen manipulation as well as code and tool calls. Automation is not complete with models alone; you need an executor with screen capture, input tools, and acceptance boundaries.

Estimated Q4_K_M memory
about 18GB
total parameters
27B
active parameter
27B
native context
256K

Start with a model

Find hardware for this model

Also running document search?

Reserve this for embedding and reranking models sharing the same GPU or unified memory. Leave it at none if they run separately on the CPU.

Matching devices 40· 5 shown

RTX 3090 24GB

Comfortable fit · Q4_K_M · Headroom 4.9GB

₩3,150,000

42.3 ~ 46.9 tok/s

RTX 4090 24GB

Comfortable fit · Q4_K_M · Headroom 4.9GB

₩4,825,000

47.4 ~ 52.5 tok/s

2× RTX 3090 48GB (NVLink)

Comfortable fit · Q4_K_M · Headroom 27.4GB

₩5,900,000

57.5 ~ 79.8 tok/s

Studio M4 Max 64GB

Comfortable fit · Q4_K_M · Headroom 38.4GB

₩6,050,000

25.3 ~ 28.1 tok/s

Studio M5 Max 64GB

Comfortable fit · Q4_K_M · Headroom 38.4GB

₩6,925,000

28.5 ~ 31.6 tok/s

Price is the range midpoint, with a basic host added for GPU-only prices. Speed is an estimate for the same model and 4K input.

If you want to choose equipment again

You can first narrow down the model configurations based on budget, noise level, operating system and whether it's acceptable used.