Quick Answer

For running local llms in Hermanus, build around RTX 4070 Super (12GB) minimum, RTX 4090 (24GB) to run larger models locally and Ryzen 7 7700 and 32-64GB of system RAM. 24GB of VRAM runs a quantised 30B-parameter model entirely on the GPU. Plan on roughly R20,000-R35,000 for a machine that handles this comfortably. Vram caps which models you can load: 12gb handles 7b-13b models, 24gb opens up 30b-class models.

What this workload actually needs

For running local llms, VRAM caps which models you can load: 12GB handles 7B-13B models, 24GB opens up 30B-class models. The headline part is RTX 4070 Super (12GB) minimum, RTX 4090 (24GB) to run larger models locally, backed by Ryzen 7 7700 and 32-64GB of system RAM. 24GB of VRAM runs a quantised 30B-parameter model entirely on the GPU. Everything here is stocked at Evetech with a South African warranty.

A sensible parts list

Start with RTX 4070 Super (12GB) minimum, RTX 4090 (24GB) to run larger models locally and Ryzen 7 7700 and 32-64GB of system RAM, add 32GB of DDR5 RAM, a 2TB NVMe SSD and a 750W power supply. That gives you a local LLM workstation that does real work rather than just looking the part. Expect roughly R25,000-R40,000 for this configuration in Hermanus.

Where Hermanus buyers should and should not spend

Resist the urge to overspend on the GPU if the workload is CPU or RAM bound. Vram caps which models you can load: 12gb handles 7b-13b models, 24gb opens up 30b-class models. A fast NVMe and adequate RAM age far better than chasing a flagship card you will not fully use.

FAQ

What GPU is best for running local llms?

Go with RTX 4070 Super (12GB) minimum, RTX 4090 (24GB) to run larger models locally. 24GB of VRAM runs a quantised 30B-parameter model entirely on the GPU, which is the level that makes this workload feel responsive rather than sluggish.

How much should I budget in Hermanus?

Plan on roughly R20,000-R40,000 depending on how heavy your projects get. Buying from Evetech keeps the whole system on one local warranty.

Do I need the most expensive CPU?

Not always. Vram caps which models you can load: 12gb handles 7b-13b models, 24gb opens up 30b-class models, so match the CPU to the workload instead of defaulting to the priciest chip.

Build your local LLM workstation at Evetech, starting with RTX 4070 Super (12GB) minimum, RTX 4090 (24GB) to run larger models locally, and check live stock before you buy.

}}