Local AI PCs
We build computers that run large language models on your premises — contracts, accounting and source code never leave the company network. For every build we say which model fits and how fast it answers.
The model runs on your hardware. Prompts and documents go nowhere — for work under NDA, with personal data or medical records this is often the only viable route.
Cloud AI is billed per user per month. Your own machine serves the whole team with no query limits and typically pays for itself within a year.
Memory and graphics cards keep getting more expensive. A build bought today is also a hedge against the next wave — and the hardware keeps its value for other workloads.
For an AI PC the decisive number is not FPS but graphics memory. That is why every build lists its total VRAM and the models that will run on it.
Pick a model, quantisation and context length. Required memory is computed from model weights and KV cache — exactly, not estimated. Only the speed is an estimate.
Fits with headroom
Out of 24 GB VRAM across 2 card(s), after driver overhead and the split between cards.
Estimated generation speed
~26 tokens per second (≈ words per second)
Rough estimate from the graphics memory bandwidth. Real speed depends on software and settings — measured values are added to the builds. Two cards add up memory, not speed — model layers run one after another.
✓ Fits: the model runs with headroom even for longer documents
! Tight: runs, but with no room for longer context or a second model
✕ Does not fit: needs more VRAM, or runs slowly from system RAM
Quantisation shrinks a model so it fits in memory: 4-bit (Q4) loses very little quality and is the standard for local deployment, 8-bit is a compromise, 16-bit is the uncompressed original.
Companies that want to use AI on their own data and cannot or will not send it to the cloud.
01
Tell us what you want AI to do and how many people will use it. We recommend a build — or adjust the configuration to fit your model.
02
We assemble the machine, burn it in and test it with the very models you will use. Speed is measured on the actual unit.
03
On request we prepare the environment (drivers, Ollama or LM Studio, web UI) so the first model runs right after power-on. VAT invoice, 24-month warranty.
Tell us what you want AI to do, how many people will use it and where the machine will live. We reply within one business day with a recommendation.