2× NVIDIA RTX 6000 Ada · 256 GB ECC RAM
AI Desktop / Separate workspace
Build local AI systems without mixing them into the normal PC page.
Deterministic model fit, multi-GPU topology, VRAM, RAM, PCIe, power, and cooling checks. All values shown here are demo fixtures.
vLLM sharding enabled
0 hidden by the no-fit gate
240 V circuit review required
AI-only component plan
Demo catalogServer build checklist
Server partSelected AI-capable optionRoleDemo price
AI benchmarks
Workload: Serving / agentsFit and throughput, with uncertainty
Estimated tokens/sec by visible model
Qwen3 8B74–110
Gemma 3 12B50–74
Qwen3 30B-A3B99–147
gpt-oss-20b99–147
Qwen3 32B19–28
VRAM residency and headroom
Qwen3 8B11.2 GB
Gemma 3 12B15.8 GB
Qwen3 30B-A3B24.2 GB
gpt-oss-20b17.6 GB
Qwen3 32B38.4 GB
Concurrent users5–8
Modeled demo range at batch-friendly load
Power efficiency77 tok/s/kW
Directional only; no measured power log
RAM headroom136 GB
After the largest visible model estimate
Models that fit
6 visible · no cloud-only modelsRunnable local model table
ModelArchitectureRequired memoryFitEstimated throughputBackend