01 / COMPUTE
M3 UltraMac Studio
256 GB unified memory, 2 TB SSD and built-in 10GbE. Quiet enough for the office; large enough for serious local inference.
One secure local AI node that runs open models privately and reaches frontier APIs only when policy allows.
256 GB unified memory, 2 TB SSD and built-in 10GbE. Quiet enough for the office; large enough for serious local inference.
Run useful quantized open-weight models locally. Smaller 8B–32B models remain fast and inexpensive for routine agents.
No server room, liquid cooling or specialist datacenter operations required. Typical mixed use will be lower.
Sources: Apple Mac Studio specifications · Apple power and thermal data
| Component | Purpose | Budget |
|---|---|---|
| Mac Studio M3 Ultra · 256 GB · 2 TB | Local AI compute | $6K–$8K |
| 8 TB encrypted Thunderbolt NVMe | Fast model storage | $700–$1.5K |
| Encrypted NAS · 20–40 TB usable | Versioned backup | $2K–$4K |
| 1,500 VA pure-sine UPS | Clean power + shutdown | $600–$1.1K |
| Firewall + managed switch | Isolated private network | $500–$1.5K |
| Complete starter system | $10K–$15K | |
Gateway reference: LiteLLM · Local runtimes: Ollama · llama.cpp · MLX
Decision: approve a $10K–$15K sovereign AI starter node and keep every model replaceable.
Security reference: Apple FileVault and Secure Enclave