Soika AI Workstation SM RTX PRO 5000
Serve a bigger model without cutting it down first.
48 GB of GDDR7 per GPU changes what fits: longer context, larger batches and higher-precision weights, with the Soika Enterprise licence and A2A network option included.
2× NVIDIA RTX PRO 5000 Blackwell
- GPU memory
- 48 GB
- Bandwidth
- 1,344 GB/s
- CUDA cores
- 14,080
- Interface
- PCIe Gen 5
GPU memory
48 GB
Bandwidth
1,344 GB/s
CUDA cores
14,080
Interface
PCIe Gen 5
48 GB per GPU for larger models and longer context.
The volume choice for departmental deployments — enough memory to serve a capable model and its retrieval index without quantising down.
- 48 GB of GDDR7 memory per GPU
- 1,344 GB/s of memory bandwidth
- 14,080 CUDA cores, 384-bit memory interface
- Agent-to-Agent network option included
The full configuration
- Parts
- 3 years
- Live service
- 3 years, engineer-led
- Support tier
- Standard 3-year; on-site service and SLA available through partners
- Model
- SOIKASMRTXPROB5000
- Device type
- Multi-GPU workstation, Xeon 60-core platform
- Graphics
- 2× NVIDIA RTX PRO 5000 Blackwell
- Processor
- Intel® Xeon® w9-3595X — 60 cores / 120 threads
- Memory
- 8× 64GB DDR5-4800 2Rx4 ECC RDIMM (512GB)
- Storage
- 8TB NVMe
- GPU memory
- 48 GB GDDR7 with ECC
- Memory interface
- 384-bit
- Memory bandwidth
- 1,344 GB/s
- Network
- 2× 10GbE RJ45 + 1× management LAN
Licensed, installed and tuned before it ships
Soika hardware exists to remove the setup problem. GPU clustering, model serving and agent design are configured at the factory, not on your time.
Soika Stack
GPU clustering, LLM deployment and inference management, pre-configured and licensed.
Soika Mockingjay
No-code AI agent design and agent-fleet management with the Enterprise licence.
Ubuntu LTS
Hardened Linux base image with the full CUDA and container toolchain already in place.
Pre-loaded model catalogue
- Qwen
- Llama
- DeepSeek
- Mistral
- MONAI
- Meditron
- BLOOM
- Calme
- 200+ optimised models
What teams run on these systems
The same box serves very different work depending on who owns it.
Healthcare
Record summarisation, diagnostic support, research and treatment planning.
Government
Citizen-service agents, smart-city systems and in-country model hosting.
Finance
Reporting, risk assessment, trading research and investment analysis.
Legal
Contract drafting, case-law summarisation and legal co-pilots.
Software
Code generation, debugging, application design and internal developer tooling.
Research & academia
Advanced data analysis, reasoning tasks, language and AI research.
E-commerce
Personalised recommendation and supply-chain intelligence.
Edge & autonomous systems
Local inference for robotics, drones and IoT without cloud access.
Other workstations
Configure a SM RTX PRO 5000
Tell us the models you plan to run and the environment it will live in. We will confirm the configuration, power and cooling budget, and introduce the partner who will deliver and commission it.