Skip to main content
Secure Australian office environment
Sovereign AI · On-Premise LLM

Your AI.
On Your Premises.
Under Your Privilege.

0 bytes egressed across active client deployments

We supply, build and integrate NVIDIA DGX Spark on-premise LLM systems for Australian legal, healthcare and finance firms. Open-weight models · SGLang routing · Open WebUI · on-prem RAG. No internet egress.

IOTAI supplies, builds and integrates NVIDIA DGX Spark on-premise AI systems for Australian legal, healthcare and finance firms, with a typical 4-week path from assessment to production. Every deployment runs open-weight models (Qwen, Gemma, Mistral or Phi-4) on hardware you own with zero data egress — no prompt or completion ever leaves your network — and hardware is supplied GST-inclusive through Australian channels with OEM ProSupport.

128 GB
Unified memory
0
Bytes egressed
~1 PFLOP
FP4 compute
4 weeks
To production

Hardware & open-source partners

NVIDIAASUSDellLenovoSGLangOpen WebUIHugging FacePostgreSQL
DGX Spark · sovereignty ledger
DGX Spark · 03 — Sydney HQ
localhost only
GB10 Grace Blackwell
128 GB unified · ~1 PFLOP FP4
qwen3.6-32b-fp8 · 11 active sessions
GPU
76%
a.shah@firm
Brief summary · 1.2k tok0 B
dr.lin@clinic
Scribe transcription0 B
compliance
CPS 230 RAG query0 B
Models
4 loaded
Egress
0 bytes
Uptime
42d 14h
~ terminal
> Where do your client briefs, patient notes, and member data go when your team uses ChatGPT?
> Let's put the model on a box you own.|
AU-channel supply, GST inclusive
OEM ProSupport hardware
Apache 2.0 software stack
4 weeks to production

Need cloud-LLM integration instead? See AI Integration →

Why Choose On-Premise AI Over Cloud AI APIs?

ConsiderationOn-Premise (DGX Spark)Cloud AI APIs
Data egressZero — prompts never leave your networkData traverses third-party services
Cost predictabilityFixed — hardware you own, no per-token feesSubject to vendor pricing and tokenizer changes
Model controlYou choose the model and every upgrade windowVendor deprecations and changes forced on you
Updates & patchingControlled mirror on your schedule (quarterly health review)Automatic, vendor-controlled
Regulatory fitSupports Privacy Act, APRA CPS 230, AHPRA and Law Council guidanceRequires third-party risk assessment

IOTAI sovereign AI deployments typically reach production in 4 weeks, sized to your firm during a sovereignty assessment.

Frequently Asked Questions