Six self-hosted models today, with more open-source models on the way — three ways to run them, served from data centers on Malaysian soil, under Malaysian jurisdiction. No data leaves the country unless you ask it to.
We run six models today, tuned and monitored on our own infrastructure, so each one performs the way its benchmark promised — not the way a shared, oversubscribed endpoint delivers it. More open-source and self-hosted models are on the way.
All six models are shown above. More open-source and self-hosted models are on the way. Tell us what you need →
Access is metered by need, not by a one-size plan. Most teams start pay-as-you-go, move to enterprise once usage stabilises, and add on-premise hardware once inference becomes part of a physical workflow.
For individuals and small teams testing use cases before committing to volume.
A GB10-based unit installed on your premises runs your agentic applications locally, while model inference is served from our Malaysian data centers over a private link. Your orchestration and sensitive context never leave the building; only model calls do.
Flat-rate access to the full roster for organisations with steady, high-volume workloads.
Primary facility. Model weights, inference, and token metering all run here — inside Malaysian borders, under Malaysian law.
Prompts, completions, and logs stay on local infrastructure unless a customer explicitly configures cross-border routing.
The Hybrid plan puts an on-premise GB10 unit inside your own building, so application logic and context never have to leave it.
SNS is not a reseller of foreign endpoints. Hardware, staff, and support are based in Malaysia.