Key facts
| API surface | OpenAI-compatible /v1/chat/completions; change base_url and model name |
| Deployment | Plugsky cloud, your VPC, on-prem or air-gapped |
| Data residency | Region selection and sovereign deployment options |
| Models | 30+ models behind one API, including long-context and multilingual options |
| Access control | Scoped API keys and request logs; enterprise RBAC and SSO |
| Pricing | Flat monthly self-serve plans; see the live pricing page |
| Free tier | Free plan with 2 free AI models (plugsky-micro, plugsky-lite), no card |
| Contracting | DPA and SLA documents available for review |
TL;DR
- Keep citizen records, case files and operational data inside a boundary you control: region choice, VPC, on-prem or air-gapped.
- Evaluate residency, records retention, auditability, model choice and exit before you commit.
- 30+ models behind one OpenAI-compatible API, so service platforms do not change with deployment.
- Start on the free plan or the 14-day full-access trial, then move the same workload to a private deployment.
- Document evidence for procurement: deployment options, logging, DPA and SLA, and a tested exit path.
How it works, step by step
- Map where citizen records, case files, correspondence and operational data live today and which services are in scope.
- Write down residency, records retention and access requirements, and who signs off on each.
- Shortlist deployment modes: managed cloud, VPC, on-prem or air-gapped, per service and data sensitivity.
- Run a pilot on the free plan or the 14-day full-access trial with a fixed set of real service questions.
- Score answers, citations, refusals, latency and cost on the same test.
- Complete security, DPA, SLA and procurement review, with exportable logs and a documented exit plan.
- Roll out service by service and re-run the evaluation after each model or deployment change.
Original data
Try it yourself
Open the sovereign AI readiness score →
What sovereign AI means for the public sector
Sovereign AI in the public sector means the models, prompts and retrieved data stay under your jurisdiction and operational control. In practice that means a deployment you can place in your own region or network, with identity, logging and retention governed by the controls your agency already runs.
The business case is not secrecy for its own sake: it is keeping citizen records, case files and operational data inside a boundary you can audit, while still using modern models through a standard API.
The five things to evaluate
Score every candidate on these five areas, and require evidence rather than claims:
- Residency and deployment: region choice, and whether VPC, on-prem and air-gapped options actually exist.
- Data handling: whether prompts and outputs are used for training, how long they are retained, and which subprocessors are involved.
- Access and audit: scoped API keys, RBAC and SSO for enterprise, and request-level logs you can export.
- Model choice and portability: how many models you can route between, and how hard it is to move away.
- Commercials and exit: predictable pricing, SLA terms, and a documented data-return path.
Deployment modes and their trade-offs
Cloud is fastest: a managed, region-based deployment with residency selection. VPC puts inference inside your own cloud account and network boundary. On-prem runs in your data center; air-gapped runs with no outbound connectivity at all. The trade-off is operational: the further you move from managed cloud, the more you own upgrades, capacity and monitoring.
Public bodies often start on managed cloud for evaluation, then move the same OpenAI-compatible workload to a private deployment when citizen or case data must stay inside a sovereign environment. Because the API surface does not change, application code usually stays put.
A scorecard approach to procurement
Run a structured evaluation: define ten questions that represent real public services, test two or three models, and score answers, citations and refusals. Add a security review covering identity, key rotation, logging and data flow, then a commercial review covering pricing predictability, SLA and exit.
Begin on the free plan or the 14-day full-access trial, document the decision, and keep the evaluation set so you can re-run it after every model or deployment change.
Honest comparison
| Evaluation area | Plugsky | Vendor cloud-only | Building in-house |
|---|---|---|---|
| Residency | Region selection plus sovereign deployment options | Usually vendor regions only | Fully under your control |
| Deployment modes | Cloud, VPC, on-prem, air-gapped | Managed cloud only | Your infrastructure |
| Model choice | 30+ models behind one API | Single vendor catalogue | You host each model |
| Auditability | Scoped keys and request logs; enterprise RBAC and SSO | Vendor-defined logging | You build logging |
| Pricing | Flat monthly self-serve plans; see live pricing | Often per-token or per-seat | GPU plus operations cost |
| Time to production | Fast on managed cloud; private options for sensitive workloads | Fast but fixed residency | Months of build and ops |
Frequently asked questions
What makes an AI deployment sovereign?
Sovereign deployment keeps models, prompts and data under your jurisdiction and control: you choose the region, own the boundary, and keep identity, logging and retention under your existing governance.
Can we keep data inside our own country or network?
Yes. Plugsky supports region selection plus VPC, on-prem and air-gapped deployments, so inference and data can sit where public-sector rules or contractual terms require.
Does Plugsky train on our data?
For strict requirements, choose a private or air-gapped deployment so content stays inside your environment. Review the current data-handling terms and DPA for cloud plans before rollout.
Can we start without a private deployment?
Yes. Evaluate on the free plan, which includes two free AI models with no card, or the 14-day full-access trial, then move the same workload to a private deployment for production.
How do we handle records retention and audit?
Keep retention under your own policy, use scoped keys and per-request logs, and export access evidence when an audit or information request arrives.
Do you provide procurement documentation?
Yes - deployment modes, data-handling terms, DPA and SLA documents are available for review. Mapping them to your own evaluation framework and scoring remains with your team.
Can it support multilingual citizen services?
Yes. Route between models on one API, including multilingual options, and keep language-specific evaluation questions in your test set before you commit.