Private LLM for sensitive documents
Contracts, HR documents or engineering data are processed with a private model instance — without content going to external model providers.
Building a private AI environment is one thing — operating it securely, up to date and economically is another. The Managed AI Runtime is our operations offering: a containerised, private AI environment with your models and your data, operated by us — for a predictable monthly fee plus transparently passed-through compute costs.
The scope of services is divided into four areas. Operations: containerised runtime, private model endpoints on modern inference servers (e.g. vLLM), updates, patches and backups. Security: single sign-on via Entra ID/OIDC, role-based permissions, secrets management and private networks. Compliance: data processing agreement, technical and organisational measures, subprocessor list, logging and deletion concept, plus a documentation package that supports you with obligations under GDPR and the AI Act. Cost control: continuous monitoring of token and GPU consumption with monthly reporting.
On compute we take an unusual approach: we pass GPU and infrastructure costs through to you transparently at cost. We earn on operations — not by reselling compute time at a markup. That makes your costs predictable and our recommendations independent: if a cheaper provider fits your workload better, we switch to it.
Our portability promise also applies to operations: your AI can leave us. Deployments are containerised, configuration and data paths are documented, and exit documentation is part of the service. On request, we test once a year in the AI Exit Drill whether your stack can actually be restored on alternative infrastructure — with us, digital sovereignty is not claimed but tested. Our support commitments are honestly calculated: we agree on response times we can organisationally keep, instead of making enterprise promises that only exist on paper.
Contracts, HR documents or engineering data are processed with a private model instance — without content going to external model providers.
RAG across technical documentation, knowledge base and file shares: employees find answers instead of documents — operated in your controlled environment.
Your applications use private inference endpoints for classification, extraction or text generation — with stable monthly costs instead of surprising API bills.
Takeover or setup of the environment according to target architecture
Hardening and integration: SSO, network, permissions
Operations start with monitoring, backup and runbooks
Regular operations with monthly cost and status report
Support and response times are agreed individually and calculated so that they can be reliably met organisationally. All prices are net; the individual quotation is binding.
Still have open questions? We're happy to clarify them in an initial call.
The structured assessment: where does your AI stand today, what does it really cost — and which operating model fits your company?
From a vague AI plan to a solid target architecture in 5–10 working days — with the AI Infrastructure Passport as a tangible result.
Finally ask your ERP questions: private AI across Business Central, documents and knowledge base — your data stays under your control.
Free initial consultation, 30–45 minutes, remote. An honest assessment — even if the answer is that you don't actually need it.