Service

Sovereign & commercial AI

From idea to production: we deploy generative AI that is useful, measurable and compliant, with market-leading models or open-weight models hosted on your premises or in Europe.

Situations we see most often

  • Promising prototypes that never make it to production.
  • Sensitive data that cannot leave your infrastructure.
  • API costs that are hard to predict as usage grows.
  • Teams already using AI, with no framework or security.

What we do

  1. 01

    Sovereign AI, on-premise or in the EU

    Deploying open-weight models (Mistral, Llama, Qwen…) on your GPUs or with a European provider, served by vLLM or Ollama on Kubernetes.

  2. 02

    RAG & internal assistants

    Assistants connected to your documents and tools, that cite their sources and respect existing access rights.

  3. 03

    Agents & automation

    Agents that chain tasks (research, drafting, data entry, tickets) through your APIs, with human approval where it is needed.

  4. 04

    Commercial AI, properly integrated

    OpenAI, Anthropic, Mistral, Google… The right model for each use case, with costs, quotas and confidentiality under control.

  5. 05

    Evaluation & guardrails

    Business test sets, answer-quality metrics, filters and logging: AI you can stand behind.

  6. 06

    AI infrastructure

    GPU sizing and sharing, scaling, and observability of models and their costs.

Technology

  • vLLM
  • Ollama
  • Mistral
  • Llama
  • Qwen
  • LangGraph
  • LlamaIndex
  • pgvector
  • Qdrant
  • Kubernetes
  • NVIDIA GPU Operator
  • OpenTelemetry

In the field

  • Public administration

    Sovereign AI assistant for staff

    Context
    Staff needed to analyse documents and draft answers from internal knowledge bases, without data leaving the organisation.
    What we did
    Local AI on on-premise hardware, RAG over internal databases and knowledge bases, MCP servers to query in-house software.
    Outcome
    A six-month project overall, with the MVP delivered at two thirds of the estimated budget.

How we engage

  • Audit & assessment

    A few days to a few weeks to assess a platform, a cluster or the feasibility of an AI project.

    You getA clear picture and a prioritised, costed roadmap.

  • Fixed-scope project

    One scope, one deliverable, one commitment: a migration, a platform, an AI use case from prototype to production.

    You getA solution in production, documented and handed over.

  • Team augmentation

    Senior engineers embedded in your teams to speed up a project or bring in expertise you are missing.

    You getImmediate capacity, and teams that level up.

  • Managed operations

    We operate and evolve your Kubernetes and AI platforms, with service levels we define together.

    You getMonitoring, upgrades and support at the agreed service level.

Frequently asked questions

What do you mean by “sovereign AI”?

AI where you control the models, the infrastructure and the data: open-weight models hosted on your premises or with a European provider, with no data sent to a third party.

Is an open-weight model as good as a commercial one?

For many business use cases (document search, summarisation, extraction, classification), the best open models are more than enough. We check that on your own data before recommending anything.

What about the AI Act?

The EU regulation sets obligations based on the risk level of each use case. We build those requirements in from the start: documentation, traceability, human oversight and appropriate hosting choices.

Contact

Have an AI use case in mind?

In a single workshop, we assess feasibility, value and the most suitable hosting option with you.

Prefer a direct line?

contact@okko.be

Our other services

  • Cloud native & Kubernetes

    Design, migrate and operate Kubernetes platforms that are reliable, secure and cost-efficient, on AWS, Azure, Google Cloud, a European cloud or your own servers.

    • Architecture & migration
    • Platform engineering & GitOps
    • Observability, security, FinOps
  • Modernisation & DevOps

    Help your development teams ship faster and with more confidence on the cloud, and bring AI into your products without rewriting what already works.

    • Containerisation & modernisation
    • CI/CD & DevOps practices
    • AI integration in your applications