AI integration

AI features that live inside your software and keep your data where you want it

Assistants, document search and automation built into what you already run. Cloud models where they fit, EU-hosted models where residency matters, or a private LLM on your own servers with Kipper AI.

$ kip ai install
preflight: 16 GiB free, storage ok
pulling qwen2.5:7b …
✓ Ollama running in kipper-ai
✓ chat UI at https://chat.example.com
Models, chat and data, all on your own server.

What we build

Four things clients ask for, and what each one involves

Support and sales assistants

A chatbot that answers from your documentation and product data, hands over to a person when it should, and logs the questions it left open.

Document search

Retrieval over contracts, manuals, tickets or wikis, with the source shown for every answer. Runs against your files where they already are.

Automation inside your software

Classification, extraction, summaries and drafting built into the tools your team already uses, with a human in the loop where it matters.

A private model in your cluster

Kipper AI installs Ollama and a chat UI on your own servers with one command. You pay for the server, keep your own keys, and your data stays inside the cluster.

Model and data residency

Three places a model can run, and when each is right

The choice is made once, in writing, before any code. It follows your data-protection requirements and your budget, in that order.

ModelsWhere it runsWhen it fits
Claude, OpenAI, GeminiProvider cloudsStrongest models, fastest to ship. Data leaves your environment under the provider's terms.
MistralEU-hostedGood models with European hosting and contracts. The default when EU residency is a requirement.
Kipper AI (Ollama)Your own serversOpen models running inside your cluster. Usable from 16 GiB RAM, fast with a GPU. Everything stays on the server.

From the work

PersoHR: AI features with everything inside the EU

Our own HR product ships assistants for onboarding and document drafting on EU-hosted models, with the data staying in German data centers. The same pattern is available for your software.

See PersoHR
Models
EU-hosted (Mistral), no US providers
Hosting
Germany, Hetzner
Stack
Nuxt, Spring Boot, PostgreSQL
Pattern
Assistant inside the product, sources shown, human in the loop

Before you call

Questions we get about AI work

Where does our data go?

Wherever you decide. Provider clouds, an EU-hosted model, or a model running inside your own cluster. We have shipped set-ups where all data stays inside the EU and set-ups where all data stays on one server.

What does Kipper AI need in hardware?

The installer refuses to run with less than 8 GiB of free memory on a node. For anything beyond a demo, plan on 16 GiB RAM and 4 vCPUs; a GPU with 16 GiB or more of VRAM makes chat fast and lets you run larger models.

Can the assistant be wrong?

Yes, so it shows its sources, stays inside the documents you give it, and hands over to a person for anything it cannot back up. We measure answer quality before and after launch rather than promising accuracy.

How is this priced?

A short assessment to pick the use case and the model, then milestones like any other software work. Running costs are either the provider's per-token bill or your own server, and we show both before you choose.

Talk to an engineer

Tell us what your people ask for most often. That is usually the first assistant.

Thirty minutes with a senior engineer. Reply within one business day.