Approval takes months
Processing agreements, impact assessments, works council, professional confidentiality. By the time that is settled, the project has lost its momentum.
AI in your own building. Not a byte leaves your network.
A preconfigured device runs the models locally and serves the applications inside your own network. Access happens from the building or through an encrypted tunnel, without a single port opened on the router. Setup, maintenance and updates are handled by us.
Patient records, client files, design data, personnel matters. That is where the biggest leverage sits, and that is where the hurdle is highest.
Processing agreements, impact assessments, works council, professional confidentiality. By the time that is settled, the project has lost its momentum.
When the official route takes too long, people use private accounts. Then exactly the data that was meant to be protected flows out, only without any control at all.
Models, containers, networking, hardening, updates. Without a team of your own that looks like a project you would rather not start. So it never starts.
The local AI server is not a set of build instructions, it is a finished appliance. It stands in your building, it belongs to your network, and it is looked after by us. As demand grows, the setup grows with it.
On the local AI server a model engine runs the language models locally and keeps the frequently used ones permanently in memory, so the answer comes immediately. In front of it sits an access gateway through which every application is reachable under a single address instead of a list of ports.
We rely on open-source models with licences that expressly permit commercial use. That is not a compromise: for the tasks that come up day to day (summarising, structuring, rephrasing, classifying, dictating) they are entirely sufficient, and on simple tasks small models are noticeably faster than large ones. Where a task genuinely needs a large model, we say so in advance instead of selling local as equivalent everywhere.
Every application runs in its own container, alongside a local database for users, roles and content. New applications get added without the existing ones being touched. Speech recognition for dictation runs locally too, so even the spoken word does not leave the building.
The setup scales in both directions. If one device is no longer enough, a second one joins it and shares the load. For sites with high availability requirements a backup server stands ready and steps in without anyone intervening. And several sites can be linked into a network in which each keeps its own data.
Support runs through an encrypted point-to-point tunnel. It needs no port forwarding on the router and grants no access to other devices in the network. Added to that is the hardening: firewall, password login switched off, network isolation, separate roles for operation and application. Setup runs through a repeatable script that only repairs what is missing, and can therefore run as often as you like.
Findings, notes and minutes are spoken and turned into text locally. Neither the recording nor the result leaves the building.
Documents are read, structured and made searchable, on your own device. Including files protected by professional confidentiality.
The specialist question goes to your own knowledge base. Not only the answer stays in the building, the question does too, and the question often reveals more than the answer.
The number of workplaces, the planned applications and the model sizes determine the hardware. Too small slows you down, too large costs without benefit.
The device arrives prepared, gets connected to the network and receives a fixed address. No port forwarding, no change to the outward-facing firewall.
A repeatable script installs every component: model engine, gateway, database, containers, speech recognition and the hardening.
Updates, model care and monitoring run through the encrypted tunnel. You can see the state at any time in the interface.
Let us work out in a conversation which of your use cases hold up locally and which do not.
Arrange a conversationHardware, operating system and every component tuned to each other. You plug in, we have prepared. If one device is no longer enough, a second joins it and shares the load.
Open-source language models run on the device graphics unit. Frequently used ones stay in memory so the first answer does not wait for loading. Small models for quick tasks, larger ones for the demanding ones.
Every application under one address instead of many ports. At the same time the place for logging and later access control.
Every application isolated. New ones get added without touching the existing ones and can be rolled back individually.
Users, roles, configuration and application data in one place in the building, with a secured backup copy.
Dictation is turned into text on the device. Neither the recording nor the transcript leaves the network.
A point-to-point tunnel for maintenance and mobile work. Without port forwarding on the router, without access to other devices in the network.
Firewall, password login disabled, telemetry switched off, separate roles for operation and application. The server is not reachable from outside.
One script installs everything and repairs only what is missing. That makes the state reproducible at any time, even years later.
Load, memory, running applications and loaded models at a glance, reachable inside your own network.
Corporate Memory, Content Creator and the other modules run on the same device, with the same roles as in the cloud.
One storage location in the building through which documents reach local processing, without a detour via an external service.
Patient data is subject to confidentiality. On the local device findings can be dictated and records analysed without a processing agreement becoming necessary.
Client files must not leave the building. Research and summarising run locally, and the question itself stays confidential.
Design data and formulations are the capital of the company. Analysis inside your own network rules out leakage technically rather than contractually.
On the left the administration, which is responsible for the state. On the right the user, who simply wants to work and should not notice the server in the basement at all. The tabs are switchable.
Load, loaded models, running applications and the backup. Reachable only from your own network or through the encrypted tunnel.
Every application under one address, no port knowledge needed, no cloud sign-in. Whoever is in the building is in.
Abstracted representation with invented values. Which applications, models and access paths run in your installation is something we define together.
On the home page this module sits under “Operation and sovereignty”, and that is exactly its role. It is not another tool, it is the place where the other tools can run when the cloud is not an option.
Not every task belongs on a local device. Where larger models are needed and the data permits it, hybrid operation is the more honest route. In the conversation we say which of your use cases hold up locally and which do not.
For this module data sovereignty is not one operating variant among several. It is the entire purpose.
One device in the building, every application on it. The normal case for practices, law firms and single premises.
One device per site, central support through the encrypted tunnel. Data stays where it arises.
Sensitive processing locally, compute-heavy tasks in European data centres. You draw the line, and it stays visible in operation.
“The decisive question in the data protection review was not which model we use, but where the data goes. The answer was nowhere, and with that the topic was settled in twenty minutes instead of four months.”
Medical Director, Specialist practice with four sites
Tiered by workplaces, model sizes and the number of applications. Setup covers sizing, installation, hardening and instruction.
| Package | For whom | Scope and hardware * | Setup * | Operation * |
|---|---|---|---|---|
| Practice | Practice, law firm, small business | 1 device · up to 15 workplaces · 1 to 3 applications · plus hardware €5,000 to €8,000 | from €4,900 | €190 / mo. |
| Business | Mid-sized site | up to 50 workplaces · larger models · up to 3 applications · plus hardware €15,000 to €20,000 | from €9,900 | €390 / mo. |
| Pro | Several sites, high availability | second device for failover · custom applications · plus hardware €30,000 to €50,000 | from €18,500 | €690 / mo. |
Anyone who does not want to buy the hardware can rent it. Device, setup, maintenance and model care then run in one monthly rate, the device stays in your building and so does the data. Sensible when an investment would break the budget or when the technology should be renewed on a predictable cycle. We will put together an individual offer for this at any time.
Hardware is passed on transparently at the purchase price plus a procurement fee, or supplied by you. The ranges above are experience values and depend on memory and compute. Operation includes maintenance, updates, model care and remote monitoring. We earn on setup and support, not on a margin on the device.
* Guide values. Scope, hardware and prices are adapted to the individual case in the quotation, because the number of workplaces, the model sizes and the availability requirement drive the effort.
All prices are net, plus statutory VAT.
For summarising, structuring, dictation and analysing your own holdings, locally run models are clearly sufficient today. With very large models and very long contexts the cloud stays ahead. We say in advance which of your use cases fall into which category, instead of selling local as equivalent everywhere.
Then the local applications stand still, as with any server in the building. That is why a backup and an agreed restore path are part of operation. Where an outage is not acceptable, the Pro package provides a second device.
We do, through the encrypted tunnel. Updates, model care and monitoring are part of operation. Nobody is needed on site except in the case of a hardware defect.
No. The maintenance tunnel is established by the device outwards, not from outside inwards. No inbound port forwarding is needed, and the tunnel grants no access to other devices in your network.
Yes. The modules are the same regardless of where they run. A move is a migration task for data and configuration, not a new project.
Tell us what you want to do with AI and which data is involved. We will tell you which cases hold up locally, which hardware that needs and where the cloud stays the better route.