AI and LLM apps, operated for you

Models and the interfaces your team talks to, inside your perimeter. Prompts and documents stay on your side of the tunnel.

Start with one workloadAll apps

Choosing among the AI apps

These apps come in layers, and most teams need one from each. The interface is what people use every day: Open WebUI and LibreChat are chat assistants with your own models behind them, and Dify builds assistants and agent workflows on top of your documents. The model server runs the model itself: Ollama for a quick start on modest hardware, vLLM when many people share a GPU and throughput matters. Qdrant stores the embeddings that let an assistant answer from your own files, and Langfuse traces what each prompt did, what it cost and where it went wrong.

If you are starting out, pick one interface and one model server. Add a vector store when the assistant has to answer from your documents, and tracing when more than one team relies on it. Read the licence on each page before you commit: Open WebUI and Dify add conditions to a permissive licence, which matter if you want to change their branding or offer them to other organisations.

Wherever they run, on your own GPUs or in a Pilae region, prompts and documents stay on machines you control. Private AI describes the whole setup as one project.

Need an app that is not listed?

We scope any open-source app: what it needs, how we run it, what it costs.