Your company's privateAI workspace.
Chat, team-scoped memory, and 200+ models on your own servers. Every conversation stays on your infrastructure, every team sees only its own knowledge, and admins see usage, cost, and audit logs across the board.
Everything the browser tab
does, on your servers.
Six things a team reaches for daily. All of them run inside your network, and none of them phone home.
Every model, one workspace
200+ models via OpenRouter, or your own Ollama and vLLM instances. Swap models without changing workflows; context length and pricing shown per model.
Team-scoped memory
Every message is embedded and searchable. New conversations start with relevant team context injected automatically, with attribution: who said it, when, and where.
Admin visibility
Usage and cost per model, team, and user. Conversation browser, audit log, and per-team model restrictions from one dashboard.
Orgs, teams, roles
Multi-tenant by design: organizations, teams, and four role levels. Access enforced at the database layer with row-level security, not just in the app.
Semantic search, Cmd+K
A command palette that searches meaning, not keywords, across your team's conversations. Keyboard-first, everywhere in the app.
Native desktop app
macOS and Windows via Tauri: system tray, native notifications, deep links for invites, auto-updates. Same UI as the web, no feature gap.
One container,
then your identity stack.
Three steps, and the last one hands over the runbook. Nothing about running Lumen depends on us afterwards.
- 01
Single container install
One Docker image, one Helm chart. Deploys to your cluster in under an hour, including SSO and storage configuration.
- 02
Identity-native setup
SAML and OIDC SSO with SCIM provisioning out of the box. Your users sign in with existing accounts; admin roles map to your IdP groups.
- 03
Observability and handoff
Prometheus metrics, OpenTelemetry traces, and structured audit logs plug into your monitoring on day one. We stay on call for the first 90 days.
Replaced unmanaged ChatGPT use with a self-hosted workspace. Shadow AI dropped to zero; adoption did the opposite.

Two more, same
deployment story.
Every product ships as a container you host. Adding a second one reuses the stack the first one already runs on.
Thirty minutes,
then a proposal in two days.
No sales deck, and no discovery phase you pay for. We look at your stack and tell you whether it fits.
- 01
Architecture call, 30 min
With our solution lead, focused on your stack and constraints, not a feature tour.
- 02
Pilot proposal in 48 hours
Scope, timeline, integration matrix, and success metrics, in writing.
- 03
Reference call, optional
Talk to a team currently running Lumen in production before you commit.
“Shadow ChatGPT is not a policy problem, it is a tooling problem. Give people a workspace that is faster than the browser tab and remembers what the team knows, and the data stops leaving on its own.”