Lumina
The agent
The agent What it is How it answers Architecture Channels Insights Who it's for Trust How you start Writing Book a call
Generation 2 · Agent as a Service · Built in the European Union

An AI agent your business can stand behind.

Anyone can deploy a chatbot. Almost nobody can defend what it said. Lumina answers from your documents — and when your documents don't cover it, it says so.

Book a call

Ask the agent when we're free — it reads our own calendar, then hands you the booking link.

Live · you can type in this one The Lumina agent, running on Lumina's own documentation
Lumina AI agent
luminawidget.xyz
Where is our conversation data actually stored?

Every document you upload and every conversation is stored in Google Firestore's eur3 multi-region — the European Union. Backend requests are processed in europe-west1, Belgium.

Since 29 July 2026 the model runs in the EU too — Vertex AI's eu endpoint. No inference call leaves the European Union.

You are talking to an AI agent.
In production today. Serving businesses across Europe. Storage, compute and inference — all in the EU
Website widget Your own interface Share link Telegram Discord WhatsApp Slack — on the roadmap
+ Generation 2 · Agent as a Service +

Meet Lumina.

You bring the knowledge. We build, tune and operate the agent, then deploy it everywhere your customers and your team already talk. One brain, every channel, one inbox.

Or keep the interface you already have. Lumina runs headless underneath your own chat UI — which is exactly what the conversation at the top of this page is doing.

Do it yourself

Live in minutes, no technical skill.

The app is built for one thing: upload what you know, choose where it answers, publish. No prompts to engineer, no pipeline to wire. If you can write a document, you can deploy an agent.

Done with you · paid add-on

Or we sit inside your account and tune it.

With your written approval we get access to your workspace and do the work ourselves: load and shape the knowledge, set the persona and guardrails, adjust the settings, test against your real questions. Access is scoped, logged and revocable at any time.

Screenshot app.luminawidget.xyz/channels
Lumina Channels — website widget, share link and messaging channelsLumina Channels — website widget, share link and messaging channels

Channels — app.luminawidget.xyz

Four capabilities, one agent.

1,000,000 tokens

The whole corpus, every question.

No chunking, no embeddings, no retrieval step. Your entire knowledge base enters the context window on every request.

One inbox

A human, whenever you want one.

Every conversation across every channel lands in one place. Take over any thread, reply on the same channel, hand it back to the agent.

Conversations →
One knowledge base

One brain, every channel.

Website, share link, Telegram direct chats and groups, Discord servers. Same knowledge, same persona, same answers. Connect in minutes.

Free/busy only

Live availability, not guesswork.

Connect Google Calendar and the agent answers scheduling questions from your real diary. It never sees event titles, guests or notes.

Calendar availability →

No chunks. No retrieval. No guessing.

Most AI assistants index your documents, split them into fragments, and pull back a handful per question. What the search misses, the model never sees — and a model that has seen only fragments will confidently fill the gaps.

Lumina takes the other path. Your entire knowledge base is injected into a one-million-token context window on every single request. The agent is not recalling your documents. It is reading them.

Demo-grade · conventional retrieval
60–80%

Accuracy we measured from conventional retrieval, over two years of testing that architecture.

Production-grade · full corpus
100%

Of your knowledge base in context, on every single question.

A support bot can afford to be wrong. A bot quoting your loan terms, your dosage guidance or your cancellation policy cannot.

The obvious objection

Doesn't sending everything, every time, cost a fortune?

It would, if the corpus were re-read from scratch on every question. It isn't. Google's implicit caching recognises the knowledge base it saw moments ago and reuses it.

Measured in production
90%

of the corpus served from cache after the first message. Only the opening question of a session pays full price — the rest are billed as though the agent were reading a few pages, while it is still reading all of them.

Stated precisely — what caching means for your data

Caching means short-lived retention of prompt content by the model provider. On Vertex AI that retention stays inside the chosen geography — the EU. It is not "no retention at all", and a client asking the question deserves the precise answer rather than a marketing sentence.

And when it can't, a person is told.

There are two things a knowledge agent must never do: guess, and quietly absorb a request it has no standing to answer. Someone asking for a human, or asking you to erase their data, is not a question — it is an obligation with your name on it.

01
Detected as it arrives

"Can I speak to someone", "delete my data" — instantly in English and Slovenian, within minutes in every other language.

02
Answered honestly

It says the request has reached you. It never claims to have deleted or exported anything, and never asks your customer for identity documents.

03
You are emailed

The conversation gets its own badge and filter. A data request is flagged separately from a support one — only one of them has a statutory clock.

04
You take it from there

Take the thread over and reply live in the same panel — or export it as JSON and erase it, which writes you a receipt: counts and field names, no message content.

Taking a conversation over answers I want a person. It does not answer delete my data — only deleting it does, and nothing is ever erased automatically. A flag is raised; a person decides. None of this is configurable, on any account, including by us.

And when it simply doesn't know — the three hard stops
Corpus over the safe ceiling the upload is refused, before answer quality can quietly degrade
Knowledge base unreachable the agent declines rather than answering ungrounded
Question outside your documents it says so, and points to a human
1,000,000
Context window
750,000
Safe working ceiling
900,000
Hard stop

A system that would rather say nothing than guess.

Your knowledge on top. European infrastructure below.

Layer 1 What you bring
Documents
PDF, Markdown, plain text
Persona
Your voice, tone, language
Live sources
Calendar free/busy, your booking link
Handoff
Where a customer finds a person
Layer 2 Lumina core
Full-corpus context engine
The entire knowledge base, injected per request
Channel adapters
Website, your own UI, Telegram, Discord
Thread store & inbox
Every conversation, server-side, backend-only
Human takeover
Pause the agent on one thread, reply, hand back
Metering & isolation
Per-tenant quotas, per-tenant data boundaries
Layer 3 Runs on
Firestore eur3
Your data at rest — European Union
Functions europe-west1
Every backend request — Belgium
Vertex AI eu
Gemini inference, inside EU member states
Implicit cache
~90% of the corpus reused between turns
AES-256-GCM
Integration credentials, dedicated keys
Where it runs, on a map
Map of Europe with Belgium marked Map of Europe with Belgium marked
Where it runs
europe-west1 St. Ghislain, Belgium

Four moving parts. Since 29 July 2026, every one of them runs inside the European Union.

Read the architecture See the sub-processor list →

Wherever the question gets asked.

Prepare your knowledge once. Every channel draws on the same agent, and every conversation returns to the same inbox.

How it connects
One brain
Back to you
Email alert
Human takeover
Lead captured
Channels
Website widget
Your own UI
Telegram · Discord
WhatsApp
Knowledge in
Documents
Google Calendar
Booking link
Your own systems
Signed-in visitors

Where the same agent answers. One brain, and every conversation returns to the same inbox.

Website widget One script tag on any page Live
Your own interface Headless — keep your UI, use Lumina as the engine. This page runs on it Live · new
Share link A full-page agent, no website required Live
Telegram Direct chats and group communities, by @mention Live
Discord Add to your server, members @mention in any channel Live
WhatsApp Your business number, answered automatically Live
Slack Internal knowledge, in the room where the work happens Roadmap

In the language it was asked in.

Lumina answers in the majority of world languages, including every European one. Set one default language, or let the agent match each customer.

Default

Match the customer

Your agent replies in whatever language each customer writes in, on every channel.

Optional

Always one language

Pick one in Agent Config and every answer comes back in it, whoever asks.

Slovenščina English Deutsch Italiano Hrvatski Français Español Polski Nederlands Magyar Čeština Português Svenska Ελληνικά + most world languages
Screenshot app.luminawidget.xyz/playground — Agent Config
Playground with the Agent Config panel — reply language set to Match the customerPlayground with the Agent Config panel — reply language set to Match the customer

Reply language sits in Agent Config, beside persona, answer length and the fallback line

Shaping how it talks

Ask your conversations a question.

Conversations answers where is that one chat? Insights answers what are my customers asking about? — over 7, 30 or 90 days, in your own vocabulary, from conversations Lumina already read as they arrived.

Screenshot app.luminawidget.xyz/insights
Insights — ask about your conversations, with the sources it based the answer onInsights — ask about your conversations, with the sources it based the answer on

It always says what it looked at.

Every answer carries its basis — based on 312 conversations, 1–27 Jul — and the conversations it read, each one clickable. You cannot tell "nobody complained" from "it only read nine of them".

Every number is a door.

"23 people asked about price" — click it, and there are the 23 conversations, same dates, filter shown as a chip. The address bar updates, so a view you check weekly can be bookmarked or sent on.

Two flags you can count.

Asked for a person and asked for their data are separate badges with separate filters. One is a staffing signal. The other is a legal one with a deadline — and when your DPO asks how many you had last quarter, you can answer.

Worth adding to your knowledge base.

The topics where customers most often ended up asking for a person — with a button straight into Knowledge. Analytics, then a better agent, then fewer people needing you.

Reading conversations
Runs in the background. Costs you nothing.
Charts, filters, search
Free. No AI runs when you open the page.
Asking a question
One message from your plan — and given back if it can't answer.

It reads the short summaries, never your customers' full chats.

How Insights works

A hairdresser and a bank have the same problem.

Both answer the same questions all day. Both lose customers to the ones they answer too late. The stakes differ; the mechanics don't.

Service businesses

Salons, clinics, restaurants, hotels, studios.

Your hours, prices, policies and real availability — answered at eleven at night, without you.

Product & community teams

Software, crypto, digital products.

The same forty questions in Telegram and Discord, every week. The agent takes them. You get your evenings back.

Regulated & enterprise

Banks, insurers, legal, pharmacy, real estate, education.

Where a wrong answer is not an inconvenience. It's a liability with your name on it.

European by architecture, not by badge.

EU data residencyEU inference — Vertex AI euPer-tenant isolationAES-256-GCM credential encryptionGDPR Art 20 export, self-serve, JSONGDPR Art 17 deletion, self-serveEU AI Act Art 50 disclosure, on by default
Published sub-processor listData Processing AgreementAudit loggingNo training on your contentBackend-only conversation storageCalendar free/busy only

Encryption, tenant isolation, secrets handling and abuse containment are documented in full on our security page — including what we haven't certified.

Your documents and every conversation are stored in the European Union, in Google Firestore's eur3 multi-region. Every backend request is processed in the European Union, in europe-west1, Belgium.

Since 29 July 2026, so does every AI inference call — on Vertex AI's eu multi-region endpoint. The reply your customer reads, the summary on your Insights screen, the answers to your own questions about your own conversations, even counting the tokens in a document you upload. The model did not change; its geography did. No inference call remains outside the European Union.

What we do not claim

That Lumina is "fully GDPR compliant" or "EU AI Act certified". Neither phrase means anything on its own, and compliance is a property of your deployment as much as of our code.

That our legal texts have been through counsel. The Data Processing Agreement is a draft, tailored per client, awaiting review — and it says so on its face.

That EU residency settles every transfer question. It settles the largest one. Payments, email delivery and whichever messaging platforms you choose to connect are separate, and are still being worked through.

That everything about one person can be erased on the website channel. A widget conversation lives in one browser tab and nothing follows a visitor between visits — so there is no profile to assemble and no "this person" to look up. Any conversation you can identify, we delete.

We are not lawyers, and we don't publish legal conclusions. We publish the architecture, in detail, and let your Data Protection Officer draw their own.

Read the compliance pack

The software is the easy part.

Uploading documents takes minutes. Writing documents an agent can answer from — clean, structured, complete, in your voice — is the actual work. Most projects fail there, quietly, and blame the model.

So we do it with you.

01

Lumina orientation

Thirty minutes. What you know, where your customers ask, what a good answer looks like — and whether you want to build it yourself or have us do it. Book it →

02

Knowledge onboarding

We shape your material into the corpus your agent will read, tune its persona and language, and test it against the questions you actually get. Knowledge →

03

Go live, everywhere

Website, Telegram, Discord, calendar. One agent, connected to every channel you use, with your inbox behind it.

Invoiced monthly against actual usage, under an agreement written with you. Self-serve subscriptions open later — deliberately, in that order. What counts as a message →

Book a call

Working notes.

We write about what we ship, including the parts that didn't work. Two pieces cover Generation 2 — the enterprise-grade, compliance-first, EU-resident product this page describes. Start with the essay; the paper is there when you want the evidence underneath it.

Research Gen 2
Context as Architecture
Full-Corpus Grounding, Orchestrated Agentic Development, and Compliance-by-Design in a Production AI Assistant Platform
Dr. Tali Režun
Read → talirezun.com
Light read Gen 1 → Gen 2 Start here
Lumina: An AI Agent Your Business Can Stand Behind
Behind the stage of a Gen-1 to Gen-2 journey
Read → talirezun.substack.com
talirezun.substack.comGen 1
Six months after I shipped Lumina
Read →
medium.com/@talirezunGen 1
From Prototype to Production: Building an AI Widget Platform in 30 Days
Read →
All writing

Ask the agent. Or ask the founder.

@lumina_community is staffed by the Lumina agent — running on Lumina's own documentation — and by the team that built it. All answer.

Join on Telegram

An AI agent your business can stand behind.

Book a call
Same conversation, still live
No cookies

This site sets none, and calls no third party. The agent keeps one key for your conversation, which is gone when you close the tab.

Privacy policy →