Free & local by default Β· premium model upgrades when you need them

Launch a Private AI Chatbot on Your Data in Minutes, Not Months

Upload your documents, configure your bot once, and get a production-ready assistant β€” embedded on any website and locked inside your own infrastructure. Runs entirely free on local AI models out of the box, and upgrades to premium cloud LLMs & embedding models with a single key whenever accuracy and speed demand it.

πŸ€–
Leo Β· Acme Assistant
● answers only from your documents
0%Data stays on-premise
0Document formats ingested
0Delivery modes (widget Β· iframe Β· page)
$0Default running cost β€” premium models optional
What We Do

A Complete Chatbot Platform,
Not Just Another Wrapper

Everything needed to run client-ready AI assistants as a service β€” ingestion, retrieval, guardrails, theming, and one-line deployment.

πŸ§™
01

One-Time Setup Wizard

Configure persona, tone, welcome message, theme colors and do-not-answer rules once. Template variables like {{company_name}} flow through every prompt automatically.

πŸ“š
02

Universal Knowledge Ingestion

Drop in PDF, DOCX, XLSX, CSV, HTML or web URLs. Scanned files are OCR'd, tables become self-describing rows, and every document can carry custom context the bot respects.

🎯
03

Precision RAG Engine

Deterministic row-level reranking puts exact matching table lines in front of the model β€” precise price lookups instead of plausible guesses, even on dense spreadsheets.

πŸ”€
04

Free Local, or Premium Cloud β€” Your Choice

Every bot starts free on local Ollama models. Premium tier: paste a key for OpenAI, Groq, OpenRouter, Gemini or premium embedding models β€” instant accuracy and speed gains, encrypted keys, zero code changes.

🎨
05

White-Label Delivery

Floating widget, inline iframe, or standalone page β€” themed with client brand colors, icons and bubble text. One script tag deploys it anywhere.

🧲
06

Lead Capture & Memory

The bot collects visitor details (name, city, email…) naturally in conversation or via a pre-chat form, remembers them across visits, personalizes every answer β€” and hands you a leads list with CSV export.

πŸ›‘οΈ
07

Guardrails & Source Control

Topic blacklists, fallback messages, and per-bot source visibility (show / filenames-only / hidden) keep every deployment compliant and on-brand.

Our Strategic Approach

From Documents to Deployed Bot
in Four Steps

1

Connect Your Data

Drag-and-drop documents or point at web pages. The pipeline parses, OCRs where needed, chunks intelligently, and embeds everything into a local vector store β€” nothing uploaded anywhere.

2

Shape the Persona

The setup wizard captures the bot's identity, tone, boundaries and template variables. The LLM even drafts document descriptions and interpretation tips for you to review.

3

Test Against Reality

Chat with the bot instantly on localhost. Ask pricing questions, edge cases and restricted topics β€” refine until answers are exactly right, then preview it on a mock client site.

4

Migrate Anywhere

A Docker bundle carries app, vectors and config to any client server or cloud VM. Swap the host in one script tag and go live β€” same behavior, their infrastructure.

Flexible Intelligence

Free Where It Counts.
Premium When It Pays.

Every deployment starts on free local models β€” private, offline-capable, zero running cost. The moment a workload demands more accuracy or speed, upgrade to premium models without changing a single line of your setup.

Included by default

Free Β· Local

$0/forever
  • βœ… Runs 100% on your machine β€” air-gap friendly
  • βœ… Open-weight LLMs via Ollama (Qwen, Llama…)
  • βœ… Local embedding models β€” unlimited ingestion
  • βœ… No data ever leaves the server
  • βœ… 2 chatbot slots Β· 2 files per knowledge base
  • βœ… Full RAG engine, guardrails, OCR, delivery modes
Start Free
One key away

Premium Models β€” Pro

$7/mo  or  $60/yr
  • ⚑ 25 chatbot slots β€” serve up to 25 clients per account
  • ⚑ 25 files per knowledge base
  • ⚑ Unlimited template variables
  • ⚑ Widget Appearance suite: icons, bubble text, theming, source control
  • ⚑ Frontier cloud LLMs + premium embedding models (BYO key)
  • ⚑ 10–50Γ— faster responses under heavy traffic
  • ⚑ Keys encrypted at rest Β· switch back to local anytime
Start with Pro

Same platform, same features core β€” only capacity and model power change. Mix and match per client.

Technology Stack

Built Entirely on Open, Free Components

No proprietary runtime, no license servers, no vendor lock-in. Every layer is inspectable and replaceable.

AI & Models

OllamaQwen 2.5Llama 3.xNomic EmbedOpenAI-compat APIsGemini

Backend

PythonFastAPISQLiteJWT AuthSSE Streaming

Retrieval

ChromaDBVector SearchRow RerankingOCR (Tesseract)

Frontend

React 18ViteTailwind CSSVanilla JS Widget

Parsers

PyMuPDFpython-docxopenpyxlTrafilaturaBeautifulSoup

Deployment

Docker ComposeBare-metalAny VPSAir-gapped OK
Industries We Support

One Engine, Every Vertical

Quick Answers

Frequently Asked Questions

Where does my data actually live?
Everything β€” documents, vector embeddings, conversations, configuration and encrypted API keys β€” stays in a folder on the machine running the platform. Nothing is sent to third parties unless you deliberately connect a cloud model provider. Fully air-gapped deployments are supported.
Do I need to pay for AI models?
No β€” that's the point. The default configuration runs open-weight LLMs and embedding models locally through Ollama at zero cost, forever. When a client's workload needs sharper answers or faster responses, you can upgrade that specific deployment to premium cloud models by pasting an API key. It's a per-client business decision, not a platform requirement.
What does the Pro plan include over Free?
Capacity and control: 25 chatbot slots (vs 2), 25 files per knowledge base (vs 2), unlimited template variables (vs 3), plus the full Widget Appearance suite (icons, bubble text, theme colors, source visibility controls) and the AI Models section for premium cloud providers. Pro is $7/month or $60/year.
If I switch a client to premium models, does their data leave our server?
Only the question text and retrieved context snippets are sent to the model provider you choose β€” the same as any cloud AI usage. Your documents, embeddings, conversation history and configuration always remain on your machine. You can also switch back to fully-local models at any time with one click.
How do clients embed it on their website?
One script tag renders a branded floating chat bubble; an iframe embeds it inline; a standalone URL works for QR codes and links. Theme color, icon and bubble text are all configurable per bot from the dashboard.
Can each client have a differently-branded bot?
Yes. Bot name, persona, welcome message, tone, theme color, icon, bubble text and rules are stored per workspace and resolved live β€” the same engine serves every client with its own identity.

Ready to Put Your Documents
to Work as an Assistant?

Spin up the studio, drop in your first document, and talk to your data before your coffee cools.

Create Free Account Preview the Widget