Skip to content
Live on self-hosted servers · Open-source edition

One inbox for every channel.
One AI assistant answering first.

Self-hosted omnichannel customer engagement platform with an AI assistant that reads your company knowledge base to reply in seconds, knowing exactly when to hand off to human agents.

Unified channels
12

channels unified into one inbox

AI providers
7

switch anytime via dashboard

AI models
23+

pre-configured or custom models

Seat licenses

zero per-seat licensing fees

Data sovereignty
100%

all data hosted on your own servers

01 Costly problems for businesses

Four leaks in traditional customer support

Not every problem needs AI. But all four issues below stem from the same root: fragmented conversations and humans repeating tasks machines can solve.

01 Fragmented channels

Messages scattered across seven disconnected channels

Facebook page handled by one agent, WhatsApp by another, email in Outlook, website live chat in a third tool. Nobody has a unified customer overview, and dropped messages happen daily.

How we solve it

Every channel routed into a single queue, one unified customer profile, automatic agent assignment.

02 Wasted human effort

70–80% of inquiries are repetitive

Shipping costs, inventory checks, warranty policies, store hours. Agents type identical answers hundreds of times a week, while complex high-value inquiries are left waiting.

How we solve it

AI assistant reads your documentation to answer repetitive queries instantly 24/7, no manual shifts needed.

03 After-hours lost sales

Off-hours create silence, and customers abandon

A customer messages at 10 PM and gets a response the next morning. For purchasing decisions, ten hours of silence usually means a lost sale.

How we solve it

Flexible scheduling: AI active 24/7 or automatically triggers during off-hours to cover staff gaps.

04 Poor chatbot experience

Rigid chatbots give wrong answers and anger customers

Rule-based bots only match exact keywords. Any slight deviation results in off-topic answers or hallucinations, with no smooth path to a human agent.

How we solve it

AI strictly answers from indexed enterprise documents, admits when uncertain, and smoothly transfers to human agents.

02 System architecture

Two layers: workspace for humans, AI handling frontline

The foundation is a full-featured omnichannel support desk: unified inbox, customer profiles, routing, labels, and reports. The top layer is an AI assistant operating frontline, escalating only when necessary.

CUSTOMER CHANNELS (12 CHANNELS) Website live chat WhatsApp Facebook Instagram Email Telegram LINE SMS / Twilio TikTok X / Twitter Mobile App Custom API Unified Inbox 1 queue · 1 history routing · tags · priority automated reply escalate to human AI answers frontline seconds · 24/7 powered by your docs Human agents handle complex cases receives complete context no repeating questions Single customer profile all channels & sessions
Figure 1 — Channel consolidation: The key advantage is having a single queue and single customer profile. Whether a customer messages on Facebook today or emails next week, both AI and human agents view the same history.
Foundation layer

Full-featured omnichannel desk

Unified inbox, customer and company profiles, auto-assignment, labels, internal notes, canned responses, automation rules, proactive campaigns, CSAT surveys, and detailed reporting by agent and channel.

AI layer

Assistant trained on your docs

Ingest website, pricing pages, warranty policies, PDF files. The assistant retrieves relevant passages before answering. Each inbox can bind an independent assistant with tailored tone and scope.

Self-service layer

Public Help Center

Multilingual, SEO-ready knowledge base enabling self-service before reaching chat. The same source serves both self-service visitors and the AI assistant knowledge base.

03 How it works

Four steps from customer message to resolution

Transparent, fully auditable, and human agents can intervene at any second.

Customer sends message from any channel Recorded in inbox triggers internal event Calls assistant engine HMAC signature · anti-replay 7-Rule Guard Filter + business hours check SKIP · ZERO INFERENCE COST assistant echo · internal notes transferred to human · off-hours valid ASSISTANT LOOP · MAX 10 INFERENCE TURNS Knowledge Base your company docs semantic search LLM Engine 7 providers · 23+ models last 20 messages context Tool Calling 5 native + HTTP APIs your internal systems passages query call result Reply to customer via channel auto-truncated by channel limits telemetry: model · tokens · latency Escalate to human agent tags updated · priority elevated AI gracefully steps back
Figure 2 — Message lifecycle: The seven-rule guard filter pre-emptively blocks unnecessary executions, ensuring the assistant never replies to its own echoes. Once escalated, the assistant steps back completely.
04 Enterprise knowledge

Feed your exact policies, products, and workflows

No hallucinations. The AI reads from verified documents you upload, extracts precise answers, and cites exact references.

ONE TIME · INGESTION Website · PDF paste URL or upload file Clean Content strip chrome & ads Chunk Text semantic boundaries Generate Embeddings cached if unchanged Vector Store PostgreSQL + pgvector top matching passages · isolated per assistant EACH INCOMING QUERY Customer inquiry "shipping policy?" Vector similarity search intent-based, beyond keywords Prompt Assembly with your custom system tone Grounded Response cites sources, admits unknowns Queries strictly scope to the bound assistant, ensuring enterprise data is never leaked across tenants.
Figure 3 — Ingestion and retrieval pipeline: The upper branch runs once when documents are added; the lower branch executes per inquiry. Unchanged content is automatically skipped during re-indexing.

This impacts response quality more than model choice

Conflicting or outdated docs produce incorrect replies even with premium models. Our built-in feedback loop lets human agents flag incorrect replies along with the exact source chunks used, making knowledge gaps easy to pinpoint.

05 Tools & integrations

Beyond answers — An AI assistant that executes actions

Look up orders, create tickets, check inventory, and send webhooks to your internal CRM.

Tool Trigger condition Observed outcome
Knowledge Base Search Questions regarding products, pricing, and policies. Document-grounded answer with source citations.
Human Escalation Customer requests human agent, or assistant lacks sufficient data. Conversation routed to human queue, AI steps back.
Tag Assignment Detects intent: complaints, pricing inquiry, warranty, return. Clean reporting without manual tagging overhead.
Priority Elevation Frustrated sentiment or critical order blocker detected. Escalated tickets prioritized at the top of the queue.
Conversation Resolution Customer satisfied, or prolonged inactivity following resolution. Queue stays tidy without unresolved stale conversations.
Custom System Webhooks Live data lookup: order tracking, inventory, booking slots, debt. AI answers with live data: "Your order is dispatched from Warehouse #3" instead of generic delays.
Operations

Operating schedules

Operate 24/7, during business hours, or strictly off-hours to cover night shifts. Many companies begin by delegating unmanned off-hours to AI.

Operations

Auto-close inactive chats

Conversations with no response after a set duration are automatically resolved with a courteous message, protected by safeguards to never close active human threads.

Continuous Learning

Automated FAQ clustering

Analyzes real conversational logs to extract recurring questions and cluster intents. Once human-approved, they enter the knowledge base so the assistant gets smarter every week.

  • Load-balanced auto assignment
  • Event-driven automation rules
  • Bulk operation macros
  • Canned responses
  • CSAT satisfaction surveys
  • Proactive broadcast campaigns
  • Custom customer attributes
  • Reports by agent / channel / tag
  • Bilingual UI (Vietnamese & English)
06 Deployment options

The entire stack runs on your private infrastructure

No proprietary SaaS lock-in. You own the keys, the database, and the entire conversation history.

YOUR PRIVATE SERVER · VPS OR ON-PREMISE Web Server frontend + API Background Workers job queue processing AI Assistant Engine plugin, clean core Vector Store PostgreSQL + pgvector Redis Cache realtime + event queue Custom Domain · HTTPS auto SSL certificates Chat logs · customer records · internal documents Never leave this private perimeter. query + passages LLM reply AI Model Providers OpenAI · Anthropic · Gemini DeepSeek · Mistral · xAI · OpenRouter hot-swappable in dashboard inbound / outbound Customer Channels website · social · email
Figure 4 — Data boundary: The boundary represents your self-hosted infrastructure. Only customer chat messages and isolated LLM inference calls cross this line; your full database is never exposed.
Measured on live production

Lighter than you think

A single-core VPS with 3.8 GB RAM runs the entire stack using roughly 1.35 GB. Base memory footprint stays lean; serving more traffic scales incrementally during document embedding.

Recommendations

Baseline specs

Mid-load single tenant: 2 vCPU / 4 GB RAM is comfortable. Multi-tenant setup for 10 light brands: 4 vCPU / 8 GB RAM. Upgrade based on verified traffic telemetry.

Setup process

Pre-packaged, live in a morning

Docker container deployment, automated custom domain & SSL, automated backup routines. The AI assistant is a pluggable add-on that leaves core files untouched for seamless upgrades.

Two operating modes

Two operational models based on who manages infrastructure

Both the support desk and AI assistant operate on open-source foundations with zero per-seat licensing fees. The only difference is where the server is hosted and who manages the AI model billing.

  Self-hosted by you Managed by Dev00
Best suited for Enterprises with strict data compliance: fintech, healthcare, education, regulated entities. Retail chains, marketing agencies managing multiple brands, software resellers.
Servers & data On your infrastructure. Dev00 provides setup and technical support. Chat logs never leave your servers. On Dev00 managed infrastructure, segregated per client. Zero sysadmin overhead for your team.
AI model billing You connect your direct API keys with model providers and set custom spending limits. Dev00 manages billing and delivers a consolidated invoice matching exact token consumption.
Readiness Production Ready running live in production. Near Ready tenant segregation ready; tenant portal in final review.
07 Cost & ROI

The economics for your business

Adjust the sliders according to your real business metrics to compare AI operational cost against human staffing.

Parameters

Your business metrics

Custom
1 Volume & Automation Rate
convs
turns
Resolved without human agent involvement
55%
0% (Min) 40% – 60% (Common average) 100% (Max)
2 Infrastructure & Token Rates
đ / 1M
đ / 1M
đ / month
3 Human Support Benchmarks
đ / month
convs / month
Projected ROI

Monthly financial impact

Auto calculated
Net monthly savings Net savings

100% deduplicated for private server hosting and actual AI token consumption.

A AI operational cost
AI response turns:
Tokens In · Out:
Model API cost:
Server hosting:
B Staff value unlocked
Fully handled chats:
FTE equivalent:
Average cost per resolved conversation:

* This calculator provides estimates for financial planning, not a binding quotation. Savings represent unlocked bandwidth for higher-value workflows. In practice, most organizations utilize it to expand service capacity and reduce turnaround time rather than trimming headcounts.

08 Current status & roadmap

What is live today, what comes next

Full engineering transparency. The platform is running live with real production LLM inference, verified end-to-end, but several roadmap modules remain underway.

Feature module Status Details
Omnichannel Desk Live 12 channels unified into a single queue and customer profile.
Autonomous AI Assistant Live Verified with production model inference: grounded responses, admits unknowns, executes tool chain (search → tag → prioritize → escalate). Ingests websites & PDFs with strict perimeter security.
Tools & External APIs Live 5 native tools live; custom HTTP webhooks fully executable, configuration dashboard in final polish.
Telemetry & Admin Portal Live Full telemetry per turn (model, tokens, latency, tool calls). Five dedicated management screens with bilingual Vietnamese & English support.
Multi-tenant Portal In Progress Platform-level isolation complete. Tenant role permissions underway. Per-tenant AI key lock and quota limits currently in staging.
Copilot for Agents Planned Decoupled module, planned for future update upon client request.
Rollout roadmap
  1. WEEK 1

    Deploy & connect channels

    Server setup, domain & SSL provisioning, channel integrations, team onboarding to unified inbox. Immediate value: zero dropped chats, accurate logs.

  2. WEEKS 2–3

    Ingest docs & dogfooding

    Review documents, ingest website & policy PDFs, fine-tune assistant tone. Staff performs internal verification before customer exposure.

  3. WEEK 4

    Live rollout, scoped

    Activate on a single channel and schedule (typically off-hours: lowest risk, highest immediate ROI). Monitor via feedback logs.

  4. MONTH 2+

    Data-driven expansion

    Review resolution rates and flagged answers to expand knowledge base. Scale across all communication channels and peak hours.

09 Next steps

Test on your own company documents

Send us sample documentation (PDF, Word, or web link). We will spin up a live demo with your data so you can verify accuracy firsthand.