Evolution Online · Belgrade · since 2003

We build AI tools that work with you.

For more than 20 years, EON has been building software that businesses rely on every day. More recently, our AI team has been turning that experience into real, production-ready AI solutions — from voice-driven 3D avatars and privacy-focused LLM platforms to autonomous AI agents.

Why EON for AI

AI that ships, not slideware

Anyone can call an API. Turning large language models, speech recognition and computer vision into dependable business software is engineering — and engineering is what EON has done since 2003. We combine deep experience in web, cloud and enterprise systems with hands-on mastery of the modern AI stack.

Production-first

Every solution below is a working system with real users in mind — authentication, admin surfaces, cost tracking, graceful degradation when AI services are unavailable.

Full AI stack

LLMs and agent frameworks (OpenAI, LangChain, LangGraph), speech (Whisper, Kokoro, TTS), vision models, vector databases (ChromaDB), hybrid retrieval and local models via Ollama.

Privacy & control

We build AI that respects your data — on-premise anonymization, encrypted vaults, department-aware access control and local model execution when data must not leave your walls.

End-to-end delivery

From discovery and business analysis to backend, frontend, deployment and support — one team, one accountable partner.

What we build

AI capabilities, proven in practice

RAG & knowledge chatbots

Chatbots grounded in your documents with citations, hybrid search and multi-stage retrieval pipelines.

Voice & speech

Real-time transcription, semantic audio search, text-to-speech and lip-synced talking avatars.

Autonomous agents

Multi-agent systems that plan, create content and react to users — with deterministic fallbacks.

Private & secure AI

PII anonymization, local LLMs, encrypted mapping vaults — frontier-model power without data exposure.

Vision & multimodal

Camera-aware assistants, image understanding and document OCR feeding into LLM pipelines.

Developer & workflow tools

AI-enriched documentation generators, in-app guided tutorials and evaluation harnesses.

Portfolio

Nine AI solutions

Every project below was designed, built and delivered by the EON AI team. Watch the demo videos to see them in action.

▶Demo video coming soon
01

Client Chatbot

Secure document-grounded assistant with cost auditing

A document-grounded chat assistant reached through encrypted, time-limited URL tokens. Users select authoritative source files and get Markdown-rendered answers constrained to that material — while IP-restricted admin surfaces manage the files and every exchange is logged with token counts and computed cost.

AI inside: a file-context retrieval pattern where the selected documents are injected directly into the prompt, so the model answers strictly from the approved material and cites which file each answer came from — with per-message cost tracking for billing visibility.

FastAPIOpenAIAES encryptionSQLiteJinja2
▶Demo video coming soon
02

Interactive Tutorial Agent

Ask a question, get a guided tour of your app

A browser-based walkthrough tool: the user asks "how do I do X in this app?" and receives an interactive, step-by-step tour rendered directly on top of the live web page — a Chrome side-panel extension navigates the tab and highlights the exact UI elements for each step.

AI inside: an agent pipeline turns free-form questions into concrete, page-aware tutorial steps over the registered application pages, with an architecture staged for LLM-driven step generation, text-to-speech narration and a talking-avatar presenter.

Django RESTChrome ExtensionReact + ViteDriver.jsManifest V3
▶Demo video coming soon
03

AI Documentation Generator

From source code to polished docs in one command

A command-line tool that turns a JavaScript source file into a single, self-contained HTML documentation page — with live search, filters, deep links and source snippets — plus watch and serve modes with automatic browser reload.

AI inside: an optional LLM enrichment layer writes summaries, descriptions, examples and per-parameter docs for every extracted API item; hand-written JSDoc always takes priority, and deterministic heuristics guarantee documentation is produced even with no AI available.

Node.jsAcorn ASTOpenAI-compatible APIZero dependencies
▶Demo video coming soon
04

LLM Pricing Calculator

Model and compare the true monthly cost of running AI

An AI cost calculator that estimates the monthly cost of running an AI application from token volumes, model choice, GPU/compute usage, subscriptions, infrastructure and labor — tunable live with sliders across six delivery methods (from token-based API calls to self-hosted compute), with a side-by-side compare view to model two scenarios bucket by bucket and export results as PDF, DOCX or XLSX.

AI inside: the cost engine itself is a deterministic, auditable formula rather than an AI call — AI is the subject matter being priced. A curated catalog ships real-world AI pricing data (LLM token rates from providers like Anthropic and OpenAI, GPU/compute rates, subscription tiers, labor rates) that users select to populate the sliders.

Django RESTSQLiteNext.jsshadcn/uiRechartsTailwind CSS v4
▶Demo video coming soon
05

Safe LLM

Frontier-model power without exposing sensitive data

Use powerful external LLMs without your sensitive data ever leaving your control. Before any text or document goes out, a local anonymization pipeline detects PII — names, institutions, IBANs, IDs and more — and replaces each with a consistent, realistic substitute. Responses are seamlessly rehydrated with the real values on the way back.

AI inside: a trusted local LLM (via Ollama) performs anonymization, complemented by Microsoft Presidio and NER/regex multi-pass PII detection; an encrypted per-project mapping vault guarantees consistent substitution across documents, with leak validation on every response.

Django RESTOllamaPresidiospaCyNext.jsOCR
▶Demo video coming soon
06

3D Avatar

Real-time voice & vision conversational avatar

A voice-controlled 3D avatar you talk to through your browser: it listens, optionally sees through your camera, thinks and answers aloud — with a lifelike character whose lips move in perfect sync with the generated speech. A full-duplex "talking head" assistant, running in real time.

AI inside: Whisper speech-to-text, SmolVLM2 vision-language model for camera understanding, an LLM for responses and Kokoro TTS with word-level timing that drives pixel-perfect lip sync — all GPU-accelerated with PyTorch and Flash Attention 2.

Next.js 15React 19FastAPIPyTorchWebSocketWhisperKokoro TTS
▶Demo video coming soon
07

Advanced RAG Chatbot

A complete platform for building secure, citation-backed RAG chatbots.

Organizations can create internal AI assistants that understand and converse with their own documents. The platform handles document ingestion, knowledge organization, retrieval, chatbot creation, and access control in one place.

Every answer includes inline citations that link directly back to the original source material.

AI inside: a seven-stage LangGraph retrieval pipeline — query rewriting, multi-query expansion, HyDE, hybrid vector + BM25 search with Reciprocal Rank Fusion, CRAG relevance checking, LLM reranking and cited answer generation with GPT-4o.

DjangoLangGraphChromaDBGPT-4oBM25Next.js 14
▶Demo video coming soon
08

RAG Chatbot Generator

Build, test and deploy custom AI chatbots on your data

A full-stack platform covering the entire chatbot lifecycle: ingest source material in almost any format (audio, images, PDFs, web pages, database schemas), organize it into collections, build vector databases, configure the agent — and generate a deployable package with an HTML embed tag for any website.

AI inside: Whisper large-v3 transcribes audio, GPT-4 Vision describes images, LLMs do prompt-guided extraction; generated bots run as LangChain/LangGraph agents over ChromaDB, and a separate LLM-as-judge module scores answers against references before you ship.

Django 5.2LangGraphCopilotKitChromaDBNext.js 15PostgreSQL
▶Demo video coming soon
09

EON FM

A fully autonomous AI radio station

A radio station with no human on the mic. An AI agent plans the weekly and daily schedule, writes and reads the news, picks and downloads music, runs listener quizzes and reacts to listener messages live — streamed as a wall-clock-synchronized broadcast, so every listener hears the same thing at the same moment.

AI inside: a LangChain/LangGraph deep agent orchestrating specialized subagents (weekly planner, daily programmer, listener-message handler), OpenAI LLMs for news and DJ patter, OpenAI or local Qwen3-TTS for the station voice — with deterministic fallbacks so the broadcast never stops.

Django 5LangGraphOpenAIQwen3-TTSNext.js 14yt-dlp
▶Demo video coming soon
10

Searchable Audio

Turn spoken audio into a searchable knowledge base

Record in the browser or upload audio files, and search across everything by meaning, not exact wording. Matching recordings surface with an AI-generated summary, and clicking a result jumps playback straight to the timestamp of the matched passage.

AI inside: Whisper produces timestamped transcriptions, OpenAI embeddings stored in ChromaDB power semantic search, voice queries are transcribed on the fly, and an LLM summarizes the matched sections in natural language.

Django RESTWhisperChromaDBOpenAINext.js

Let's talk

Bring AI into the way your business already works.

Whether you need a chatbot that understands your documents, a voice interface, an AI agent, or a secure way to use the latest models, we help turn it into a real working solution.

office@eonsystem.com

or call +381 (0)11 24 74 364