Deploy & Operate
Run Turing on your own infrastructure under Apache 2.0 — no vendor lock-in for your data, your models or your search engine. Install it, point it at any LLM provider (OpenAI, Anthropic, Gemini, Azure OpenAI or local Ollama), and configure per-instance behavior; API keys are stored encrypted. Scale to multiple isolated tenants behind a single identity realm when you need to, or keep a single-tenant deployment simple.
Operating Turing in production is first-class: Micrometer metrics, conversation and chat analytics, cache-hit gauges, token usage accounting and cost governance give you the visibility to run it confidently. This section covers installation, the full configuration reference, provider and instance setup, assets, multi-tenancy, administration, logging and observability.
Installation Guide
Viglet Turing ES Installation Guide
Configuration Reference
Complete reference for the Turing ES application.yaml configuration file.
Generative AI & LLM Configuration
Overview of GenAI capabilities in Turing ES — LLM providers, RAG architecture, tool calling, MCP servers, and AI agents.
LLM Instances
Configure and manage Language Model instances in Viglet Turing ES — 11 vendor types, per-vendor authentication, and the capability matrix.
Assets
Manage files and train the RAG knowledge base with Viglet Turing ES Assets.
Multi-Tenancy
Run one Turing ES JVM that serves many fully isolated tenants — discriminator-column data isolation, a single Keycloak realm with a tenant claim, tenant-owned infrastructure, plan quotas, and tenant lifecycle.
Administration Guide
Manage users, groups, roles, API tokens, global settings, and system diagnostics in Turing ES.
Observability — Prometheus & Grafana
The metrics, dashboards, and traces that turn Turing ES from a black box into a glass box. Prometheus scrapes Spring Boot Actuator, Grafana renders the story, and you sleep through the night.
Logging
Monitor server activity, indexing pipeline, and AEM connector logs from the Turing ES admin console.
Chat Analytics
Turn thousands of chat conversations into the voice of your customer. Intent, goal achievement, sentiment trajectory, funnels, conversation replay, tool-latency p95, router-decision and SSE diagnostics — engine-agnostic, cluster-safe, AI-on-AI classified.
Token Usage
Monitor and analyse LLM token consumption in Viglet Turing ES.
Cost Governance
See and control your AI spend. Turing freezes a USD cost on every turn from an editable price table, rolls it up into a live dashboard by agent / model / stage, and lets each agent enforce a soft monthly budget that downgrades to a cheaper model instead of overspending.