Sovereign LLM Workbench
Capabilities Assurance Try the workbench
v2.3.0 Local-first AI Enterprise governance Offline sovereign

Sovereign LLM Workbench

The enterprise-grade, local-first AI platform for confidential documents — grounded Q&A, executive reports, diagrams, and unified assurance without sending data to public cloud AI.

0
Data egress default
98%
Unified assurance*
<5 min
Policy answers*
25+
GRC frameworks

*Reference pilot metrics — see live capabilities on your deployment.

The problem

Teams need AI speed, but regulated and confidential work cannot use public chatbots.

Data sovereignty

Client IP, PHI, classified material, and board materials cannot leave the perimeter — yet analysts still reach for ChatGPT.

Compliance & audit

No citations, no audit trail, no acceptable-use enforcement — manual PDF search does not scale.

Unpredictable cost

Per-token cloud bills grow with every analyst; air-gapped sites pay twice (blocked AI + manual work).

Tool sprawl

Separate RAG tools, report writers, and governance spreadsheets — none integrated with SSO or RBAC.

The solution

One workbench: ingest, ask, deliver, and prove — on hardware you control.

Your documents → Local embeddings (ChromaDB) → RAG with citations → Chat · Reports (DOCX/PPTX) · Diagrams (Mermaid/PlantUML) → Audit log · Assurance dashboard · GRC repository ↑ Local LLM (LM Studio via Docker llmster · OpenAI-compatible · failover URL)

100% local by default

Web search and URL scrape off in offline/air-gap profile. Optional hybrid mode when policy allows.

Grounded answers

RAG with source citations, anti-hallucination prompts, hybrid search and reranking.

Production ready

Portal auth, RBAC, Prometheus, Docker, SAML/OIDC, unified assurance APIs.

Why Sovereign

Built for defense, BFSI, healthcare, and government — not repurposed from a consumer chatbot.

vs public cloud AIvs generic RAG tools
No mandatory data egressIntegrated reports + diagrams + chat
Predictable infra cost (your GPU/CPU)Single analyst workbench + API
Offline / air-gap capableAudit log + RBAC + portal
Unified assurance & GRC repositoryOWASP LLM + EU AI Act workflows
IDE integration (Cursor / VS Code)Enterprise ML & SLHP scheduler

Product capabilities

Full stack from SLM inference to board-ready assurance evidence.

LayerWhat you getEvidence
LLM / SLM runtimeLM Studio (primary); failover URL; streaming SSE; OpenAI gateway; model routing (auto/fast/quality); autostart & recoverGET /health
RAG & knowledgeChromaDB + local embeddings; PDF/DOCX/XLSX/CSV; token-aware chunking; hybrid + rerank; evaluate APIPOST /rag_query
Documents & visualsDOCX/PPTX reports; Mermaid/PlantUML → PNG/SVG; spreadsheet & documentation assistPOST /generate_report
User portalSign-in, plans, OAuth, OTP/TOTP, API keys, IDE bundle, usage ledger/app · /login
SecurityRBAC, API keys, OIDC/SAML, PII scan, prompt guard, SSRF-safe scrape, tenant encryptionGET /governance/assurance/posture
Governance & GRC25+ frameworks offline; OWASP LLM; EU AI Act; audit chain; WORM archive; SLHP scheduler/governance
Enterprise MLDatasets, training jobs, model registry, learning profiles, stack completion tracker/enterprise
ObservabilityHealth, readiness, Prometheus metrics, structured audit log, session tracking/metrics
CyberShield CVA AI pentestPlaybook, next-gen STRIDE/design review, 19-agent roster, orchestrator, evidence replay, continuous validation — CPA/CVA/XDR gatedGET /assist/ai-pentest/status

View live capability catalog →

Feature highlights

Core analyst workflow

  • Multi-turn chat with conversation store and optional KB context
  • RAG query with hybrid search, reranking, and source citations
  • Knowledge base — PDF, DOCX, TXT, MD, CSV, XLSX; folder ingest, rebuild
  • Streaming chat (SSE), prompt library, and file/image attachments
  • Batch summarize and OpenAI-compatible API gateway
  • Agent tools — KB search, stats, diagram brief (function calling)

Integrations & government

  • Cursor & VS Code extension via MCP (sovereign_chat, sovereign_rag_query)
  • Chat modes & auto-routing — Agent, Plan, Debug, Multitask, Ask, Auto
  • Government productivity — multilingual Indian, GR summary, meeting minutes
  • Sovereign readiness API — IP, OSS, platform independence
  • LM Studio runtime with autostart, recover, and model routing (fast/quality)
  • MahaTraining catalog for government employee AI training

Enterprise & assurance

  • Portal plans (Free / Pro / Enterprise) + OAuth enrollment
  • Vision & PDF OCR pipelines (optional MLLM)
  • Multi-tenant vector stores + encryption
  • GRC webhook + local export · Enterprise ML stack builder
  • OT/IoT deployment blueprints (IEC 62443)

Governance automation

  • Self-Learning & Healing Program (SLHP) cycles
  • Production scheduler (daily / weekly / monthly)
  • Security engineering gap assessment API
  • Red team program registry
  • CyberShield CVA AI pentest — playbook, orchestrator, governed execution plane
  • Regulatory KB ingest (offline corpus)
  • Unified assurance composite score

Deployment flexibility

  • Windows / Linux — Python or Docker
  • Offline sovereign profile (start.bat)
  • Production scripts + Prometheus alerts
  • Approved SLM model catalog
  • HA runbook + LLM URL fallback
  • Team install bundle for IDE + MCP

Use cases

Proven patterns from pilots and reference deployments.

IDScenarioPrimary outcomeModules
UC-01Internal policy & compliance Q&ASourced answers in minutes, audit-readyRAG · Portal · Audit
UC-02Consulting / client confidentialityZero document egress; faster exec summariesReports · Air-gap · API keys
UC-03Air-gapped government labAI assist with web features disabledOffline mode · Approved models
UC-04OT / industrial operationsAssurance evidence + deployment templatesGovernance · OT blueprints
UC-05Financial services model riskGrounded policy interpretationRAG · EU AI Act workflow
UC-06Developer productivity (sovereign)IDE chat without cloud API keysMCP · Continue · Aider bundle
UC-07Executive reportingDOCX/PPTX from natural languageReports · Diagrams
UC-08Security & GRC evidenceLive assurance dashboard for auditors/governance/assurance
UC-09CyberShield CVA continuous validationGoverned AI pentest planning with CPA proxy and evidence replayCVA runtime · MCP · CPA

Case studies

Reference scenarios from internal pilots — customize with your logos and signed quotes.

01

Consulting — client confidentiality

Challenge: Consultants used personal cloud AI on client contracts — GDPR and contract risk.

Results: 0 egress incidents · ~40% faster first-draft summaries · full audit log for partners.

Stack: Docker · LM Studio 7B-class · web search off · API key per analyst.

02

Financial compliance — policy library

Challenge: Policy answers took 2–4 hours of manual PDF search.

Results: Median answer time under 5 minutes with citations; silent when context insufficient.

Stack: RAG over internal policy corpus · RBAC · enrollment.

03

Air-gapped government

Challenge: No outbound network; analysts blocked from AI assist entirely.

Results: Full workbench offline after model install; web scrape/search disabled by policy.

Stack: Offline profile · approved SLM catalog · local embeddings.

04

OT / industrial assurance

Challenge: Ungoverned AI tools on plant floor documentation.

Results: 98% unified assurance composite · 97 AI trust index · OT deploy templates.

Stack: IEC 62443 frameworks · assurance APIs · SLHP scheduler.

About us

Sovereign LLM Workbench

We build offline-capable, enterprise-grade AI workbenches for organizations that cannot rely on public cloud LLMs. Our platform combines local inference (LM Studio and OpenAI-compatible runtimes), production RAG, document generation, IDE integration, and an integrated governance command center — so legal, security, and engineering teams share one source of truth.

The product ships with an offline GRC repository (NIST AI RMF, ISO 42001, OWASP LLM Top 10, EU AI Act, IEC 62443, and 20+ industry packs), unified assurance scoring, and a user portal with IDE integrations for Cursor and VS Code — no cloud API keys required in the editor.

Mission: Make sovereign AI practical on a laptop or private server — with evidence auditors and CISOs can trust.

5-minute demo

  1. Ingest sample policy
  2. Ask retention policy question
  3. Generate DOCX summary
  4. Show /health + audit log

Ideal customer

Consulting · Compliance · Government · Defense · BFSI · Healthcare · OT/industrial · Any org blocking public ChatGPT.

Requirements

16GB+ RAM · Windows/Linux · Python 3.10+ or Docker · LM Studio with approved SLM (e.g. google/gemma-3-1b).

Editions & plans

EditionDeploymentHighlights
CommunityLocal installCore chat, RAG, reports, diagrams
ProDocker + portalEnrollment, API keys, governance viewer, IDE integration
EnterpriseSSO + private serverFull GRC, assurance admin, enterprise ML, multi-tenant, SLA-ready ops

Cost illustration (10 analysts, year 1): hardware $2K–5K amortized vs $6K–24K avoided cloud API spend — see pilot metrics.