Skip to content

Enterprise RAG Logo

GPT-RAG Solution Accelerator

GPT-RAG is an enterprise-grade accelerator for building conversational AI assistants on Azure, powered by intelligent agents that understand questions, find the right information, and deliver clear, accurate answers using trusted enterprise data.

Designed with Zero-Trust security and Infrastructure as Code (IaC) principles from the ground up, GPT-RAG accelerates production deployments while ensuring consistency, governance, and operational excellence. It supports text, image, and voice scenarios, enabling organizations to rapidly create rich multimodal experiences.

Latest Stable Release v3.7.0

Architecture at a glance

GPT-RAG can start as a Basic deployment and expand into Zero Trust, public ingress, existing-platform integration, or optional AI capabilities. The approved target makes Microsoft Foundry hosted/no-panel the fresh-deployment chat runtime and retains the Container Apps orchestrator as an explicit fallback. See the Architecture page for the mode contract and required-vs-configurable deployment table.

Exact matrix pinned; runtime validation and release remain blocked

The umbrella integration pins UI v2.6.0, orchestrator v4.0.0, ingestion v2.7.0, and AILZ v2.5.0. Explicit hosted-panel topology selection is supported, but continuity, user-history, owner-binding validation, and operator-surface gates remain deployment-published false. The latest runtime attempt activated the agent version but session readiness returned HTTP 424, so the integration is not validated or shipped. See the hosted-agent integration matrix for the exact runtime, identity, RBAC, panel, data, and rollback contracts.

Chat runtime modes and hosted deployment lifecycle

Modular architecture layers

Zero Trust Architecture Full Zero Trust reference architecture. This is the complete network-isolated view, not the minimum Basic deployment. The focused SVG above describes the component-level hosted design; the matrix page records its umbrella release gate.

Governance

Before connecting enterprise data or relying on telemetry as evidence, review Governance and responsible operation. It explains intended use, limitations, data and telemetry responsibilities, retention, access, and the exact audit trail contract available since GPT-RAG v3.7.0, disabled by default.

Runtime Services

Services Description
Orchestrator Manages agentic workflows with Microsoft Agent Framework, Azure AI, and strategy-specific integrations.
Web UI User interface for chat interactions, supports streaming and custom themes.
Data Ingestion Extracts, chunks, and indexes enterprise data for optimized retrieval.
MCP Server Optional Model Context Protocol service for tool hosting and business logic integration.

Contributing

We welcome contributions from the community! Check our Contribution Guidelines for CLA, code of conduct, and PR guidelines.

© 2025 GPT-RAG — powered by ❤️ and coffee ☕