
GPT-RAG Solution Accelerator
GPT-RAG is an enterprise-grade accelerator for building conversational AI assistants on Azure, powered by intelligent agents that understand questions, find the right information, and deliver clear, accurate answers using trusted enterprise data.
Designed with Zero-Trust security and Infrastructure as Code (IaC) principles from the ground up, GPT-RAG accelerates production deployments while ensuring consistency, governance, and operational excellence. It supports text, image, and voice scenarios, enabling organizations to rapidly create rich multimodal experiences.
Architecture at a glance
GPT-RAG can start as a Basic deployment and expand into Zero Trust, public ingress, existing-platform integration, or optional AI capabilities. The approved target makes Microsoft Foundry hosted/no-panel the fresh-deployment chat runtime and retains the Container Apps orchestrator as an explicit fallback. See the Architecture page for the mode contract and required-vs-configurable deployment table.
Exact matrix pinned; runtime validation and release remain blocked
The umbrella integration pins UI v2.6.0, orchestrator v4.0.0, ingestion
v2.7.0, and AILZ v2.5.0. Explicit hosted-panel topology selection is
supported, but continuity, user-history, owner-binding validation, and
operator-surface gates remain deployment-published false. The latest
runtime attempt activated the agent version but session readiness returned
HTTP 424, so the integration is not validated or shipped. See the
hosted-agent integration matrix for
the exact runtime, identity, RBAC, panel, data, and rollback contracts.
Full Zero Trust reference architecture. This is the complete network-isolated
view, not the minimum Basic deployment. The focused SVG above describes the
component-level hosted design; the matrix page records its umbrella release
gate.
Governance
Before connecting enterprise data or relying on telemetry as evidence, review Governance and responsible operation. It explains intended use, limitations, data and telemetry responsibilities, retention, access, and the exact audit trail contract available since GPT-RAG v3.7.0, disabled by default.
Runtime Services
| Services | Description |
|---|---|
| Orchestrator | Manages agentic workflows with Microsoft Agent Framework, Azure AI, and strategy-specific integrations. |
| Web UI | User interface for chat interactions, supports streaming and custom themes. |
| Data Ingestion | Extracts, chunks, and indexes enterprise data for optimized retrieval. |
| MCP Server | Optional Model Context Protocol service for tool hosting and business logic integration. |
Contributing
We welcome contributions from the community! Check our Contribution Guidelines for CLA, code of conduct, and PR guidelines.