Skip to content
Beta

Architecture Asset: Generative AI with Retrieval-Augmented Generation on STACKIT

Last updated on

This pattern delivers grounded generative AI. Documents are embedded into an OpenSearch vector store; at query time the app retrieves relevant context and calls a served LLM, keeping data and inference on sovereign infrastructure.

  • Knowledge assistants: answer questions grounded in internal documents.
  • Support automation: draft responses from trusted knowledge sources.
  • Sovereign GenAI: run retrieval and inference with control over data and residency.
Knowledge IngestionQuery & GenerationPlatform ServicesDocuments (Object Storage)Embedding Pipeline (Kubernetes)Vector Store (OpenSearch)API Gateway / LBRAG Application (Kubernetes)Model Serving (LLM)Secret ManagerObservability
  • Ground every answer: retrieve from the vector store before generation to reduce hallucination.
  • Keep data sovereign: run embedding, retrieval, and inference on STACKIT.
  • Evaluate continuously: track answer quality, safety, and cost before and after release.
  • Secure prompts and secrets: isolate keys and never log sensitive context.
Asset historyActive 4 of the last 12 weeksTMUpdatedNo updates · 1 bar = 1 week i
Maintainers
TMTobias M.Head of STACKIT Cloud Framework · STACKITOwnerActive 12 of the last 12 weeks · 168 updatesSTACKITwww.linkedin.com/in/tobias-müller-011304172Contributed in STACKIT
Show full history (3 more)