About Course
AI Engineering and LLMOps is an advanced, hands-on course for learners who want to understand how AI systems are engineered, deployed and operated in production.
You will move beyond building individual AI applications to explore what happens underneath them. The course covers LLM inference, model serving, KV caching, batching, quantisation, deployment architecture, evaluation, observability, security and cost optimisation.
Through practical exercises, you will learn how to turn an AI prototype into a reliable production system. You will examine the engineering decisions behind latency, throughput, scalability and model quality, and learn how those decisions affect users and operating costs.
By the end of the course, you will have designed and implemented a production-ready AI system, complete with evaluation, monitoring, deployment and operational considerations.
