Class Central is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

Coursera

Red Hat AI Inference Technical Overview

Red Hat via Coursera

Overview

Google, IBM & Meta Certificates – 40% Off
One plan covers every Professional Certificate on Coursera.
Unlock All Certificates
Gain essential insights into AI deployment with this Red Hat AI Inference technical overview. Learn how to address the complexities and costs of running AI models in production. Discover how Red Hat's solution, powered by vLLM, optimizes performance and delivers significant cost savings across cloud, on-premise, virtualized, and edge environments. Dive into advanced techniques like quantization and speculative decoding to enhance your AI inference capabilities. This course demonstrates seamless model deployment and management within OpenShift AI, showcasing how you can achieve unparalleled efficiency and flexibility for your AI workloads.

Syllabus

  • Foundations of Enterprise Inference and Operational Challenges
  • Inference Runtime Ecosystem and Model Compression
  • Core vLLM Runtimes and Acceleration Techniques
  • Parallelism Strategies and Distributed Inference Using llm-D
  • Inference Optimizations in Context
  • Installation Options and Single Server Deployment Demo
  • Graded Assessment

Taught by

Red Hat Training

Reviews

Start your review of Red Hat AI Inference Technical Overview

Never Stop Learning.

Get personalized course recommendations, track subjects and courses with reminders, and more.

Someone learning on their laptop while sitting on the floor.