К содержимому
learnspaceYOUR NEXT CHAPTER
ПРОСТРАНСТВО ОБУЧЕНИЯ
ГлавнаяКаталог курсовМоё обучениеCoursera

Знания без границ

Учитесь у лучших университетов и компаний мира.

Открыть Coursera
Интеграция
Пространство университета
Моё пространствоСтраница курса
↵
ЯЛичный кабинетСтудент
© 2026 LearnSpaceКаждый день — возможность узнать больше.Помощь
Analyze & Deploy Scalable LLM Architectures · LearnSpace
Назад в каталог
courseraПрограммирование

Analyze & Deploy Scalable LLM Architectures

Курс от Coursera
Средний≈ 2.6 чАнглийский
О курсеНавыкиПрограммаПреподаватели

О курсе

Analyze & Deploy Scalable LLM Architectures is an intermediate course for ML engineers and AI practitioners tasked with moving large language model (LLM) prototypes into production. Many powerful models fail under real-world load due to architectural flaws. This course teaches you to prevent that. You will learn to analyze multi-stage architectures such as RAG to diagnose and quantify performance bottlenecks with evidence, not assumptions. You will then master the tools of production-grade operations, designing and writing declarative Helm charts to deploy containerized LLM applications on Kubernetes. The curriculum focuses on building resilient, scalable systems by implementing Horizontal Pod Autoscaling (HPA) to handle unpredictable traffic and managing the full deployment lifecycle with controlled rollouts and rapid rollbacks. By the end of this course, you will be able to transform fragile prototypes into robust, reliable, and scalable production services.

Навыки, которые вы освоите

ContainerizationApplication Performance ManagementKubernetesDebuggingPerformance TuningPerformance AnalysisApplication DeploymentRetrieval-Augmented GenerationLLM ApplicationConfiguration ManagementSystems AnalysisRelease ManagementAnalysisModel DeploymentContinuous DeliveryMLOps (Machine Learning Operations)Large Language ModelingScalabilityCloud-Native Computing

Программа курса

3 модулей · 20 учебных материалов

01Architecture Performance Analysis7 материалов
Course Introduction: The Weekend OutageDIALOGUEWhy Performance is a Pipeline ProblemВидеоDeconstructing a RAG ArchitectureЧтениеHow to Trace a Request and Spot BottlenecksВидео

Учитесь у экспертов

Professionals in the Industry

Преподаватель курса

Analyze & Deploy Scalable LLM Architectures
В каталоге вашей программы

Инвестируйте в себя

Новые знания — в удобное для вас время.

Начать на Coursera

Обучение откроется на Coursera
в новой вкладке

Обучение на Coursera

≈ 2.6 ч

3 модулей

Язык: Английский

Субтитры: Арабский, Французский, Итальянский, Бразильский португальский, Корейский, Немецкий, Индонезийский, Испанский, Японский

Часть программы вашего университета
Hands-On Learning (HOL): Analyze the Architecture DiagramЗадание
Presenting Your Architectural FindingsDIALOGUE
Scenario-Based Question: Architectural AnalysisЗадание
02Performance Tuning and Optimization6 материалов
Evidence Replaces Assumption: The Power of ProfilingЧтениеHow to Quantify Latency from LogsВидеоInterpreting Performance DashboardsЧтениеHands-On Learning (HOL): Analyzing Production Logs to Identify Performance BottlenecksЗаданиеDesigning a Causal ExperimentDIALOGUEEvidence-Based Performance Tuning QuizЗадание
03Container Orchestration and Deployment7 материалов
Why Prototypes Fail in ProductionВидеоDeclarative Deployments with Helm and KubernetesЧтениеHow to Write a Helm Chart with AutoscalingВидеоAnatomy of a Production Helm ChartЧтениеHands-On Learning (HOL): Review and Correct the Helm ManifestЗаданиеSimulating a Production RolloutDIALOGUEFinal Project: Scalable LLM Deployment PortfolioЗадание