К содержимому
learnspaceYOUR NEXT CHAPTER
ПРОСТРАНСТВО ОБУЧЕНИЯ
ГлавнаяКаталог курсовМоё обучениеCoursera

Знания без границ

Учитесь у лучших университетов и компаний мира.

Открыть Coursera
Интеграция
Пространство университета
Моё пространствоСтраница курса
↵
ЯЛичный кабинетСтудент
© 2026 LearnSpaceКаждый день — возможность узнать больше.Помощь
Evaluate & Optimize LLM Performance · LearnSpace
Назад в каталог
courseraАнализ данных

Evaluate & Optimize LLM Performance

Курс от Coursera
Средний≈ 4.5 чАнглийский
О курсеНавыкиПрограммаПреподаватели

О курсе

You've integrated a powerful Large Language Model (LLM) into your application. The initial results are impressive, and your team is excited. But then the hard questions start. Is the new prompt really better than the old one, or does it just "feel" better? How do you prove to stakeholders that switching from GPT-3.5 to GPT-4 is worth the extra cost? When you have two models that give slightly different answers, how do you decide which one is objectively superior? After completing this course, you will have the confidence to lead your team in making smart, evidence-based decisions that measurably improve your AI applications. Ready to Become an LLM Expert? It's time to bring scientific rigor to the art of AI. Enroll in Evaluate & Optimize LLM Performance and gain the essential skills to build, validate, and perfect the next generation of language models.

Навыки, которые вы освоите

Test Script DevelopmentStatistical MethodsStatistical InferenceStatistical AnalysisModel EvaluationLLM ApplicationScriptingProbability & StatisticsPrompt EngineeringData-Driven Decision-MakingLarge Language ModelingStatistical Hypothesis TestingEmbeddingsNatural Language ProcessingModel OptimizationPerformance Metric

Программа курса

3 модулей · 19 учебных материалов

01Build Automated LLM Evaluation Systems6 материалов
Welcome: Why Does "Good" AI Content Go Bad?DIALOGUEA Guide to LLM Evaluation: Lexical and Semantic MetricsЧтениеHow to Compute Lexical Metrics: BLEU & ROUGE-L in Python?ВидеоHow to Compute Semantic Similarity with Embeddings?Видео

Учитесь у экспертов

Professionals in the Industry

Преподаватель курса

Evaluate & Optimize LLM Performance
В каталоге вашей программы

Инвестируйте в себя

Новые знания — в удобное для вас время.

Начать на Coursera

Обучение откроется на Coursera
в новой вкладке

Обучение на Coursera

≈ 4.5 ч

3 модулей

Язык: Английский

Субтитры: Арабский, Французский, Итальянский, Бразильский португальский, Корейский, Немецкий, Испанский, Японский

Часть программы вашего университета
Knowledge Check: Choosing Your MetricsЗадание
Building Your First Automated Evaluation ScriptЛабораторная
02Statistical Significance Testing7 материалов
Why Guess When You Can Know? The Case of the "Better" PromptВидеоThe Language of Experimentation: Hypotheses, P-Values, and PowerВидеоDesigning a Fair Race: A/B Testing for LLMsЧтениеRunning the Numbers: A/B Test Analysis in PythonВидеоStatistical Significance TestingЛабораторнаяKnowledge Check: Statistical Testing ConceptsЗаданиеCoach Dialogue: Interpreting Your A/B Test ResultsDIALOGUE
03Performance Analysis and Optimization6 материалов
From Report to Action: The Optimization LoopВидеоCase Study: Benchmarking a Sentiment AnalyzerВидеоBuilding a Reproducible Evaluation WorkflowЧтениеScripting Your First Evaluation ReportВидеоPlanning Your Optimization StrategyЛабораторнаяFinal Project: Build Your LLM Evaluation ToolkitЗадание