К содержимому
learnspaceYOUR NEXT CHAPTER
ПРОСТРАНСТВО ОБУЧЕНИЯ
ГлавнаяКаталог курсовМоё обучениеCoursera

Знания без границ

Учитесь у лучших университетов и компаний мира.

Открыть Coursera
Интеграция
Пространство университета
Моё пространствоСтраница курса
↵
ЯЛичный кабинетСтудент
© 2026 LearnSpaceКаждый день — возможность узнать больше.Помощь
LLM Benchmarking and Evaluation Training · LearnSpace
Назад в каталог
courseraАнализ данных

LLM Benchmarking and Evaluation Training

Курс от Simplilearn
Начальный≈ 5.6 чАнглийский
О курсеНавыкиПрограммаПреподаватели

О курсе

This comprehensive course on Evaluating and Applying LLM Capabilities equips you with the skills to analyze, implement, and assess large language models in real-world scenarios. Begin with core capabilities, learn summarization, translation, and how LLMs power industry-relevant content generation. Progress to interactive and analytical applications—explore chatbots, virtual assistants, and sentiment analysis with hands-on demos using LangChain and ChromaDB. Conclude with benchmarking and evaluation—master frameworks like ROUGE, GLUE, SuperGLUE, and BIG-bench to measure model accuracy, relevance, and performance. To be successful in this course, you should have a basic understanding of LLMs, Python, and NLP fundamentals. By the end of this course, you will be able to: - Explore LLM Capabilities: Understand summarization, translation, and their applications - Build LLM Applications: Create chatbots and sentiment analysis tools using real-world tools - Evaluate Model Performance: Use ROUGE, GLUE, and BIG-bench to benchmark LLMs - Analyze Use Cases: Assess benefits, limitations, and deployment of LLM-powered solutions Ideal for AI developers, ML engineers, and GenAI professionals.

Навыки, которые вы освоите

LLM ApplicationLarge Language ModelingLangChainModel EvaluationEmbeddingsBenchmarkingGenerative AIAnalytical SkillsNatural Language Processing

Программа курса

3 модулей · 29 учебных материалов

01Core Capabilities of LLMs10 материалов

Introduction to LLM Capabilities

Course SyllabusЧтениеLearning ObjectivesВидеоFour Major Capabilities of LLMВидеоQuiz on Introduction to LLM CapabilitiesЗадание

Учитесь у экспертов

Priyanka Mehta

Преподаватель курса

LLM Benchmarking and Evaluation Training
В каталоге вашей программы

Инвестируйте в себя

Новые знания — в удобное для вас время.

Начать на Coursera

Обучение откроется на Coursera
в новой вкладке

Обучение на Coursera

≈ 5.6 ч

3 модулей

Язык: Английский

Субтитры: Арабский, Французский, Итальянский, Бразильский португальский, Корейский, Немецкий, Индонезийский, Испанский, Японский, Венгерский

Часть программы вашего университета

Introduction to Summarization

Overview, Benefits, Limitations, and Industrial Applications of SummarizationВидеоDemo: Text SummarizerВидеоQuiz on Introduction to SummarizationЗадание

Introduction to Content Translation

Overview, Benefits, Limitations, and Industrial Applications of Content TranslationВидеоQuiz on Introduction to Content TranslationЗаданиеAssessment on Core Capabilities of LLMsЗадание
02Interactive and Analytical LLM Applications7 материалов

Chatbots and Virtual Assistants

Overview, Benefits, Limitations, and Industrial Applications of Chatbots and Virtual AssistantsВидеоDemo: MultiPDF QA Retriever with ChromaDB and LangChainВидеоQuiz on Chatbots and Virtual AssistantsЗадание

Introduction to Sentiment Analysis

Overview, Benefits, and Limitations of Sentiment AnalysisВидеоDemo: Sentiment AnalysisВидеоQuiz on Introduction to Sentiment AnalysisЗаданиеAssessment on Interactive and Analytical LLM ApplicationsЗадание
03LLM Evaluation and Benchmarking12 материалов

Introduction to Benchmarking

Benchmarking and Its StepsВидеоQuiz on Introduction to BenchmarkingЗадание

Benchmarks for Evaluating LLMs

Benchmarks for Language ModelsВидеоDemo: ROUGE BenchmarkВидеоNeed for New BenchmarksВидеоGLUE Benchmark TasksВидеоSuperGLUE Benchmark Tasks: Part 1ВидеоSuperGLUE Benchmark Tasks: Part 2ВидеоBeyond the Imitation Game Benchmark (BIG-bench)ВидеоKey TakeawaysВидеоQuiz on Benchmarks for Evaluating LLMsЗаданиеAssessment on LLM Evaluation and BenchmarkingЗадание