К содержимому
learnspaceYOUR NEXT CHAPTER
ПРОСТРАНСТВО ОБУЧЕНИЯ
ГлавнаяКаталог курсовМоё обучениеCoursera

Знания без границ

Учитесь у лучших университетов и компаний мира.

Открыть Coursera
Интеграция
Пространство университета
Моё пространствоСтраница курса
↵
ЯЛичный кабинетСтудент
© 2026 LearnSpaceКаждый день — возможность узнать больше.Помощь
Prompt Engineering for Vision Models · LearnSpace
Назад в каталог
courseraПрограммирование

Prompt Engineering for Vision Models

Курс от DeepLearning.AI
Начальный≈ 1.5 чАнглийский
О курсеНавыкиПрограммаПреподаватели

О курсе

Prompt engineering is used not only in text models but also in vision models. Depending on the vision model, they may use text prompts, but can also work with pixel coordinates, bounding boxes, or segmentation masks. In this course, you’ll learn to prompt different vision models like Meta’s Segment Anything Model (SAM), a universal image segmentation model, OWL-ViT, a zero-shot object detection model, and Stable Diffusion 2.0, a widely used diffusion model. You’ll also use a fine-tuning technique called DreamBooth to tune a diffusion model to associate a text label with an object of your preference. In detail, you’ll explore: 1. Image Generation: Prompt with text and by adjusting hyperparameters like strength, guidance scale, and number of inference steps. 2. Image Segmentation: Prompt with positive or negative coordinates, and with bounding box coordinates. 3. Object detection: Prompt with natural language to produce a bounding box to isolate specific objects within images. 4. In-painting: Combine the above techniques to replace objects within an image with generated content. 5. Personalization with Fine-tuning: Generate custom images based on pictures of people or places that you provide, using a fine-tuning technique called DreamBooth. 6. Iterating and Experiment Tracking: Prompting and hyperparameter tuning are iterative processes, and therefore experiment tracking can help to identify the most effective combinations. This course will use Comet, a library to track experiments and optimize visual prompt engineering workflows.

Навыки, которые вы освоите

Prompt EngineeringFine-tuningGenerative Model ArchitecturesPrompt Engineering ToolsVision Transformer (ViT)Multimodal PromptsImage AnalysisAI PersonalizationComputer VisionGenerative AI

Программа курса

1 модулей · 2 учебных материалов

01Prompt Engineering for Vision Models2 материалов
Prompt Engineering for Vision ModelsВнешний инструментQuiz: Prompt Engineering for Vision ModelsЗадание

Учитесь у экспертов

Abigail Morgan

Machine Learning Engineer

Jacques Verre

Head of Product

Caleb Kaiser

Machine Learning Engineer

Prompt Engineering for Vision Models
В каталоге вашей программы

Инвестируйте в себя

Новые знания — в удобное для вас время.

Начать на Coursera

Обучение откроется на Coursera
в новой вкладке

Обучение на Coursera

≈ 1.5 ч

1 модулей

Язык: Английский

Часть программы вашего университета