Visual data is becoming a critical asset for modern organisations. This course explores how to build intelligent applications that can interpret, analyse, and reason over images and documents using Azure AI services and multimodal models.
We believe organisations that can combine data, AI, and cloud technologies will unlock faster and more accurate decision-making. This course focuses on applying multimodal and agent-based AI patterns to extract structured insights from visual inputs and integrate them into real-world workflows. Learners will gain practical experience in designing solutions that combine visual understanding with language models, enabling applications that move from perception to action within Azure environments.
Prerequisites
Participants should have:
- Basic programming experience in a language such as Python
- A general understanding of cloud computing or artificial intelligence concepts
- Familiarity with data handling and application development workflows is beneficial
- No prior experience in computer vision is required
Target audience
This course is designed for:
- Developers building intelligent and data-driven applications
- AI engineers working with multimodal and generative AI solutions
- Technical professionals implementing Azure-based AI services
- Teams looking to integrate visual data into decision-making workflows
























