Course Name
Course Code : SZW49
Venue Details
Postal Code : 30-644
Session Dates
Duration: 3 days (21 hours)
This program introduces participants to Vision-Language Models
(VLMs) using fully open-source, modern models like Qwen 3 VL, Gemma 3
Vision, and Kimi-VL.
Learners will understand how multimodal AI works, prepare simple image/video
datasets, run open-source VLMs, perform fine-tuning using LoRA,
and deploy multimodal applications.
The course focuses on intuitive explanations, hands-on guided exercises, and easy-to-use tools (Google Colab, Hugging Face Spaces, Gradio).
By the end, learners will be able to build and deploy their own simple VLM applications.
Introduction to VLMs & Core Foundations
Module 1: Understanding Vision-Language Models
Module 2: Modern Open-Source VLMs
Module 3: Architecture
Module 4: Data Fundamentals
Module 5: Simple Dataset Preparation
Module 6: Running Your First VLM (Hands-on)
Fine-Tuning (Images + Simple Video)
Module 7: Introduction to Fine-Tuning
Module 8: Fine-Tuning on Images
Module 9: Evaluating Fine-Tuned Models
Module 10: Intro to Video Fine-Tuning
Module 11: Fine-Tuning Kimi-VL or Qwen 3 VL on Small Video Tasks
Evaluation, Deployment & Responsible AI
Module 12: Model Evaluation
Module 13: Error Analysis & Debugging Basics
Module 14: Deployment Basics
Module 15: Building a Simple End-to-End Demo
Module 16: Responsible AI Essentials
Module 17: Wrap-Up & Learning Path
Mode of Delivery : The event can be attended both online and at nearby ProgNXT classroom by Individual Professionals and Corporate Employees as per the seat availability. Please Contact Us at [email protected] for checking the seat availability
Audience : We have a global audience that logs in to using their own computers to work hand in hand with our world-class instructors.
Assessment : Each training course will have ProgNXT Assessment at the end.
Certification : After successful passing of ProgNXT Assessment, ProgNXT Certification will be provided, which has got acceptance in 55+ Countries.
| Global Region | Location | Start Date | End Date | Action |
|---|---|---|---|---|
| | | | | |
| | | | | |
| | | | | |
| | | | | |
| | | | | |