How do LLMs work with Vision AI? | OCR, Image & Video Analysis
How do LLMs work with Vision AI? | OCR, Image & Video Analysis

How do LLMs work with Vision AI? | OCR, Image & Video Analysis

AMEN@12

7 min0 plays0 favorites
News
Play

Description

<p>Combine vision and language in an AI model with the latest vision AI model in Azure Cognitive Services. Use natural language to fetch visual content in images and videos without needing metadata or location, generate automatic and detailed descriptions of images using the model’s knowledge of the world, and use a verbal description to search video content.</p> <p>Cognitive Service for Vision AI combines both natural language models (LLM) with computer vision and is part of the Azure Cognitive Services suite of pre-trained AI capabilities. It can carry out a variety of vision-language tasks including automatic image classification, object detection, and image segmentation. Similar to GPT, the foundational language model, Project Florence, used in this case infuses deeper language skill with vision analytics to make training, inferencing and interacting with your image and video content simpler using natural language. </p> <p>Azure Expert, Matt McSpirit shares how to customize the model and use these capabilities in your own apps.</p> <p>► QUICK LINKS:</p> <p>00:00 - Introduction</p> <p>00:48 - Project Florence</p> <p>01:52 - Open-world recognition</p> <p>03:19 - Dense captioning</p> <p>04:23 - Run frame analysis</p> <p>05:02 - Train a custom model</p> <p>06:29 - Build custom apps</p> <p>07:41 - Wrap up</p> <p>► Link References:</p> <p>Check out <a href= "https://aka.ms/CognitiveVision">https://aka.ms/CognitiveVision</a></p> <p>► Unfamiliar with Microsoft Mechanics?</p> <p>As Microsoft's official video series for IT, you can watch and share valuable content and demos of current and upcoming tech from the people who build it at Microsoft.</p> <p>• Subscribe to our YouTube: <a href= "https://www.youtube.com/c/MicrosoftMechanicsSeries">https://www.youtube.com/c/MicrosoftMechanicsSeries</a></p> <p>• Talk with other IT Pros, join us on the Microsoft Tech Community: <a href= "https://techcommunity.microsoft.com/t5/microsoft-mechanics-blog/bg-p/MicrosoftMechanicsBlog"> https://techcommunity.microsoft.com/t5/microsoft-m

Creators

BlakeWave

BlakeWave

Creator

How do LLMs work with Vision AI? | OCR, Image & Video Analysis - Listen Free | WowFM