
VISION & MULTIMODEL
Computer VisionOCRVideo UnderstandingMultimodal AI
ABOUT THIS ROLE
Provides visual capabilities to machines. This division experiments with image and video processing, object recognition, and the development of multimodal models that combine image and text analysis.
REQUIREMENT
- Understand basic image processing and Computer Vision concepts.
- Familiar with Python and OpenCV.
- Understand the concepts of image classification or object detection.
- Plus Point: Experience with YOLO or CNN architectures, as well as knowledge regarding OCR, segmentation, or video processing.
