Requirements
Key Responsibilities
• Develop AI solutions for image/video understanding, generation, analysis and transformation.
• Work with VLMs, multimodal LLMs, diffusion models and GenAI APIs.
• Build image/video pipelines, multimodal RAG and AI-agent workflows.
• Evaluate and optimize models for quality, latency and cost.
• Take solutions from PoC to production using cloud and modern AI engineering practices.
Required Skills
• Strong Python, PyTorch and Computer Vision experience.
• Hands-on experience with GenAI, VLMs/LLMs, image/video AI.
• Knowledge of OpenCV, Hugging Face, OCR, embeddings and RAG.
• Experience with Azure/AWS/GCP, Docker and APIs.
• Exposure to image/video generation, diffusion models and AI agents is a plus.
Experience: 4–8+ years in AI/ML, Computer Vision or GenAI, with strong hands-on GenAI experience.