T

Gen AI Engineer

Tata Consultancy Services

Dhahran, Eastern Province, Saudi Arabia · Full Time

Be the first to apply

Experience
Any
Salary
Openings
1
Posted
4 ಗಂಟೆಗಳು ಹಿಂದೆ
Work mode
In office
Resume
Required to apply

Where you'll work

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

About Tata Consultancy Services

Tata Consultancy Services (TCS) is a global leader in IT services, consulting, and business solutions, partnering with some of the world’s largest enterprises for over 50 years. It provides consulting-driven, cognitively powered, integrated business, technology, and engineering services through its unique Location Independent Agile™ delivery model, recognized for outstanding software development standards. Part of the Tata Group, India's largest multinational conglomerate, TCS employs over 616,000 consultants across 53 countries representing 157 nationalities.

Job Overview

This position is based in Dhahran, Saudi Arabia, and offers a full-time, onsite role as a Gen AI Engineer specializing in computer vision and real-time video analytics.

Key Responsibilities

  • Architect, create, and deploy computer vision models to process multiple real-time video streams simultaneously.
  • Develop and enhance deep learning workflows utilizing CNNs, YOLO, Auto-Encoders, Vision-Language Models, and OCR technologies.
  • Fine-tune and train cutting-edge models including YOLO, VLMs, and Paddle OCR for visual and textual recognition tasks in images and videos.
  • Work closely with data engineering teams to set up robust annotation pipelines and ensure the quality of large-scale datasets, utilizing tools such as Label Studio and CVAT.
  • Implement scalable production deployments leveraging NVIDIA Deep Stream SDK, containerization with Docker, and microservices design patterns.
  • Create and maintain backend APIs for inference and integration using Fast API.
  • Engage in continuous experimentation to optimize model accuracy and system performance.
  • Document, scale, and support vision system deployments in real-world production environments.

Required Skills and Experience

  • Expertise in video analytics including object detection, tracking, segmentation, and comprehensive scene interpretation.
  • Hands-on experience with edge computing, GPU performance tuning, and model quantization to achieve real-time results.
  • Proficiency with continuous integration and continuous deployment (CI/CD) pipelines and MLOps best practices tailored to vision systems.
  • Competence handling extensive vision datasets and data curation workflows.
  • Advanced Python programming skills with practical knowledge of deep learning frameworks such as PyTorch, TensorFlow, or Paddle Paddle.
  • Experience training and deploying YOLO models and familiarity with CNNs, Auto-Encoders, and Vision-Language Models for multimedia understanding.
  • Operational knowledge of OCR frameworks like Paddle OCR for extracting text from images and video sequences.
  • Experience with NVIDIA Deep Stream SDK for multi-stream video inference acceleration.
  • Use of annotation platforms including Label Studio and CVAT for dataset labeling and maintenance.
  • Development and deployment of microservices using Fast API and Docker technology.
  • Strong grasp of data preprocessing, augmentation strategies, and pipeline formulation for computer vision tasks.

Additional Skills

  • Outstanding problem-solving capability and adaptability in a dynamic, fast-paced setting.
  • Excellent communication skills to facilitate smooth collaboration across interdisciplinary teams.

Application Details

Please note the application deadline is June 30, 2026. Candidates should be prepared to work onsite in Dhahran, Saudi Arabia.

Tools & software

PyTorch required TensorFlow required

How they work

Communication Teamwork & Collaboration Problem Solving Adaptability

Leave it if you'd like a reply — we won't use it for anything else.

Click to browse, drag & drop, or paste a screenshot

PNG, JPG, GIF, MP4, WebM, MOV · Max 20MB each · Up to 5 files

🤖
Online · instant AI help