Job Title
– Gen AI Engineer
Company
– TCS (MEA)
Location
– Dhahran, Saudi Arabia
Job type
– Full time
About Us:
Tata Consultancy Services (TCS) is an IT services, consulting and business solutions organization that has been partnering with many of the world’s largest businesses in their transformation journeys for over 50 years. TCS offers a consulting-led, cognitive powered, integrated portfolio of business, technology and engineering services and solutions. This is delivered through its unique Location Independent Agile™ delivery model, recognized as a benchmark of excellence in software development.
A part of the Tata group, India's largest multinational business group, TCS has over 616,171 of the world’s best-trained consultants with 157 nationalities in 53 countries. For more information, visit www.tcs.com and follow TCS news at @TCS_News.
Job Description:
Must Have Technical/Functional Skills;
-
Experience with
video analytics
,
object detection
,
tracking
,
segmentation
, and
scene understanding
.
-
Exposure to
edge computing
,
GPU optimization
, and
model quantization
for real-time deployment.
-
Familiarity with
CI/CD pipelines
and MLOps practices in vision systems.
-
Experience with
large-scale vision datasets
and data curation tools.
-
Proficiency in Python
, with hands-on experience in
PyTorch
,
TensorFlow
, or
Paddle Paddle
.
-
Experience with YOLO models
, including training, fine-tuning, and deployment.
-
Working knowledge of CNNs, Auto-Encoders, and Vision-Language Models (VLMs)
for both images and videos.
-
Experience using Paddle OCR or similar OCR frameworks
for text detection and recognition in images and video frames.
-
Experience with NVIDIA Deep Stream SDK
for multi-stream video processing and inference acceleration.
-
Familiarity with annotation tools
such as
Label Studio
or
CVAT
, including managing and cleaning annotated datasets.
-
Experience building and deploying microservices
with
Fast API
and
Docker
.
-
Strong understanding of
data preprocessing
, augmentation, and pipeline design for vision tasks.
Roles & Responsibilities;
-
Design, develop, and deploy
computer vision models
for processing
multiple real-time camera streams
.
-
Build and optimize
deep learning pipelines
using
CNNs
,
YOLO
,
Auto-Encoders
,
Vision-Language Models (VLMs)
, and
OCR models
.
-
Fine-tune and train
YOLO
,
VLM
, and
OCR models
(such as
Paddle OCR
) for both
image and video
understanding, including
text detection and recognition
.
-
Collaborate with data engineering teams to develop data annotation pipelines and clean large-scale datasets.
-
Utilize tools such as
Label Studio
or
CVAT
for annotation tasks and ensure high-quality labeled data.
-
Deploy models using
NVIDIA Deep Stream
,
Docker
, and
microservices architecture
for scalable production environments.
-
Develop APIs and backend services using
Fast API
for model inference and integration.
-
Continuously improve model performance through experimentation, evaluation, and feedback loops.
-
Contribute to the documentation, scaling, and maintenance of vision systems in production.
Generic Managerial Skills, If any
-
Excellent problem-solving skills and the ability to work in a fast-paced, dynamic environment.
-
Strong communication skills and the ability to collaborate effectively with cross-functional teams.
Application Deadline: 30-June-2026
Privacy Note:
https://www.tcs.com/connect-with-tcs/privacy-policy