A

Associate MLOps Engineer

AppliedAI

Abu Dhabi, United Arab Emirates · Full Time

Be the first to apply

Experience
1–2 yrs
Salary
—
Openings
1
Posted
5 days ago
Work mode
In office
Resume
Required to apply

Where you'll work

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

About AppliedAI

AppliedAI is an innovative AI technology firm based in Abu Dhabi, UAE, specializing in delivering advanced artificial intelligence solutions to highly regulated sectors including healthcare, insurance, government, and financial services. Our mission is to transform traditional workflows by automating complex, document-intensive processes, enhancing both efficiency and precision. We integrate human expertise with AI to drive success in regulated industries.

Job Overview

We are looking for an Associate MLOps Engineer to join our expanding team in Abu Dhabi. This role bridges the gap between operational and developmental functions, ensuring platform reliability, performance, and security. You will support scalable solutions development with guidance from senior team members—a valuable opportunity for those early in their SRE/MLOps careers.

Key Responsibilities

  • Supervise and sustain reliability for production, staging, and development environments, guaranteeing high availability and optimal system performance.
  • Work collaboratively with DevOps, MLOps, and development teams to diagnose and address issues.
  • Maintain and enhance the observability infrastructure under senior engineers’ mentorship.
  • Participate in refining incident management procedures.
  • Support capacity planning and enhance performance managing up to 10,000 requests per minute.
  • Ensure compliance with security protocols and regulatory standards throughout the infrastructure.
  • Join regular business hours on-call rotations with escalation support from senior engineers.
  • Implement and sustain service level objectives (SLOs), indicators (SLIs), and agreements (SLAs).
  • Contribute to cloud cost-efficiency alongside system reliability and performance efforts.
  • Assist in maintaining CI/CD pipelines and deployment workflows.

Required Skills & Experience

  • Familiarity with infrastructure best practices, including AWS and Azure Well-Architected Frameworks.
  • Basic grasp of resilient infrastructure patterns such as high availability, fault tolerance, and disaster recovery.
  • Understanding fundamental infrastructure security concepts like least privilege, network segmentation, and encryption.
  • Awareness of compliance and governance frameworks.
  • Interest in financial operations and cost minimization methodologies.
  • Exposure to Infrastructure as Code with modularity, reuse, and versioning concepts.
  • Knowledge of observability including logging, metrics, and tracing.
  • 1-2 years experience in SRE, DevOps, or related roles; internships and projects considered.
  • Hands-on experience with AWS services including Lambda, ECS, Fargate, ALB, ELB, API Gateway, Route53, CloudFront, AppSync, DynamoDB, RDS (PostgreSQL), Aurora, EventBridge, SNS, SQS, Security Groups, Secrets Manager, Systems Manager, IAM, ECR, CodeBuild, and CodeDeploy.
  • Experience with monitoring tools and observability platforms.
  • Basic infrastructure as code skills using CDK and/or Terraform.
  • Understanding of event-driven architectures.
  • Some exposure to containerization and microservices.
  • Competent in scripting and automation.
  • Strong analytical and problem-solving capabilities with a methodical debugging approach.
  • Experience working in Agile development environments is advantageous.

Preferred Qualifications

  • Professional certifications in AWS, Azure, or GCP.
  • Experience with Next.js, Node.js, and Python programming languages.
  • Familiarity with authentication frameworks such as Auth0 and SSO.
  • Knowledge of regulatory compliance standards including SOC 2, HIPAA, GDPR, PCI DSS.
  • Interest in machine learning and large language model operations.
  • Experience with multi-region AWS deployments and handling high-traffic systems.
  • Basic database management and optimization skills.
  • Understanding of caching mechanisms and content delivery networks.
  • Familiarity with data lifecycle management, ETL processes, and vector/graph databases is a plus.

What We Offer

  • Exposure to state-of-the-art technologies.
  • Mentoring by senior SRE, architecture, DevOps, and MLOps professionals.
  • Supportive collaboration across multiple technical teams.
  • Opportunities for career advancement in a swiftly growing startup.
  • Work with a globally distributed team.
  • Regular hours with flexibility for occasional urgent support requests.
  • A defined pathway to increase responsibility within reliability engineering.

Required Personal Attributes

  • Effective communication skills.
  • Analytical and solution-oriented approach.
  • Strong team collaboration mindset.
  • Self-driven with a proactive attitude.
  • Comfortable working in a fast-evolving startup environment.
  • Commitment to continuous learning and skill development.

Additional Benefits

  • Access to a leading AI technology company environment.
  • Innovative, entrepreneurial work culture.
  • Professional growth and development prospects.
  • Based in a thriving technological ecosystem at our Abu Dhabi headquarters.
  • 21 days of paid annual leave.
  • Comprehensive health insurance provided by the company.
  • Visa sponsorship available for international candidates.

Tools & software

AWS Amazon Web Services AWS required

How they work

Communication Teamwork & Collaboration Problem Solving Learning Agility Motivation
🤖
Online · instant AI help
Broxer