Apply now »

Principal Job

Date:  19 Jul 2026
Custom Field 1:  650015
Location: 

Riyadh, SA

Facility:  Technology & Engineering

Job Description

OVERVIEW

Job Title

Principal

Job Code

 

Grade

E1

Group

Government Product

Division

Technology

Department

Technology Operations Center

Unit

-

 

ROLE PURPOSE

The aim is to state the overall significance of the job from the organization’s perspective.

To lead the strategy, architecture, and advancement of enterprise observability, monitoring, and AIOps capabilities across Elm’s technology platforms. The role is responsible for establishing scalable monitoring frameworks, enhancing service reliability and operational resilience, driving intelligent automation initiatives, and enabling proactive identification and resolution of technology issues. The position serves as the organization's subject matter expert in observability and operational intelligence, ensuring technology services remain reliable, secure, and aligned with business objectives.

 

KEY ACCOUNTABILITIES & ACTIVITIES

This section describes the principal outputs required from the job.

Key Accountabilities

Key Activities

  1. Enterprise Observability Strategy & Governance
  • Define and lead the enterprise observability strategy across infrastructure, platforms, applications, and cloud environments.
  • Establish monitoring standards, governance frameworks, and operating models that support organizational objectives.
  • Develop strategic roadmaps for observability and monitoring capabilities.
  • Evaluate emerging technologies and industry best practices to continuously enhance monitoring services.
  1. Monitoring Platform Architecture & Engineering
  • Lead the design, development, and optimization of enterprise monitoring and observability platforms.
  • Establish architecture standards covering metrics, logs, traces, events, and telemetry collection.
  • Ensure monitoring platforms are scalable, resilient, secure, and capable of supporting future growth.
  • Drive continuous enhancement of platform capabilities and monitoring coverage.
  1. Service Reliability & Availability Engineering
  • Define and monitor service level indicators (SLIs), service level objectives (SLOs), and operational performance metrics.
  • Drive initiatives that improve technology service reliability, resilience, and availability.
  • Identify recurring operational issues and establish preventive improvement plans.
  • Support the reduction of service outages through proactive monitoring practices.
  1. AIOps & Intelligent Automation Leadership
  • Develop and lead the strategic roadmap for AIOps adoption across technology operations.
  • Drive automation initiatives leveraging artificial intelligence and machine learning capabilities.
  • Establish intelligent alerting, anomaly detection, and predictive monitoring capabilities.
  • Evaluate and implement advanced operational technologies that improve efficiency and operational decision-making.
  1. Operational Intelligence & Advanced Analytics
  • Lead the development of dashboards, operational analytics, and executive reporting capabilities.
  • Transform monitoring data into actionable insights that improve operational performance.
  • Analyze technology trends, service performance indicators, and operational risks.
  • Enable data-driven decision-making through advanced observability insights.

 

  1. Platform Engineering Integration & Enablement
  • Integrate observability capabilities into platform engineering and application delivery processes.
  • Establish observability-by-design practices across technology platforms.
  • Enable development and operations teams through reusable monitoring frameworks and standards.
  • Promote self-service monitoring capabilities and operational best practices.
  1. Incident Intelligence & Operational Excellence
  • Enhance incident detection, correlation, and root cause analysis capabilities across technology environments.
  • Improve incident response effectiveness through advanced monitoring and automation solutions.
  • Lead continuous improvement initiatives focused on operational excellence and service optimization.
  • Strengthen operational resilience through proactive risk identification and mitigation practices.
  1. Technical Leadership, Innovation & Knowledge Management
  • Act as the organization's subject matter expert for observability, monitoring, and AIOps disciplines.
  • Provide expert guidance and influence technology strategy through thought leadership and innovation.
  • Lead capability development initiatives and promote knowledge sharing across technology teams.
  • Reduce operational dependency on individual resources through standardization, documentation, and knowledge transfer practices.
  1. Policies, Processes & Procedures
  • Follow and enforce all relevant departmental policies, processes, and standard operating procedures.
  • Ensure work is carried out in a controlled and consistent manner.
  • Comply with safety, quality, and environmental policies.
  1. Information Security
  • Ensure compliance with information security policies and standards.
  • Maintain secure infrastructure environments and access controls.
  • Support implementation of security controls and risk mitigation measures.

 

 

JOB SPECIFICATIONS

Academic and professional qualifications

  • Bachelor’s Degree in Computer Science, Information Technology, Software Engineering, Computer Engineering, or a related discipline.

Years and Nature of Experience

  • 8+ years of experience in Technology Operations, Platform Engineering, Infrastructure Operations, Observability, Monitoring, or related domains.

 


Job Segment: Information Security, Software Engineer, Cloud, Computer Science, Engineer, Technology, Engineering

Apply now »