Job description
Principal Job
Date:
19 Jul 2026
Custom Field 1:
650015
Location:
Riyadh, SA
#job-location.job-location-inline { display: inline; }
Facility:
Technology & Engineering
Job Description
OVERVIEW
Job Title
Principal
Job Code
Grade
E1
Group
Government Product
Division
Technology
Department
Technology Operations Center
Unit
-
ROLE PURPOSE
The aim is to state the overall significance of the job from the organization’s perspective.
To lead the strategy, architecture, and advancement of enterprise observability, monitoring, and AIOps capabilities across Elm’s technology platforms. The role is responsible for establishing scalable monitoring frameworks, enhancing service reliability and operational resilience, driving intelligent automation initiatives, and enabling proactive identification and resolution of technology issues. The position serves as the organization's subject matter expert in observability and operational intelligence, ensuring technology services remain reliable, secure, and aligned with business objectives.
KEY ACCOUNTABILITIES & ACTIVITIES
This section describes the principal outputs required from the job.
Key Accountabilities
Key Activities
- Enterprise Observability Strategy & Governance
- Define and lead the enterprise observability strategy across infrastructure, platforms, applications, and cloud environments.
- Establish monitoring standards, governance frameworks, and operating models that support organizational objectives.
- Develop strategic roadmaps for observability and monitoring capabilities.
- Evaluate emerging technologies and industry best practices to continuously enhance monitoring services.
- Monitoring Platform Architecture & Engineering
- Lead the design, development, and optimization of enterprise monitoring and observability platforms.
- Establish architecture standards covering metrics, logs, traces, events, and telemetry collection.
- Ensure monitoring platforms are scalable, resilient, secure, and capable of supporting future growth.
- Drive continuous enhancement of platform capabilities and monitoring coverage.
- Service Reliability & Availability Engineering
- Define and monitor service level indicators (SLIs), service level objectives (SLOs), and operational performance metrics.
- Drive initiatives that improve technology service reliability, resilience, and availability.
- Identify recurring operational issues and establish preventive improvement plans.
- Support the reduction of service outages through proactive monitoring practices.
- AIOps & Intelligent Automation Leadership
- Develop and lead the strategic roadmap for AIOps adoption across technology operations.
- Drive automation initiatives leveraging artificial intelligence and machine learning capabilities.
- Establish intelligent alerting, anomaly detection, and predictive monitoring capabilities.
- Evaluate and implement advanced operational technologies that improve efficiency and operational decision-making.
- Operational Intelligence & Advanced Analytics
- Lead the development of dashboards, operational analytics, and executive reporting capabilities.
- Transform monitoring data into actionable insights that improve operational performance.
- Analyze technology trends, service performance indicators, and operational risks.
- Enable data-driven decision-making through advanced observability insights.
- Platform Engineering Integration & Enablement
- Integrate observability capabilities into platform engineering and application delivery processes.
- Establish observability-by-design practices across technology platforms.
- Enable development and operations teams through reusable monitoring frameworks and standards.
- Promote self-service monitoring capabilities and operational best practices.
- Incident Intelligence & Operational Excellence
- Enhance incident detection, correlation, and root cause analysis capabilities across technology environments.
- Improve incident response effectiveness through advanced monitoring and automation solutions.
- Lead continuous improvement initiatives focused on operational excellence and service optimization.
- Strengthen operational resilience through proactive risk identification and mitigation practices.
- Technical Leadership, Innovation & Knowledge Management
- Act as the organization's subject matter expert for observability, monitoring, and AIOps disciplines.
- Provide expert guidance and influence technology strategy through thought leadership and innovation.
- Lead capability development initiatives and promote knowledge sharing across technology teams.
- Reduce operational dependency on individual resources through standardization, documentation, and knowledge transfer practices.
- Policies, Processes & Procedures
- Follow and enforce all relevant departmental policies, processes, and standard operating procedures.
- Ensure work is carried out in a controlled and consistent manner.
- Comply with safety, quality, and environmental policies.
- Information Security
- Ensure compliance with information security policies and standards.
- Maintain secure infrastructure environments and access controls.
- Support implementation of security controls and risk mitigation measures.
JOB SPECIFICATIONS
Academic and professional qualifications
- Bachelor’s Degree in Computer Science, Information Technology, Software Engineering, Computer Engineering, or a related discipline.
Years and Nature of Experience
- 8+ years of experience in Technology Operations, Platform Engineering, Infrastructure Operations, Observability, Monitoring, or related domains.
Job Segment:
Information Security, Software Engineer, Cloud, Computer Science, Engineer, Technology, Engineering
Preferred candidate
Years of experience
8+ years
Degree
Bachelor's degree / higher diploma
وصف الوظيفة
الوظيفة الرئيسية
التاريخ:
19 يوليو 2026
الحقل المخصص 1:
650015
الموقع:
الرياض، المملكة العربية السعودية
#job-location.job-location-inline { display: inline; }
المنشأة:
التكنولوجيا والهندسة
وصف العمل
نظرة عامة
المسمى الوظيفي
المسؤول
الرمز الوظيفي
الدرجة
E1
المجموعة
Government Product
القسم
التكنولوجيا
الإدارة
مركز عمليات التكنولوجيا
الوحدة
-
الغرض من الدور
يُقصد بيان الأهمية العامة للوظيفة من منظور المنظمة.
قيادة الاستراتيجية والهندسة وتطور قدرات الرصد المؤسسي والمراقبة و-AIOps عبر منصات Elm التقنية. يتحمل الدور مسؤولية وضع أطر مراقبة قابلة للتوسع، تعزيز موثوقية الخدمات والمرونة التشغيلية، قيادة مبادرات الأتمتة الذكية، وتمكين التعرف والحل الاستباقي لمشكلات التكنولوجيا. يشغل المنصب كبائع خبرة المؤسسة في الرصد والذكاء التشغيلي، لضمان أن تظل الخدمات التقنية موثوقة وآمنة ومتوافقة مع أهداف الأعمال.
المساءلة الأساسية والأنشطة
هذا القسم يوضح المخرجات الرئيسية المطلوبة من الوظيفة.
المسؤوليات الأساسية
الأنشطة الأساسية
- استراتيجية الحوسبة الرصدية المؤسسية والحوكمة
- تحديد وقيادة استراتيجية الرصد المؤسسي عبر البنية التحتية والمنصات والتطبيقات وبيئات السحابة.
- إرساء معايير المراقبة وأطر الحوكمة ونماذج التشغيل التي تدعم أهداف المنظمة.
- تطوير خرائط طريق استراتيجية لقدرات الرصد والمراقبة.
- تقييم التقنيات الناشئة وأفضل الممارسات الصناعية لتحسين خدمات الرصد باستمرار.
- هندسة عمارة منصة الرصد
- قيادة تصميم وتطوير وتحسين منصات الرصد والرصد المؤسسي.
- وضع معايير الهندسة تشمل المقاييس والسجلات والتتبع والفصول والأحداث وجمع القياس.
- التأكد من أن منصات الرصد قابلة للتوسع ومرنة وآمنة وقادرة على دعم النمو المستقبلي.
- قيادة التحسين المستمر لقدرات المنصة وتغطية الرصد.
- هندسة الاعتمادية والتوافر للخدمات
- تعريف ومراقبة مؤشرات مستوى الخدمة (SLIs)، أهداف مستوى الخدمة (SLOs)، ومقاييس الأداء التشغيلية.
- قيادة مبادرات تعزز موثوقية الخدمات التقنية ومرونتها وتوافرها.
- تحديد المشكلات التشغيلية المتكررة ووضع خطط تحسين وقائية.
- دعم تقليل انقطاعات الخدمة من خلال ممارسات المراقبة الاستباقية.
- قيادة AIOps والأتمتة الذكية
- تطوير وقيادة خارطة طريق استراتيجية لتبني AIOps عبر عمليات التكنولوجيا.
- دفع مبادرات الأتمتة باستخدام قدرات الذكاء الاصطناعي وتعلم الآلة.
- إرساء الإنذار الذكي واكتشاف الشذوذ والرصد التنبؤي.
- تقييم وتنفيذ تقنيات تشغيلية متقدمة تحسن الكفاءة واتخاذ القرار التشغيلي.
- الذكاء التشغيلي والتحليلات المتقدمة
- قيادة تطوير لوحات المعلومات والتحليلات التشغيلية وتقارير تنفيذية.
- تحويل بيانات الرصد إلى رؤى قابلة للتنفيذ وتحسين الأداء التشغيلي.
- تحليل اتجاهات التكنولوجيا ومؤشرات أداء الخدمات والمخاطر التشغيلية.
- تمكين اتخاذ القرار المستند إلى البيانات من خلال رؤى الرصد المتقدم.
- تكامل وتمكين هندسة المنصة
- دمج قدرات الرؤية في عمليات هندسة المنصة وتوصيل التطبيقات.
- إرساء ممارسات الرصد بالتصميم عبر المنصات التقنية.
- تمكين فرق التطوير والعمليات من خلال أطر ومقاييس الرصد القابلة لإعادة الاستخدام.
- تعزيز قدرات الرصد الذاتي وأفضل الممارسات التشغيلية.
- ذكاء الحوادث والتفوق التشغيلي
- تعزيز اكتشاف الحوادث وربطها وتحليل السبب الجذري عبر بيئات التكنولوجيا.
- تحسين فاعلية الاستجابة للحوادث من خلال حلول الرصد والأتمتة المتقدمة.
- قيادة مبادرات التحسين المستمر مركّزة على التميز التشغيلي وتحسين الخدمة.
- تعزيز المرونة التشغيلية من خلال التعرف والمخاطر الاستباقية والحد منها.
- القيادة الفنية والابتكار وإدارة المعرفة
- التصرف كخبير موضوعي للمراقبة والرصد وAIOps داخل المنظمة.
- تقديم التوجيه الخبيري وتأثير استراتيجية التكنولوجيا من خلال التفكير القيادي والابتكار.
- قيادة مبادرات تطوير القدرات وتشجيع تبادل المعرفة عبر فرق التكنولوجيا.
- تقليل الاعتماد التشغيلي على موارد فردية من خلال التوحيد والوثائق ونقل المعرفة.
- السياسات والعمليات والإجراءات
- اتباع وتطبيق جميع السياسات والعمليات والإجراءات التشغيلية القياسية ذات الصلة بالأقسام.
- ضمان تنفيذ العمل بطريقة محكومة ومتسقة.
- الامتثال لسياسات السلامة والجودة والبيئة.
- أمن المعلومات
- ضمان الامتثال لسياسات ومعايير أمن المعلومات.
- الحفاظ على بيئات بنية تحتية آمنة وحدود وصول.
- دعم تطبيق ضوابط الأمن وتدابير الحد من المخاطر.
مواصفات الوظيفة
المؤهلات الأكاديمية والمهنية
- درجة البكالوريوس في علوم الحاسوب أو تكنولوجيا المعلومات أو هندسة البرمجيات أو الهندسة الحاسوبية أو تخصص ذي صلة.
سنوات ونوع الخبرة
- 8+ سنوات من الخبرة في عمليات التكنولوجيا وهندسة المنصات وعمليات البنية التحتية والرصد والمراقبة أو مجالات ذات صلة.
فئة الوظيفة:
أمن المعلومات، مهندس برمجيات، سحابة، علوم حاسوب، مهندس، تكنولوجيا، هندسة
مرشح مفضل
سنوات الخبرة
8+ سنوات
المؤهل
درجة البكالوريوس / دبلومة أعلى