نحن نبحث عن مهندس عمليات تطبيقات ذو خبرة لدعم الاستقرار والتوافر والأداء والتحسين المستمر لتطبيقات المؤسسات. يغطي الدور عمليات التطبيق عبر الحوادث والمشكلات والتغيير والإصدار والنسخ الاحتياطي/الاسترداد، والتعافي من الكوارث، وأنشطة تحسين الخدمة.
المسؤوليات الرئيسية
- إدارة الحوادث الوظيفية من المستوى L2/L3، بما في ذلك المشاركة في جسور الحوادث وإدارة التصعيد.
- تنفيذ إدارة الحوادث والمشكلات، بما في ذلك تحليل اتجاه الحوادث وتنفيذ الإصلاحات الدائمة.
- تطوير وإدارة خطط تحسين الخدمة.
- دعم جداول الإصدار وأنشطة التخطيط.
- التخطيط وتنسيق طلبات التغيير (CRs)، بما في ذلك تقييم المخاطر والتأثير.
- تنفيذ تغييرات على مستوى التطبيق وإجراء مراجعات ما بعد التنفيذ.
- التعامل مع طلبات تكوين التطبيق.
- رصد أداء وتوافر التطبيق.
- إجراء ترابط الأحداث والتحقيق في الأحداث المتعلقة بالتطبيق.
- إدارة سعة وأداء التطبيق.
- تطوير واختبار إجراءات استرداد التطبيق.
- تكوين النسخ الاحتياطي للتطبيق وتنفيذ استعادة التطبيق والبيانات.
- المشاركة في اختبارات التعافي من الكوارث (DR) وتنفيذها.
- إجراء تعزيزات الحماية لأنظمة التشغيل وقواعد البيانات ووسيطات البرمجيات.
- تطبيق التصحيحات والتحديثات لمكونات التطبيق.
- الحفاظ على سجلات CMDB دقيقة لمكونات التطبيق.
- التعاون مع الفرق الفنية ذات الصلة لضمان استقرار التطبيق وتوفر الخدمة.
الملف الشخصي المرغوب للمرشح
- درجة البكالوريوس في علوم الحاسوب أو تكنولوجيا المعلومات أو الهندسة أو مجال ذي صلة.
- خبرة ذات صلة في عمليات التطبيقات / دعم التطبيقات في بيئة مؤسسية.
- خبرة عملية قوية مع إدارة الحوادث والمشكلات لتطبيقات المستوى L2/L3.
- خبرة في إدارة التغيير، وتخطيط الإصدار، وتطبيق تغييرات على مستوى التطبيق.
- خبرة في إجراء تقييمات المخاطر والتأثير لتغيّرات التطبيقات.
- خبرة عملية في رصد أداء وتوافر التطبيق والترابط بين الأحداث.
- خبرة في عمليات النسخ الاحتياطي والاستعادة والتعافي للتطبيق.
- خبرة في المشاركة في اختبارات التعافي من الكوارث (DR) أو تنفيذها.
- فهم جيد لتعزيز أمان نظام التشغيل وقواعد البيانات ووسيطات البرمجيات وتحديثاتها.
- خبرة في إدارة السعة والأداء.
- خبرة في الحفاظ على سجلات مكونات تطبيق CMDB.
- مهارات قوية في استكشاف الأخطاء وتحليلها وإدارة التصعيد والاتصال.
- القدرة على العمل مع فرق تقنية متعددة خلال الحوادث الكبرى وجسور الحوادث.
We are looking for an experienced Application Operations Engineer to support the stability, availability, performance, and continuous improvement of enterprise applications. The role covers application operations across incident, problem, change, release, backup/recovery, disaster recovery, and service improvement activities.
Key Responsibilities
- Manage L2/L3 functional incidents , including participation in incident bridges and escalation management.
- Perform incident and problem management , including incident trend analysis and implementation of permanent fixes.
- Develop and manage Service Improvement Plans .
- Support release schedule and planning activities.
- Plan and coordinate Change Requests (CRs) , including risk and impact assessments.
- Perform application-layer change implementation and conduct post-implementation reviews.
- Handle application configuration requests.
- Monitor application performance and availability .
- Perform event correlation and investigate application-related events.
- Manage application capacity and performance .
- Develop and test application recovery procedures .
- Configure application backups and execute application and data restores .
- Participate in and execute Disaster Recovery (DR) tests .
- Perform OS, database, and middleware hardening activities.
- Apply patches and updates to application components.
- Maintain accurate CMDB records for application components.
- Collaborate with relevant technical teams to ensure application stability and service availability.
Desired Candidate Profile
- Bachelor s degree in Computer Science, Information Technology, Engineering, or a related field .
- Relevant experience in Application Operations / Application Support in an enterprise environment.
- Strong hands-on experience with L2/L3 application incident and problem management .
- Experience with change management, release planning, and application-layer change implementation .
- Experience conducting risk and impact assessments for application changes.
- Hands-on experience with application performance and availability monitoring and event correlation.
- Experience with application backup, restore, and recovery procedures .
- Experience participating in or executing Disaster Recovery (DR) tests .
- Good understanding of OS, database, and middleware hardening and patching .
- Experience with capacity and performance management .
- Experience maintaining CMDB application component records .
- Strong troubleshooting, analytical, escalation management, and communication skills.
- Ability to work with multiple technical teams during major incidents and incident bridges .