اتصل بفرصتك ادخل في دور عالي التأثير حيث ستساعدنا في بناء وتشغيل المركز القيادي المركزي الجديد للإدارة العامة للمالية في القاهرة. كمهندس موثوقية موقع (SRE) ستضمن بقاء بنية المؤسسة التحتية عالية التوفر وقابلة للتوسع ومُحسّنة بشكل كامل. ستتجاوز مجرد استكشاف الأخطاء لحل التحديات التقنية المعقدة بشكل استباقي عبر بيئات سحابية متنوعة. ستكوّن فريق مستوى الدعم 2 و3 لمنتج على مستوى وطني تم نشره في 3 دول مجلس التعاون الخليجي، وسيتم نقل معرفة المنتج من خلال جلسات تدريبية.
المراقبة والاستكشاف الاستباقي لمنع الانقطاعات والتوقفات الخدمية
التحقيق وحل الحوادث البنيوية والتطبيقية القادمة من دعم المستوى 1 أو من أدوات الملاحظية
تحليل السبب الجذري وإدارة المشكلة، إجراء مراجعات ما بعد الحوادث وتحديث دفاتر التشغيل ومقالات قاعدة المعرفة
إدارة الأداء والسعة، مراقبة استغلال السحابة وتحسين بيئات متعددة السحابات لضمان موثوقية النظام المستمرة ووقت التشغيل
الأتمتة والكفاءة، تطوير نصوص لتقليل الجهد وتحقيق التوحيد القياسي، الحفاظ على قوالب البنية التحتية كرمز وإجراءات قابلة للتكرار
الأمن والامتثال، المساعدة في تدقيقات العملاء وردود الحوادث
النسخ الاحتياطي والاسترداد، التحقق من وظائف النسخ الاحتياطي ومراقبتها، اختبار إجراءات الاسترداد
التوثيق وتبادل المعرفة، الحفاظ على مستندات دقيقة للتهيئة، تدريب دعم المستوى 1 لزيادة معدل الحل من أول اتصال
التواصل مع العملاء وتتبع وتقرير اتفاقيات مستوى الخدمة/مؤشرات الأداء الرئيسية
التنسيق مع البائعين والأدوات، التنسيق مع مقدمي الخدمات السحابية وتتبع التذاكر المفتوحة لضمان الحل في الوقت المناسب
مسؤوليات النوبة أثناء الاتصال، المشاركة في دورات الدوران خلال النوبات
الملف المرشح المرغوب
اتصل بمهاراتك وخبرتك المهنية- التفكير التحليلي يمكّنك من تفكيك سلوك النظام المعقد لإيجاد السبب الجذري المباشر لمشاكل البنية التحتية.
- القدرة على التكيف تتيح لك الانتقال بسلاسة بين منصات وأدوات سحابية مختلفة في بيئة سريعة الإيقاع.
- حل المشكلات بشكل تعاوني يضمن عملك بفعالية مع فرق هندسية متخصصة لبناء حلول تقنية مرنة وطويلة الأمد.
الأساسيات:
- خبرة واسعة ومتقدمة في عمليات السحابة المتعددة وهندسة موثوقية الموقع، مع وجود خبرة صناعية على مستوى كبير.
- خبرة مثبتة في السحابة المتعددة مع قدرات تقنية عملية عبر المنصات.
- امتلاك واحد على الأقل من الشهادات التالية (ملاحظة: إذا حصلت على اثنتين، ستقوم Deloitte بتدريبك ودعمك للحصول على الثالثة):
- شهادة خدمات أمازون ويب AWS
- شهادة Google Cloud Platform GCP
- شهادة Oracle Cloud Infrastructure OCI
- خبرة مع أدوات الرصد والملاحظية
- خبرة مع تدفقات التكامل
- خبرة مع أدوات DevOps
- الإلمام بممارسات هندسة موثوقية الموقع الأساسية وأطر الأتمتة
- 3 إلى 7 سنوات على الأقل من الخبرة السابقة في العمل ضمن مركز قيادة IT على مستوى المؤسسة.
Connect to your opportunity Step into a highly impactful role where you will help us build and operate our brand-new centralized command center for public finance in Cairo. As a Site Reliability Engineer (SRE) you will ensure our enterprise infrastructure remains highly available, scalable, and fully optimized. You will go beyond basic troubleshooting to proactively manage and resolve complex technical challenges across diverse cloud environments. You will compose our support level 2 and 3 team for a national scale product that has been deployed in 3 GCC counties, the knowledge of the product will be transferred in training sessions.
Monitor and proactively troubleshooting to prevent outage and service interruption
Investigate and resolve infrastructure and application incidents coming from support level 1 or from the observability tools.
Root cause analysis and problem management, conduct post incident reviews and update runbooks and knowledge base articles
Performance and capacity management, monitor cloud utilization and optimize multi-cloud environments to ensure continuous high system reliability and uptime.
Automation and efficiencies, develop scripts to generate effort reductions and standardization, maintain infrastructure as a code templates and repeatable procedures.
Security and compliance, assist with client s audits and incidents responses
Back up and recovery, validate and monitor back up jobs, test recovery procedures
Document and knowledge sharing, maintain accurate documents for configuration, train support level 1 to increase first call resolution
Customer communication and SLAs/KPIs tracking and reporting
Vendor and tool liaison, coordinate with cloud providers and track open tickets to ensure timely resolution
On call shift responsibilities, participate in on call rotations
Desired Candidate Profile
Connect to your skills and professional experience- Analytical Thinking enables you to break down complex system behaviors to find the direct root cause of infrastructure issues.
- Adaptability allows you to seamlessly transition between different cloud platforms and tools in a fast-paced environment.
- Collaborative Problem-Solving ensures you work effectively with specialized engineering teams to build resilient, long-term technical solutions.
Essentials:
- Extensive, highly advanced expertise in multi-cloud operations and Site Reliability Engineering, reflecting senior-level industry tenure.
- Demonstrated multi-cloud expertise with hands-on technical capabilities across platforms.
- Possession of at least two of the following certifications (Note: if you hold two, Deloitte will train and support you in acquiring the third):
- Amazon Web Services (AWS) Certification
- Google Cloud Platform (GCP) Certification
- Oracle Cloud Infrastructure (OCI) Certification
- Experience with observability tools
- Experience with integration flows
- Experience with DevOps tools
- Familiarity with foundational Site Reliability Engineering (SRE) practices and automation frameworks.
- 3 to 7 years minimum prior experience operating within an enterprise-level IT command center.