Job description
Influences and leads regular administration and conducts highly complex performance trend analyses and manages server capacity. Facilitates the optimization of system configurations and backups. Analyzes system performance data and insights to drive enhancements to improve the performance reliability, and security, of systems and environments. Serves as a subject matter expert in the investigation of highly complex system issues and facilitates high-severity incident triage. Designs disaster recovery solutions.
Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.
True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.
We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing [Click to show email] or by calling 1-888-404-2494 in the United States.
Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
Responsibilities:
Key Responsibilities
System Installation &Configuration – Software Administration:
-Evaluates andsets standards for the performance and installation requirements of operatingsystems to ensure optimal installation and performance.
-Provides guidanceon the administration of middleware products in environments.
-Facilitates andimplements strategies for the deployment, maintenance, and operation ofinternal applications, ensuring the efficiency and performance of thesesystems.
-Influences andleads regular administration and conducts highly complex performance trendanalyses and manages server capacity to ensure service performance meets andexceeds standards.
-Serves as asubject matter expert in utilizing application monitoring tools to optimize andensure efficiency.
System Installation &Configuration – Installation and Configuration:
-Leads the teaminstalling and configuring servers, cloud infrastructure, and all software andenvironments.
-Facilitates theoptimization of system configurations and backups to ensure the optimalperformance and stability of the server infrastructure.
-Leads hardwaremaintenance, auditing, installation, and provisioning, ensuring all tasks areperformed efficiently and effectively.
-Partners withinternal technical experts and third-party vendors to resolve integrationchallenges, providing expert guidance and industry insights for innovativesolutions.
System Installation &Configuration – Identity & Access Management:
-Providesadditional support and guidance for the administration of access privileges inthe identity and access management system, ensuring accurate and secure accessto IT resources.
-Interprets useractivity data in the identity and access management system and leveragesexpertise to provide insights on system activity and recommend improvements.
-Designs andoptimizes access management systems.
Service Lifecycle Management– Batch Processing:
-Drives themonitoring and assessment of the batch process to ensure updates are appliedand proactively resolves any issues that arise.
-Leveragesindustry insights to drive improvements in batch management techniques usingdifferent work schedulers to configure jobs and job streams, definedependencies, and report job performance.
-Influences andcollaborates with teams to ensure scheduling and budgets of batch monitoringservices align with and support Service Level Agreements (SLA).
Service Lifecycle Management– Security Maintenance:
-Drivesstrategic improvements to procedures to ensure that compute and storage devicesare secure.
-Influences andcollaborates with teams to maintain privileged accounts/secrets integrity ofsystems and compute and file system security for the compute and storageenvironment.
-Analyzes andevaluates highly complex service and infrastructure dashboards, taking the leadin addressing identified anomalies.
-Coordinates andmanages long-term implementation strategies for monthly, quarterly, or hotfixpatches to address security vulnerabilities or bugs.
Service Lifecycle Management– System & Security Improvements:
-Analyzes systemperformance data and insights to drive enhancements to improve the performance,reliability, and security of systems and environments.
-Influences andcollaborates with Service teams to proactively identify, address, and predictgaps in operational capabilities, enhancing scalability and resiliency.
Incident Management &Support – Incident Management:
-Leadsend-to-end incident management lifecycle to ensure systems are stable, secure,and performing accurately.
-Evaluatesresults from incident-based data analyses for team metrics and key performanceindicators (KPIs) to identify patterns, root causes, and solutions to preventsystem and network incidents.
-Facilitatesincident review meetings to provide strategic oversight for operationalperformance and long-term solution implementation.
-Influencesthird party vendors and cross-functional teams (e.g., Development, CloudEngineering, Product Engineering, other IT teams) to develop and implementlong-term solutions for high-severity incidents, risks, or migrations.
-Serves as asubject matter expert in the investigation of highly complex system issues andfacilitates high-severity incident triage by designing Corrective andPreventative Action plans (CAPA) to drive incident resolution and prevention.
Incident Management &Support – Escalation Cases:
-Providesexpertise for escalated support cases by collaborating with internal technicalteams and third party vendors to drive issue resolution for a wide range ofproduction environment problems (e.g., immense growth, scaling, leveraging thecloud, extremely high performance, high availability requirements).
Incident Management &Support – Technical Support:
-Drives andevaluates the production environment by analyzing system error logs and ticketqueues, and coordinating with multiple teams involved in maintaining theenvironments.
-Adheres to teamschedule to drive ongoing technical support and service objectives.
-Facilitatesstrategic resolution and long-term solutions for highly complex, criticalcustomer system issues and develops and documents comprehensive technicalsolutions.
Incident Management &Support – Backups and Disaster Recovery:
-Drives theexecution and effectiveness of backup, restore, and disaster recovery processes.
-Influences thestrategic planning and coordination of disaster recovery drills.
-Designsdisaster recovery solutions to ensure preparedness and regulatory compliance.
Communication &Documentation – Technical Communication:
-Communicateshighly complex technical information to both technical and nontechnicalpersonnel including management.
-Develops andimplements training programs to ensure personnel are well-versed indomain-specific knowledge and practices.
-Drivestechnical strategies and solutions for cross-organization projects, programs,and activities by leveraging domain-specific expertise and interpreting highlycomplex technical information.
Communication &Documentation – Documentation & Reporting:
-Drives thecreation of documentation on ticket updates, code contributions,infrastructure, configurations, processes, and procedures (e.g., DisasterRecovery plans, Standard Operating Procedures, Corrective and PreventativeAction Plans).
-Analyzes andinterprets weekly and monthly reports on system performance and incidentprogress to provide insights on operational and management outcomes andbusiness impacts.
-Reviews andrefines technical documentation standards and best practices for internal use.
Additional Responsibilities(as needed)
Cloud InfrastructureSupport:
-Serves as asubject matter expert in collaborations with DevOps and Site ReliabilityEngineer (SRE) teams to manage large-scale infrastructure.
-Managescontinuous integration and continuous deployment (CI/CD) pipelines.
-Outlines andplans for patching and version upgrades to support cloud infrastructure.
Automation:
-Approvesrecommendations and determines plans for improvements to reduce incidents andproblems with automation and simplify server management.
-Owns and leadsthe implementation of reusable frameworks, standards, and automation to supportOracle Cloud Infrastructure.
-Establishes andoversees Workload Automation tools through design support, administration, andoptimization efforts.
الوصف الوظيفي
يؤثر ويقود الأعمال الإدارية الاعتيادية ويجري تحليلات اتجاهات الأداء المعقدة للغاية ويدير سعة الخوادم. يسهل تحسين تكوينات النظام والنسخ الاحتياطي. يحلل بيانات أداء النظام والرؤى لدفع التحسينات بهدف تعزيز الموثوقية والأمان والأداء للأنظمة والبيئات. يعمل كخبير موضوعي في التحقيق في مشكلات النظام المعقدة للغاية ويسهل فرز الحوادث عالية الخطورة. يصمم حلول التعافي من الكوارث.
أوراكل وحدها هي التي تجمع بين البيانات والبنية التحتية والتطبيقات والخبرة لتمكين كل شيء بدءًا من الابتكارات الصناعية ووصولاً إلى الرعاية المنقذة للحياة. ومع دمج الذكاء الاصطناعي عبر منتجاتنا وخدماتنا، نساعد العملاء على تحويل هذا الوعد إلى مستقبل أفضل للجميع. اكتشف إمكاناتك في شركة تقود الطريق في حلول الذكاء الاصطناعي والسحابة التي تؤثر على مليارات الأرواح.
يبدأ الابتكار الحقيقي عندما يتم تمكين الجميع من المساهمة. لهذا السبب نحن ملتزمون بتنمية القوى العاملة التي تعزز الفرص للجميع مع مزايا تنافسية تدعم موظفينا بخيارات التأمين الطبي والتأمين على الحياة والتقاعد المرنة. كما نشجع الموظفين على العطاء لمجتمعاتهم من خلال برامج التطوع لدينا.
نحن ملتزمون بإشراك الأشخاص ذوي الإعاقة في جميع مراحل عملية التوظيف. إذا كنت بحاجة إلى مساعدة بشأن إمكانية الوصول أو التسهيلات بسبب إعاقة في أي مرحلة، فيرجى إعلامنا عن طريق إرسال بريد إلكتروني إلى [Click to show email] أو بالاتصال بالرقم 1-888-404-2494 في الولايات المتحدة.
شركة أوراكل هي صاحب عمل يضمن تكافؤ الفرص. سيحصل جميع المتقدمين المؤهلين على الاعتبار للتوظيف دون النظر إلى العرق أو اللون أو الدين أو الجنس أو الأصل القومي أو التوجه الجنسي أو الهوية الجنسية أو الإعاقة أو وضع المحاربين القدامى المحميين أو أي خاصية أخرى يحميها القانون. ستنظر أوراكل في توظيف المتقدمين المؤهلين ممن لديهم سجلات اعتقال وإدانة وفقًا للقانون المعمول به.
المسؤوليات:
المسؤوليات الرئيسية
تثبيت النظام وتكوينه – إدارة البرامج:
-يقيم ويحدد المعايير لأداء ومتطلبات تثبيت أنظمة التشغيل لضمان التثبيت والأداء الأمثل.
-يقدم الإرشادات بشأن إدارة منتجات البرمجيات الوسيطة في البيئات.
-يسهل وينفذ استراتيجيات نشر التطبيقات الداخلية وصيانتها وتشغيلها، مما يضمن كفاءة وأداء هذه الأنظمة.
-يؤثر ويقود الإدارة الروتينية ويجري تحليلات اتجاهات الأداء المعقدة للغاية ويدير سعة الخادم لضمان تلبية أداء الخدمة وتجاوزه للمعايير.
-يعمل كخبير موضوعي في استخدام أدوات مراقبة التطبيقات لتحسين الكفاءة وضمانها.
تثبيت النظام وتكوينه – التثبيت والتكوين:
-يقود الفريق الذي يقوم بتثبيت وتكوين الخوادم والبنية التحتية السحابية وجميع البرامج والبيئات.
-يسهل تحسين تكوينات النظام والنسخ الاحتياطية لضمان الأداء الأفضل والاستقرار للبنية التحتية للخادم.
-يقود صيانة الأجهزة والتدقيق والتثبيت والتهيئة، مع ضمان تنفيذ جميع المهام بكفاءة وفاعلية.
-يتعاون مع الخبراء الفنيين الداخليين والبائعين الخارجيين لحل تحديات التكامل، وتقديم التوجيه الخبير والرؤى القطاعية للحلول المبتكرة.
تثبيت النظام وتكوينه – إدارة الهوية والوصول:
-يقدم دعمًا وإرشادات إضافية لإدارة صلاحيات الوصول في نظام إدارة الهوية والوصول، مما يضمن الوصول الدقيق والآمن إلى موارد تكنولوجيا المعلومات.
-يفسر بيانات نشاط المستخدم في نظام إدارة الهوية والوصول ويستفيد من الخبرة تقديم رؤى حول نشاط النظام والتوصية بالتحسينات.
-يصمم ويحسن أنظمة إدارة الوصول.
إدارة دورة حياة الخدمة – المعالجة بالدفعة:
-يقود عملية مراقبة وتقييم المعالجة بالدفعة لضمان تطبيق التحديثات وحل أي مشكلات تنشأ بشكل استباقي.
-يستفيد من الرؤى الصناعية لدفع التحسينات في تقنيات إدارة الدفعات باستخدام أدوات جدولة عمل مختلفة لتكوين المهام وتدفقات المهام، وتحديد التبعيات، وإعداد تقارير أداء المهام.
-يؤثر ويتعاون مع الفرق لضمان توافق الجدولة والميزانيات الخاصة بخدمات مراقبة الدفعات مع اتفاقيات مستوى الخدمة (SLA) ودعمها.
إدارة دورة حياة الخدمة – الصيانة الأمنية:
-يقود التحسينات الاستراتيجية للإجراءات لضمان أمان أجهزة الحوسبة والتخزين.
-يؤثر ويتعاون مع الفرق للحفاظ على سلامة الحسابات/الأسرار ذات الصلاحيات العالية للأنظمة وأمان نظام الملفات والحوسبة لبيئة الحوسبة والتخزين.
-يحلل ويقيم لوحات معلومات الخدمات والبنية التحتية المعقدة للغاية، مع أخذ زمام المبادرة في معالجة الحالات الشاذة المحددة.
-ينسق ويدير استراتيجيات التنفيذ طويلة المدى للتحديثات الشهرية أو الفصلية أو الإصلاحات العاجلة لمعالجة الثغرات الأمنية أو الأخطاء البرمجية.
إدارة دورة حياة الخدمة – تحسينات النظام والأمان:
-يحلل بيانات وأفكار أداء النظام لدفع التحسينات لتعزيز أداء وموثوقية وأمان الأنظمة والبيئات.
-يؤثر ويتعاون مع فرق الخدمات لتحديد الثغرات في القدرات التشغيلية ومعالجتها والتنبؤ بها بشكل استباقي، مما يعزز القابلية للتوسع والمرونة.
إدارة الحوادث والدعم – إدارة الحوادث:
-يقود دورة حياة إدارة الحوادث الشاملة لضمان استقرار الأنظمة وأمانها وأدائها بدقة.
-يقيم نتائج تحليلات البيانات المستندة إلى الحوادث لمقاييس الفريق ومؤشرات الأداء الرئيسية (KPIs) لتحديد الأنماط والأسباب الجذرية والحلول لمنع حوادث النظام والشبكة.
-يسهل اجتماعات مراجعة الحوادث لتوفير الإشراف الاستراتيجي للأداء التشغيلي وتنفيذ الحلول طويلة الأجل.
-يؤثر على البائعين الخارجيين والفرق متعددة الوظائف (مثل التطوير، وهندسة السحابة، وهندسة المنتجات، وفرق تكنولوجيا المعلومات الأخرى) لتطوير وتنفيذ حلول طويلة الأجل للحوادث عالية الخطورة أو المخاطر أو عمليات الانتقال.
-يعمل كخبير موضوعي في التحقيق في مشكلات النظام المعقدة للغاية ويسهل فرز الحوادث عالية الخطورة من خلال تصميم خطط الإجراءات التصحيحية والوقائية (CAPA) لدفع حل الحوادث ومنعها.
إدارة الحوادث والدعم – حالات التصعيد:
-يقدم الخبرة لحالات الدعم المرفوعة عن طريق التعاون مع الفرق الفنية الداخلية والبائعين الخارجيين لدفع حل المشكلات لمجموعة واسعة من مشكلات بيئة الإنتاج (مثل النمو الهائل، التوسع، الاستفادة من السحابة، الأداء العالي للغاية، متطلبات التوافر العالي).
إدارة الحوادث والدعم – الدعم الفني:
-يقود ويقيم بيئة الإنتاج من خلال تحليل سجلات أخطاء النظام وصفوف التذاكر، والتنسيق مع فرق متعددة مشاركة في صيانة البيئات.
-يلتزم بجدول الفريق لدفع الدعم الفني المستمر وأهداف الخدمة.
-يسهل الحل الاستراتيجي والحلول طويلة الأجل لمشكلات أنظمة العملاء الحرجة والمعقدة للغاية، ويطور ويوثق حلولاً فنية شاملة.
إدارة الحوادث والدعم – النسخ الاحتياطي والتعافي من الكوارث:
-يقود تنفيذ وفاعلية عمليات النسخ الاحتياطي والاستعادة والتعافي من الكوارث.
-يؤثر في التخطيط الاستراتيجي والتنسيق لتمارين التعافي من الكوارث.
-يصمم حلول التعافي من الكوارث لضمان الجاهزية والامتثال التنظيمي.
التواصل والتوثيق – التواصل الفني:
-ينقل المعلومات الفنية المعقدة للغاية إلى الموظفين الفنيين وغير الفنيين بما في ذلك الإدارة.
-يطور وينفذ برامج تدريبية لضمان دراية الموظفين الجيدة بالمجالات والممارسات الخاصة بالنطاق.
-يقود الاستراتيجيات والحلول الفنية للمشاريع والبرامج والأنشطة المشتركة بين المؤسسات من خلال الاستفادة من الخبرة المتخصصة وتفسير المعلومات الفنية المعقدة للغاية.
التواصل والتوثيق – التوثيق وإعداد التقارير:
-يقود إنشاء التوثيق الخاص بتحديثات التذاكر، ومساهمات البرمجة، والبنية التحتية، والتكوينات، والعمليات والإجراءات (مثل خطط التعافي من الكوارث، وإجراءات التشغيل القياسية، وخطط الإجراءات التصحيحية والوقائية).
-يحلل ويفسر التقارير الأسبوعية والشهرية حول أداء النظام وتطور الحوادث لتقديم رؤى حول النتائج التشغيلية والإدارية والآثار التجارية.
-يراجع ويصقل معايير التوثيق الفني وأفضل الممارسات للاستخدام الداخلي.
مسؤوليات إضافية (حسب الحاجة)
دعم البنية التحتية السحابية:
-يعمل كخبير موضوعي في التعاون مع فرق DevOps ومهندسي موثوقية الموقع (SRE) لإدارة البنية التحتية واسعة النطاق.
-يدير خطوط أنابيب التكامل المستمر والنشر المستمر (CI/CD).
-يحدد ويخطط للتصحيحات وترقيات الإصدار لدعم البنية التحتية السحابية.
الأتمتة:
-يعتمد التوصيات ويحدد خطط التحسين لتقليل الحوادث والمشكلات المتعلقة بالأتمتة وتبسيط إدارة الخادم.
-يمتلك ويقود تنفيذ الأطر والمقاييس القابلة لإعادة الاستخدام والأتمتة لدعم البنية التحتية لسحابة أوراكل (Oracle Cloud Infrastructure).
-ينشئ ويشرف على أدوات أتمتة عبء العمل من خلال جهود دعم التصميم والإدارة والتحسين.