On-site Full Time
Oracle - Egypt -
Egypt , Cairo
--
Oracle - Egypt

Job Details

Job description

Influences and leads regular administration and conducts highly complex performance trend analyses and manages server capacity. Facilitates the optimization of system configurations and backups. Analyzes system performance data and insights to drive enhancements to improve the performance reliability, and security, of systems and environments. Serves as a subject matter expert in the investigation of highly complex system issues and facilitates high-severity incident triage. Designs disaster recovery solutions.

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.


True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.


We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing [Click to show email] or by calling 1-888-404-2494 in the United States.


Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.



Responsibilities:

Key Responsibilities


System Installation &Configuration – Software Administration:


-Evaluates andsets standards for the performance and installation requirements of operatingsystems to ensure optimal installation and performance.


-Provides guidanceon the administration of middleware products in environments.


-Facilitates andimplements strategies for the deployment, maintenance, and operation ofinternal applications, ensuring the efficiency and performance of thesesystems.


-Influences andleads regular administration and conducts highly complex performance trendanalyses and manages server capacity to ensure service performance meets andexceeds standards.


-Serves as asubject matter expert in utilizing application monitoring tools to optimize andensure efficiency.


System Installation &Configuration – Installation and Configuration:


-Leads the teaminstalling and configuring servers, cloud infrastructure, and all software andenvironments.


-Facilitates theoptimization of system configurations and backups to ensure the optimalperformance and stability of the server infrastructure.


-Leads hardwaremaintenance, auditing, installation, and provisioning, ensuring all tasks areperformed efficiently and effectively.


-Partners withinternal technical experts and third-party vendors to resolve integrationchallenges, providing expert guidance and industry insights for innovativesolutions.


System Installation &Configuration – Identity & Access Management:


-Providesadditional support and guidance for the administration of access privileges inthe identity and access management system, ensuring accurate and secure accessto IT resources.


-Interprets useractivity data in the identity and access management system and leveragesexpertise to provide insights on system activity and recommend improvements.


-Designs andoptimizes access management systems.


Service Lifecycle Management– Batch Processing:


-Drives themonitoring and assessment of the batch process to ensure updates are appliedand proactively resolves any issues that arise.


-Leveragesindustry insights to drive improvements in batch management techniques usingdifferent work schedulers to configure jobs and job streams, definedependencies, and report job performance.


-Influences andcollaborates with teams to ensure scheduling and budgets of batch monitoringservices align with and support Service Level Agreements (SLA).


Service Lifecycle Management– Security Maintenance:


-Drivesstrategic improvements to procedures to ensure that compute and storage devicesare secure.


-Influences andcollaborates with teams to maintain privileged accounts/secrets integrity ofsystems and compute and file system security for the compute and storageenvironment.


-Analyzes andevaluates highly complex service and infrastructure dashboards, taking the leadin addressing identified anomalies.


-Coordinates andmanages long-term implementation strategies for monthly, quarterly, or hotfixpatches to address security vulnerabilities or bugs.


Service Lifecycle Management– System & Security Improvements:


-Analyzes systemperformance data and insights to drive enhancements to improve the performance,reliability, and security of systems and environments.


-Influences andcollaborates with Service teams to proactively identify, address, and predictgaps in operational capabilities, enhancing scalability and resiliency.


Incident Management &Support – Incident Management:


-Leadsend-to-end incident management lifecycle to ensure systems are stable, secure,and performing accurately.


-Evaluatesresults from incident-based data analyses for team metrics and key performanceindicators (KPIs) to identify patterns, root causes, and solutions to preventsystem and network incidents.


-Facilitatesincident review meetings to provide strategic oversight for operationalperformance and long-term solution implementation.


-Influencesthird party vendors and cross-functional teams (e.g., Development, CloudEngineering, Product Engineering, other IT teams) to develop and implementlong-term solutions for high-severity incidents, risks, or migrations.


-Serves as asubject matter expert in the investigation of highly complex system issues andfacilitates high-severity incident triage by designing Corrective andPreventative Action plans (CAPA) to drive incident resolution and prevention.


Incident Management &Support – Escalation Cases:


-Providesexpertise for escalated support cases by collaborating with internal technicalteams and third party vendors to drive issue resolution for a wide range ofproduction environment problems (e.g., immense growth, scaling, leveraging thecloud, extremely high performance, high availability requirements).


Incident Management &Support – Technical Support:


-Drives andevaluates the production environment by analyzing system error logs and ticketqueues, and coordinating with multiple teams involved in maintaining theenvironments.


-Adheres to teamschedule to drive ongoing technical support and service objectives.


-Facilitatesstrategic resolution and long-term solutions for highly complex, criticalcustomer system issues and develops and documents comprehensive technicalsolutions.


Incident Management &Support – Backups and Disaster Recovery:


-Drives theexecution and effectiveness of backup, restore, and disaster recovery processes.


-Influences thestrategic planning and coordination of disaster recovery drills.


-Designsdisaster recovery solutions to ensure preparedness and regulatory compliance.


Communication &Documentation – Technical Communication:


-Communicateshighly complex technical information to both technical and nontechnicalpersonnel including management.


-Develops andimplements training programs to ensure personnel are well-versed indomain-specific knowledge and practices.


-Drivestechnical strategies and solutions for cross-organization projects, programs,and activities by leveraging domain-specific expertise and interpreting highlycomplex technical information.


Communication &Documentation – Documentation & Reporting:


-Drives thecreation of documentation on ticket updates, code contributions,infrastructure, configurations, processes, and procedures (e.g., DisasterRecovery plans, Standard Operating Procedures, Corrective and PreventativeAction Plans).


-Analyzes andinterprets weekly and monthly reports on system performance and incidentprogress to provide insights on operational and management outcomes andbusiness impacts.


-Reviews andrefines technical documentation standards and best practices for internal use.


Additional Responsibilities(as needed)


Cloud InfrastructureSupport:


-Serves as asubject matter expert in collaborations with DevOps and Site ReliabilityEngineer (SRE) teams to manage large-scale infrastructure.


-Managescontinuous integration and continuous deployment (CI/CD) pipelines.


-Outlines andplans for patching and version upgrades to support cloud infrastructure.


Automation:


-Approvesrecommendations and determines plans for improvements to reduce incidents andproblems with automation and simplify server management.


-Owns and leadsthe implementation of reusable frameworks, standards, and automation to supportOracle Cloud Infrastructure.


-Establishes andoversees Workload Automation tools through design support, administration, andoptimization efforts.


Similar Jobs

About Oracle - Egypt
Egypt, Cairo