Ansible Jobs in Egypt
87 Jobs Found
Responsibilities Linux Systems Administration Install, configure, manage, and harden Linux systems (RHEL). Monitor system performance, availability, and log data. Perform patching, upgrades, and capacity planning. Manage users, permissions, storage, and backups. Troubleshoot OS-level issues and performance bottlenecks. Enforce security policies and system hardening standards. Administer Red Hat Identity Manager. Administer Red Hat Satellite Automation & Configuration Management Develop and maintain Ansible playbooks and roles for system provisioning and configuration. Automate repetitive tasks and manage configuration with Ansible Dev Ops Deploy, maintain, and operate Kubernetes clusters. Manage workloads, namespaces, networking, and storage. Administer CI/CD pipelines and Tools<br>Requirements Technical Skills Advanced Linux administration experience. Strong Ansible automation skills. Solid Kubernetes administration knowledge. Bash and/or Python scripting skills. Experience with Docker or Podman. Familiarity with CI/CD tools and pipelines. Experience with monitoring and logging tools. Strong networking fundamentals (TCP/IP, Layer 2/3, etc..). Experience with Git-based workflows. Qualifications2 years in Linux systems administration or Dev Ops roles. Hands-on experience managing production systems. Ability to troubleshoot complex multi-layer issues. Documentation and operational discipline. Experience with RHEL-based environments. Preferred (Bonus Points) RHCSA (Red Hat Certified System Administrator) RHCE (Red Hat Certified Engineer) Soft Skills Strong problem-solving ability. Clear documentation and communication. Ability to work under pressure. Flexibility to work late hours and on weekends
<ul><li><p>Design, implement, and maintain scalable CI/CD pipelines to automate software delivery processes.</p></li><li><p>Collaborate with development and IT teams to streamline deployment, monitoring, and operational workflows.</p></li><li><p>Manage and optimize cloud infrastructure (AWS, Azure, or GCP) for high availability and cost efficiency.</p></li><li><p>Monitor system performance, troubleshoot issues, and proactively resolve bottlenecks or failures.</p></li><li><p>Implement and enforce security best practices across infrastructure and deployment pipelines.</p></li><li><p>Automate configuration management and infrastructure provisioning using tools like Terraform, Ansible, or similar.</p></li><li><p>Develop and maintain scripts for automation of routine tasks and system maintenance.</p></li><li><p>Participate in incident response, root cause analysis, and post-mortem reviews to improve system reliability.</p></li><li><p>Document processes, configurations, and architectural decisions for knowledge sharing and compliance.</p></li><li><p>Stay up-to-date with emerging DevOps tools, trends, and best practices to drive continuous improvement.</p></li></ul><p></p><p><strong>Requirements</strong></p><ul><li><p>Bachelor’s degree in Computer Science, Information Technology, or a related field.</p></li><li><p>3-5 years of professional experience in a DevOps, Infrastructure, or related engineering role.</p></li><li><p>Proven experience with CI/CD tools such as Jenkins, GitLab CI, or CircleCI.</p></li><li><p>Hands-on expertise with cloud platforms like AWS, Azure, or Google Cloud Platform.</p></li><li><p>Strong scripting skills in languages such as Bash, Python, or PowerShell.</p></li><li><p>Experience with configuration management and infrastructure-as-code tools (e.g., Terraform, Ansible, Chef, Puppet).</p></li><li><p>Solid understanding of containerization technologies such as Docker and orchestration tools like Kubernetes.</p></li><li><p>Familiarity with monitoring and logging solutions (e.g., Prometheus, Grafana, ELK Stack).</p></li><li><p>Excellent problem-solving skills and the ability to work independently in a fast-paced environment.</p></li><li><p>Strong communication and collaboration skills, with a proactive approach to cross-functional teamwork.</p></li></ul><p></p>
<section><p class="heading jdMain">Job Description</p><p class="heading">Roles & Responsibilities</p><div class="paragraph"><p>At ServerHub, we power businesses with high-performance cloud and hosting solutions. Our mission is to provide customers with reliable, scalable, and secure infrastructure worldwide. As an L3 Linux System Engineer, you will be at the forefront of our operations, ensuring our hosting platforms are optimized, secure, and always online.</p><p><strong>Job Responsibilities:</strong></p><ul><li><strong>Escalation & Troubleshooting:</strong> Act as the final escalation point (L3) for complex server, hosting, and network-related issues. Diagnose and resolve critical system failures, network outages, and performance issues.</li><li><strong>Linux Server Administration:</strong> Manage, optimize, and secure Linux-based hosting environments (CentOS, Ubuntu, RHEL). Administer and fine-tune web servers (Apache, Nginx, LiteSpeed), databases (MySQL, PostgreSQL), and caching layers (Redis, Memcached). Build and deploy servers from the ground up, ensuring optimal configurations for performance and security.</li><li><strong>Automation & DevOps:</strong> Develop and maintain automation scripts (Bash, Python, Perl) for server provisioning and configuration management. Utilize Ansible, Terraform, or similar tools for automating infrastructure deployments.</li><li><strong>Cloud & Virtualization:</strong> Deploy and manage KVM, OpenStack, VMware, or containerized environments (Docker, Kubernetes). Support cloud-based hosting solutions (AWS, Google Cloud, Azure).</li><li><strong>Security & Compliance:</strong> Implement security best practices, including firewall rules, SELinux/AppArmor, IDS/IPS. Perform vulnerability assessments and patch management to secure customer environments.</li><li><strong>Monitoring & Incident Response:</strong> Set up and manage monitoring tools like Prometheus, Grafana, Zabbix, Nagios. Participate in 24/7 on-call rotation for urgent system issues.</li><li><strong>Collaboration & Mentoring:</strong> Work closely with NOC, DevOps, and Engineering teams to ensure smooth operations. Provide guidance and mentorship to L1 & L2 support engineers.</li></ul><p><strong>Requirements:</strong></p><ul><li>11+ years of hands-on Linux system administration experience in a web hosting or cloud environment.</li><li>7+ years of experience in a Senior position or Level 3 Engineer role.</li><li>MUST be a KVM and Livbirt expert with a minimum of 7+ yrs of experience.</li><li>Expert knowledge of web hosting technologies (cPanel, Plesk, WHM, LAMP/LEMP stacks).</li><li>Strong scripting ability in Perl, Bash, and Python.</li><li>Experience with configuration management tools (Ansible, Puppet, Chef).</li><li>Networking expertise understanding of TCP/IP, DNS, VPN, Firewalls, Load Balancing.</li><li>Familiarity with RAID, SAN, NAS, and distributed storage systems.</li><li>Experience working in a 24/7 production environment with on-call duties.</li><li>Ability to build and configure servers from the ground up.</li><li>Certifications such as RHCE, AWS Certified SysOps Administrator, or Kubernetes (CKA) are a plus.</li><li>Good English Communication Skills</li></ul><p><strong>What ServerHub Offers:</strong></p><ul><li>A fast-paced, innovative environment in a growing cloud hosting company.</li><li>Cutting-edge technologies and challenging projects.</li><li>Career growth opportunities and professional development.</li><li>Paid leaves and fully remote set-up</li></ul><p>Join ServerHub and be part of a team that keeps the internet running! Apply today!</p></div></section><section><p class="heading">Desired Candidate Profile</p><p class="paragraph"></p><ul><li>11+ years of hands-on Linux system administration experience in a web hosting or cloud environment.</li><li>7+ years of experience in a Senior position or Level 3 Engineer role.</li><li>MUST be a KVM and Livbirt expert with a minimum of 7+ yrs of experience.</li><li>Expert knowledge of web hosting technologies (cPanel, Plesk, WHM, LAMP/LEMP stacks).</li><li>Strong scripting ability in Perl, Bash, and Python.</li><li>Experience with configuration management tools (Ansible, Puppet, Chef).</li><li>Networking expertise understanding of TCP/IP, DNS, VPN, Firewalls, Load Balancing.</li><li>Familiarity with RAID, SAN, NAS, and distributed storage systems.</li><li>Experience working in a 24/7 production environment with on-call duties.</li><li>Ability to build and configure servers from the ground up.</li><li>Certifications such as RHCE, AWS Certified SysOps Administrator, or Kubernetes (CKA) are a plus.</li><li>Good English Communication Skills</li></ul><p></p></section>
A Linux & Dev Ops Engineer is responsible for managing Linux-based infrastructure, automating system administration tasks, and building reliable deployment pipelines to support software development and operations. The role focuses on maintaining secure, scalable, and highly available environments while improving operational efficiency through automation and Dev Ops best practices. Key Responsibilities:Install, configure, and maintain Linux servers (Ubuntu, Cent OS, RHEL). Monitor system performance, availability, and security. Manage users, permissions, storage, and networking. Automate routine administration tasks using Bash or Python scripting. Configure and maintain web servers such as Nginx and Apache. Implement and manage CI/CD pipelines using Jenkins, Git Hub Actions, or Git Lab CI. Build, deploy, and manage containerized applications using Docker. Orchestrate containerized workloads using Kubernetes. Provision and manage infrastructure using Infrastructure as Code (Terraform or Ansible). Deploy and manage cloud infrastructure on AWS, Azure, or Google Cloud Platform. Configure monitoring and logging solutions such as Prometheus, Grafana, and ELK Stack. Troubleshoot infrastructure, deployment, and application issues. Implement backup, disaster recovery, and security best practices. Collaborate with development and operations teams to ensure reliable software delivery. Required Skills:Strong knowledge of Linux administration. Good understanding of networking fundamentals (TCP/IP, DNS, DHCP, HTTP, SSH). Experience with Git and version control. Hands-on experience with Docker and Kubernetes. Knowledge of CI/CD tools. Experience with C Panel & CWPExperience with cloud platforms (AWS, Azure, or GCP). Scripting experience (Bash and/or Python). Familiarity with Infrastructure as Code tools (Terraform, Ansible). Strong troubleshooting and problem-solving skills. Preferred Qualifications:Bachelor's degree in Computer Science, Information Technology, or a related field. Relevant certifications such as RHCSA, RHCE, AWS Certified, Kubernetes (CKA), or Terraform Associate are a plus.
????Freelancing opportunity – Wintel System Administrator – 2‑year contract???? - Remote - EMEA<br>Duties and responsibilities:<br>Infrastructure management: administer and maintain Windows Server Infrastructure across multi-national environments & manage Active Directory, Group Policy Objects (GPOs), DNS, DHCP and Certificate Services Troubleshooting & Support: investigate and resolve complex issues related to server performance, security and services in Windows environments & provide L3 support for escalated incidents Documentation: create and maintain detailed runbooks, operational procedures and technical documentation Virtualization: Support VMware v Sphere / ESXi environments including v Center administration (clusters, HA, DRS, storage policies, templates, snapshots) Automation & Scripting: Develop and maintain automation scripts using Powershell, Bash and Python to streamline system and virtualization tasks & utilize tools like Ansible for configuration management and task automation across environments<br>Requirements:<br>Minimum of 5 years of hands-on experience managing Windows Server infrastructure in enterprise environments Good understanding of English – spoken and written;Proven ability in documenting operational procedures, creating runbooks, and maintaining internal knowledge bases Solid background in Windows Server administration, including configuration and maintenance Strong troubleshooting skills in enterprise Windows environments, covering performance, security, and service-related issues Advanced knowledge of Active Directory (AD), Group Policy Objects (GPOs), DNS, DHCP, and Certificate Services Basic proficiency in VMware v Sphere/ESXi administration, configuration, and troubleshooting, including VMware v Center (clusters, HA, DRS, storage policies, templates, snapshots) Familiarity with Nutanix Prism Central / Element and AHV hypervisor (including N2C) Experience using Power Shell, Bash, and Python for automation of virtualization tasks (e.g., via Ansible) Working knowledge of backup tools for virtual environments, such as Commvault and Networker
▸ Design, deploy, and manage VMware environments including v Sphere, NSX-T, v SAN, and ESXi for enterprise clients<br><br>▸ Build and maintain infrastructure automation pipelines using Terraform and Ansible across public, private, and hybrid cloud environments<br><br>▸ Deploy and optimize Kubernetes clusters for containerized workloads, supporting CI/CD integration and day-to-day operations<br><br>▸ Architect and implement Google Cloud Platform (GCP) solutions, including compute, networking, storage, and identity management<br><br>▸ Develop and maintain IaC templates for consistent, repeatable deployments across staging and production environments<br><br>▸ Integrate VMware infrastructure with CI/CD pipelines to streamline Dev Ops workflows and reduce manual operations<br><br>▸ Provide L2/L3 support for complex issues across virtualization, networking, and hybrid cloud environments<br><br>▸ Lead incident resolution, root cause analysis, and problem management for critical infrastructure issues<br><br>▸ Collaborate with cross-functional teams including Dev Ops, security, and project management to ensure delivery excellence<br><br>▸ Produce and maintain technical documentation, architecture diagrams, and operational runbooks<br><br>▸ Support data center migration and infrastructure resiliency planning initiatives<br><br>Requirements<br><br>▸ 7+ years of hands-on experience in cloud infrastructure, virtualization, and enterprise IT environments<br><br>▸ Strong proficiency in VMware stack: v Sphere, NSX-T, v SAN, and ESXi administration and troubleshooting<br><br>▸ Proven experience with Terraform for infrastructure as code and Ansible for configuration management<br><br>▸ Hands-on experience deploying and managing Kubernetes clusters and containerized workloads<br><br>▸ Working knowledge of Google Cloud Platform or equivalent public cloud environment (AWS, Azure)<br><br>▸ Experience with CI/CD pipeline design and integration with infrastructure workflows<br><br>▸ Strong incident management skills with experience in L2/L3 support and shift leadership<br><br>▸ Excellent communication skills; able to collaborate with both technical teams and client stakeholders<br><br>▸ Bachelor's degree in Communications Engineering, Computer Science, or a related field
<section><p class="heading jdMain">Job Description</p><p class="heading">Roles & Responsibilities</p><div class="paragraph"><p>At ServerHub, we power businesses with high-performance cloud and hosting solutions. Our mission is to provide customers with reliable, scalable, and secure infrastructure worldwide. As an L3 Linux System Engineer, you will be at the forefront of our operations, ensuring our hosting platforms are optimized, secure, and always online.</p><p><strong>Job Responsibilities:</strong></p><ul><li><strong>Escalation & Troubleshooting:</strong> Act as the final escalation point (L3) for complex server, hosting, and network-related issues. Diagnose and resolve critical system failures, network outages, and performance issues.</li><li><strong>Linux Server Administration:</strong> Manage, optimize, and secure Linux-based hosting environments (CentOS, Ubuntu, RHEL). Administer and fine-tune web servers (Apache, Nginx, LiteSpeed), databases (MySQL, PostgreSQL), and caching layers (Redis, Memcached). Build and deploy servers from the ground up, ensuring optimal configurations for performance and security.</li><li><strong>Automation & DevOps:</strong> Develop and maintain automation scripts (Bash, Python, Perl) for server provisioning and configuration management. Utilize Ansible, Terraform, or similar tools for automating infrastructure deployments.</li><li><strong>Cloud & Virtualization:</strong> Deploy and manage KVM, OpenStack, VMware, or containerized environments (Docker, Kubernetes). Support cloud-based hosting solutions (AWS, Google Cloud, Azure).</li><li><strong>Security & Compliance:</strong> Implement security best practices, including firewall rules, SELinux/AppArmor, IDS/IPS. Perform vulnerability assessments and patch management to secure customer environments.</li><li><strong>Monitoring & Incident Response:</strong> Set up and manage monitoring tools like Prometheus, Grafana, Zabbix, Nagios. Participate in 24/7 on-call rotation for urgent system issues.</li><li><strong>Collaboration & Mentoring:</strong> Work closely with NOC, DevOps, and Engineering teams to ensure smooth operations. Provide guidance and mentorship to L1 & L2 support engineers.</li></ul><p><strong>Requirements:</strong></p><ul><li>11+ years of hands-on Linux system administration experience in a web hosting or cloud environment.</li><li>7+ years of experience in a Senior position or Level 3 Engineer role.</li><li>MUST be a KVM and Livbirt expert with a minimum of 7+ yrs of experience.</li><li>Expert knowledge of web hosting technologies (cPanel, Plesk, WHM, LAMP/LEMP stacks).</li><li>Strong scripting ability in Perl, Bash, and Python.</li><li>Experience with configuration management tools (Ansible, Puppet, Chef).</li><li>Networking expertise understanding of TCP/IP, DNS, VPN, Firewalls, Load Balancing.</li><li>Familiarity with RAID, SAN, NAS, and distributed storage systems.</li><li>Experience working in a 24/7 production environment with on-call duties.</li><li>Ability to build and configure servers from the ground up.</li><li>Certifications such as RHCE, AWS Certified SysOps Administrator, or Kubernetes (CKA) are a plus.</li><li>Good English Communication Skills</li></ul><p><strong>What ServerHub Offers:</strong></p><ul><li>A fast-paced, innovative environment in a growing cloud hosting company.</li><li>Cutting-edge technologies and challenging projects.</li><li>Career growth opportunities and professional development.</li><li>Paid leaves and fully remote set-up</li></ul><p>Join ServerHub and be part of a team that keeps the internet running! Apply today!</p></div></section><section><p class="heading">Desired Candidate Profile</p><p class="paragraph"></p><p><strong>Requirements:</strong></p><ul><li>11+ years of hands-on Linux system administration experience in a web hosting or cloud environment.</li><li>7+ years of experience in a Senior position or Level 3 Engineer role.</li><li>MUST be a KVM and Livbirt expert with a minimum of 7+ yrs of experience.</li><li>Expert knowledge of web hosting technologies (cPanel, Plesk, WHM, LAMP/LEMP stacks).</li><li>Strong scripting ability in Perl, Bash, and Python.</li><li>Experience with configuration management tools (Ansible, Puppet, Chef).</li><li>Networking expertise understanding of TCP/IP, DNS, VPN, Firewalls, Load Balancing.</li><li>Familiarity with RAID, SAN, NAS, and distributed storage systems.</li><li>Experience working in a 24/7 production environment with on-call duties.</li><li>Ability to build and configure servers from the ground up.</li><li>Certifications such as RHCE, AWS Certified SysOps Administrator, or Kubernetes (CKA) are a plus.</li><li>Good English Communication Skills</li></ul><p></p></section>
<section><p class="heading jdMain">Job Description</p><p class="heading">Roles & Responsibilities</p><div class="paragraph"><p>About ServerHub: At ServerHub, we power businesses with high-performance cloud and hosting solutions. Our mission is to provide customers with reliable, scalable, and secure infrastructure worldwide. As an L3 Linux System Engineer, you will be at the forefront of our operations, ensuring our hosting platforms are optimized, secure, and always online.</p><p>Job Responsibilities:</p><ul><li>Escalation & Troubleshooting: Act as the final escalation point (L3) for complex server, hosting, and network-related issues. Diagnose and resolve critical system failures, network outages, and performance issues.</li><li>Linux Server Administration: Manage, optimize, and secure Linux-based hosting environments (CentOS, Ubuntu, RHEL). Administer and fine-tune web servers (Apache, Nginx, LiteSpeed), databases (MySQL, PostgreSQL), and caching layers (Redis, Memcached). Build and deploy servers from the ground up, ensuring optimal configurations for performance and security.</li><li>Automation & DevOps: Develop and maintain automation scripts (Bash, Python, Perl) for server provisioning and configuration management. Utilize Ansible, Terraform, or similar tools for automating infrastructure deployments.</li><li>Cloud & Virtualization: Deploy and manage KVM, OpenStack, VMware, or containerized environments (Docker, Kubernetes). Support cloud-based hosting solutions (AWS, Google Cloud, Azure).</li><li>Security & Compliance: Implement security best practices, including firewall rules, SELinux/AppArmor, IDS/IPS. Perform vulnerability assessments and patch management to secure customer environments.</li><li>Monitoring & Incident Response: Set up and manage monitoring tools like Prometheus, Grafana, Zabbix, Nagios. Participate in 24/7 on-call rotation for urgent system issues.</li><li>Collaboration & Mentoring: Work closely with NOC, DevOps, and Engineering teams to ensure smooth operations. Provide guidance and mentorship to L1 & L2 support engineers.</li></ul></div></section><section><p class="heading">Desired Candidate Profile</p><p class="paragraph"></p><ul><li>11+ years of hands-on Linux system administration experience in a web hosting or cloud environment.</li><li>7+ years of experience in a Senior position or Level 3 Engineer role.</li><li>MUST be a KVM and Livbirt expert with a minimum of 7+ yrs of experience.</li><li>Expert knowledge of web hosting technologies (cPanel, Plesk, WHM, LAMP/LEMP stacks).</li><li>Strong scripting ability in Perl, Bash, and Python.</li><li>Experience with configuration management tools (Ansible, Puppet, Chef).</li><li>Networking expertise understanding of TCP/IP, DNS, VPN, Firewalls, Load Balancing.</li><li>Familiarity with RAID, SAN, NAS, and distributed storage systems.</li><li>Experience working in a 24/7 production environment with on-call duties.</li><li>Ability to build and configure servers from the ground up.</li><li>Certifications such as RHCE, AWS Certified SysOps Administrator, or Kubernetes (CKA) are a plus.</li><li>Good English Communication Skills</li></ul><p></p></section>
Role Purpose<br><br>The Senior Infrastructure & Multi-Cloud Automation / IaC Engineer will lead infrastructure automation across private cloud, OCI, GCP, SIT, virtualization, and systems operations.<br><br>The role focuses on standardizing infrastructure provisioning, improving cloud asset inventory, eliminating orphaned resources, reducing manual operational effort, and enabling scalable multi-cloud operations through Infrastructure-as-Code and automation.<br><br>This is not primarily an application Dev Ops role. Dev Ops knowledge is useful, but the main focus is infrastructure automation, operational control, cloud hygiene, lifecycle governance, and managed operations efficiency.<br><br>Key Responsibilities<br><br>Infrastructure & Multi-Cloud Automation<br><br> Design and implement Infrastructure-as-Code templates for OCI, GCP, SIT, private cloud, VMware/VCF, and systems environments Automate infrastructure provisioning across cloud, virtualization, and systems platforms Build reusable automation templates for compute, network, IAM, storage integration, monitoring onboarding, and backup onboarding Standardize automation patterns using Terraform, Ansible, scripts, and platform-native tools Reduce manual provisioning and repetitive operational tasks across infrastructure towers Support automation for VM lifecycle, cloud resource provisioning, patching workflows, and operational handovers Support limited CI/CD or pipeline usage only where required for infrastructure automation workflows<br><br>Cloud Asset Management & Inventory<br><br> Own cloud infrastructure asset management automation across OCI, GCP, SIT, and private cloud environments Maintain automated cloud inventory and resource discovery workflows Identify orphaned, unused, unattached, or underutilized cloud resources Automate reporting and controlled cleanup workflows for orphaned objects, subject to approval and change control Implement tagging standards for ownership, environment, cost center, service, criticality, lifecycle status, and business owner Support cloud resource lifecycle management from provisioning to decommissioning Integrate cloud asset data with CMDB / Service Now where applicable Generate regular cloud inventory, utilization, hygiene, and optimization reports Support Fin Ops reporting by identifying waste, idle resources, and optimization opportunities Ensure automation templates enforce standard naming, tagging, and governance policies<br><br>Governance, Documentation & Operational Integration<br><br> Define automation standards, reusable templates, coding practices, and operational runbooks Support integration between automation workflows, ITSM, monitoring, and operational processes Work with Storage, Backup, Virtualization, Cloud, and Systems towers to identify automation opportunities Review and approve automation code, scripts, and templates Guide junior automation engineers and tower engineers on automation practices Ensure automation activities follow change management and approval processes Maintain version-controlled repositories for infrastructure automation assets<br><br><br>Requirements<br><br> 7+ years of experience in infrastructure, cloud, virtualization, systems, or automation roles Strong hands-on experience with Terraform and Ansible Experience automating infrastructure across at least two of the following: OCI, GCP, VMware/VCF, private cloud, Windows, Linux Strong understanding of infrastructure operations, provisioning, patching, monitoring, lifecycle management, and operational governance Experience with cloud landing zones, IAM, networking, compute, storage, and monitoring automation Experience in cloud asset inventory, tagging governance, and resource lifecycle management Experience identifying orphaned cloud resources such as unused disks, unattached IPs, idle VMs, orphan snapshots, unused buckets, stale accounts/projects, and unused network objects Good scripting skills using Python, Power Shell, Bash, or similar Familiarity with Git and version control for automation templates Understanding of ITSM, Service Now, change management, and operational approval workflows Ability to design reusable automation frameworks, not only standalone scripts<br><br><br>Preferred Experience<br><br> Multi-cloud automation experience VMware / VCF automation exposure Open Shift or Kubernetes infrastructure automation exposure Service Now workflow or CMDB integration experience Basic Fin Ops awareness Basic CI/CD awareness for managing infrastructure automation pipelines, not application delivery<br><br><br>Preferred Certifications<br><br> Hashi Corp Terraform Associate Red Hat Ansible certification Google Cloud certification Oracle Cloud Infrastructure certification VMware certification Red Hat / Open Shift certification is a plus ITIL Foundation
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Role Purpose The Senior Infrastructure & Multi-Cloud Automation / IaC Engineer will lead infrastructure automation across private cloud, OCI, GCP, SIT, virtualization, and systems operations.<br> The role focuses on standardizing infrastructure provisioning, improving cloud asset inventory, eliminating orphaned resources, reducing manual operational effort, and enabling scalable multi-cloud operations through Infrastructure-as-Code and automation.<br> This is not primarily an application DevOps role.<br> DevOps knowledge is useful, but the main focus is infrastructure automation, operational control, cloud hygiene, lifecycle governance, and managed operations efficiency.<br> Key Responsibilities Infrastructure & Multi-Cloud Automation • Design and implement Infrastructure-as-Code templates for OCI, GCP, SIT, private cloud, VMware/VCF, and systems environments.<br> • Automate infrastructure provisioning across cloud, virtualization, and systems platforms.<br> • Build reusable automation templates for compute, network, IAM, storage integration, monitoring onboarding, and backup onboarding.<br> • Standardize automation patterns using Terraform, Ansible, scripts, and platform-native tools.<br> • Reduce manual provisioning and repetitive operational tasks across infrastructure towers.<br> • Support automation for VM lifecycle, cloud resource provisioning, patching workflows, and operational handovers.<br> • Support limited CI/CD or pipeline usage only where required for infrastructure automation workflows.<br> Cloud Asset Management & Inventory • Own cloud infrastructure asset management automation across OCI, GCP, SIT, and private cloud environments.<br> • Maintain automated cloud inventory and resource discovery workflows.<br> • Identify orphaned, unused, unattached, or underutilized cloud resources.<br> • Automate reporting and controlled cleanup workflows for orphaned objects, subject to approval and change control.<br> • Implement tagging standards for ownership, environment, cost center, service, criticality, lifecycle status, and business owner.<br> • Support cloud resource lifecycle management from provisioning to decommissioning.<br> • Integrate cloud asset data with CMDB / ServiceNow where applicable.<br> • Generate regular cloud inventory, utilization, hygiene, and optimization reports.<br> • Support FinOps reporting by identifying waste, idle resources, and optimization opportunities.<br> • Ensure automation templates enforce standard naming, tagging, and governance policies.<br> Governance, Documentation & Operational Integration • Define automation standards, reusable templates, coding practices, and operational runbooks.<br> • Support integration between automation workflows, ITSM, monitoring, and operational processes.<br> • Work with Storage, Backup, Virtualization, Cloud, and Systems towers to identify automation opportunities.<br> • Review and approve automation code, scripts, and templates.<br> • Guide junior automation engineers and tower engineers on automation practices.<br> • Ensure automation activities follow change management and approval processes.<br> • Maintain version-controlled repositories for infrastructure automation assets.<br> • 7+ years of experience in infrastructure, cloud, virtualization, systems, or automation roles.<br> • Strong hands-on experience with Terraform and Ansible.<br> • Experience automating infrastructure across at least two of the following: OCI, GCP, VMware/VCF, private cloud, Windows, Linux.<br> • Strong understanding of infrastructure operations, provisioning, patching, monitoring, lifecycle management, and operational governance.<br> • Experience with cloud landing zones, IAM, networking, compute, storage, and monitoring automation.<br> • Experience in cloud asset inventory, tagging governance, and resource lifecycle management.<br> • Experience identifying orphaned cloud resources such as unused disks, unattached IPs, idle VMs, orphan snapshots, unused buckets, stale accounts/projects, and unused network objects.<br> • Good scripting skills using Python, PowerShell, Bash, or similar.<br> • Familiarity with Git and version control for automation templates.<br> • Understanding of ITSM, ServiceNow, change management, and operational approval workflows.<br> • Ability to design reusable automation frameworks, not only standalone scripts.<br> Preferred Experience • Multi-cloud automation experience.<br> • VMware / VCF automation exposure.<br> • OpenShift or Kubernetes infrastructure automation exposure.<br> • ServiceNow workflow or CMDB integration experience.<br> • Basic FinOps awareness.<br> • Basic CI/CD awareness for managing infrastructure automation pipelines, not application delivery.<br> Preferred Certifications • HashiCorp Terraform Associate.<br> • Red Hat Ansible certification.<br> • Google Cloud certification.<br> • Oracle Cloud Infrastructure certification.<br> • VMware certification.<br> • Red Hat / OpenShift certification is a plus.<br> • ITIL Foundation.<br></span> </div>
Are you a passionate Dev Ops Engineer with a solid foundation in automation, CI/CD, and cloud infrastructure? We are looking for a **Dev Ops Engineer** with around **2 years of experience** to join our technical team and help optimize, automate, and secure our deployment and delivery processes. --- ## ???? Role Overview * **Position:** Dev Ops Engineer * **Experience Level:** 2+ Years * **Employment Type:** Full-time * **Location:** Hybrid --- ## ???? Key Responsibilities * **CI/CD Pipelines:** Build, maintain, and optimize continuous integration and deployment pipelines using tools like Git Hub Actions, Git Lab CI, or Jenkins. * **Infrastructure Management:** Provision and manage cloud infrastructure using Infrastructure as Code (IaC) tools like Terraform or Ansible. * **Containerization & Orchestration:** Package applications using Docker and manage deployments using Kubernetes or Docker Swarm. * **System Monitoring & Alerting:** Monitor system health, performance, and application metrics using tools like Prometheus, Grafana, or ELK Stack. * **Collaboration & Support:** Work closely with software developers and QA teams to automate deployments, streamline releases, and resolve build/deployment issues. * **Security & Best Practices:** Assist in implementing security best practices across CI/CD pipelines and cloud environments (Dev Sec Ops). * **Operating systems Management:** Can Analyze the issues with operating systems (Windows, Linux) * **on-Premise environment:** Can deals with on-Premise environment --- ## ????️ Key Qualifications & Requirements * **Experience:** 2+ years of hands-on experience in a Dev Ops, Sys Admin, or Cloud Engineering role. * **Cloud Platforms:** Experience with at least one major cloud provider (**AWS**, **Azure**, or **GCP**). * **Containerization:** Strong knowledge of **Docker** and exposure to **Kubernetes**. * **CI/CD Tools:** Hands-on experience with **Git Lab CI**, **Git Hub Actions**, **Jenkins**, or Azure Dev Ops. * **Scripting Skills:** Proficiency in scripting languages such as **Bash**, **Python**, or **Power Shell**. * **IaC & Config Management:** Basic to intermediate experience with **Terraform** or **Ansible**. * **Operating Systems:** Solid understanding of Linux administration and networking fundamentals. --- ## ???? Nice-to-Have (Bonus) * Knowledge of web servers (Nginx, Apache) and SSL management. * Experience with database administration or basic database performance tuning. * Relevant certifications (e.g., AWS Certified Cloud Practitioner / Solutions Architect Associate, CKA).
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Design and Architecture: Create scalable and secure cloud architectures that meet business needs, translating requirements into technical specifications.<br> Implementation and Management: Deploy and manage cloud infrastructure using GCP services, ensuring high availability and performance.<br> Automation: Utilize tools like Terraform and Ansible to automate deployment processes and manage cloud resources efficiently.<br> Collaboration: Work closely with development, operations, and security teams to deliver integrated cloud solutions and provide technical support.<br> Monitoring and Optimization: Monitor cloud performance, identify bottlenecks, and implement strategies for cost optimization and resource management.<br> Security Compliance: Ensure that cloud solutions adhere to security best practices and industry standards, implementing necessary security measures.<br> 5+ Years of experience.<br></span> </div>
Role Purpose:<br><br>The Cloud Dev Sec Ops Engineer is responsible for designing, implementing, and securing cloud-based infrastructure and CI/CD pipelines. The role ensures seamless automation, robust security integration, and efficient monitoring of systems across OCI and GCP environments to enable secure, scalable, and high-performing applications.<br><br>Key Responsibilities:<br><br> CI/CD Pipeline & Dev Sec Ops Design, build, and secure CI/CD pipelines using Git Hub Actions Integrate automated security testing (SAST/DAST) using Sonar Qube and related tools Implement and manage Git Ops workflows for continuous delivery using Argo CD Cloud Infrastructure & Automation Design, build, and secure cloud infrastructure on OCI and GCP using Terraform Automate configuration and security hardening tasks using Ansible Develop and maintain automation scripts and perform system administration using Bash Secure and manage container orchestration platforms including Oracle Kubernetes Engine (OKE) and Google Kubernetes Engine (GKE) QA Automation Enablement Collaborate with the QA team to integrate automated test suites into CI/CD pipelines Enable automated execution of QA tests to act as quality gates before deployment Observability & Monitoring Build and manage the enterprise observability and monitoring stack using Splunk, Grafana, and Prometheus Develop dashboards, alerts, and incident response playbooks for performance and security monitoring<br><br>Key Interactions:<br><br>Internal: Infrastructure, Developers, Security Teams External: Technology vendors and solution partners (as needed)<br><br>Requirements<br><br>Education<br><br>Bachelor's degree in Computer Engineering, Information Technology, or a related field (preferred)<br><br>Experience<br><br>3+ years of hands-on experience in Dev Ops or Cloud Engineering Proven expertise in Dev Sec Ops principles and Software Development Lifecycle (SDLC) Strong proficiency in Bash scripting and automation with Ansible Deep understanding of cloud security architecture and native security services on OCI and GCPExtensive experience with Infrastructure as Code (Terraform) Advanced skills in CI/CD design using Git Hub Actions and Git Ops tools like Argo CDHands-on experience in Kubernetes (OKE/GKE) setup and security Practical experience with security scanning tools (Sonar Qube, SAST, DAST) Strong knowledge of observability platforms (Splunk, Grafana, Prometheus) Familiarity with automated functional testing frameworks and integration into Dev Ops pipelines<br><br>Core Competencies:<br><br>Strong analytical and problem-solving skills Attention to detail and focus on security best practices Ability to collaborate effectively across cross-functional teams Excellent communication and documentation skills Continuous learning mindset and adaptability to new technologies
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span> What you´ll do </span><ul> <li>Design, build, and optimize cloud infrastructure on Microsoft Azure and Google Cloud Platform, including compute, networking, IAM, and storage services.</li> <li>Develop and maintain infrastructure automation solutions using Terraform and Ansible following Infrastructure as Code (IaC) practices.</li> <li>Build and maintain CI/CD deployment pipelines supporting reliable multi-environment platform delivery and GitOps workflows.</li> <li>Support the provisioning and lifecycle management of cloud projects and resources such as VPCs, service accounts, landing zones, and platform blueprints.</li> <li>Contribute to the team's Internal Developer Platform (IDP) initiative by engineering self-service capabilities and reusable platform blueprints for application teams.</li> <li>Support API and agent gateway design, including model routing and FinOps tracking for LLM token consumption.</li> <li>Support platform security, identity management, and compliance capabilities across cloud-native environments.</li> <li>Collaborate with global and cross-functional teams including Platform Owners globally to deliver scalable and reliable enterprise platform solutions.</li> <li>Participate in disaster recovery planning and execution. • Optimize cloud-based solutions and infrastructure as part of the company's digital strategy.</li> <li>Contribute to the development of best practices and documentation for IT operations.</li> </ul>
<br><br>
What makes you a good fit <ul> <li>2–3 years of hands-on experience in cloud engineering or DevOps roles</li> <li>Solid working knowledge of GCP or Azure - compute, networking, IAM, and storage fundamentals; experience with both is a plus.</li> <li>Hands on experince with infrastructure-as-code practices, preferably Terraform; Ansible is an advantage.</li> <li>Experience building and maintaining CI/CD pipelines and multi-environment deployment workflows.</li> <li>Familiarity with Linux environments, shell scripting (Bash), and container-based workloads.</li> <li>Understanding of Kubernetes or OpenShift and GitOps practices is an advantage.</li> <li>Basic familiarity with AI/LLM APIs such as Azure OpenAI or Vertex AI; hands-on experience is a strong plus.</li> <li>Experience with API gateway solutions (Azure APIM, Apigee, or similar) is a plus.</li> <li>Scripting experience with Python is considered a plus.</li> <li>Strong analytical, problem-solving, and communication skills.</li> <li>Ability to work independently and collaborate effectively across distributed, cross-functional teams.</li> <li>Proactive mindset with a passion for automation, engineering quality, and continuously learning new technologies.</li> </ul> Some perks of joining Henkel <ul> <li>Flexible work scheme with flexible hours, hybrid work model, and work from anywhere policy for up to 30 days per year</li> <li>Diverse national and international growth opportunities</li> <li>Global wellbeing standards with health and preventive care programs</li> <li>Gender-neutral parental leave for a minimum of 8 weeks</li> <li>Employee Share Plan with voluntary investment and Henkel matching shares</li> <li>Comprehensive Health Insurance for employee + dependents</li> <li>Employee Assistance Programme provides a wide range of mental health and wellbeing benefits</li> </ul> <p>At Henkel, we come from a broad range of backgrounds, perspectives, and life experiences. We believe the uniqueness of all our employees is the power in us. Become part of the team and bring your uniqueness to us! We look for a diverse team of individuals who possess different backgrounds, experiences, personalities and mindsets.</p><br>
<br> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>Adree is seeking a DevSecOps, Architect to support our product development by leveraging technical and analytical expertise to articulate the value of secure, scalable digital solutions.<br> In this role, you will work closely with stakeholders to understand business goals and technical requirements, and clearly demonstrate how target-state DevSecOps architecture and golden path standards can address their needs.<br> You will be responsible for bridging the gap between secure supply chain infrastructure and seamless delivery execution, collaborating across teams to craft exceptional enterprise platforms.<br> By blending infrastructure-as-code, GitOps workflows, automated security gates, and container orchestration, you will ensure our digital products are both highly resilient and fully auditable.<br> Key Roles and Responsibilities Engage with clients and stakeholders to gather requirements and understand their digital product goals.<br> Deliver impactful presentations, reference architectures, and documentation showcasing CI/CD pipelines, artifact lifecycles, and environment strategies.<br> Support the product and engineering teams in developing secure, high-availability Kubernetes and OpenShift clusters across private cloud environments.<br> Provide technical insights and DevSecOps expertise throughout the product lifecycle to drive pipeline patterns, Fortify gates, SBOM/SCA expectations, and Sigstore signing.<br> Collaborate with cross-functional teams (including security, testing, and product managers) to foster a DevOps culture and capture feedback to improve infrastructure offerings.<br> Stay current with industry trends, market changes, IaC standards using Terraform, and configuration standards via Ansible to position solutions effectively.<br> Conduct workshops, architectural reviews, and technical discussions internally and with clients to lead adoption waves and deprecate non-compliant delivery paths.<br> Participate in Agile development processes and team alignments, providing technical mentorship to engineers and ensuring smooth integrations with third parties.<br> Education Bachelor's degree in Computer Science, Software Engineering, Information Technology, or a related field.<br> Experience 10–15+ years of overall IT experience, with at least 5+ years dedicated to platform engineering or DevSecOps architecture.<br> Proven track record designing enterprise GitOps models using Argo CD into OpenShift or Kubernetes, including multi-environment promotion.<br> Demonstrated experience working within government, heavily regulated, or complex enterprise sectors is highly preferred.<br> Skills & Competencies (Technical & Analytical + Soft) Deep proficiency in enterprise DevSecOps platform design, secure supply chain management, and policy-as-code concepts.<br> Strong hands-on architectural knowledge of Azure DevOps Server, JFrog Artifactory, HashiCorp Vault, Terraform, and Ansible.<br> Expert command over automated testing integrations (e.<br>g., OpenText Service Virtualization, UFT One, UFT Digital Lab, LoadRunner).<br> Analytical skills to design monitoring and operational readiness architectures aligned with AppDynamics, BMC, and Azure Monitoring.<br> Strategic thinking, adoption leadership, and outstanding documentation skills to drive stakeholder alignment and pragmatic architectural trade-offs.<br> Experience (summary) Target-state DevSecOps reference architecture and golden path design.<br> Multi-environment GitOps deployment modeling using Argo CD into K8s/OpenShift.<br> End-to-end secure software supply chain implementation (Fortify, Sigstore, Vault).<br> Automated testing gate integration and virtualization orchestration.<br> Infrastructure-as-Code (IaC) and configuration management standardization.<br> Enterprise observability and monitoring architecture deployment.<br> Skills & Competencies (summary) Azure DevOps Server & JFrog Artifactory ecosystem.<br> Container orchestration (Kubernetes & Red Hat OpenShift architecture).<br> Infrastructure-as-Code (Terraform & Ansible automation).<br> Application security automation (SCA, SBOM, code signing).<br> Enterprise monitoring (AppDynamics / BMC / Azure Monitoring).<br> Stakeholder relationship management & governance alignment.<br> Prioritization and adoption leadership.<br> Travel for client-facing activities.<br> Job Location: HQ</span> </div>
<h2 class="h5">Job description</h2>
<div class="t-break" data-jb-field="description">
<span>The Cloud Engineer – OCI is responsible for providing advanced technical support and troubleshooting for Oracle Cloud Infrastructure (OCI) environments.<br> The role acts as a key escalation point for complex technical issues raised by Level 1 support engineers, ensuring timely resolution while maintaining high service standards and adherence to operational processes.<br> Working closely with customers, internal engineering teams, and Oracle support services, the engineer assists in diagnosing and resolving issues related to cloud infrastructure, networking, security, and platform services.<br> The role also contributes to the development of technical documentation, knowledge sharing across the support team, and the continuous improvement of support processes to enhance service delivery and customer satisfaction.<br> Responsibilities: Handle escalated technical issues from L1 engineers.<br> Provide advanced troubleshooting for OCI services and architectures.<br> Create and maintain technical documentation and runbooks.<br> Mentor L1 engineers and conduct knowledge transfer sessions.<br> Identify patterns in customer issues and propose solutions.<br> Collaborate with OCI service teams on complex customer cases.<br> Participate in customer Well-Architected Reviews.<br> Oracle Cloud Infrastructure Certified Architect Professional (required) and one additional OCI certification.<br> 2+ years of experience with OCI services in another MSP or vendor support role.<br> Strong knowledge of networking concepts and security best practices.<br> Experience with infrastructure as code (Terraform/Ansible).<br> Proficiency in at least one programming/scripting language (Python, Shell scripting, etc.<br>). Advanced Linux/Windows troubleshooting skills.<br> Bachelor's degree in Computer Science or related field.<br> Strong problem-solving and analytical abilities.<br> Excellent customer service orientation.<br> Ability to work in a fast-paced environment.<br> Good time management and prioritization skills.<br> Team player with strong collaboration abilities.<br></span> </div>
<section><p class="heading jdMain">Job Description</p><p class="heading">Roles & Responsibilities</p><div class="paragraph"><p>The Cloud Engineer OCI is responsible for providing advanced technical support and troubleshooting for Oracle Cloud Infrastructure (OCI) environments. The role acts as a key escalation point for complex technical issues raised by Level 1 support engineers, ensuring timely resolution while maintaining high service standards and adherence to operational processes. Working closely with customers, internal engineering teams, and Oracle support services, the engineer assists in diagnosing and resolving issues related to cloud infrastructure, networking, security, and platform services. The role also contributes to the development of technical documentation, knowledge sharing across the support team, and the continuous improvement of support processes to enhance service delivery and customer satisfaction.</p><p>Responsibilities:</p><ul><li>Handle escalated technical issues from L1 engineers.</li><li>Provide advanced troubleshooting for OCI services and architectures.</li><li>Create and maintain technical documentation and runbooks.</li><li>Mentor L1 engineers and conduct knowledge transfer sessions.</li><li>Identify patterns in customer issues and propose solutions.</li><li>Collaborate with OCI service teams on complex customer cases.</li><li>Participate in customer Well-Architected Reviews.</li></ul></div></section><section><p class="heading">Desired Candidate Profile</p><p class="paragraph"></p><ul><li>Oracle Cloud Infrastructure Certified Architect Professional (required) and one additional OCI certification.</li><li>2+ years of experience with OCI services in another MSP or vendor support role.</li><li>Strong knowledge of networking concepts and security best practices.</li><li>Experience with infrastructure as code (Terraform/Ansible).</li><li>Proficiency in at least one programming/scripting language (Python, Shell scripting, etc.).</li><li>Advanced Linux/Windows troubleshooting skills.</li><li>Bachelor's degree in Computer Science or related field.</li><li>Strong problem-solving and analytical abilities.</li><li>Excellent customer service orientation.</li><li>Ability to work in a fast-paced environment.</li><li>Good time management and prioritization skills.</li><li>Team player with strong collaboration abilities.</li></ul><p></p></section>
We are looking for a Senior System Administrator to manage and scale our multi-cloud and on-prem environments, including air-gapped deployments in highly secure banking environments. You will handle Windows and Linux servers, web and database platforms, networking and VPNs, firewalls, and Cloudflare services. You’ll also build CI/CD pipelines, implement Infrastructure as Code, manage observability with Prometheus/Grafana/Wazuh, and ensure reliability, security, and compliance (ISO 27001, PCI DSS), including backup and disaster recovery. Key Responsibilities Infrastructure & Systems Administer Windows & Linux servers in cloud and isolated on-prem environments. Manage web servers (IIS, Apache/Nginx) and databases (MS SQL, Postgre SQL, MySQL). Manage domain controllers and related services, including Active Directory roles, replication, and group policy enforcement. Handle directory services: Active Directory and Entra ID (Azure AD). Oversee snapshots and backups. Networking & Security Design and maintain secure networks: VLANs, DNS/DHCP, load balancers. Configure VPNs (IPSec/Open VPN/Wire Guard) for site-to-site and remote access. Administer firewalls and WAF; manage Cloudflare (DNS, CDN, WAF, Zero Trust). Apply Zero Trust, MFA/SSO, and secrets management. Manage security solutions such as Kaspersky or equivalent tools to ensure device protection and compliance. Dev Ops & Automation Build and maintain CI/CD pipelines using open-source tools. Implement Infrastructure as Code (Terraform, Ansible) for provisioning and configuration. Manage and troubleshoot containerized environments using Docker and Kubernetes, including deployment, scaling, and monitoring. Monitoring & Compliance Deploy Prometheus and Grafana for metrics and dashboards. Operate Wazuh for SIEM/XDR (agent rollout, alert tuning, integrations). Ensure compliance with ISO 27001 and PCI DSS standards. Maintain audit-ready documentation and enforce security best practices. Reliability & Disaster Recovery Architect for high availability and fault tolerance. Design and maintain Backup & Disaster Recovery plans (RPO/RTO, immutable backups, DR drills). Support air-gapped environments with strict security and operational controls. Required Qualifications7–10 years of experience in IT Infrastructure, Dev Ops, or related roles. Strong Windows & Linux administration experience. Expertise in web servers (IIS, Apache/Nginx) and databases (MS SQL, Postgre SQL, MySQL). Hands-on with VPNs, firewalls, and Cloudflare. Experience with CI/CD, IaC (Terraform, Ansible), and containerization (Docker, Kubernetes). Ability to work with air-gapped environments and strict security controls. Scripting skills (Power Shell, Bash, or Python). Preferred Qualifications Familiarity with Prometheus, Grafana, and SIEM tools (e.g., Wazuh). Advanced Cloudflare (Zero Trust, tunnels, WAF tuning). Backup/DR tooling (Veeam, restic, Bacula). Endpoint management (Intune/MDM), CIS hardening benchmarks. Experience in regulated industries (finance, banking, e-commerce). Knowledge of ISO 27001 and PCI DSS implementation. Interpersonal & Language Skills Strong sense of accountability and ownership of tasks and outcomes. Proven ability to work effectively in team environments, collaborating across functions. Excellent problem-solving and adaptability in dynamic and complex situations. Clear and professional communication skills, both written and verbal. Mid to high English proficiency required for documentation, meetings, and cross-team collaboration. Ability to manage priorities, meet deadlines, and contribute to a culture of continuous improvement. What We Offer Competitive compensation package. Ownership of infrastructure and Dev Ops strategy. A collaborative environment with high impact and autonomy.
Project Overview:<br>Our customer is a multinational corporation in tobacco Industry with more than a century of history and offices in over 180 countries. Their most ambitious goal at the time is to introduce a range of Reduced-Risk Products (RRPs). The target audience is more than 1 billion consumers around the globe. IT platform hosts 700+ applications. Intellia's mission is to help the client with the engineering of a comprehensive software ecosystem for a game-changing IoT product on the margin of innovative consumer experience and cutting-edge technology. Our teams are involved in the engineering of core platform components for best-in-class e Commerce, Digital Marketing and IoT solutions. As a Dev Ops engineer, you will become a part of Core Architecture Team and be responsible for the architecture, implementation of best practices in our Digital Engineering Enterprise Platform. The Platform is a set of services and internet applications that accelerate the development and delivery of software applications by taking care of common SDLC challenges. The Platform provides access and consumption for engineering teams to a set of services, technologies, practices for their development and for operating their application, ensuring a set of compliance and best practices. Project is in production for 2+ years, being supported by multiple teams. Our technical domains are:- AWS cloud, partially Azure- SSO, Organizations, Service control policies, access models.- IAAC: terraform enterprise, terratest, chalice- Serverless: lambda, step functions, wide range of misc automations, fargate- System, Application, Network and security architectures- Orchecstration: k8s (eks)- SRE activities (logging, tracing, monitoring), Ops Genie, Splunk- Hashicorp Vault- Hybrid Networking<br>Responsibilities:<br>Design and implement observability frameworks for agent-based and distributed systems Build and maintain monitoring, logging, and tracing pipelines Develop dashboards and alerts to ensure system health and performance visibility Analyze system behavior and identify performance bottlenecks and anomalies Ensure high availability and reliability of runtime components Integrate observability tools with AWS infrastructure and CI/CD pipelines Support incident response, troubleshooting, and root cause analysis Collaborate with platform and AI teams to improve system transparency and operability<br><br>5+ years of experience working as a Dev Ops / Platform Engineer2+ years of experience building AI agents or agent-based systems (Agentic AI) Strong experience with AWS (EKS, EC2, VPC, RDS, Route53, API Gateway, Lambda) Hands-on experience with Terraform (AWS, Kubernetes/Helm, Hashicorp Vault) Hands-on experience with Observability tools: New Relic, Open Telemetry Strong knowledge of Kubernetes Strong programming skills in Python (scripting, Fast API, Swagger) and Bash / Power Shell Solid understanding of monitoring, logging, and distributed tracing concepts Experience with containerization (Docker, Kubernetes) Experience with CI/CD tools (Jenkins, Git Lab) Experience with configuration management tools (Ansible, Chef, Puppet)<br>Requirements:<br>5+ years of experience working as a Dev Ops / Platform Engineer2+ years of experience building AI agents or agent-based systems (Agentic AI) Strong experience with AWS (EKS, EC2, VPC, RDS, Route53, API Gateway, Lambda) Hands-on experience with Terraform (AWS, Kubernetes/Helm, Hashicorp Vault) Hands-on experience with Observability tools: New Relic, Open Telemetry Strong knowledge of Kubernetes Strong programming skills in Python (scripting, Fast API, Swagger) and Bash / Power Shell Solid understanding of monitoring, logging, and distributed tracing concepts Experience with containerization (Docker, Kubernetes) Experience with CI/CD tools (Jenkins, Git Lab) Experience with configuration management tools (Ansible, Chef, Puppet)
Sarmad is on the lookout for a Junior Dev Ops Engineer who will play a pivotal role in optimizing our development and operational processes. This position is ideal for someone who thrives in a collaborative environment and is passionate about utilizing cutting-edge technology to streamline workflows. If you have a strong background in Dev Ops practices and are eager to lead initiatives that drive efficiency, we want you on our team!<br><br>Key Responsibilities:<br><br>Design and implement robust CI/CD pipelines to automate processes Collaborate closely with development teams to enhance system performance and reliability Automate deployment processes and system monitoring Provide technical guidance and mentorship to junior team members Conduct regular system audits and performance tuning Document infrastructure and process workflows effectively Respond to and troubleshoot multiple production environments during on-call duties<br><br>Requirements<br><br>Minimum of 1 year of experience in a Dev Ops or similar role Familiar with cloud services such as AWS, Azure, OCI, or GCPExperience with container orchestration (Docker, Kubernetes) Familiarity with CI/CD tools like Jenkins, Argo CD, Git Lab, or Circle CIKnowledge of Infrastructure as Code principles using Terraform, Cloud Formation, or Ansible Strong scripting skills in languages such as Python, Ruby, or Bash Experience with monitoring and logging solutions (e.g., ELK Stack, Sentry, Datadog, New Relic, etc...) Familiar with Dev Sec Ops (concept and tools) Solid understanding of networking concepts and protocols Strong analytical and problem-solving abilities Bachelor's degree in Computer Science, Engineering, or a relevant field<br><br>Benefits<br><br>Hybrid work model Healthy working environment Medical Insurance Social Insurance