في المكتب دوام كامل
--
Intellias

تفاصيل الوظيفة

Project Overview:
Our customer is a multinational corporation in tobacco Industry with more than a century of history and offices in over 180 countries. Their most ambitious goal at the time is to introduce a range of Reduced-Risk Products (RRPs). The target audience is more than 1 billion consumers around the globe. IT platform hosts 700+ applications. Intellia's mission is to help the client with the engineering of a comprehensive software ecosystem for a game-changing IoT product on the margin of innovative consumer experience and cutting-edge technology. Our teams are involved in the engineering of core platform components for best-in-class e Commerce, Digital Marketing and IoT solutions. As a Dev Ops engineer, you will become a part of Core Architecture Team and be responsible for the architecture, implementation of best practices in our Digital Engineering Enterprise Platform. The Platform is a set of services and internet applications that accelerate the development and delivery of software applications by taking care of common SDLC challenges. The Platform provides access and consumption for engineering teams to a set of services, technologies, practices for their development and for operating their application, ensuring a set of compliance and best practices. Project is in production for 2+ years, being supported by multiple teams. Our technical domains are:- AWS cloud, partially Azure- SSO, Organizations, Service control policies, access models.- IAAC: terraform enterprise, terratest, chalice- Serverless: lambda, step functions, wide range of misc automations, fargate- System, Application, Network and security architectures- Orchecstration: k8s (eks)- SRE activities (logging, tracing, monitoring), Ops Genie, Splunk- Hashicorp Vault- Hybrid Networking
Responsibilities:
Design and implement observability frameworks for agent-based and distributed systems Build and maintain monitoring, logging, and tracing pipelines Develop dashboards and alerts to ensure system health and performance visibility Analyze system behavior and identify performance bottlenecks and anomalies Ensure high availability and reliability of runtime components Integrate observability tools with AWS infrastructure and CI/CD pipelines Support incident response, troubleshooting, and root cause analysis Collaborate with platform and AI teams to improve system transparency and operability

5+ years of experience working as a Dev Ops / Platform Engineer2+ years of experience building AI agents or agent-based systems (Agentic AI) Strong experience with AWS (EKS, EC2, VPC, RDS, Route53, API Gateway, Lambda) Hands-on experience with Terraform (AWS, Kubernetes/Helm, Hashicorp Vault) Hands-on experience with Observability tools: New Relic, Open Telemetry Strong knowledge of Kubernetes Strong programming skills in Python (scripting, Fast API, Swagger) and Bash / Power Shell Solid understanding of monitoring, logging, and distributed tracing concepts Experience with containerization (Docker, Kubernetes) Experience with CI/CD tools (Jenkins, Git Lab) Experience with configuration management tools (Ansible, Chef, Puppet)
Requirements:
5+ years of experience working as a Dev Ops / Platform Engineer2+ years of experience building AI agents or agent-based systems (Agentic AI) Strong experience with AWS (EKS, EC2, VPC, RDS, Route53, API Gateway, Lambda) Hands-on experience with Terraform (AWS, Kubernetes/Helm, Hashicorp Vault) Hands-on experience with Observability tools: New Relic, Open Telemetry Strong knowledge of Kubernetes Strong programming skills in Python (scripting, Fast API, Swagger) and Bash / Power Shell Solid understanding of monitoring, logging, and distributed tracing concepts Experience with containerization (Docker, Kubernetes) Experience with CI/CD tools (Jenkins, Git Lab) Experience with configuration management tools (Ansible, Chef, Puppet)

وظائف مشابهة

حول Intellias
مصر, القاهرة
تكنولوجيا المعلومات والخدمات