وصف الوظيفة
حول Mozn MOZN هي شركة ذكاء اصطناعي مؤسسي رائدة تُمكن المؤسسات من اتخاذ قرارات مستنيرة في مجالين حاسمين: الوقاية من الجرائم المالية والذكاء المعرفي المؤسسي.
نحن فريق متنوع ومتعاون من المبتكرين نتشاطر هدفًا واحدًا: بناء ذكاء اصطناعي يقدم قيمة عمل ملموسة، يبني الثقة، ويمكّن الأشخاص والمنظمات من خلال الذكاء المعزز.
ثقافتنا مبنية على السعي المستمر للتفوق والتأثير ذو المعنى.
إذا كنت شغوفًا بالعمل بجانب مواهب استثنائية في ذكاء اصطناعي من الطراز العالمي، وتريد الاستقلالية والإمكانيات للقيام بأفضل عمل في مسيرتك المهنية، انضم إلينا لنشكل مستقبل المؤسّسات الذكية.
حول الدور نحن نبحث عن مهندس منصة بيانات عالي الدافع III للانضمام إلى فريق هندسة السحابة لدينا.
المرشح المثالي شغوف ببناء وتشغيل منصات بيانات حديثة تدعم التحليلات، ومعالجة البيانات في الوقت الفعلي، وأعباء العمل الخاصة بالذكاء الاصطناعي.
يتركز هذا الدور على تصميم، تشغيل، أتمتة، وتحسين بنية بيانات كبيرة النطاق بما في ذلك MySQL، PostgreSQL، Kafka، خطوط ETL/ELT، منصات CDC، قواعد البيانات التحليلية (StarRocks، ClickHouse)، وأطر معالجة البيانات الموزعة مثل Flink أو Spark عبر بيئات سحابية أصلية وهجينة.
سيملك المرشح المثالي خبرة عميقة في منصات هندسة البيانات وأنظمة البيانات الموزعة.
خبرة في Kubernetes وبيئة المنصة السحابية مهمة، لكن التركيز الأساسي على تمكين منصات بيانات قابلة للتوسع وموثوقة وذات أداء عالٍ تدعم فرق الهندسة والتحليلات وعلوم البيانات.
ما ستفعله هندسة منصة البيانات تصميم، نشر، تشغيل، وتحسين منصات بيانات كبيرة النطاق تدعم أعباء عمل معاملية وتحليلية.
بناء وصيانة خطوط ETL/ELT موثوقة للدفعات والوقت الحقيقي.
تصميم وتشغيل منصات CDC باستخدام Debezium، Kafka Connect، أو تقنيات مماثلة.
نشر، تشغيل، وتحسين قواعد البيانات التحليلية مثل StarRocks، ClickHouse، Apache Doris، أو منصات OLAP مماثلة.
تصميم معماريات Kafka قابلة للتوسع، بما في ذلك الموضوعات، التقسيمات، النسخ، مجموعات المستهلكين، وخطوط التدفق.
إدارة عناقيد MySQL و PostgreSQL، بما في ذلك النسخ الاحتياطي، الاستعادة، التعافي من الكوارث، وتحسين الأداء.
تحسين أداء الاستعلام الموزع، تخطيطات التخزين، استراتيجيات الفهرسة، وإدارة دورة حياة البيانات.
دعم عمليات الاستيعان، التحويل، والأعباء التحليلية كبيرة النطاق مع ضمان موثوقية المنصة وقابليتها للتوسع.
أنظمة البيانات الموزعة تصميم وتشغيل منصات البيانات المتدفقة في الوقت الفعلي باستخدام Kafka، Flink، Spark، أو تقنيات مماثلة.
تحسين الأنظمة الموزعة من ناحية الطاقة، الكمون، قابلية التوسع، والمرونة.
استكشاف مشكلات معقدة عبر قواعد البيانات الموزعة وأنظمة الرسائل وخطوط معالجة البيانات.
التعاون الوثيق مع فرق هندسة البيانات والتحليلات وعلوم البيانات لبناء قدرات منصة بيانات موثوقة.
أتمتة المنصة أتمتة التزويد، النشر، الترقيات، التوسع، وإدارة دورة حياة منصات البيانات.
بناء قدرات الخدمة الذاتية لفرق الهندسة.
تحسين الرصد، الموثوقية، والتميز التشغيلي عبر منصة البيانات.
ستكون في طليعة وقت مثير للشرق الأوسط، ملتحقًا بركوب صاروخي عالي النمو في مساحة مثيرة ستُعطىك قدرًا كبيرًا من المسؤولية والثقة.
نؤمن بأن أفضل النتائج تأتي عندما يُمنح الأشخاص المسؤولون عن وظيفة ما الحرية للقيام بما يعتقدون أنه الأفضل. ستتم العناية بالأساسيات: تعويض تنافسي، تأمين صحي من المستوى الأول، وثقافة ممكّنة حتى تتمكن من التركيز على ما تفعل أفضل ما لديك.
ستستمتع بمكان عمل ممتع وديناميكي تعمل فيه بجانب عقول عظيمة في الذكاء الاصطناعي. نؤمن أن القوة تكمن في الاختلاف، مع الاحتضان للجميع كما هم ومكّونين لأن يكونوا أفضل إصدار لأنفسهم.
من 4-7 سنوات من الخبرة في هندسة منصة البيانات، هندسة قواعد البيانات، هندسة المنصات، أو بنية تحتية للبيانات.
خبرة قوية في الإنتاج مع MySQL و PostgreSQL.
خبرة عميقة في تشغيل Kafka في بيئات الإنتاج.
خبرة قوية في بناء خطوط ETL/ELT.
خبرة عملية مع StarRocks، ClickHouse، Apache Doris، أو قواعد بيانات OLAP مماثلة.
خبرة مع منصات التدفق مثل Kafka، Flink، Spark Streaming، أو Kafka Streams.
خبرة قوية في تحسين SQL وتحسين أداء قواعد البيانات.
خبرة مع تقنيات CDC مثل Debezium أو Kafka Connect.
خبرة في نشر وتشغيل أحمال البيانات ذات الطبيعة الحية باستخدام Kubernetes.
خبرة مع AWS، GCP، OCI، أو Azure.
خبرة مع Terraform، Helm، GitOps، وأطر الأتمتة.
مهارات برمجة قوية في Python، Bash، أو Go.
خبرة مع منصات الرصد مثل Prometheus، Grafana، ELK/OpenSearch، أو LGTM.
المؤهلات المفضلة خبرة مع Apache Iceberg، Delta Lake، أو Apache Hudi.
خبرة مع Trino، Presto، Pinot، أو Druid.
خبرة دعم منصات Data Science والتحليلات.
خبرة في تصميم معماريات حديثة لـ data lakehouse.
مساهمات في تقنيات منصة البيانات مفتوحة المصدر.
شهادات Cloud، Kubernetes، Kafka، أو Database تعتبر ميزة إضافية.
Job description
About Mozn MOZN is a leading Enterprise AI company enabling organizations to make informed decisions in two critical domains: Financial Crime Prevention and Enterprise Knowledge Intelligence.
We’re a diverse, collaborative team of innovators united by a shared purpose: to build AI that delivers tangible business value, builds trust, and empowers people and organizations with augmented intelligence.
Our culture is built on the relentless pursuit of excellence and meaningful impact.
If you’re passionate about working alongside exceptional talent on world-class AI, and you want the autonomy and runway to do the best work of your career, join us in shaping the future of intelligent enterprises.
About the role We are looking for a highly motivated Data Platform Engineer III to join our Cloud Engineering team.
The ideal candidate is passionate about building and operating modern data platforms that power analytics, real-time data processing, and AI workloads.
This role focuses on designing, operating, automating, and continuously improving large-scale data infrastructure including MySQL, PostgreSQL, Kafka, ETL/ELT pipelines, CDC platforms, analytical databases (StarRocks, ClickHouse), and distributed data processing frameworks such as Flink or Spark across cloud-native and hybrid environments.
The ideal candidate will possess deep expertise in data engineering platforms and distributed data systems.
Kubernetes and cloud platform experience are important, but the primary focus is on enabling scalable, reliable, and high-performance data platforms that support engineering, analytics, and data science teams.
What you'll do Data Platform Engineering Design, deploy, operate, and optimize large-scale data platforms supporting transactional and analytical workloads.
Build and maintain reliable batch and real-time ETL/ELT pipelines.
Design and operate CDC platforms using Debezium, Kafka Connect, or similar technologies.
Deploy, operate, and optimize analytical databases such as StarRocks, ClickHouse, Apache Doris, or similar OLAP platforms.
Design scalable Kafka architectures, including topics, partitions, replication, consumer groups, and streaming pipelines.
Administer MySQL and PostgreSQL clusters, including replication, backup, recovery, disaster recovery, and performance tuning.
Optimize distributed query performance, storage layouts, indexing strategies, and data lifecycle management.
Support large-scale ingestion, transformation, and analytical workloads while ensuring platform reliability and scalability.
Distributed Data Systems Design and operate streaming and real-time data platforms using Kafka, Flink, Spark, or similar technologies.
Optimize distributed systems for throughput, latency, scalability, and resilience.
Troubleshoot complex issues across distributed databases, messaging systems, and data processing pipelines.
Collaborate closely with Data Engineering, Analytics, and Data Science teams to build reliable data platform capabilities.
Platform Automation Automate provisioning, deployment, upgrades, scaling, and lifecycle management of data platforms.
Build self-service capabilities for engineering teams.
Improve observability, monitoring, reliability, and operational excellence across the data platform.
You will be at the forefront of an exciting time for the Middle East, joining a high-growth rocket-ship in an exciting space You will be given a lot of responsibility and trust.
We believe that the best results come when the people responsible for a function are given the freedom to do what they think is best The fundamentals will be taken care of: competitive compensation, top-tier health insurance, and an enabling culture so that you can focus on what you do best You will enjoy a fun and dynamic workplace working alongside some of the greatest minds in AI We believe strength lies in difference, embracing all for who they are and empowered to be the best version of themselves 4-7 years of experience in Data Platform Engineering, Database Engineering, Platform Engineering, or Data Infrastructure.
Strong production experience with MySQL and PostgreSQL.
Deep expertise operating Kafka in production environments.
Strong experience building ETL/ELT pipelines.
Hands-on experience with StarRocks, ClickHouse, Apache Doris, or similar OLAP databases.
Experience with streaming platforms such as Kafka, Flink, Spark Streaming, or Kafka Streams.
Strong SQL optimization and database performance tuning experience.
Experience with CDC technologies such as Debezium or Kafka Connect.
Kubernetes experience deploying and operating stateful data workloads.
Experience with AWS, GCP, OCI, or Azure.
Experience with Terraform, Helm, GitOps, and automation frameworks.
Strong scripting skills in Python, Bash, or Go.
Experience with observability platforms such as Prometheus, Grafana, ELK/OpenSearch, or LGTM.
Preferred Qualifications Experience with Apache Iceberg, Delta Lake, or Apache Hudi.
Experience with Trino, Presto, Pinot, or Druid.
Experience supporting Data Science and Analytics platforms.
Experience designing modern data lakehouse architectures.
Contributions to open-source data platform technologies.
Cloud, Kubernetes, Kafka, or Database certifications are a plus.