وصف العمل
الأدوار والمسؤوليات
تطوير خط أنابيب البيانات وتصميم البنية التحتية
- بناء وصيانة خطوط أنابيب بيانات قابلة للتوسع
- تطوير أطر المعالجة الآنية والدفعية للبيانات المنظمة وغير المنظمة.
- تنفيذ سير عمل ETL/ELT لاستيعاب البيانات من مصادر متنوعة، مع ضمان التوفر العالي والأداء الممتاز.
- تحسين تخزين البيانات واسترجاعها في الوقت الفعلي القريب والعمليات الدفعية
- ضمان بنية تحتية للبيانات ذات كفاءة تكلفية وعالية الأداء تتسع لاحتياجات العمل.
حلول البيانات والتطبيقات المعتمدة على الذكاء الاصطناعي
- تطوير خطوط أنابيب البيانات لخوارزميات التوصية ML، ووظائف البحث، وميزات معززة بالذكاء الاصطناعي.
- تطوير وصيانة نماذج البيانات وواجهات البرمجة والتكاملات لدعم التحليلات وتطبيقات العملاء.
- دعم حلول البيانات المتعلقة بالتجارة الإلكترونية، بما في ذلك توصيات المنتجات، تقسيم العملاء، ونماذج التخصيص.
التعاون والتحسين المستمر
- العمل بشكل وثيق مع مهندسي البيانات والمحللين وفِرق المنتجات لفهم متطلبات البيانات وتقديم حلول من الدرجة الأولى.
- مراقبة وتتبع مشكلات الأداء، مع ضمان التوفر العالي وكفاءة خطوط البيانات.
- التحسين المستمر للتكلفة، والأداء، وقابلية التوسع في حلول هندسة البيانات.
في Sana Commerce نحن ملتزمون ببيئة شاملة وندرك أن تنوع فريق العمل هو أحد أقوى أصولنا. بدأ كل شيء في عام 2007، مع بيتزا وخطة. Sana Commerce هي منصة تجارة إلكترونية مصممة لمساعدة المصنعين والموزعين وتجار الجملة على النجاح من خلال بناء علاقات طويلة الأمد مع العملاء الذين يعتمدون عليهم. نحن شركة SaaS سريعة النمو تتيح لك امتلاك مسارك المهني. نحن نبحث عن مهندس بيانات أول لتصميم وبناء وتحسين بنية البيانات المستندة إلى Azure وDatabricks. ستلعب دوراً حاسماً في تطوير خطوط ETL قابلة للتوسع، وأطر استيعاب البيانات، والمساهمة في حلول مدعومة بالذكاء الاصطناعي/التعلم الآلي تدفع قرارات داخلية ورؤى أمام العملاء. هذا دور عملي يتطلب خبرة في خدمات بيانات Azure، وDatabricks، وبايثون/SQL. ستعمل عن كثب مع مهندسي البيانات، والمهندسين، وفِرق المنتجات لتطوير حلول بيانات قوية تدعم التحليلات، والتخصيص، وتجارب العملاء المدعومة بالذكاء الاصطناعي.
المرشح المثالي
- >5+ سنوات خبرة كمهندس بيانات، العمل مع Apache Spark أو حملة Spark.
- خبرة قوية في تطوير خطوط أنابيب البيانات، سير عمل ETL، ومعالجة البيانات في الوقت الفعلي/الدفعي.
- خبرة في تقنيات البيانات الضخمة والتخزين على نطاق واسع.
- إلمام بممارسات حوكمة البيانات، الأمن، والامتثال وأدواتها
- إتقان Python وSQL، لتحويل البيانات، جودة البيانات والأتمتة.
- خبرة قوية في تحسين أداء البيانات وتقليل التكاليف في بيئات السحابة.
- مُفضل وجوده
- خبرة مع حلول بيانات التجارة الإلكترونية، مثل تقسيم العملاء، محرّكات التوصية، ونماذج التخصيص.
- معرفة بهندسة الحدث المعتمدة، ومعالجة البيانات المتدفقة والتحليلات في الوقت الفعلي.
- الإلمام بتطبيقات الذكاء الاصطناعي القائمة على LLM وواجهات برمجة OpenAI أو ما يماثلها.
Job Description
Roles & Responsibilities
Data Pipeline Development & Infrastructure Design
- Build, and maintain scalable data pipelines
- Develop real-time and batch data processing frameworks for structured and unstructured data.
- Implement ETL/ELT workflows to ingest data from various sources, ensuring high availability and performance .
- Optimize data storage and retrieval of data in (near) real time and batch processes
- Ensure cost-efficient and high-performance data infrastructure that scales with business needs.
Data Solutions & AI-Driven Applications
- Develop data pipelines for ML recommenders, search functionality, and AI-enhanced features .
- Develop and maintain data models, APIs, and integrations to support analytics and customer applications.
- Support eCommerce-related data solutions , including product recommendations, customer segmentation, and personalization models.
Collaboration & Continuous Improvement
- Work closely with Data Architects, Analysts, and Product Teams to understand data requirements and deliver best-in-class solutions.
- Monitor and troubleshoot performance issues , ensuring high availability and efficiency of data pipelines.
- Continuously optimize cost, performance, and scalability of data engineering solutions.
At Sana Commerce we're committed to an inclusive environment and recognize that our diverse work orce is one of our greatest strengths. It all started in 2007, with a pizza and a plan. Sana Commerce is an e-commerce platform designed to help manufacturers, distributors and wholesalers succeed by fostering lasting relationships with customers who depend on them. We re a fast-growing SaaS company that allows you to take ownership of your career. We are looking for a Senior Data Engineer to design, build, and optimize our Azure and Databricks-based data infrastructure . You will play a critical role in developing scalable ETL pipelines, data ingestion frameworks, and contribute to AI/ML-powered solutions that drive internal decision-making and customer-facing insights. This is a hands-on role that requires expertise in Azure Data Services, Databricks, and Python/SQL . You will work closely with Data Architects, Engineers, and Product teams to develop robust data solutions that power analytics, personalization, and AI-driven customer experiences .
Desired Candidate Profile
- 5+ years of experience as a Data Engineer , working with Apache Spark or bySpark.
- Strong expertise in data pipeline development, ETL workflows, and real-time/batch data processing .
- Experience in big data technologies, and large-scale data storage .
- Familiarity with data governance, security, and compliance best practices and tools
- Proficiency in Python, and SQL, for data transformation, data quality and automation.
- Strong experience in optimizing data performance and cost-efficiency in cloud environments .
- Nice to Have
- Experience with eCommerce data solutions , such as customer segmentation, recommendation engines, and personalization models.
- Knowledge of event-driven architectures, streaming data processing, and real-time analytics .
- Familiarity with LLM-based AI applications and OpenAI or similar APIs .