Job description
Company Description
What started in 2007 with a pizza and a plan has grown into a fast-moving SaaS company empowering manufacturers, distributors, and wholesalers to thrive in complex B2B commerce.
Our mission is simple: help businesses build stronger relationships through seamless digital commerce.
At Sana Commerce, we're looking for a Data Engineer (ML/AI) to design, build, and scale data systems that power our analytics and machine learning initiatives. Your work will ensure high-quality, reliable, and ML-ready data pipelines that enable both traditional analytics and advanced AI-driven solutions across the business.
Job Description
What you'll be doing
- Designing and maintaining data pipelines optimized for ML/AI workloads, including handling of large-scale, unstructured, and semi-structured data.
- Building feature pipelines and feature stores that ensure reusability and consistency of data used by machine learning models.
- Collaborating with Data Scientists and ML Engineers to understand data requirements for training, validation, and production deployment.
- Ensuring data quality, lineage, and governance meet standards required for AI/ML applications.
- Supporting MLOps practices by integrating data pipelines with model training, monitoring, and deployment workflows.
- Leveraging distributed processing frameworks (e.g., Spark, Databricks, Azure Synapse) for scalable ML data processing.
Qualifications
What you bring
- 5+ years of experience as a Data Engineer, working with Azure and Databricks, ideally with exposure to ML/AI-related data workflows.
- College degree that demonstrates your analytic abilities, such as Econometrics, Computer Sciences, Mathematics or similar;
- Excellent analytical and problem-solving skills;
- Experience with data preparation for ML/AI: managing large datasets, feature engineering, and real-time or batch data pipelines.
- Familiarity with MLOps concepts and how data engineering supports model lifecycle management.
- Experience with orchestration frameworks (Airflow, Prefect, or Azure Data Factory) for complex ML pipelines.
- Knowledge of unstructured data processing (text, images, logs) is a plus.
- Strong SQL and Python skills; experience with distributed data processing (PySpark, Dask, etc.) is a plus.
Why you’ll love working here
- Impact from day one – Join a scale-up where your ideas shape how global businesses operate online.
- Continuous learning – Access a structured onboarding rated 9.1/10 by previous hires, mentorship, and feedback culture.
- Hybrid flexibility – Work from our office in Alexandria 3 days per week and from home 2 days.
- Career growth – Expand your technical and leadership scope in a company built for long-term success.
Our values
At Sana Commerce, our values drive everything we do:
- Champions of Our League – We deliver lasting success, balancing quick wins and long-term value
- Supercharge Our Customers – We’re revolutionizing B2B commerce together, helping our customers to lead and succeed.
- Determined to Grow – We embrace challenges, growing and raising the bar for ourselves and our industry.
- Bold Together – We dare to be bold because we have each other’s back.
Ready to build reliability that scales?
Apply now and help shape the foundation of our next-generation SaaS platform.
Additional Information
#LI-Hybrid
وصف الوظيفة
وصف الشركة
ما بدأ في 2007 مع بيتزا وخطة أصبح شركة SaaS سريعة الحركة تمكّن المصنعين والموزعين وتجار الجملة من الازدهار في تجارة B2B المعقدة.
مهمتنا بسيطة: مساعدة الشركات على بناء علاقات أقوى من خلال التجارة الرقمية السلسة.
في Sana Commerce، نبحث عن مهندس بيانات (ML/AI) لتصميم وبناء وتوسيع أنظمة البيانات التي تدعم تحليلاتنا ومبادرات التعلم الآلي. سيضمن عملك خطوط أنابيب بيانات عالية الجودة وموثوقة وجاهزة لـ ML تمكّن كل من التحليلات التقليدية وحلول الذكاء الاصطناعي المتقدمة عبر الأعمال.
وصف الوظيفة
ما ستقوم به
- تصميم وصيانة خطوط أنابيب البيانات المحسّنة لأعباء عمل ML/AI، بما في ذلك التعامل مع بيانات كبيرة الحجم وغير منظمة وشبه منظمة.
- بناء خطوط أنابيب الميزات ومتاجر الميزات التي تضمن قابلية إعادة الاستخدام واتساق البيانات المستخدمة في نماذج التعلم الآلي.
- التعاون مع علماء البيانات ومهندسي ML لفهم متطلبات البيانات للتدريب والاختبار ونشر الإنتاج.
- ضمان جودة البيانات وسلسلة الأصل والحوكمة بما يتماشى مع معايير تطبيقات AI/ML.
- دعم ممارسات MLOps من خلال دمج خطوط أنابيب البيانات مع تدريب النماذج ومراقبتها وتد workflows النشر.
- الاستفادة من أطر المعالجة الموزعة (مثل Spark وDatabricks وAzure Synapse) لـ معالجة بيانات ML قابلة للتوسع.
المؤهلات
ما تجلبه
- 5+ سنوات من الخبرة كمهندس بيانات، العمل مع Azure وDatabricks، ويفضل التعرض لـخطوط بيانات ML/AI.
- درجة جامعية تُظهر قدراتك التحليلية، مثل الاقتصاد القياسي، علوم الكمبيوتر، الرياضيات أو ما يماثلها;
- مهارات تحليلية ومهارات حل المشكلات ممتازة؛
- خبرة في إعداد البيانات لـ ML/AI: إدارة مجموعات بيانات كبيرة، هندسة الميزات، وخطوط بيانات في الوقت الفعلي أو الدفعة.
- إلمام بمفاهيم MLOps وكيف تدعم هندسة البيانات إدارة دورة حياة النماذج.
- خبرة مع أطر التنظيم (Airflow وPrefect أو Azure Data Factory) لـ خطوط أنابيب ML معقدة.
- معرفة بمعالجة البيانات غير المهيكلة (النصوص، الصور، السجلات) ميزة إضافية.
- مهارات SQL وPython قوية؛ خبرة في معالجة البيانات الموزعة (PySpark، Dask، إلخ) ميزة إضافية.
لماذا ستستمتع بالعمل هنا
- التأثير من اليوم الأول – انضم إلى شركة ناشئة متوسطة الحجم حيث تشكل أفكارك كيفية عمل الشركات العالمية عبر الإنترنت.
- التعلم المستمر – وصول إلى عملية توجيه منظّمة بتقييم 9.1/10 من قبل من سبقك في العمل، والإرشاد، وثقافة التغذية الراجعة.
- المرونة الهجينة – العمل من مكتبنا في الإسكندرية 3 أيام في الأسبوع ومن المنزل يومين.
- نمو وظيفي – توسيع نطاقك التقني والقيادي في شركة مبنية للنجاح على المدى الطويل.
قيمنا
في Sana Commerce، تقود قيمنا كل ما نقوم به:
- أبطال دوريونا –نحن نحقق نجاحاً مستداماً، موازنين بين الانتصارات السريعة والقيمة الطويلة الأجل
- نعزز عملائنا –نحن نعيد تشكيل تجارة B2B معاً، ونساعد عملاءنا على القيادة والنجاح.
- عازمون على النمو –نقبل التحديات، ننمو ونرفع مستوى أنفسنا وصناعتنا.
- جرأة معاً –نجرؤ على أن نكون جريئين لأن لدينا ظهر بعضنا البعض.
هل أنت مستعد لبناء موثوقية قابلة للتوسع؟
قدّم الآن وساعد في تشكيل أساس منصة SaaS من الجيل التالي.
معلومات إضافية
#LI-Hybrid