نظرة عامة على الوظيفة
نحن نبحث عن مهندس بيانات AI متحفز وذو طابع تحليلي للانضمام إلى قسم AI4ALL بدوام جزئي. يوفر هذا التدريب خبرة عملية في هندسة البيانات ضمن بيئة بحث وتطوير AI ديناميكية. سيعمل المرشح الناجح مع فرق علوم البيانات وهندسة البرمجيات لتصميم وتنفيذ وتحسين خطوط البيانات والبنية التحتية التي تُمكّن من حلول AI قابلة للتوسع.
عن الشركة
منظمتنا مكرسة لتعزيز الذكاء الاصطناعي المتاح والمسؤول. نُعِد ثقافة شاملة وتعاونية تقدر الفضول وحل المشكلات بدقة والأثر العملي. مقيمة في مصر، نتعاون مع فرق إقليمية وعالمية لتقديم مشاريع مدفوعة بالذكاء الاصطناعي عبر صناعات مختلفة، مع تركيز على التعليم والرعاية الصحية والحلول المؤسسية.
المسؤوليات الرئيسية
- تصميم وبناء والحفاظ على خطوط بيانات قوية لجمع واستيعاب وتحويل وتخزين مجموعات بيانات كبيرة من مصادر متنوعة.
- التعاون مع علماء البيانات لفهم متطلبات البيانات واحتياجات إنشاء الميزات وتدفقات تدريب النماذج.
- تنفيذ فحوصات جودة البيانات ومراقبتها وعمليات التحقق لضمان موثوقية وسلامة البيانات.
- تحسين عمليات ETL/ELT للأداء وقابلية التوسع وكفاءة التكلفة في بيئات السحابة.
- المساعدة في تطوير مخطط البيانات وفهارس البيانات ووثائق سلاسل البيانات.
- المساهمة في الحوكمة الأمنية وخصوصية البيانات وفق سياسات المنظمة.
- المشاركة في مراجعات الشفرة والاختبار والتوثيق لدعم ممارسات هندسية قابلة للح maintenance.
- دعم النشر والتشغيل للبنية التحتية للبيانات، بما في ذلك خطوط CI/CD لعمليات البيانات.
المؤهلات والمتطلبات
- درجة البكالوريوس في علوم الحاسب أو الهندسة أو علوم البيانات أو مجال ذي صلة؛ قبول برامج التدريب التي تركز على التدريب الداخلي كخيار.
- سنة واحدة من الخبرة المتعلقة في هندسة البيانات أو خطوط البيانات أو تكامل البيانات.
- الإلمام بلغات البرمجة الشائعة في هندسة البيانات (بايثون، SQL؛ المعرفة بجافا/سكالا تعتبر إضافة).
- خبرة مع أطر وأدوات معالجة البيانات (مثل Apache Spark، Apache Airflow، منصات ETL/ELT).
- فهم لقواعد البيانات العلائقية وNoSQL، ومفاهيم نمذجة البيانات، وخدمات البيانات السحابية (AWS، GCP، أو Azure).
- مهارات تحليلية قوية وحل المشكلات مع الانتباه للتفاصيل والدقة.
- مهارات تواصل ممتازة والقدرة على العمل بتعاون في فريق متعدد التخصصات.
- القدرة على العمل في مصر والتفرغ بدوام جزئي وفق جدولة التدريب.
المهارات المطلوبة
- تصميم وهندسة خطوط البيانات
- تحسين استعلامات SQL وطلبات البيانات
- برمجة بايثون لمهام هندسة البيانات
- خبرة في أدوات ETL/ELT وتنظيم تدفقات العمل (يفضل Airflow)
- معرفة بخدمات البيانات السحابية وممارسات أمان البيانات الأساسية
- مهارات اتصال والعمل الجماعي القوية
الفوائد والمزايا
- خبرة عملية وتطبيقية مع مشاريع بيانات AI حقيقية
- توجيه من مهندسين وعلماء بيانات متمرسين
- جدول عمل جزئي ومرن لاستيعاب الالتزامات الأكاديمية
- فرص تواصل داخل منظمة AI في نمو مستمر
- مسار محتمل إلى فرص بدوام كامل بناءً على الأداء
Job Overview
We are seeking a motivated and analytical AI Data Engineer to join the AI4ALL Department on a part-time basis. This internship position offers hands-on experience in data engineering within a dynamic AI research and development environment. The successful candidate will collaborate with data science and software engineering teams to design, implement, and optimize data pipelines and infrastructure that enable scalable AI solutions.
About the Company
Our organization is dedicated to advancing accessible and responsible artificial intelligence. We foster an inclusive, collaborative culture that values curiosity, rigorous problem solving, and practical impact. Based in Egypt, we partner with regional and global teams to deliver AI-driven projects across industries, with a focus on education, healthcare, and enterprise solutions.
Key Responsibilities
- Design, build, and maintain robust data pipelines to collect, ingest, transform, and store large-scale datasets from diverse sources.
- Collaborate with data scientists to understand data requirements, feature engineering needs, and model training workflows.
- Implement data quality checks, monitoring, and validation processes to ensure data reliability and integrity.
- Optimize ETL/ELT processes for performance, scalability, and cost-efficiency in cloud environments.
- Assist in the development of data schemas, metadata catalogs, and data lineage documentation.
- Contribute to data governance, security, and privacy compliance in accordance with organizational policies.
- Participate in code reviews, testing, and documentation to support maintainable engineering practices.
- Support deployment and operations of data infrastructure, including CI/CD pipelines for data workflows.
Qualifications and Requirements
- Bachelor’s degree in Computer Science, Engineering, Data Science, or a related field; pursuing or recently completed internship-focused program acceptable.
- 1 year of related experience in data engineering, data pipelines, or data integration.
- Familiarity with programming languages commonly used in data engineering (Python, SQL; knowledge of Java/Scala a plus).
- Experience with data processing frameworks and tools (e.g., Apache Spark, Apache Airflow, ETL/ELT platforms).
- Understanding of relational and NoSQL databases, data modeling concepts, and cloud-based data services (AWS, GCP, or Azure).
- Strong analytical and problem-solving skills with attention to detail and accuracy.
- Excellent communication skills and ability to work collaboratively in a cross-functional team.
- Eligibility to work in Egypt and availability for part-time commitment as per the internship schedule.
Required Skills
- Data pipeline design and engineering
- SQL and data querying optimization
- Python programming for data engineering tasks
- Experience with ETL/ELT tools and workflow orchestration (Airflow preferred)
- Knowledge of cloud data services and basic data security practices
- Strong communication and teamwork abilities
Benefits and Perks
- Practical, hands-on experience with real-world AI data projects
- Mentorship from experienced engineers and data scientists
- Flexible, part-time schedule to accommodate academic commitments
- Networking opportunities within a growing AI organization
- Potential pathway to full-time opportunities based on performance