نحن نبحث عن مهندس بيانات وبنية تحتية أول (ذكر/أنثى/مُتعدِّد الجنسين) للانضمام إلى فريق البيانات لدينا والمساعدة في تحويل شركتنا إلى منظمة تعتمد البيانات حقًا. في هذا الدور، ستقوم بتصميم وبناء حلول بيانات قابلة للتوسع تمكّن فرق الأعمال من بيانات موثوقة عالية الجودة ورؤى قابلة للتنفيذ. إذا كنت تستمتع بحل تحديات البيانات المعقدة، وبناء منصات قوية، والعمل في بيئة دولية وتعاونية، سنسعد بسماع أخبارك.
المسؤوليات
- تصميم وتطوير ونشر وصيانة حلول بيانات قابلة للتوسع ومفتوحة المصدر تدعم أعمالنا الدولية سريعة النمو.
- بناء وتطوير هندسة البيانات لدينا تمكين التحليلات المتقدمة والتقارير واتخاذ القرار المستند إلى البيانات عبر المؤسسة.
- تطوير وتحسين خطوط أنظمة CDC / ETL/ELT لمعالجة البيانات دفعة وبيانات قريبة من الحدث.
- دعم سير عمل التحليل والتعلم الآلي من خلال توفير مجموعات بيانات موثوقة ومهيأة جيداً.
- دمج البيانات من نطاق واسع من المصادر، بما في ذلك قواعد البيانات المعاملات، والمنصات الطرف الثالث، وواجهات برمجة التطبيقات الخارجية (مثل PostgreSQL، MySQL، Google Analytics، BigQuery، Google Ads، Adjust، وغيرها).
- التعاون عن كثب مع مهندسي البيانات والمحللين وعلماء البيانات وأصحاب المصلحة الأعمال لضمان جودة البيانات واتساقها ومعايير الحوكمة.
- مراقبة واستكشاف الأخطاء وتحسين أداء منصة البيانات وموثوقيتها وقابليتها للتوسع باستخدام حلول مُدارة ومفتوحة المصدر.
- تقييم وتوصية بتقنيات جديدة (مفتوحة المصدر/ مُدارة)، وأدوات، ونهج معماري لتعزيز نظام البيانات لدينا.
- المساهمة في الوثائق التقنية وتبادل المعرفة وأفضل الممارسات داخل فريق البيانات والمؤسسة بشكل أوسع.
المرشح المثالي
أكثر من 4 سنوات من الخبرة المهنية في هندسة البيانات، مخازن البيانات، عمليات البيانات أو مجال ذي صلة.
خبرة عملية جيدة في إحدى السُحُب الكبرى. يُفضل AWS، مطلوب خبرة عملية متقدمة في أدوات لينكس
يجب أن be able to deploy التطبيقات باستخدام Docker على خوادم عارية.
درجة البكالوريوس أو أعلى في علوم الكمبيوتر، الرياضيات، الفيزياء، الهندسة، الاقتصاد، أو تخصص كمي آخر.
خبرة قوية في SQL وPSQL ومفاهيم قواعد البيانات العلائقية، بما في ذلك نمذجة البيانات وتحسين الاستعلامات.
خبرة في تصميم وصيانة بنى مخازن البيانات.
خبرة عملية مع قواعد البيانات العمودية وتقنيات تحسين الأداء.
خبرة في العمل مع أدوات تنظيم سير العمل وجدولة المهام الموزعة، ويفضل Apache Airflow.
مهارات برمجة قوية في بايثون أو جافا، مع تفضيل بايثون.
فهم جيد لبيئات لينكس والبرمجة النصية (Bash).
خبرة مع منصات البيانات الحديثة القائمة على السحابة وقواعد البيانات التحليلية مثل ClickHouse، StarRocks، أو تقنيات مشابهة.
لدى المرشح خبرة في العمل مع حلول مفتوحة المصدر متوافقة مع S3 وحلول مُدارة.
المألوف مع تكامل البيانات، أطر CDC/ETL/ELT، ومعالجة البيانات على نطاق واسع.
مهارات تواصل قوية باللغة الإنجليزية.
We are looking for a Senior Data and infrastructure Engineer (f/m/d) to join our Data Team and help transform our company into a truly data-driven organization. In this role, you will design and build scalable data solutions that empower business teams with reliable, high-quality data and actionable insights. If you enjoy solving complex data challenges, building robust platforms, and working in an international and collaborative environment, we'd love to hear from you.
Responsibilities
- Design, develop, deploy and maintain scalable and open source data solutions that support our rapidly growing international business.
- Build and evolve our data Architecture to enable advanced analytics, reporting, and data-driven decision-making across the organization.
- Develop and optimize CDC / ETL/ELT pipelines for batch and near real-time data processing.
- Support analytical and machine learning workflows by providing reliable and well-structured datasets.
- Integrate data from a wide range of sources, including transactional databases, third-party platforms, and external APIs (e.g., PostgreSQL, MySQL, Google Analytics, BigQuery, Google Ads, Adjust, and others).
- Collaborate closely with data engineers, analysts, data scientists, and business stakeholders to ensure data quality, consistency, and governance standards.
- Monitor, troubleshoot, and continuously improve data platform performance, reliability, and scalability using managed and open source solutions.
- Evaluate and recommend new technologies (Open source/ Managed), tools, and architectural approaches to enhance our data ecosystem.
- Contribute to technical documentation, knowledge sharing, and best practices within the Data Team and the wider organization.
Desired Candidate Profile
4+ years of professional experience in Data Engineering, Data Warehousing, Data Ops or a related field.
Good hands on experience in one of the major clouds . AWS preferred, Expert Hands on experience required on linux tools
Should be able to deploy applications using docker on bare metal servers.
Bachelor's degree or higher in Computer Science, Mathematics, Physics, Engineering, Economics, or another quantitative discipline.
Strong expertise in SQL PSQL and relational database concepts, including data modeling and query optimization.
Experience designing and maintaining data warehouse architectures.
Hands-on experience with columnar databases and performance optimization techniques.
Experience working with workflow orchestration and distributed job scheduling tools, preferably Apache Airflow.
Solid programming skills in Python or Java, with a preference for Python.
Good understanding of Linux environments and scripting (Bash).
Experience with modern cloud-based data platforms and analytical databases such as ClickHouse, StarRocks, or similar technologies.
Have experience working s3 compatible open source solution and managed solutions.
Familiarity with data integration, CDC/ETL/ELT frameworks, and large-scale data processing.
Strong communication skills in English.