شركة Jobgether تعلن عن وظيفة مهندس بيانات - Web Scraping في السعودية
تفاصيل الوظيفة
يُعلن Jobgether، بالنيابة عن شركة شريكة، عن توفر وظيفة مهندس بيانات - استخراج بيانات الويب (Data Engineer - Web Scraping) في السعودية. تتولى الشركة الشريكة إدارة جميع الطلبات والخطوات التالية.
نبذة عن الوظيفة
هذه فرصة مثيرة لبناء حلول قابلة للتطوير لاستخراج بيانات الويب وخطوط أنابيب البيانات التي تدعم الرؤى الحيوية للأعمال. ستعمل عن كثب مع المحللين والمهندسين والفرق متعددة الوظائف لتطوير مجموعات بيانات موثوقة تدعم اتخاذ القرارات الاستراتيجية. مع مستوى عالٍ من الملكية على مشاريعك، ستقوم بتصميم وأتمتة وتحسين سير عمل جمع البيانات مع ضمان جودة البيانات وموثوقيتها. تشجع البيئة على الابتكار والتعاون والتعلم المستمر، مما يمنحك حرية تجربة تقنيات جديدة وتحسين العمليات الحالية. إذا كنت تستمتع بحل تحديات البيانات المعقدة وبناء الأتمتة على نطاق واسع، فهذه الوظيفة توفر منصة ممتازة لإحداث تأثير ذي معنى.
المهام والمسؤوليات
- التعاون مع المحللين وأصحاب المصلحة لفهم متطلبات البيانات وتقديم حلول بيانات مخصصة.
- تصميم وتطوير وصيانة أدوات استخراج بيانات الويب لمجموعة واسعة من مصادر البيانات المنظمة وغير المنظمة.
- تنظيف وتحويل والتحقق من صحة ومعالجة مجموعات البيانات الكبيرة باستخدام Python وPandas.
- بناء وصيانة خطوط إدخال البيانات إلى قواعد البيانات أو مستودعات البيانات.
- جدولة ومراقبة وتحسين سير عمل الاستخراج باستخدام أدوات التنسيق مثل Apache Airflow.
- تطوير فحوصات مراقبة الجودة لضمان سلامة البيانات واتساقها وتوفرها.
- التحقيق في مشكلات خطوط أنابيب البيانات والحوادث الإنتاجية الحساسة للوقت وحلها.
- تصميم وتحسين الأدوات الداخلية وأطر الأتمتة وقدرات المنصة لتحسين الكفاءة التشغيلية.
- العمل عن كثب مع فرق الهندسة متعددة الوظائف لتنفيذ سير عمل معالجة بيانات قابلة للتطوير وقابلة للصيانة.
الشروط والمتطلبات
- درجة البكالوريوس أو الماجستير في علوم الحاسب أو تخصص تقني ذي صلة.
- 2-4 سنوات من الخبرة المهنية في تطوير البرمجيات.
- مهارات برمجة قوية في Python ومعرفة متينة بـ SQL / قواعد البيانات.
- خبرة متقدمة في استخدام مكتبة Pandas لتنظيف البيانات وتحويلها وتحليلها.
- خبرة في العمل مع تقنيات الويب بما في ذلك HTML وJavaScript وAPIs والبروتوكولات ذات الصلة.
- خبرة مثبتة في معالجة وتنظيف وتحويل مجموعات البيانات الكبيرة.
- الإلمام بأطر وأدوات استخراج بيانات الويب مثل Selenium وScrapy وXPath وFiddler أو Postman.
- خبرة مع أدوات تنسيق سير العمل مثل Apache Airflow أو منصات مماثلة.
- معرفة بـ Docker للحاويات؛ خبرة في Kubernetes تعتبر ميزة إضافية.
- الإلمام بأدوات CI/CD مثل Jenkins أو GitLab CI/CD.
- خبرة العمل مع خدمات السحابة، خاصة تقنيات AWS مثل S3 وRDS وLambda وSNS أو SQS (يفضل).
- تفكير تحليلي قوي، الاهتمام بالتفاصيل، مهارات تواصل، وشغف بالأتمتة والتحسين المستمر.
المزايا
- فرصة العمل على مشاريع تحدٍ تدعم بيئة إدارة أصول عالمية رائدة.
- مستوى عالٍ من الملكية والاستقلالية في ثقافة تعاونية موجهة نحو الفريق.
- التعرض لتقنيات حديثة في هندسة البيانات واستخراج بيانات الويب والحوسبة السحابية والأتمتة.
- بيئة تعاونية مع مهندسين وفرق منتجات وبيانات ذوي خبرة.
- فرص للتعلم المستمر والنمو المهني وتطوير المهارات التقنية.
- ثقافة قائمة على الجدارة تقدر الابتكار والمبادرة والمساهمات الفردية.
- بيئة مرنة تركز على التكنولوجيا مع فرص للعمل على منتجات بيانات ذات تأثير.
عرض النص الأصلي للإعلان
This role is an exciting opportunity to build scalable web scraping solutions and data pipelines that power business-critical insights. You will work closely with analysts, engineers, and cross-functional teams to develop reliable datasets that support strategic decision-making. With significant ownership over your projects, you'll design, automate, and optimize data collection workflows while ensuring data quality and reliability. The environment encourages innovation, collaboration, and continuous learning, giving you the freedom to experiment with new technologies and improve existing processes. If you enjoy solving complex data challenges and building automation at scale, this role offers an excellent platform to make a meaningful impact.
Accountabilities
As a Data Engineer specializing in web scraping, you will design, develop, and maintain automated data collection systems while ensuring the quality, accuracy, and availability of large-scale datasets. You will collaborate across teams to build efficient, reliable, and scalable data solutions.
- Collaborate with analysts and stakeholders to understand data requirements and deliver tailored data solutions.
- Design, develop, and maintain web scrapers for a wide range of structured and unstructured data sources.
- Clean, transform, validate, and manipulate large datasets using Python and Pandas.
- Build and maintain data ingestion pipelines into databases or data warehouses.
- Schedule, monitor, and optimize scraping workflows using orchestration tools such as Apache Airflow.
- Develop quality control checks to ensure data integrity, consistency, and availability.
- Investigate and resolve data pipeline issues and time-sensitive production incidents.
- Design and enhance internal tools, automation frameworks, and platform capabilities to improve operational efficiency.
- Work closely with cross-functional engineering teams to implement scalable and maintainable data processing workflows.
The ideal candidate combines strong software engineering skills with hands-on experience in web scraping, data processing, and automation. Success in this role requires both technical expertise and a proactive, problem-solving mindset.
- Bachelor's or Master's degree in Computer Science or a related technical discipline.
- 2-4 years of professional software development experience.
- Strong programming skills in Python and solid SQL/database knowledge.
- Advanced experience using the Pandas library for data cleaning, transformation, and analysis.
- Experience working with web technologies, including HTML, JavaScript, APIs, and related protocols.
- Proven experience processing, cleaning, and transforming large datasets.
- Familiarity with web scraping frameworks and tools such as Selenium, Scrapy, XPath, Fiddler, or Postman.
- Experience with workflow orchestration tools such as Apache Airflow or similar platforms.
- Knowledge of Docker containerization; Kubernetes experience is an advantage.
- Familiarity with CI/CD tools such as Jenkins or GitLab CI/CD.
- Experience working with cloud services, particularly AWS technologies such as S3, RDS, Lambda, SNS, or SQS, is preferred.
- Strong analytical thinking, attention to detail, communication skills, and a passion for automation and continuous improvement.
- Opportunity to work on challenging projects supporting a leading global asset management environment.
- High level of ownership and autonomy in a collaborative, team-oriented culture.
- Exposure to modern data engineering, web scraping, cloud, and automation technologies.
- Collaborative environment with experienced engineering, product, and data professionals.
- Opportunities for continuous learning, professional growth, and technical skill development.
- Merit-driven culture that values innovation, initiative, and individual contributions.
- Flexible, technology-focused environment with opportunities to work on impactful data products.
We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
Why Apply Through Jobgether?
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
وظائف أخرى لدى Jobgether
وظيفة مهندس DevOps لدى Jobgether في السعودية
Jobgether تعلن عن وظيفة مدير تسويق في السعودية