📍 المملكة العربية السعودية تحديث مستمر على مدار الساعة وظائف تناسب سيرتك الذاتيةمجاناً قناة تيليجرام

شركة Soar تعلن عن وظيفة مدير هندسة موثوقية المواقع (SREM) في الرياض

Site Reliability Engineer Manager (SREM)
🏢 Soar
🕒 نُشرت: (منذ 7 أيام) 📍 الرياض وظائف الهندسة والتقنية

تفاصيل الوظيفة

تعلن شركة Soar عن توفر فرصة وظيفية لشغل منصب مدير هندسة موثوقية المواقع (SREM) في الرياض، حيث سيكون المرشح المسؤول عن تعزيز موثوقية وأداء وكفاءة الأنظمة عبر المؤسسة، وسد الفجوة بين فرق التطوير والعمليات من خلال تطبيق مبادئ هندسة البرمجيات على إدارة الأنظمة.

نبذة عن الوظيفة

نبحث عن مهندس موثوقية مواقع أول (Senior SRE) ليكون الركيزة الأساسية في ضمان موثوقية وأداء وكفاءة الأنظمة. ستعمل على بناء بنية تحتية مقاومة للأعطال، ووضع أهداف مستوى الخدمة (SLOs)، وضمان توفر وأمان منصاتنا بما يتوافق مع معايير الامتثال المالي وحماية البيانات في المملكة العربية السعودية.

المهام والمسؤوليات

  • قيادة الموثوقية: تحديد وقياس وإدارة مؤشرات مستوى الخدمة (SLIs) وأهداف مستوى الخدمة (SLOs) وميزانيات الأخطاء لتحقيق التوازن بين سرعة إطلاق الميزات واستقرار النظام.
  • تصميم البنية التحتية: تصميم وبناء وصيانة البنية التحتية السحابية الأصلية على AWS/GCP/Azure باستخدام أدوات البنية التحتية كرمز (IaC) مثل Terraform/Pulumi.
  • أتمتة كل شيء: التخلص من المهام اليدوية المتكررة عن طريق بناء أدوات أتمتة للتزويد والتوسع والتعافي من الأعطال والنشر.
  • إدارة الحوادث: قيادة عملية إدارة الحوادث، والمشاركة في المناوبات، وإجراء تحليلات ما بعد الحادث دون لوم لضمان التحسين المستمر ومنع تكرار المشكلات.
  • تعزيز قابلية المراقبة: تطبيق أنظمة مراقبة وتسجيل وتنبيه شاملة باستخدام أدوات مثل Prometheus وGrafana وDatadog أو ELK Stack لتحديد اختناقات النظام بشكل استباقي.
  • تحسين CI/CD: التعاون مع فرق التطوير لتحسين خطوط أنابيب النشر من حيث السرعة والأمان والموثوقية.
  • ضمان الامتثال والأمان: العمل مع فريق الأمن لتطبيق ممارسات أمان قوية وضمان امتثال البنية التحتية للوائح السعودية (مثل متطلبات SAMA وNCA وPDPL الخاصة بتوطين البيانات).
  • الإرشاد: توجيه المهندسين المبتدئين وتعزيز ثقافة الموثوقية والتميز التشغيلي والملكية المشتركة عبر فريق الهندسة.

الشروط والمتطلبات

  • خبرة لا تقل عن 5 سنوات في هندسة البرمجيات أو DevOps أو أدوار SRE، منها سنتان على الأقل في منصب أول.
  • خبرة عملية عميقة مع مزودي الخدمات السحابية الرئيسيين (يفضل AWS أو GCP).
  • إتقان قوي لتنسيق الحاويات، خصوصًا Kubernetes وDocker.
  • مهارات برمجية قوية في لغة واحدة على الأقل من Go أو Python أو Java (لا يقتصر الأمر على scripting فقط).
  • خبرة مثبتة مع Terraform أو Ansible أو أدوات IaC مماثلة.
  • فهم عميق لأنظمة تشغيل Linux والشبكات (TCP/IP، DNS، التوجيه) وهندسة الأنظمة الموزعة.
  • مهارات تواصل ممتازة باللغة الإنجليزية كتابةً وتحدثًا (العربية ميزة إضافية قوية).

المهارات المطلوبة

  • خبرة سابقة في بيئة مالية (Fintech) أو عقارية (Proptech) أو بيئة شديدة التنظيم.
  • الإلمام بلوائح المملكة المالية وخصوصية البيانات (أطر SAMA وNCA).
  • خبرة في موثوقية قواعد البيانات (PostgreSQL، Redis، Kafka).
عرض النص الأصلي للإعلان
About The Role

Role Summary:

We are looking for a Senior Site Reliability Engineer (SRE) to champion reliability, performance, and efficiency across our engineering organization. In this role, you will bridge the gap between development and operations by applying software engineering principles to system administration. You will be instrumental in building fault-tolerant infrastructure, establishing SLOs, and ensuring our platforms remain highly available and secure, adhering to strict financial and data compliance standards in KSA.

Key Responsibilities:

  • Drive Reliability: Define, measure, and manage Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets to balance feature velocity with system stability.
  • Architect Infrastructure: Design, build, and maintain our cloud-native infrastructure on [AWS / GCP / Azure] using Infrastructure as Code (IaC) tools like [Terraform / Pulumi].
  • Automate Everything: Ruthlessly eliminate manual toil by building automation tools for provisioning, scaling, failover, and deployment.
  • Manage Incidents: Lead the incident management process, participate in on-call rotations, and conduct blameless post-mortems to ensure continuous improvement and prevent recurring issues.
  • Enhance Observability: Implement comprehensive monitoring, logging, and alerting systems using tools like [Prometheus, Grafana, Datadog, or ELK stack] to proactively identify system bottlenecks.
  • Optimize CI/CD: Partner with development teams to optimize deployment pipelines for speed, safety, and reliability.
  • Ensure Compliance & Security: Collaborate with the security team to implement robust security practices and ensure infrastructure complies with KSA regulations (e.g., SAMA, NCA, and PDPL data localization requirements).
  • Mentorship: Mentor junior engineers and cultivate a strong culture of reliability, operational excellence, and shared ownership across the engineering team.

What We Are Looking For:

  • Experience: 5+ years in Software Engineering, DevOps, or SRE roles, with at least 2 years in a senior capacity.
  • Cloud Expertise: Deep, hands-on experience with major cloud providers, preferably [AWS / GCP].
  • Containerization: Strong proficiency in container orchestration, specifically Kubernetes and Docker.
  • Coding Skills: Strong programming skills in at least one language such as Go, Python, or Java (not just bash scripting).
  • Infrastructure as Code: Proven experience with Terraform, Ansible, or similar IaC tools.
  • Systems Knowledge: Deep understanding of Linux operating systems, networking (TCP/IP, DNS, routing), and distributed systems architecture.
  • Communication: Excellent verbal and written communication skills in English (Arabic is a strong plus).

Nice to Have:

  • Prior experience working in a Fintech, Proptech, or highly regulated environment.
  • Familiarity with Saudi Arabian financial and data privacy regulations (SAMA, NCA frameworks).
  • Experience with database reliability (PostgreSQL, Redis, Kafka).
المصدر: LinkedIn - أُضيفت للموقع في 3 سبتمبر 2026
رقم الإعلان لدى المصدر: 4462793869