تفاصيل الوظيفة
تعلن شركة عزم السعودية عن توفر وظيفة مهندس DevOps في مدينة الرياض.
المهام والمسؤوليات
- تصميم وتنفيذ وإدارة خطوط التكامل المستمر والتوزيع المستمر (CI/CD).
- أتمتة توفير البنية التحتية باستخدام البنية التحتية ككود (IaC).
- نشر وصيانة التطبيقات في بيئات سحابية ومحلية.
- إدارة أحمال العمل المحوسبة باستخدام Docker و Kubernetes.
- مراقبة توفر النظام والأداء والأمان والسعة.
- تحديد وحل مشكلات البنية التحتية والتطبيقات والنشر.
- تنفيذ حلول التسجيل والتنبيه والنسخ الاحتياطي واستعادة البيانات بعد الكوارث.
- تطبيق أفضل ممارسات الأمان طوال دورة حياة تطوير البرمجيات.
- إدارة استراتيجيات التحكم بالمصادر وعمليات الإصدار وإعدادات البيئة.
- تحسين موثوقية المنصة وقابلية التوسع وتكرار النشر.
- تطوير سكريبتات وأدوات لتقليل المهام التشغيلية اليدوية.
- صيانة التوثيق الفني وأدلة التشغيل ومخططات البنية.
- المشاركة في الاستجابة للحوادث وتحليل الأسباب الجذرية والمناوبات عند الطلب.
- التعاون مع الفرق متعددة التخصصات لتعزيز ممارسات DevOps والتحسين المستمر.
الشروط والمتطلبات
- درجة البكالوريوس في علوم الحاسب أو تقنية المعلومات أو الهندسة أو ما يعادلها من الخبرة العملية.
- خبرة مثبتة في أدوار DevOps أو هندسة السحابة أو هندسة موثوقية المواقع أو إدارة الأنظمة.
- خبرة مع منصة سحابية كبرى واحدة على الأقل: AWS أو Microsoft Azure أو Google Cloud.
- إلمام قوي بإدارة Linux وأساسيات الشبكات.
- خبرة مع أدوات CI/CD مثل GitHub Actions أو GitLab CI/CD أو Jenkins أو Azure DevOps.
- إتقان أدوات البنية التحتية ككود مثل Terraform أو CloudFormation أو Pulumi.
- خبرة مع أدوات إدارة التهيئة مثل Ansible أو Puppet أو Chef.
- معرفة عملية بـ Docker ومنصات تنسيق الحاويات.
- خبرة في البرمجة النصية باستخدام Python أو Bash أو PowerShell.
- إلمام بأدوات المراقبة والملاحظة مثل Prometheus أو Grafana أو Datadog أو Splunk أو ELK Stack.
- فهم مبادئ إدارة الأسرار وإدارة الهوية والوصول وأمن السحابة.
- مهارات قوية في استكشاف الأخطاء وإصلاحها والتواصل والتعاون.
المهارات المطلوبة
- شهادات احترافية ذات صلة في السحابة أو Kubernetes (مفضلة).
- خبرة في إدارة بيئات Kubernetes الإنتاجية (مفضلة).
- إلمام بأدوات GitOps مثل Argo CD أو Flux (مفضلة).
- معرفة بشبكات الخدمات والمنصات الخالية من الخوادم وبنية الخدمات المصغرة (مفضلة).
- خبرة في عمليات قواعد البيانات وضبط الأداء واستعادة الكوارث (مفضلة).
- فهم مفاهيم SRE بما في ذلك SLIs و SLOs وميزانيات الأخطاء وإدارة الحوادث (مفضلة).
- خبرة في دعم الأنظمة الموزعة عالية التوفر وكبيرة النطاق (مفضلة).
عرض النص الأصلي للإعلان
Description
Key Responsibilities
Design, implement, and manage CI/CD pipelines.
Automate infrastructure provisioning using Infrastructure as Code (IaC).
Deploy and maintain applications across cloud and on-premises environments.
Manage containerized workloads using Docker and Kubernetes.
Monitor system availability, performance, security, and capacity.
Identify and resolve infrastructure, application, and deployment issues.
Implement logging, alerting, backup, and disaster-recovery solutions.
Apply security best practices throughout the software development lifecycle.
Manage source-control strategies, release processes, and environment configurations.
Improve platform reliability, scalability, and deployment frequency.
Develop scripts and tools to reduce manual operational tasks.
Maintain technical documentation, runbooks, and architecture diagrams.
Participate in incident response, root-cause analysis, and on-call rotations.
Collaborate with cross-functional teams to promote DevOps practices and continuous improvement.
Requirements
Bachelor’s degree in Computer Science, Information Technology, Engineering, or equivalent practical experience.
Proven experience in DevOps, cloud engineering, site reliability engineering, or systems administration.
Experience with at least one major cloud platform: AWS, Microsoft Azure, or Google Cloud.
Strong knowledge of Linux administration and networking fundamentals.
Experience with CI/CD tools such as GitHub Actions, GitLab CI/CD, Jenkins, or Azure DevOps.
Proficiency with Infrastructure as Code tools such as Terraform, CloudFormation, or Pulumi.
Experience with configuration-management tools such as Ansible, Puppet, or Chef.
Hands-on knowledge of Docker and container-orchestration platforms.
Scripting experience using Python, Bash, or PowerShell.
Familiarity with monitoring and observability tools such as Prometheus, Grafana, Datadog, Splunk, or the ELK Stack.
Understanding of secrets management, identity and access management, and cloud security principles.
Strong troubleshooting, communication, and collaboration skills.
Preferred Qualifications
Relevant cloud or Kubernetes certifications.
Experience managing production Kubernetes environments.
Familiarity with GitOps tools such as Argo CD or Flux.
Knowledge of service meshes, serverless platforms, and microservices architecture.
Experience with database operations, performance tuning, and disaster recovery.
Understanding of SRE concepts, including SLIs, SLOs, error budgets, and incident management.
Experience supporting high-availability and large-scale distributed systems.
Key Performance Indicators
Deployment frequency and lead time for changes
Application and infrastructure availability
Mean time to detect and recover from incidents
Change-failure rate
Percentage of infrastructure managed through automation
Security and compliance findings
Reduction in manual operational work
رقم الإعلان لدى المصدر: 2732044