Site Reliability Engineer (SRE)

Wipro
الرياض, الرياض دوام كامل
نشر: 1448/1/14 | 2026/06/29 ينتهي: 1448/2/15 | 2026/07/29 ✨ وصف بالذكاء الاصطناعي
تقدم للوظيفة الآن

الوصف الوظيفي

Role Summary

Ensure high availability, performance, and reliability of production systems through automation, cloud operations, monitoring, and incident management.

Key Responsibilities

  • Build and manage scalable AWS cloud infrastructure using Terraform / CloudFormation
  • Implement monitoring and observability using Prometheus, Grafana, Splunk, or Datadog
  • Handle incidents, perform root cause analysis, provide on-call support, manage service level objectives (SLOs), service level indicators (SLIs), and error budgets
  • Automate operations to improve efficiency, reliability, and mean time to recovery (MTTR)
  • Manage Kubernetes, Docker, CI/CD pipelines, runbooks, and ITIL processes

Mandatory Skills

  • Proven experience in SRE / DevOps / Production Support with AWS cloud experience
  • Strong proficiency in Kubernetes, Docker, Linux, and networking fundamentals
  • Expertise in scripting using Python, Bash, or Go
  • Hands-on experience with monitoring tools and incident / RCA management

Preferred Skills

  • Terraform / CloudFormation, Jenkins / GitLab CI/CD, ServiceNow / ITSM
  • Cloud security tools and enterprise environment experience

Experience

5+ years in SRE / DevOps / Cloud / Production Support

يمكن أن يرتكب الذكاء الاصطناعي أخطاءً.

المصدر: لينكد إن ↗ • 4 مشاهدة

وظائف مشابهة

تقدم للوظيفة الآن