Site Reliability Engineer (SRE)
Wipro
الوصف الوظيفي
Role Summary
Ensure high availability, performance, and reliability of production systems through automation, cloud operations, monitoring, and incident management.
Key Responsibilities
- Build and manage scalable AWS cloud infrastructure using Terraform / CloudFormation
- Implement monitoring and observability using Prometheus, Grafana, Splunk, or Datadog
- Handle incidents, perform root cause analysis, provide on-call support, manage service level objectives (SLOs), service level indicators (SLIs), and error budgets
- Automate operations to improve efficiency, reliability, and mean time to recovery (MTTR)
- Manage Kubernetes, Docker, CI/CD pipelines, runbooks, and ITIL processes
Mandatory Skills
- Proven experience in SRE / DevOps / Production Support with AWS cloud experience
- Strong proficiency in Kubernetes, Docker, Linux, and networking fundamentals
- Expertise in scripting using Python, Bash, or Go
- Hands-on experience with monitoring tools and incident / RCA management
Preferred Skills
- Terraform / CloudFormation, Jenkins / GitLab CI/CD, ServiceNow / ITSM
- Cloud security tools and enterprise environment experience
Experience
5+ years in SRE / DevOps / Cloud / Production Support
يمكن أن يرتكب الذكاء الاصطناعي أخطاءً.
المصدر:
لينكد إن ↗
• 4 مشاهدة
وظائف مشابهة
K
Executive Assistant
Kinetic Business Solutions
A
Sales Manager
AVEVA Select Gulf
L
Communications Manager
Lucid Motors Middle …
S
E-Commerce Specialist
Segadty