IT Operations Engineer

Datamatics Technologies
الرياض, الرياض دوام كامل
نشر: 1448/1/25 | 2026/07/10 ينتهي: 1448/2/26 | 2026/08/09 ✨ وصف بالذكاء الاصطناعي
تقدم للوظيفة الآن

الوصف الوظيفي

Position Overview

We are looking for a skilled and proactive IT Operations Engineer to join our dynamic team. In this role, you will be responsible for designing, implementing, automating, and maintaining enterprise-grade IT infrastructure across both on-premises and cloud environments. Your expertise will be pivotal in ensuring the reliability, scalability, security, and high performance of our systems. The ideal candidate will have a strong background in infrastructure automation, configuration management, virtualization, cloud technologies, and monitoring solutions, enabling seamless operations and proactive issue resolution.

Key Responsibilities

As an IT Operations Engineer, your core responsibilities will include:

  • Infrastructure Design and Deployment: Architect, deploy, and maintain robust IT infrastructure solutions that span on-premises data centers and cloud platforms, ensuring alignment with business and technical requirements.
  • Automation and Orchestration: Develop and implement automation frameworks to streamline infrastructure provisioning, configuration management, and deployment processes, reducing manual intervention and enhancing operational efficiency.
  • Monitoring and Observability: Implement and manage enterprise-grade monitoring, logging, and observability solutions to proactively track system health, performance metrics, and availability, enabling rapid incident detection and resolution.
  • Virtualization and Cloud Management: Oversee virtualization platforms and private cloud environments, optimizing resource utilization, scalability, and performance while ensuring seamless integration with public cloud services.
  • Incident Management and Troubleshooting: Conduct thorough root cause analysis and troubleshoot complex infrastructure issues to minimize downtime and restore services promptly, adhering to best practices in incident management.
  • Infrastructure as Code (IaC) and Configuration Management: Champion IaC methodologies and configuration management tools to maintain consistency, reproducibility, and scalability across environments, while fostering collaboration with cross-functional teams.
  • Collaboration and Stakeholder Engagement: Work closely with DevOps, Cloud, Security, and Application teams to ensure cohesive system operations, align on technical strategies, and support business objectives through reliable and secure infrastructure.
  • Disaster Recovery and High Availability: Design, implement, and maintain robust backup, disaster recovery, and high availability solutions to safeguard critical systems and data against disruptions.
  • Compliance and Governance: Ensure adherence to enterprise security policies, governance frameworks, and operational standards, while maintaining comprehensive documentation, runbooks, and technical procedures for knowledge sharing and compliance audits.

Required Skills and Qualifications

To excel in this role, you must possess the following technical competencies and experience:

  • Log Management and Monitoring: Hands-on experience with ELK Stack (Elasticsearch, Logstash, Kibana) or similar solutions for centralized log management, monitoring, and observability.
  • Infrastructure as Code (IaC): Proficiency in Terraform or Ansible for automating infrastructure provisioning, configuration, and management.
  • Virtualization Platforms: Experience with VMware vCenter or OpenStack for managing virtualized environments and private cloud infrastructure.
  • Operating Systems: Strong knowledge of Linux and Windows Server administration, including system configuration, patch management, and performance optimization.
  • Scripting and Automation: Expertise in Shell Scripting or Python for developing automation scripts and tools to enhance operational efficiency.
  • Networking Fundamentals: Deep understanding of networking concepts such as TCP/IP, DNS, DHCP, VPN, and Load Balancing, with the ability to troubleshoot network-related issues.
  • Monitoring and Incident Management: Experience with system monitoring tools, performance tuning, and incident management processes to ensure optimal system performance and rapid issue resolution.
  • Troubleshooting and Problem-Solving: Exceptional analytical and troubleshooting skills, with the ability to diagnose complex infrastructure issues and implement effective solutions.

Preferred Skills and Qualifications

While the following skills are not mandatory, candidates with experience in these areas will be highly valued:

  • Containerization and Orchestration: Familiarity with Docker or Kubernetes for containerized application deployment and management.
  • Cloud Platforms: Experience with Microsoft Azure, AWS, or Google Cloud Platform (GCP) for deploying and managing cloud-based infrastructure.
  • Advanced Monitoring Tools: Knowledge of Prometheus, Grafana, or Zabbix for enhanced infrastructure monitoring and visualization.
  • Version Control and CI/CD: Proficiency with Git and continuous integration/continuous deployment (CI/CD) pipelines for streamlined software delivery and infrastructure updates.
  • IT Service Management: Understanding of ITIL processes and service management best practices to align IT operations with business needs.
  • Security and Compliance: Awareness of security hardening techniques, identity management solutions, and compliance frameworks such as ISO 27001 or NIST to ensure robust security postures.

Education and Experience

To qualify for this role, you should meet the following criteria:

  • A Bachelor’s or Master’s degree in Computer Science, Information Technology, Information Systems, or a related field.
  • A minimum of 6+ years of hands-on experience in IT Operations, Infrastructure Engineering, Systems Administration, Platform Engineering, or Cloud Operations.

Personal Attributes

We are seeking a candidate who demonstrates:

  • Technical Proficiency: A strong aptitude for learning new technologies and methodologies to stay ahead in a rapidly evolving IT landscape.
  • Collaboration and Communication: Excellent interpersonal skills with the ability to articulate technical concepts clearly to both technical and non-technical stakeholders.
  • Problem-Solving Mindset: A proactive approach to identifying and resolving operational challenges, coupled with a commitment to continuous improvement.
  • Documentation and Knowledge Sharing: The ability to create and maintain comprehensive operational documentation, ensuring knowledge transfer and operational transparency.
  • Adaptability: Flexibility to work in dynamic environments, including Agile, DevOps, or Site Reliability Engineering (SRE) teams, and adapt to evolving business and technical requirements.

يمكن أن يرتكب الذكاء الاصطناعي أخطاءً.

المصدر: لينكد إن ↗ • 16 مشاهدة

وظائف مشابهة

تقدم للوظيفة الآن