Site Reliability Engineer (SRE) – AWS, DevOps & Cloud Operations | Riyadh, Saudi Arabia

Website kbctechnologies
2. Location
Riyadh, Saudi Arabia
3. Job Category
- Information Technology (IT) & Software
- Telecommunications
- Engineering & Technical
- Others / Miscellaneous
4. Job Overview
We are hiring an experienced Site Reliability Engineer (SRE) to ensure high availability, performance, scalability, and reliability of enterprise production systems in Riyadh, Saudi Arabia. The role focuses on cloud operations, infrastructure automation, observability, incident management, and continuous improvement of production environments.
The ideal candidate will have strong SRE, DevOps, or production support experience with AWS, together with hands-on expertise in Kubernetes, Docker, Linux, networking, scripting, monitoring, and incident management. With 8+ years of experience preferred, this position offers opportunities for career growth, professional development, technical training, and cloud and SRE certification.
5. Key Responsibilities
- Build and manage scalable AWS cloud infrastructure using Terraform and CloudFormation.
- Implement and maintain monitoring and observability solutions using Prometheus, Grafana, Splunk, or Datadog.
- Manage production incidents and coordinate incident response activities.
- Conduct root cause analysis (RCA) and implement corrective and preventive measures.
- Define and monitor SLOs, SLIs, and error budgets.
- Participate in on-call support and production operations.
- Automate repetitive operational processes to improve system reliability and reduce MTTR.
- Manage Kubernetes and Docker environments.
- Administer and troubleshoot Linux-based production systems.
- Support CI/CD pipelines and deployment processes.
- Develop and maintain operational runbooks and documentation.
- Follow ITIL processes and established production support practices.
- Identify opportunities to improve system uptime, scalability, performance, and operational efficiency.
- Collaborate with technology teams to strengthen cloud infrastructure and production reliability.
6. Requirements & Qualifications
Experience
- 8+ years of experience preferred in SRE, DevOps, production support, or related infrastructure roles.
- Strong hands-on experience with AWS and enterprise production environments.
- Proven experience managing reliable and scalable cloud infrastructure.
Mandatory Technical Skills
- AWS cloud infrastructure.
- Kubernetes.
- Docker.
- Linux.
- Networking fundamentals.
- Python, Bash, or Go scripting.
- Monitoring and observability.
- Incident management and RCA.
- Production support and reliability engineering.
Preferred Skills
- Terraform and/or CloudFormation.
- Jenkins and/or GitLab CI/CD.
- ServiceNow and ITSM.
- Cloud security tools.
- Cisco and enterprise technology environments.
- SLOs, SLIs, error budgets, and reliability engineering practices.
Professional Skills
- Strong analytical and troubleshooting capabilities.
- Effective incident response and problem-solving skills.
- Strong communication and collaboration abilities.
- Ability to work effectively in production and on-call environments.
- Strong focus on automation, reliability, and continuous improvement.
7. Salary, Benefits & Career Growth
Career Growth & Professional Development
Salary information was not provided for this position.
The role offers opportunities for:
- Career growth in Site Reliability Engineering, DevOps, AWS, and cloud operations.
- Professional development through hands-on experience with enterprise production environments.
- Technical upskilling in Kubernetes, infrastructure as code, observability, automation, and cloud security.
- Opportunities to strengthen expertise in reliability engineering practices, incident management, and production operations.
- Continued development toward relevant cloud, DevOps, Kubernetes, and SRE certifications.
Benefits: Specific benefits were not provided in the job announcement.
8. Application Process
Interested candidates should submit their updated CV through the official recruitment contact.
Application Process:
- Prepare an updated CV highlighting SRE, DevOps, AWS, cloud infrastructure, and production support experience.
- Highlight relevant experience with Kubernetes, Docker, Linux, Terraform, CI/CD, monitoring, and incident management.
- Apply through the official job link and click Apply Now, or send your CV directly to the HR email below.
9. HR Email for Application
Send your updated CV directly
To apply for this job email your details to Hashim@kbctechnologies.com
