We are seeking a dedicated Site Reliability Engineer to enhance our team in Durban. The successful candidate will be responsible for ensuring the reliability and performance of our cloud infrastructure, striving for seamless service delivery.
Responsibilities
Ensure system reliability and availability through comprehensive monitoring.
Develop automation scripts to simplify operations.
Work closely with developers to identify and resolve code issues impacting reliability.
Develop and refine incident response protocols.
Assist in capacity planning and performance optimization.
Create documentation for system processes and incident handling.
Review and propose improvements to existing infrastructure.
Requirements
Education
Bachelor's degree in Information Technology or related field
Relevant certifications in cloud or DevOps are highly regarded
Experience
3-5 years of experience in a similar role
Technical Skills
Docker
Infrastructure as Code (Terraform)
Soft Skills
Problem-solving
Communication
Certifications
Google Professional Cloud Architect
Certified ScrumMaster (CSM)
Languages
English: Fluent
Advantageous
Familiarity with CI/CD tools: Experience with tools to enhance continuous integration and deployment processes.
Knowledge of security best practices: Understanding of security protocols and best practices for system integrity.
Benefits
Comprehensive medical aid and insurance
Retirement fund contributions
Flexible work schedules
Professional development opportunities
Company Culture
Team Spirit: We believe in fostering a collaborative environment where teamwork drives success.
Innovation Driven: Our culture encourages creativity and innovative thinking among employees.
Emphasis on Continuous Learning: Employees have access to resources for ongoing professional development.
Status: Open
Other Jobs in Information Technology (IT) and Software Development