Join us as a Site Reliability Engineer, where you will ensure the reliability of our services by leveraging best practices in disciplines like DevOps and cloud computing. Apply your skills in troubleshooting and system design.
Responsibilities
Establishing SRE standards and best practices for the engineering team.
Automating infrastructure as code using tools such as Terraform or CloudFormation.
Creating and maintaining documentations for operational procedures.
Collaborating with security teams to ensure secure cloud environments.
Monitoring system metrics and improving system reliability.
Deploying and maintaining CI/CD pipelines in cloud environments.
Supporting capacity planning for cloud resources.
Engaging in performance testing and tuning across production systems.
Requirements
Education
Bachelor's degree in Computer Science or related field
Master's degree is an advantage
Experience
3+ years of experience in site reliability engineering or DevOps.
Technical Skills
Linux Systems Administration
Cloud Platforms
Python
Docker
Kubernetes
CI/CD Tools
Soft Skills
Analytical Thinking
Team Collaboration
Effective Communication
Languages
English: Fluent
Advantageous
DevOps Certifications: Certifications such as Google Professional Cloud Architect or AWS Certified SysOps Administrator are preferred.
Experience with Microservice Architecture: Knowledge of microservices and service-oriented architecture design.
Benefits
Competitive salary package with performance bonuses
Health and wellness benefits
Flexible work hours with remote work options
Career growth opportunities and training
Company Culture
Innovation: Encouraging creative thinking and new ideas to solve problems.
Collaboration: Promoting teamwork and open communication amongst colleagues.
Diversity: Fostering an inclusive environment where diverse perspectives are valued.
Status: Open
Other Jobs in Information Technology (IT) and Software Development