As a Site Reliability Engineer, you'll play a critical role in managing services and maintaining system health. You will work closely with cross-functional teams to automate processes and ensure seamless operations.
Responsibilities
Design and implement monitoring solutions to track system health.
Leverage automation to reduce time spent on repetitive manual tasks.
Engage in capacity planning to facilitate system growth.
Work with on-call teams to troubleshoot and resolve live issues swiftly.
Contribute to the development of internal tools for improved team workflows.
Requirements
Education
Bachelor's degree in Information Technology or related field
Experience
4+ years of relevant experience in cloud environments
Technical Skills
Monitoring Tools (Prometheus, Grafana)
Scripting (Python, Bash)
Soft Skills
Teamwork
Certifications
Google Cloud Professional DevOps Engineer
Languages
English: Fluent
Advantageous
Knowledge of Infrastructure as Code tools: Experience with tools such as Terraform or CloudFormation.
Familiarity with CI/CD solutions: Understanding of Continuous Integration and Continuous Deployment practices.
Benefits
Competitive salary package
Health and wellness benefits
Flexible working arrangements
Opportunity for growth and development
Company Culture
Collaborative Work Environment: We promote a culture of collaboration and open communication.
Innovation-driven: We encourage innovative thinking and support professional growth.
Diversity and Inclusion: A diverse workforce and inclusive environment are at the core of our values.
Status: Open
Other Jobs in Information Technology (IT) and Software Development