Join us as a Site Reliability Engineer, where you'll be hands-on in building and maintaining scalable systems. You will analyze system performance, implement improvements, and support the operational growth of our services in an agile environment.
Responsibilities
Design and implement infrastructure solutions.
Act as a point of contact for reliability-related incidents.
Work with platform teams to improve overall system accessibility.
Lead initiatives to reduce latency and improve response times.
Engage in capacity planning and forecasting.
Collect and analyze metrics to improve service performance.
Collaborate with security teams on compliance and best practices.
Requirements
Education
Bachelor's degree in Information Technology or related field
Experience
5+ years experience in DevOps.
Technical Skills
Terraform
Monitoring Tools (Prometheus, Grafana)
Soft Skills
Team Leadership
Adaptability
Certifications
Certified Kubernetes Administrator (CKA)
Languages
English: Fluent
Advantageous
Understanding of security practices: Knowledge in securing cloud and on-premise applications against vulnerabilities.
Project Management experience: Experience in managing projects, timelines, and team deliverables.
Benefits
Health, dental, and vision coverage
Pension fund with employer contributions
Flexible working arrangements
Employee assistance program
Company Culture
Continuous Development: We provide opportunities for professional growth and continued education.
Work-Life Balance: Ensure that employees have a healthy work-life balance while achieving their career goals.
Accountability: Every team member is accountable for their roles and contributes to the organization's success.
Status: Open
Other Jobs in Information Technology (IT) and Software Development