You'll play a key role in enhancing our infrastructure as a Site Reliability Engineer. Collaborating with cross-functional teams, you will implement best practices and automate processes to ensure the reliability and efficiency of our systems.
Responsibilities
Manage and optimize cloud infrastructure.
Collaborate with software engineers on deployment.
Enhance system observability through telemetry.
Automate processes to reduce manual intervention.
Conduct risk assessments and vulnerability management.
Work on scaling systems meet growing demands.
Craft and maintain documentation for infrastructure changes.
Participate in on-call rotations for production support.
Requirements
Education
Bachelor's degree in Computer Science or related field
Experience
5+ years of experience in site reliability engineering or related areas
Technical Skills
Infrastructure as Code (Terraform)
Database Management (SQL, NoSQL)
Soft Skills
Problem-solving
Interpersonal Skills
Certifications
Google Professional Cloud Architect
Certified Kubernetes Administrator (CKA)
Languages
English: Fluent
Advantageous
Hands-on experience with monitoring systems: Experience configuring Prometheus, Grafana, or similar tools.
Experience with Agile methodologies: Familiarity with Kanban and Scrum processes.
Benefits
Health, dental, and vision insurance
Employee stock purchase plan
Remote work options where applicable
Wellness programs for employees
Company Culture
Innovation: We encourage creativity and innovative thinking across all levels.
Community Engagement: Our team actively participates in CSR initiatives and community service.
Employee Empowerment: We empower our employees by providing them with autonomy and accountability.
Status: Open
Other Jobs in Information Technology (IT) and Software Development