As a Site Reliability Engineer, your mission will be to enhance our systems with a focus on availability and performance. Bring your strong problem-solving skills to diagnose issues and automate solutions, ensuring seamless operation.
Responsibilities
Develop strategies for reliability and system efficiency.
Conduct system health checks and identify areas for improvement.
Assist in capacity planning and performance analysis.
Be a key point of contact during incident management.
Continuously evaluate and integrate new technologies.
Requirements
Education
Bachelor's degree in Information Technology or a related field
Relevant advanced degree preferred
Experience
4+ years of experience in system administration
Technical Skills
Cloud Technologies
Monitoring Tools (Nagios, Prometheus)
Soft Skills
Analytical Thinking
Teamwork
Certifications
Google Professional Cloud Architect
Languages
English: Fluent
Advantageous
Experience in Site Reliability best practices: Familiarity with SRE practices to enhance system reliability.
Knowledge of load balancing and caching techniques: Understanding of how to implement these techniques for performance.
Benefits
Health, dental, and vision coverage
Employer matching contributions to retirement savings
Remote work opportunities available
Access to ongoing training and professional development programs
Company Culture
Continuous Learning: Encouragement to continuously develop skills and knowledge.
Focus on Work-Life Balance: We prioritize a healthy work-life balance for all employees.
Status: Open
Other Jobs in Information Technology (IT) and Software Development