Join our innovative team as a Site Reliability Engineer. You will be responsible for maintaining cloud service performance, reliability, and availability while fostering a culture of collaboration and proactive issue resolution.
Responsibilities
Architect, deploy, and maintain cloud infrastructure that meets performance and availability goals.
Automate critical incident response tasks to minimize downtime.
Utilize metrics and logs for troubleshooting and performance tuning.
Research and introduce innovative solutions for ongoing reliability improvements.
Facilitate training on infrastructure best practices for other teams.
Ensure compliance with security and operational standards.
Promote a site reliability culture across development teams.
Requirements
Education
Bachelor's degree in Information Technology, Computer Science, or equivalent
Master’s degree is preferred
Experience
5+ years of experience in cloud systems engineering or SRE
Technical Skills
Docker
Monitoring Tools
Soft Skills
Team-oriented mindset
Adaptability
Certifications
AWS Certified DevOps Engineer
Languages
English: Fluent
Advantageous
Understanding of networking and cybersecurity principles: Knowledge of how networking functions and principles of safeguarding systems.
Familiarity with microservices architecture: Experience working in environments that utilize microservices.
Benefits
Comprehensive health insurance plans
Training and certification reimbursements
Work-from-home options
Paid time off for personal days
Company Culture
Supportive Environment: We maintain a supportive atmosphere where employees feel valued and empowered.
Continuous Learning: We support ongoing training and skill development to promote career growth.
Work-life Balance: We understand the importance of balancing work commitments with personal life.
Status: Open
Other Jobs in Information Technology (IT) and Software Development