Overview
The role involves designing, building, scaling, and maturing the multi-cloud platform for hosting internal and external services. The Senior Site Reliability Engineer will work within the SRE team to ensure the reliability of the global Elastic infrastructure.
Key Responsibilities
- Lead technical initiatives to automate system engineering efforts for infrastructure reliability.
- Grow global Platform infrastructure to meet scaling demands by developing and maintaining software, tooling, and automations.
- Promote collaboration and operational excellence within the team.
- Respond to and prevent customer impacts during major incidents, participating in an on-call rotation.
Requirements
- Experience focusing on platform reliability with a customer-first approach.
- Background in software engineering with knowledge of public cloud and managed Kubernetes services.
- Ability to communicate inclusively and build strong partner and team relationships.
Bonus Points
- Experience operating a SaaS product in a public cloud using Infrastructure-as-Code tooling.
- Experience with Kubernetes infrastructure across multiple cloud providers.
- Proficiency in programming, particularly with Golang.
- Experience with containerized services like Docker.
- Familiarity with alerting and incident management systems.
- Professional skills in Linux system administration.
- Experience with the Elastic Stack.
- Experience working in a globally distributed team.
Benefits
- Competitive pay based on work rather than previous salary.
- Health coverage for employees and their families in many locations.
- Flexible working locations and schedules.
- Generous vacation days.
- Financial donation matching and volunteer project support.
- Minimum of 16 weeks parental leave.
Location
Remote within Ireland.
How to Apply
Please apply through the application process outlined on the company's career page.
Deadline
No specific deadline mentioned.