Overview
The role is for a Senior Cloud Site Reliability Engineer at NiCE, where you will be responsible for running the production environment, improving reliability, and optimizing system performance.
Key Responsibilities
- Monitor availability and take a holistic view of system health.
- Build software and systems to manage platform infrastructure and applications.
- Measure and optimize system performance while improving services through rigorous testing and release procedures.
- Gather and analyze metrics to assist in performance tuning and fault finding.
- Create sustainable systems through automation and participate in system design consulting.
- Provide primary operational support for large distributed software applications.
Requirements
- 3-6 years of experience in systems engineering, automation, and reliability.
- Proficiency in at least one programming language (e.g., Python, Go, Java, C#) and experience with scripting languages (e.g., Bash, PowerShell).
- Deep understanding of cloud computing platforms (e.g., AWS) and their services (e.g., EC2, Lambda).
- Experience with infrastructure as code tools like CloudFormation and Terraform.
- Knowledge of CI/CD concepts and tools (e.g., Jenkins, GitLab CI/CD).
- Strong knowledge of containerization technologies (e.g., Docker, Kubernetes).
- Experience with monitoring tools (e.g., Prometheus, Grafana).
- Excellent problem-solving skills and incident management experience.
Benefits
NICE offers a hybrid work model (2 days in office, 3 days remote) to provide flexibility and promote collaboration.
Location
United Kingdom - remote work allowed.
How to Apply
Interested candidates should apply directly through the company’s careers page.
Deadline
No specific deadline mentioned.