Overview
The Lead Site Reliability Engineer role at Juniper Square is focused on driving the technical direction for infrastructure systems, ensuring reliability, and improving operational excellence within a cloud environment. The company aims to unlock the potential of private markets through technology and data services.
Key Responsibilities
- Own and drive the technical direction for infrastructure systems; make architectural decisions balancing reliability, scalability, and cost.
- Design systems of moderate to high complexity and conduct architectural reviews while advancing organization-wide design patterns.
- Establish reliability postures for team-owned services, setting SLOs and monitoring SLAs.
- Lead incident response, debugging complex multi-service issues to prevent recurrence.
- Act as DRI for SRE projects involving cross-team collaboration, ensuring timely delivery and effective scope management.
- Identify and resolve project risks proactively.
Requirements
- 7-10 years of experience in Site Reliability Engineering or related fields in a production cloud environment.
- 5+ years with AWS cloud services and managing Linux environments.
- Proficiency in Infrastructure-as-Code (Terraform, CloudFormation) and cloud security best practices.
- Experience with Kubernetes and PostgreSQL in production.
- Strong programming skills, preferably in Python or Go.
- Experience with observability tools and CI/CD pipeline management.
Benefits
- High-impact role influencing financial technology.
- Growth potential in shaping the SRE practice.
- Collaborative culture valuing quality and ownership.
- Competitive compensation and benefits package.
Location
This role is based in India, offering remote work opportunities.
How to Apply
Please submit your application through the company’s career page.