Overview
tvScientific is the first and only CTV advertising platform purpose-built for performance marketers. They are seeking a Senior Site Reliability Engineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven GitOps workflows.
Key Responsibilities
- Ensuring the reliability, availability, and performance of production infrastructure and platform services.
- Operating and scaling Kubernetes platforms, including governance and support for multi-tenant workloads.
- Managing GitOps-based deployment workflows using ArgoCD and Helm.
- Driving infrastructure provisioning and change management through Terraform/Terragrunt.
- Building and supporting CI/CD automation and deployment workflows using GitHub Actions.
- Leading incident response efforts, root cause analysis, and post-incident improvement initiatives.
- Reducing operational toil through scripting, tooling, and process automation.
- Advancing observability practices across logs, metrics, traces, dashboards, and alerting.
- Supporting secure secrets integration, IAM-aware operations, and platform guardrails.
- Partnering closely with application, security, and platform teams to improve reliability and delivery outcomes.
Requirements
- 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure.
- Strong hands-on experience operating AWS in production environments.
- Deep expertise in Kubernetes, cluster operations, and platform administration.
- Proven experience with Kubernetes multi-tenancy, including namespaces and RBAC.
- Experience implementing and operating ArgoCD in a GitOps delivery model.
- Strong hands-on experience with Helm.
- Strong experience with Terraform/Terragrunt for infrastructure provisioning.
- Solid scripting and automation skills using Bash and/or Python.
- Experience building and maintaining CI/CD pipelines, ideally using GitHub Actions.
- Strong troubleshooting skills across Linux, containers, IAM, networking, and distributed systems.
- Experience with monitoring, alerting, and observability in production environments.
- Bachelor’s degree in computer science, engineering, or a related field.
- Demonstrated ability to use AI to improve workflow quality and speed.
- High integrity and ownership regarding sensitive data and accountability.
Benefits
The listed base salary range for this position is $139,764 — $287,749 USD, and the position is eligible for equity.
Location
This is a remote position available to US-based applicants only.
How to Apply
Details on how to apply are not provided in the listing.
Deadline
No deadline is mentioned.