Senior Site Reliability Engineer

Movable Ink · Movable Ink - Toronto (Remote) · Engineering

Posted 2026-07-15

Apply for this role →

As one of our Senior Site Reliability Engineers, you will be 100% hands-on across infrastructure and software development. You will support and help inform the evolution of major systems within our multi-cloud, multi-region, active-active content serving platform that serves upwards of 25 Billion requests daily. Through a combination of technical expertise and  cross-team collaboration, you will help support the reliability initiatives and collaborate on the technical strategy that scales our platform to 50 Billion requests per day and beyond.

Responsibilities:

Improve the tooling and automation of our infrastructure to minimize manual work, increase performance, and decrease the frequency and severity of incidents

Build, maintain, and support core applications

Monitor our systems for capacity, performance, and troubleshoot issues

Partner with the rest of the SRE team to ensure smooth, continued delivery of our service to clients

Demonstrate a high level of  autonomy in anticipating, identifying, and addressing systemic weaknesses and opportunities for platform improvement

Qualifications:

Experience in Site Reliability or Software Engineering, building and maintaining scalable, resilient services.

Building the tooling and automation to manage those services, as well as investigating system and application metrics to diagnose and resolve performance issues.

4+ years experience as an SRE or Software Engineer, with a focus on Cloud platforms (AWS/GCP)

Experience architecting and leading large-scale observability platforms, including defining observability standards and SLO frameworks. We use Prometheus and Thanos with Grafana Alloy, Loki and Tempo

Experience and willingness to operate in an on-call environment, evaluating and improving monitoring and alerting systems, and developing run books to investigate and debug issues

Strong experience with infrastructure as code tools.  Terraform experience is a major plus

Kubernetes experience, including cluster operations, multi-tenancy strategies, and supporting teams on container orchestration best practices. We use EKS and GKE

Experience with one or more high level programming languages; NodeJS, Go, Ruby, Python, in addition Shell Scripting

Linux experience is a must

The base pay range for this position is $140K - 182K/year CAD, which can include additional bonus depending on the position ultimately offered, in addition to a full range of medical, financial, and/or other benefits. The base pay offered may vary depending on job-related knowledge, skills, and experience.

Apply for this role →

← Back to all jobs