Senior Site Reliability Engineer

Zocdoc · USA Remote · Engineering

Posted 2026-09-16

Apply for this role →

Your Impact on our Mission:

Zocdoc is looking for a Senior Site Reliability Engineer to help develop, monitor, and maintain our distributed production systems. You’ll be challenged with building frameworks and processes for ensuring uptime for our patients and providers in a constantly changing environment. You’ll work with distributed systems and microservices, leveraging many interconnected services in AWS Cloud. We’re looking for someone who loves challenging the status quo and strives to make everything they touch safer, more secure, faster, and easier to maintain.

You’ll enjoy this role if you are…

Passionate about ensuring complex systems never skip a beat

Motivated to learn new technologies, design patterns, and work in the cloud

Comfortable in an outage situation and believe in blameless post-mortems

Excited to work in a highly collaborative environment with diverse individuals and numerous product development teams to improve future uptime

Enforce a culture around strong DevOps and where product teams share a big role in site reliability and first response

Autonomous, individually accountable, and always pushing to improve

A believer that diverse and inclusive teams and cultures are non-negotiable

Your day to day is…

Monitoring and maintaining complex cloud-based infrastructure, systems, and services and ensuring their uptime to help millions of patients get the care they need

Automating and developing our tooling, processes, and infrastructure to speed up development and make them repeatable and error-proof

Supporting our large product engineering org with their scaling, performance, and uptime needs as well as helping diagnose and debug production related issues

Analyzing and performance tuning systems, code, and networking for scaling and optimal operation

Working with cutting edge GenAI tools and technology

You’ll be successful in this role if you have…

5+ years of supporting consumer facing web application production environments and systems in a Site Reliability Engineering or Production Engineering role

2+ years of on-call experience in a 24/7 cloud-based production environment

2+ years of experience in managing and supporting modern cloud-based environments and infrastructure like AWS/GCP, Docker, Kubernetes, etc.

Experience with edge technologies such as load balancers, reverse proxies, web application firewalls, routing, etc.

Deep understanding of protocols such as TCP/IP, HTTP/HTTPS, TLS, DNS, NTP

A Bachelor’s degree in Computer Science, Computer Engineering, or equivalent engineering experience is a plus, but not required

Benefits

Flexible, hybrid work environment at our convenient Soho location (If based in NYC)

Unlimited Vacation

100% paid employee health benefit options (including medical, dental, and vision)

Commuter Benefits

401(k) with employer funded match

Corporate wellness program with Wellhub

Sabbatical leave (for employees with 5+ years of service)

Competitive paid parental leave and fertility/family planning reimbursement

Cell phone reimbursement

Catered lunch everyday along with beverages and snacks

Employee Resource Groups and ZocClubs to promote shared community and belonging

Great Place to Work Certified

Zocdoc is committed to fair and equitable compensation practices. Salary ranges are determined through alignment with market data. Base salary offered is determined by a number of factors including the candidate’s experience, qualifications, and skills. Certain positions are also eligible for variable pay and/or equity; your recruiter will discuss the full compensation package details.

NYC Base Salary Range

$180,000—$220,000 USD

Apply for this role →

← Back to all jobs