Site Reliability Engineer

Raydar · United States

SeniorDevOps & InfraRemotePosted today
Pay — up front
$160k–$210k
Level
Senior
Location
Remote
Apply via
Workable
Day 1 of 7 on the board — 7 days of runway left
[RECON REPORT]

The employer, decoded.

What they do, how they make money, why this role exists.

Reading up on Raydar…

[HOW TO WIN THIS JOB 💡]

Show them you get it.

One proof-of-work idea, an hour to build, that shows this employer you get their business.

Sketching a proof-of-work idea…

[SKIP THE PILE]

Reach the hiring manager.

Who likely hires for this role, a LinkedIn search pre-filled to find them, and a short intro written for this exact job — you send it yourself.

About the company

Our client is a Series B conversational AI company building voice agents for customer service. Its platform has been running in production since 2020 and handles millions of calls each month for enterprise customers. A small, senior engineering team works remotely across North America to build and improve the systems behind these interactions.

The role / why it matters

As Site Reliability Engineer, you will build and operate the platform that helps engineering teams deliver reliable AI products at scale. You will own critical infrastructure domains across CI/CD, developer experience, observability, and agent harness engineering, with room to shape how the platform evolves.

What you'll do

  • Own and improve CI/CD pipelines, including caching, architecture, developer self-service, and deployment workflows.
  • Build developer tools and platform capabilities that reduce toil and help engineering teams ship faster.
  • Extend the agent harness, including continuous integration, sandboxes, guardrails, and validation for autonomous agents.
  • Operate Kubernetes-based cloud infrastructure and improve cost management, reliability, and observability.
  • Develop monitoring, alerting, and incident-response practices across logs, metrics, and traces.
  • Participate in an on-call rotation for core infrastructure; non-business-hours pages are rare.

What we're looking for

  • Six or more years in software development enablement roles such as SRE, platform engineering, or DevEx.
  • Experience owning CI/CD platforms end to end, including caching, system design, and developer self-service.
  • Hands-on experience with TypeScript or Node.js, Python, Terraform, and Kubernetes; familiarity with Helm.
  • Practical experience with observability, including logs, metrics, tracing, monitoring, alerting, and incident management.
  • Understanding of LLMs and experience using AI development tools with sound judgment about where they help.
  • Experience on fully remote teams and the ability to pass a software-engineer-oriented technical screen.

Bonus points

  • Experience building platforms for autonomous AI agents, including harness engineering.

Compensation and benefits

The base salary range is $160,000 to $210,000 USD, plus competitive equity. Canadian compensation bands are lower. Benefits include medical, dental, and vision coverage, flexible vacation, a monthly wellness stipend, and a technology and learning stipend.

Location / work model

This is a fully remote role for candidates based in Canada or the United States who can work U.S. time zones. New visa sponsorship is not available; visa transfers may be considered for exceptional senior candidates.

Straight from the source — this role comes from Raydar's own hiring system (Workable), not a scraped repost. iRocket links you directly; we don't host the posting or the application.