Raydar Logo

Raydar

Site Reliability Engineer

Posted 3 Days Ago
Remote
Hiring Remotely in United States
160K-210K Annually
Senior level
Remote
Hiring Remotely in United States
160K-210K Annually
Senior level
Own and improve CI/CD, developer tooling, agent harnesses, Kubernetes infrastructure, observability, cost management, and incident-response practices. Build platform capabilities that reduce engineering toil and improve delivery speed for conversational AI products. Participate in an on-call rotation for core infrastructure and help shape reliable, scalable systems across a fully remote engineering organization.
The summary above was generated by AI

About the company

Our client is a Series B conversational AI company building voice agents for customer service. Its platform has been running in production since 2020 and handles millions of calls each month for enterprise customers. A small, senior engineering team works remotely across North America to build and improve the systems behind these interactions.

The role / why it matters

As Site Reliability Engineer, you will build and operate the platform that helps engineering teams deliver reliable AI products at scale. You will own critical infrastructure domains across CI/CD, developer experience, observability, and agent harness engineering, with room to shape how the platform evolves.

What you'll do

- Own and improve CI/CD pipelines, including caching, architecture, developer self-service, and deployment workflows.

- Build developer tools and platform capabilities that reduce toil and help engineering teams ship faster.

- Extend the agent harness, including continuous integration, sandboxes, guardrails, and validation for autonomous agents.

- Operate Kubernetes-based cloud infrastructure and improve cost management, reliability, and observability.

- Develop monitoring, alerting, and incident-response practices across logs, metrics, and traces.

- Participate in an on-call rotation for core infrastructure; non-business-hours pages are rare.

What we're looking for

- Six or more years in software development enablement roles such as SRE, platform engineering, or DevEx.

- Experience owning CI/CD platforms end to end, including caching, system design, and developer self-service.

- Hands-on experience with TypeScript or Node.js, Python, Terraform, and Kubernetes; familiarity with Helm.

- Practical experience with observability, including logs, metrics, tracing, monitoring, alerting, and incident management.

- Understanding of LLMs and experience using AI development tools with sound judgment about where they help.

- Experience on fully remote teams and the ability to pass a software-engineer-oriented technical screen.

Bonus points

- Experience building platforms for autonomous AI agents, including harness engineering.

Compensation and benefits

The base salary range is $160,000 to $210,000 USD, plus competitive equity. Canadian compensation bands are lower. Benefits include medical, dental, and vision coverage, flexible vacation, a monthly wellness stipend, and a technology and learning stipend.

Location / work model

This is a fully remote role for candidates based in Canada or the United States who can work U.S. time zones. New visa sponsorship is not available; visa transfers may be considered for exceptional senior candidates.

Similar Jobs

10 Days Ago
Easy Apply
Remote
United States
Easy Apply
223K-380K Annually
Expert/Leader
223K-380K Annually
Expert/Leader
Cloud • Security • Software • Cybersecurity • Automation
Provide technical direction for GitLab Dedicated, a managed single-tenant SaaS platform. Lead architecture and transformation across resilience, failover, tenant orchestration, change management, automation, and platform integrations. Identify systemic reliability and scalability risks, establish reusable platform patterns, strengthen service ownership, and guide cross-team technical decisions. Mentor senior engineers and advance engineering excellence across the organization.
Top Skills: Cloud InfrastructureDevsecopsDistributed SystemsGoInfrastructure As CodeObservabilityPythonRuby
Yesterday
In-Office or Remote
Entry level
Entry level
Information Technology • Legal Tech • Analytics
Designs, builds, and operates reliable cloud infrastructure across AWS and Azure. Responsibilities include developing Terraform infrastructure as code, managing EKS and AKS Kubernetes clusters, maintaining CI/CD pipelines, improving security and cost efficiency, troubleshooting distributed infrastructure issues, implementing monitoring and alerting, and collaborating with engineering and architecture teams on platform standards and documentation.
Top Skills: Amazon CloudwatchAmazon EksAWSAzureAzure AdAzure AksAzure MonitorBashCi/CdDnsGithub ActionsGoGrafanaIamKubernetesLoad BalancersNetwork Security GroupsOidcPrometheusPythonSecurity GroupsTerraformVnetsVpc
4 Days Ago
In-Office or Remote
80K-110K Annually
Mid level
80K-110K Annually
Mid level
Software • PropTech
Responds to production incidents across AWS and Kubernetes, diagnosing and remediating infrastructure, networking, database, cache, and deployment issues. Builds monitoring, observability, synthetic and load tests, SLOs, and reliability improvements. Maintains Terraform, Kubernetes, CI/CD, IAM, autoscaling, and security configurations; leads postmortems and incident follow-up. Participates in on-call operations and collaborates with engineering teams to improve application performance and reliability.
Top Skills: .NetAmazon CloudwatchAmazon Ec2Amazon EksAmazon LinuxAmazon RdsAmazon S3AnsibleApplication Load BalancerAWSAzureBashCloudflareGithub ActionsGitlab CiGoogle Cloud PlatformGrafanaJavaScriptKubernetesLokiMemcachedMimirMySQLNetwork Load BalancerNginxOpentelemetryPagerdutyPHPPostgresPrometheusPythonRedisRuby On RailsRustTempoTerraformTraefikTypescriptUbuntu

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account