BeyondTrust Logo

BeyondTrust

Staff Site Reliability Engineer

Posted 2 Days Ago
Remote
Hiring Remotely in United States
Senior level
Remote
Hiring Remotely in United States
Senior level
Lead architecture and delivery of highly available, resilient cloud and on‑prem systems for the Password Safe platform. Own platform engineering, CI/CD pipelines, IaC/GitOps, release orchestration, observability (metrics/logs/traces), chaos engineering, SLO/SLI definition, and core services. Mentor engineers, define SRE strategy, and drive reliability, security, and automation improvements.
The summary above was generated by AI

BeyondTrust is a place where you can bring your purpose to life through the work that you do, creating a safer world through our cybersecurity SaaS portfolio.

Our culture of flexibility, trust, and continual learning means you will be recognized for your growth, and for the impact you make on our success. You will be surrounded by people who challenge, support, and inspire you to be the best version of yourself.

The Role

We are seeking a Staff Site Reliability Engineer (SRE) to lead the evolution of the Password Safe platform, infrastructure, and deployment ecosystem. In this role, you will bridge the gap between systems engineering and software development, architecting highly availability highly resilient systems across both cloud and on-premises environments.

As a Staff level engineer, you will tackle complex technical challenges and serve as a technical leader, mentoring engineers and influencing our broader engineering roadmap. You will own the reliability, scalability, and efficiency of shared services, CI/CD pipelines, and platform engineering initiatives via an AI first operating model.

What You’ll Do

Infrastructure & Platform Engineering

  • Design, scale, and maintain highly available, secure, and resilient systems spanning both cloud (AWS/Azure) and on premises environments.
  • Champion platform engineering initiatives that reduce cognitive load for software engineers and accelerate velocity.
  • Own and optimize foundational services, including api gateways, service meshes, caches, configuration management, and secrets management.
  • Drive a "Everything as Code" culture, ensuring that all cloud and on-prem infrastructure, CI/CD pipelines, and configurations are declaratively defined, version-controlled, and deployed via automated GitOps workflows.

CI/CD & Automation

  • Standardize, secure, and optimize modern CI/CD pipelines to ensure safe, repeatable, and rapid code deployments.
  • Treat infrastructure as code (Terraform, OpenTofu, or Ansible), driving architectural patterns that eliminate configuration drift.

Reliability, Testing & Observability

  • Design and implement automated chaos engineering frameworks and disaster recovery simulations to proactively identify system weaknesses.
  • Architect and mature our telemetry stack (metrics, logs, traces) using tools like Grafana Cloud, Datadog, and OpenTelemetry to ensure deep system visibility.
  • Define, implement, and enforce Service Level Objectives (SLOs) and Service Level Indicators (SLIs) across critical applications.

Leadership

  • Partner with engineering leadership to define the long-term SRE strategy.
  • Raise the bar through mentorship and standards. Coach engineers on reliability practices, run design and incident reviews, and build documentation and tooling that makes reliability knowledge accessible.

What You’ll Bring

  • Experience: 7+ years of experience in SRE, DevOps, or Platform Engineering, with at least 2 years in a Senior or Staff level.
  • Systems Architecture: Proven track record of managing automation and infrastructure across both cloud and on-premises environments.
  • Containerization: Experience with Docker and Kubernetes (including cluster administration, networking, and security primitives).
  • Release Orchestration: Prior experience with release orchestration strategies such Canary or Blue Green models.
  • Software Engineering: Proficiency in at least one systems language (e.g. Go, Java, or C#) to build automation, tooling, and internal APIs.
  • Observability: Deep understanding of open telemetry and monitoring and observability best practices (e.g. Golden vs Red Signals, and App monitoring).
  • Everything as Code: Familiarity with GitOps workflows, Infrastructure and Config as Code best practices.

Nice To Have

  • Knowledge of UI automation testing.
  • Experience designing and testing microservice based applications.
  • Experience working with virtual machines and managing test environments.
  • Experience working in a continuous integration environment.
  • Deep understanding of Linux and Windows internals.
  • Experience migrating workloads seamlessly across on-prem and cloud environments.
  • Understanding of modern DevSec Ops practices.

Better Together

Diversity. Inclusion. They’re more than just words for us. They are the guiding values of how we build our teams, cultivate leaders, and create a culture where people feel connected.

We take care of our employees so they can take care of our customers. Customers who come from all walks of life just like us. We hire incredible people from diverse backgrounds because when we are different together, we are stronger together.

About Us

BeyondTrust is the global identity security leader protecting Paths to Privilege™. Our identity-centric approach goes beyond securing privileges and access, empowering organizations with the most effective solution to manage the entire identity attack surface and neutralize threats, whether from external attacks or insiders.

BeyondTrust is leading the charge in transforming identity security to prevent breaches and limit the blast radius of attacks, while creating a superior customer experience and operational efficiencies. We are trusted by 20,000 customers, including 75 of the Fortune 100, and our global ecosystem of partners.

Learn more at www.beyondtrust.com. 

#LI-DF1

BeyondTrust Aliso Viejo, California, USA Office

Aliso Viejo, United States

Similar Jobs

22 Days Ago
In-Office or Remote
Senior level
Senior level
Big Data • Information Technology • Software • Analytics • Energy
Manage and scale Enverus' global AWS infrastructure, automate deployments and CI/CD, ensure high uptime, collaborate with developers to enable zero-downtime releases, participate in on-call rotations, and improve operational practices.
Top Skills: AWSAzureC#Ci/CdCloudFormationGoKubernetesLinuxPythonTerraformWindows
Yesterday
Easy Apply
Remote
Easy Apply
126K-314K Annually
Senior level
126K-314K Annually
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Maintain and improve reliability, scalability, and automation for user-facing production systems. Build infrastructure tooling, operate Kubernetes-based services, write IaC, participate in on-call and incident response, and advance observability and runbooks to reduce toil and improve platform reliability.
Top Skills: AWSCi/CdGCPGitopsGoInfrastructure As Code (Iac)KubernetesKubernetes Operators/ControllersLoggingMetricsRubySlos/SlisTerraform
58 Minutes Ago
In-Office or Remote
Expert/Leader
Expert/Leader
Natural Language Processing • Software • Conversational AI
Design, build, and operate highly available, scalable infrastructure on GCP; automate CI/CD and IaC; implement monitoring and observability; drive incident response and post-mortems; optimize performance, cost, and reliability; eliminate operational toil and lead compliance initiatives.
Top Skills: Ci/CdClojureClojurescriptCloud MonitoringCloud RunCompute EngineContainersDatadogGkeGoogle Cloud Platform (Gcp)GrafanaKubernetesPciPrometheusPub/SubPulumiService MeshSocTerraform

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account