GuideWell Logo

GuideWell

Cloud Operations Engineer

Posted Yesterday
Remote or Hybrid
Hiring Remotely in United States
109K-178K Annually
Senior level
Remote or Hybrid
Hiring Remotely in United States
109K-178K Annually
Senior level
Operate and improve production AWS workloads through automation, monitoring, patching, backups, certificate renewal, disaster recovery, and restore testing. Define operational readiness standards, manage SLOs and error budgets, respond to incidents, participate in on-call rotation, troubleshoot Windows and Linux compute, and create self-healing workflows. Partner with infrastructure teams to transition migrated workloads into steady-state support while maintaining runbooks and operational documentation.
The summary above was generated by AI

 

We're moving workloads to AWS, and this is the person who keeps them running afterward: patched, backed up, monitored, and recoverable. The Cloud Operations Engineer owns that ongoing operational work and builds automation so it doesn't keep growing headcount as the estate grows.

This role is a standard business-hours on eastern standard time zone engineering role with a shared on-call rotation, not a NOC or shift job. The focus is building automation and operational standards, not watching dashboards.

What You Will Be Doing

 

  • Build and maintain monitoring and alerting for migrated workloads. Alert on conditions that require action.

  • Automate patching, backup validation, certificate renewal, and other recurring operational tasks.

  • Define and enforce operational readiness criteria a workload must meet before it goes live in cloud.

  • Build and test disaster recovery runbooks. Validate recovery time and recovery point targets through failover exercises.

  • Respond to incidents, participate in the on-call rotation, and drive root cause to closure.

  • Convert manual procedures into executable automation and self-healing where the task is repeatable.

  • Track availability and operational health against defined targets. Report on recurring issues.

  • Manage backup, retention, and restore testing for cloud workloads.

  • Troubleshoot and maintain AWS compute, including Windows Server and Linux instances.

  • Partner with infrastructure operations to transition migrated workloads into steady-state support.

  • Maintain runbooks and operational documentation a new on-call engineer can use unaided.

What We Require

  • 6+ years related work experience in infrastructure operations or systems engineering with 4+ years experience operating cloud workloads 
  • 1+ years direct supervisory/management experience 
  • Related Bachelor's degree required 
  • Strong hands-on AWS experience running production workloads. 
  • Proficiency in a programming or scripting language such as Python or Go. 
  • Hands-on Terraform and CI/CD pipeline experience. 
  • Experience defining and operating against SLOs and error budgets. 
  • Observability tooling experience with CloudWatch, Prometheus, Grafana, Datadog, or equivalent. 
  • Incident management and on-call experience in a production environment. 
  • Strong hands-on AWS experience running production workloads. 
  • Proficiency in a programming or scripting language such as Python or Go. 
  • Hands-on Terraform and CI/CD pipeline experience. 
  • Experience defining and operating against SLOs and error budgets. 
  • Observability tooling experience with CloudWatch, Prometheus, Grafana, Datadog, or equivalent. 
  • Incident management and on-call experience in a production environment. 
  • Proficiency in a programming or scripting language such as Python or Go. 
  • Hands-on Terraform and CI/CD pipeline experience. 
  • Experience defining and operating against SLOs and error budgets. 
  • Observability tooling experience with CloudWatch, Prometheus, Grafana, Datadog, or equivalent. 
  • Incident management and on-call experience in a production environment. 
  • Kubernetes or EKS in production.

 

What We Prefer 

 

  • Multi-region or disaster recovery design, including failover testing. 

  • Healthcare, financial services, or other regulated industry experience. 

  • AWS GovCloud experience. Chaos engineering or resilience testing. 

  • AWS certification such as DevOps Engineer, SysOps Administrator, or Solutions Architect. 
     

General Physical Demands

 

  • Exerting up to 10 pounds of force occasionally to move objects. 

  • Jobs are sedentary if traversing activities are required only occasionally. 

  • This position offers a hybrid work arrangement which combines flexibility with in office engagement. Team schedules are determined based on role requirements and business needs. Travel to and from our corporate offices may be required.
     

What We Offer
As a Florida Blue employee, you will be at the heart of GuideWell’s vision – to lead the nation in transforming health through compassionate, connected, and technology-enabled care that delivers personalized value and empowered living. 
To support your wellbeing, comprehensive benefits are offered. As an employee, you will have access to: 
 

  • Medical, dental, vision, life and global travel health insurance
  • Income protection benefits: life insurance, short- and long-term disability programs
  • Leave programs to support personal circumstances
  • Retirement Savings Plan including employer match
  • Paid time off, volunteer time off, 10 holidays and 2 well-being days
  • Additional voluntary benefits available; and a comprehensive wellness program

Employee benefits are designed to align with federal and state employment laws. Benefits may vary based on the state in which work is performed. Benefits for intern, part-time and seasonal employees may differ.
To support your financial wellbeing, we offer competitive pay as well as opportunities for incentive or commission compensation. We also conduct regular annual reviews with pay for performance considerations for base pay increases. 

Typical Annualized Offer/Hiring Range: $109,300 - $136,600
Annualized Salary Range: $109,300 - $177,600
Final pay will be determined with consideration of market competitiveness, internal equity, and the job-related knowledge, skills, training, and experience you bring.
We are an Equal Employment Opportunity employer committed to cultivating a work experience where everyone feels like they belong and can perform at their best in pursuit of our mission. All qualified applicants will receive consideration for employment.

Similar Jobs

21 Days Ago
In-Office or Remote
113K-193K Annually
Senior level
113K-193K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Leads the architecture, development, deployment, and operation of enterprise AI Ops solutions, including RAG pipelines, agentic workflows, and AI-powered applications. Improves SRE practices, leads incident response, automates cloud infrastructure with Terraform and GitHub Actions, and ensures reliability, scalability, security, and responsible AI. Designs and operates software using Python and Node.js, manages Kubernetes environments, mentors engineers, and leads globally distributed technical teams.
Top Skills: Agentic AiAWSAzureCi/CdGCPGenerative AiGithub ActionsInfrastructure As CodeKubernetesNode.jsObservabilityPythonRetrieval Augmented Generation (Rag)Site Reliability Engineering (Sre)TerraformVector Search
3 Days Ago
In-Office or Remote
92K-164K Annually
Senior level
92K-164K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Manage and improve Azure cloud infrastructure and Windows/Linux server environments. Responsibilities include architecture changes, application installation, performance troubleshooting, maintenance scheduling, security compliance, centralized logging, access controls, penetration-test remediation, disaster recovery validation, monitoring, network administration, automation, reporting, and operational documentation. The role supports 24x7 production environments and requires collaboration across cloud, application, database, and infrastructure operations.
Top Skills: AclsApplication InsightsAzureAzure DevopsAzure Front DoorAzure Key VaultAzure MonitorAzure Traffic ManagerAzure Web ServicesDnsFirewallsGitIdentity ServicesJenkinsKubernetesLinuxNatNetwork WatcherOracle LinuxPerlPythonRed Hat LinuxSharepointShellTcp/IpTerraformVMwareWindows DesktopWindows Server
Yesterday
Remote or Hybrid
United States
109K-178K Annually
Mid level
109K-178K Annually
Mid level
Healthtech
Owns cloud financial management across AWS and Azure by analyzing spend, designing allocation and tagging models, building showback and forecasting, managing commitments, and executing optimization. Partners with engineering, finance, and executive teams to embed cost efficiency into architecture, provisioning, and CI/CD. Responsibilities include anomaly detection, rightsizing, storage tiering, idle resource cleanup, budgeting, migration planning, and cost reporting.
Top Skills: AWSAws Cost And Usage Report (Cur)Aws Cost ExplorerAws OrganizationsAws Savings PlansAzureAzure SubscriptionsCi/CdCloudabilityFlexeraFocus SpecificationHarnessPythonReserved InstancesSQLTerraform

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account