Private Health Management Logo

Private Health Management

AI Infrastructure Operations Engineer

Posted 5 Days Ago
Remote
Hiring Remotely in USA
120K-140K Annually
Mid level
Remote
Hiring Remotely in USA
120K-140K Annually
Mid level
Operate and scale an Azure-based AI platform (Companion) by maintaining AKS infrastructure, improving observability, handling incidents, ensuring security and operational hygiene, and building runbooks and processes to support production AI agent workloads and deployments.
The summary above was generated by AI

AI Infrastructure & Operations Engineer 

Location: Remote (U.S.) 
Reports To: Juan Sandoval-Tobias 

About Private Health Management 

Private Health Management (PHM) supports people with serious and complex medical conditions, helping them obtain the best possible medical care. We guide individuals and families to top specialists, advanced diagnostics, and personalized care. Trusted by healthcare providers and businesses, PHM offers independent, science-backed insights to help clients make informed decisions and access the best care. 

About the Role 

PHM is building and scaling Companion, an AI-enabled clinical platform operating in a high-trust healthcare environment where reliability, observability, and security are foundational requirements. The platform includes headless AI agents designed to support clinical and operational professionals by acting as intelligent workstations that integrate with enterprise applications and workflows. 

The AI Infrastructure & Operations Engineer will operationalize the platform so it runs reliably at production scale, helping ensure the systems behind Companion are observable, recoverable, secure, and maintainable as adoption grows. 

This role sits at the intersection of Kubernetes operations, AI platform reliability, observability engineering, and operational security. You will help evolve and maintain the Azure-based infrastructure stack while partnering closely with technology leadership, AI architects, and security stakeholders. This is a high-ownership role for someone who thrives in fast-moving environments, is comfortable operating with incomplete information, and enjoys building operational discipline around emerging AI systems. 

What You’ll Accomplish 

  • Establish operational reliability for Companion across AKS infrastructure, AI agent workloads, monitoring systems, and deployment pipelines.  
  • Build meaningful observability practices that help PHM understand platform behavior, usage trends, and operational risks before they become incidents.  
  • Create sustainable operational hygiene around patching, CVE remediation, secrets rotation, dependency management, and cloud maintenance cycles.  
  • Strengthen platform resilience, documentation, and operational processes so the environment can scale without relying on tribal knowledge.  

How You’ll Spend Your Days 

Operate and Improve Platform Reliability 

  • Monitor and maintain AKS infrastructure, AI agent workloads, deployment pipelines, and support Azure services.  
  • Investigate incidents, troubleshoot production issues, and improve platform resilience through better operational patterns and tooling.  
  • Support release operations and help ensure deployments remain stable, observable, and recoverable.  

Build Observability and Operational Insight 

  • Develop dashboards, alerts, logging patterns, and operational baselines using Azure Log Analytics and Application Insights.  
  • Identify system trends, performance bottlenecks, and emerging operational risks across infrastructure and AI workloads.  
  • Improve visibility into AI agent behavior, enterprise workflow integrations, latency patterns, and system health under real user load.  

Strengthen Security and Operational Hygiene 

  • Maintain operational cadence for dependency updates, CVE remediation, image signing, secrets rotation, and cluster patching.  
  • Support security-first infrastructure practices across Kubernetes, CI/CD pipelines, and Azure environments.  
  • Partner with security and engineering stakeholders to maintain compliance-aware operational practices in a HIPAA-regulated environment.  

Collaborate Across a Small, High-Ownership Team 

  • Work closely with technology leadership, platform engineers, security stakeholders, and AI architects to evolve the operational maturity of Companion.  
  • Contribute documentation, operational runbooks, and shared knowledge that reduce platform fragility over time.  
  • Help establish practical operational patterns for AI systems where industry best practices are still emerging.  

What You Bring to the Table 

Required 

  • Strong hands-on Kubernetes operations experience, including troubleshooting workloads, admission controllers, cluster networking, and production incidents.  
  • Experience supporting cloud-native infrastructure in Azure environments, particularly AKS and related operational tooling.  
  • Demonstrated strength in monitoring, observability, and incident response using structured logging and metrics platforms.  
  • SRE mindset with experience handling on-call responsibilities, operational prioritization, and post-incident analysis.  
  • Comfort operating in fast-moving environments with incomplete documentation, evolving processes, and broad ownership areas.  
  • Strong communication and collaboration skills with the ability to explain technical issues clearly across technical and non-technical audiences.  

Nice to Have 

  • Experience with CI/CD pipeline tooling including GitHub Actions, Kaniko, cosign, image signing, or Actions Runner Controller.  
  • Familiarity with Infrastructure as Code practices using Bicep or Azure resource automation tooling.  
  • Exposure to HIPAA, SOC2, or other compliance-aware operational environments.  
  • Experience supporting AI or LLM-backed applications in production environments.  

Compensation 

The target base salary for this position is $120000 - $140000 

This base salary is only a part of a total compensation package that also includes health/dental/vision benefits, annual cash incentive program, 401k with match, flexible PTO, PHM for PHM — our services for you and your dependents — and other benefits. Individual pay may vary from the target range as several factors including market forces, experience, location, disparities in market data, and other relevant business considerations may all factor into final compensation. 

Location 

This is a remote role requiring that you live in and physically perform all work in the United States. 

Next Steps 

Private Health Management is a remote company with employees around the United States. We’re committed to providing a thoughtful, transparent interview experience and meaningful opportunities to get to know our company, mission, and wonderful teammates through fully remote interviews. 

If your application is selected for interviews, you’ll hear from a member of our recruiting team to schedule next steps. Interviews will also include the hiring manager, peers, and often an executive from the department. 

PHM uses AI-enabled tools at certain points in the recruiting process to help identify and evaluate top talent; however, all hiring decisions are made by human reviewers. 

Have a quick question about the role? Email [email protected] or simply apply here. 

Anticipated Pay Range
$120,000$140,000 USD

Private Health Management Los Angeles, California, USA Office

10877 Wilshire Blvd, 1100, , Los Angeles, California , United States, 90024

Similar Jobs

6 Minutes Ago
Remote or Hybrid
112K-204K Annually
Senior level
112K-204K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Software • Biotech
Lead NORAM Customer Success to drive retention, growth, and revenue for genomic solutions. Define customer experience strategy, coach teams, execute success plans, monitor usage, surface upsell opportunities, and manage people operations to ensure adoption and customer outcomes.
Top Skills: AICloud-Native PlatformSophia Ddm Platform
2 Hours Ago
Remote or Hybrid
140K-208K Annually
Senior level
140K-208K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Manage ServiceNow partner relationships for the US public sector to drive co-sell/co-deliver revenue. Develop joint go-to-market and capacity plans, coordinate with Sales, Pre-sales, Marketing, and Customer Outcome teams, accelerate pipeline and sourced NNACV, ensure partner certifications and successful implementations, and meet quarterly and annual sales quotas.
Top Skills: AINow PlatformServicenow
2 Hours Ago
Remote or Hybrid
Senior level
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Lead market success for ServiceNow CRM in Transportation & Logistics: drive territory and account strategy, support digital transformation value articulation, coach account teams, align roadmap with specialists, and manage full sales cycle to close enterprise CRM SaaS deals. Travel 30-50%.
Top Skills: AICrm SaasServicenow

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account