Kong Logo

Kong

Senior Site Reliability Engineer - Volcano

Posted 3 Days Ago
Remote
Hiring Remotely in United States
150K-170K Annually
Senior level
Remote
Hiring Remotely in United States
150K-170K Annually
Senior level
Own reliability for Kong’s Volcano internal developer platform by defining SLOs, incident practices, and observability. Design multi-region Kubernetes infrastructure, GitOps deployment automation, preview environments, managed PostgreSQL, Redis, and object storage. Lead reliability and compliance initiatives with engineering, security, and OCTO leadership while evaluating emerging edge, serverless, vector database, and AI infrastructure technologies.
The summary above was generated by AI

Are you ready to unlock intelligence?

If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box - we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.

About The Role:

Kong is building Project Volcano, an internal developer platform purpose-built for Kong's engineering ecosystem. Volcano will provide teams with on-demand preview environments, edge deployments, managed PostgreSQL, auth, realtime, and storage APIs all deeply integrated with Kong products.

As the Senior SRE for Volcano, you will be the reliability voice for this platform. This role is a strategic initiative driven by the Office of the CTO (OCTO). You will partner directly with engineering leadership to define the platform's reliability posture. This is a high-visibility, high-impact role with direct influence on Kong's next generation developer platform.

 

What You'll Do:

  • Own reliability for Volcano end-to-end: Define and drive SLOs, error budgets, and incident response practices for all Volcano services — edge deployments, managed Postgres, auth, realtime, storage, and the control plane.

  • Contribute to the platform's infrastructure: Design and build the multi-region Kubernetes infrastructure, networking, and data plane that powers Volcano's edge deployment pipeline and backend-as-a-service capabilities.

  • Build the GitOps and CI/CD backbone: Establish deployment automation, canary pipelines, and preview environment provisioning using ArgoCD, Helm, and Terraform/Terragrunt — setting patterns the broader team will follow.

  • Scale managed data services: Design, operate, and harden multi-tenant PostgreSQL clusters, Redis caching layers, and object storage — with a focus on data isolation, performance, and disaster recovery.

  • Drive observability from day one: Instrument every Volcano service with meaningful SLIs; build dashboards, alerts, and runbooks using Datadog, Prometheus, and Grafana before services go live, not after incidents.

  • Lead cross-functional reliability work: Collaborate with the OCTO team, product engineering, and security to bake reliability and compliance into Volcano's architecture — not bolt it on later.

  • Evaluate and adopt emerging technologies: Given Volcano's greenfield nature, evaluate and make architectural decisions on edge runtimes, serverless compute, vector databases, and AI-native infrastructure components.

 

What You'll Bring:

  • BS in Computer Science or equivalent; substantial experience at Staff or Principal IC level in SRE/Platform Engineering.

  • Proven track record building SRE or platform engineering practices for developer-facing platforms or PaaS/SaaS products — ideally at greenfield stage.

  • Kubernetes expertise: multi-tenant cluster design, networking (CNI, service mesh, ingress), autoscaling, and security hardening.

#LI-BR2

About Kong:

Kong Inc., the AI Connectivity Company, is building the connectivity layer of AI. Trusted by the Fortune 500® and AI-native startups alike, Kong’s unified API and AI platform enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI traffic — on any model, any cloud. For more information, visit www.konghq.com.

Similar Jobs

An Hour Ago
Remote or Hybrid
286K-392K Annually
Senior level
286K-392K Annually
Senior level
Fintech • Machine Learning • Payments • Software • Financial Services
Designs, develops, deploys, and supports large-scale AI systems, including foundation models, LLM inference, agentic workflows, similarity search, guardrails, and model evaluation. Defines enterprise AI architecture, optimizes model performance, cost, latency, and throughput, establishes AI safety and governance standards, leads multi-year platform initiatives, and mentors senior technical leaders across engineering and research.
Top Skills: Agentic AiAi GovernanceAi ObservabilityAWSAws UltraclustersAzureC#C++CudaFoundation ModelsGoGCPHugging FaceJavaLarge Language ModelsModel EvaluationMulti-Agent WorkflowsPythonPyTorchScalaSimilarity SearchVectordbs
An Hour Ago
In-Office or Remote
135K-231K Annually
Senior level
135K-231K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Leads executive and enterprise storytelling for senior Optum leaders, including the CEO. Develops speeches, presentations, articles, thought leadership, employee communications, social content, video scripts, and other high-visibility materials. Translates complex healthcare and business topics into compelling narratives, maintains consistent executive voice and editorial standards, manages multiple priorities, responsibly uses AI, and leads two communications professionals.
Top Skills: Artificial IntelligenceSocial Media
An Hour Ago
In-Office or Remote
Cypress, CA, USA
20-36 Hourly
Junior
20-36 Hourly
Junior
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Provides telephonic after-hours care management for inpatient discharge planning, safe transfers, emergency department diversion, referrals, authorizations, and utilization management. Coordinates with physicians, hospitals, facilities, patients, families, and care teams while documenting interventions, managing caseloads, meeting productivity standards, identifying high-risk patients, and supporting weekend and holiday coverage.
Top Skills: HipaaPc Applications

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account