Apex Logo

Apex

Site Reliability Engineer

Posted 6 Days Ago
Be an Early Applicant
In-Office
Los Angeles, CA, USA
155K-195K Annually
Senior level
In-Office
Los Angeles, CA, USA
155K-195K Annually
Senior level
Design, build, and operate highly available ground and site network infrastructure across commercial and secured environments. Architect Kubernetes platforms, enforce security and compliance controls, develop Terraform-based infrastructure as code, and establish monitoring, logging, alerting, CI/CD, and automation. Serve as a site reliability owner for mission-critical systems, driving uptime and incident response. Partner with security and mission operations teams, define architectural standards, and mentor engineers.
The summary above was generated by AI

Spacecraft represent the most pressing unmet need across the entire aerospace industry. As more launch vehicles come online and the cost to orbit decreases, more companies launching payloads to space continue to emerge.

For the first time in history, this influx of payload companies combined with reduced launch costs has resulted in a massive increase in need for commercial spacecraft platforms, known as satellite buses. These buses hold the payloads of our customers and are flown on launch vehicles.

Apex manufactures these satellite buses at scale using a combination of software, vertical integration, and hardware that is designed for manufacturing. Our spacecraft enable the future of society: ranging from earth observation to communications and more.

We’d love for you to join us on our mission of providing humankind access to the galaxy beyond our planet. 

About the Role

We are seeking a Site Reliability Engineer to join our Ground Software team. As a Site Reliability Engineer, you will design, build, and operate the ground and site network infrastructure that connects Apex facilities and mission operations across both commercial and secured environments. This is a hands-on engineering role for someone who is equally comfortable architecting a highly available Kubernetes platform, and defining the security posture that keeps it all compliant. You will be a primary technical owner, setting the standards that the rest of the team builds on.

Responsibilities:
  • Architect and build ground and site network infrastructure spanning commercial and secured or classified environments through the platform layer.

  • Design, deploy, and scale highly available Kubernetes clusters that support workloads across multiple security and classification levels.

  • Own the security posture of ground infrastructure, defining and enforcing controls, hardening, and compliance for sensitive government and commercial programs.

  • Build and maintain infrastructure as code using Terraform, Terragrunt, and similar tooling so that environments are repeatable, reviewable, and fast to stand up.

  • Establish observability across the ground network, including monitoring, metrics, logging, and alerting, so issues are caught before they reach a mission.

  • Design and run CI/CD pipelines and automation that streamline deployments and reduce manual, error-prone work.

  • Act as a primary site reliability engineer for ground systems, driving uptime, incident response, and reliability standards on programs where downtime is not an option.

  • Partner with security, mission operations, and program teams to translate requirements into infrastructure that is both compliant and operable.

  • Set architectural standards and mentor other engineers, raising the bar for how ground infrastructure is built at Apex.

Requirements:
  • Applicants must be U.S. persons as defined by U.S. export-control law.

  • 5+ years of experience in site reliability, infrastructure, or network engineering, with a meaningful portion in aerospace, defense, satellite, or another mission-critical domain.

  • Hands-on experience architecting and operating ground network or large-scale production network infrastructure.

  • Deep expertise with Kubernetes and containers, including building and scaling high availability clusters in production.

  • Strong networking fundamentals across routing, switching, segmentation, and secure network design.

  • Proven experience with infrastructure as code with CI/CD, automation, and observability tooling such as Prometheus, Grafana, or similar.

  • Working knowledge of security and compliance for regulated or classified environments, and the judgment to design a defensible security posture.

  • Proficiency scripting and building tooling in Python, Go, or a comparable language.

Nice-to-Haves
  • Active Top Secret or Top Secret/SCI clearance

  • Prior experience supporting classified or government space programs or standing up infrastructure across multiple classification levels.

  • Familiarity with GitOps workflows and tools such as ArgoCD, and with modern observability stacks including Mimir.

#LI-AB1

Why Join Apex?

Apex believes in creating a work environment that you look forward to embracing every day. Our employees love working at Apex, and we want you to love it too. We're a fast-growing startup that has raised more than $500M in funding, and we invest heavily in our people from day one.

What We Offer For Full-time Employees:
  • Shared upside: Receive equity in Apex, letting you benefit from the work you create

  • Best-in-class healthcare: 100% company-paid medical, dental, and vision for you and your dependents, plus $100k life insurance at no cost

  • Comprehensive PTO package to reset and recharge - starting at 15 days vacation, growing to 20+ days annually, plus 10 paid holidays

  • Competitive 401(k) plan with generous matching - 100% match on first 3%, 50% on next 2%

  • 8 weeks paid parental leave plus childcare reimbursement up to $350/day for work-related travel

  • Daily catered lunch and unlimited snacks to keep you fueled throughout the day

  • Vibrant community: Monthly office socials, pickleball tournaments, run club, and gatherings for you and your family

  • Your dream desk setup and all the tools you need to be your most productive self

  • World-class Playa Vista office with the benefit of in-person collaboration with amazing coworkers and flexibility to integrate work and life

  • Real impact opportunity: Work alongside experts from aerospace, new space, and other cutting-edge industries to make a lasting difference

Ready to join a team where your contributions matter and your future is bright? Let's build something extraordinary together.

Equal Opportunity Employer

Apex Technology, Inc. is an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. Candidates and employees are always evaluated based on merit, qualifications, and performance. We will never discriminate on the basis of race, color, gender, national origin, ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability, or any other legally protected status.

Similar Jobs

Yesterday
Hybrid
172K-300K Annually
Senior level
172K-300K Annually
Senior level
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
Found and establish GM Vehicle Autonomy’s centralized SRE practice. Build shared reliability tooling, service catalogs, SLOs, error budgets, readiness automation, standards, and production infrastructure. Partner with service owners to assess criticality, dependencies, failure modes, and operational health. Improve incident response, post-incident practices, on-call sustainability, and toil reduction while driving adoption across infrastructure teams. The role requires strong software engineering and distributed-systems expertise across Linux, networking, Kubernetes, cloud, CI/CD, and observability.
Top Skills: C++Ci/CdCloud InfrastructureGoJavaKubernetesLinuxNetworkingObservabilityPolicy-As-CodePython
3 Days Ago
In-Office
190K-280K Annually
Senior level
190K-280K Annually
Senior level
Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
Lead the establishment and maturation of the SRE function across cloud infrastructure and platform services. Define reliability targets, build observability, lead incident response and root-cause analysis, improve resilience through automation and testing, and develop operational tooling. Partner with engineering teams on reliability requirements, establish incident practices, mentor engineers, and manage the short- and long-term SRE roadmap.
Top Skills: AWSGoKubernetesPython
3 Days Ago
In-Office
220K-340K Annually
Senior level
220K-340K Annually
Senior level
Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
Lead the establishment and maturation of the SRE function across cloud infrastructure and platform services. Define reliability targets, build observability and automation, lead incident response and root-cause analysis, improve resilience and recovery, and reduce operational toil. Partner with product and platform teams on reliability-focused design, establish incident practices, manage the SRE roadmap, and mentor engineers across Cloud Engineering and Reliability.
Top Skills: AWSCloud-Native ObservabilityGoInfrastructure As CodeKubernetesPython

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account