Stability AI Logo

Stability AI

Senior Site Reliability Engineer

Posted 5 Days Ago
Remote
Hiring Remotely in United States
Senior level
Remote
Hiring Remotely in United States
Senior level
The Senior Site Reliability Engineer will enhance system reliability, manage cloud infrastructure, and enforce best SRE practices while mentoring juniors.
The summary above was generated by AI

< Remote - United States >

Job Description:
Stability AI’s Engineering Operations team is looking for a Senior Site Reliability Engineer (SRE) to join our growing team and play a pivotal role in improving and shaping our cloud infrastructure. The person will closely work with engineering, IT, security, and product teams to drive innovation and reliability in an evolving environment. Candidates should have the initiative to build and improve a maturing cloud landscape.

Responsibilities:
  • Developing and enforcing SRE best practices and standards across the organization.
  • Architecting and managing scalable systems in AWS and other cloud environments, focusing on high availability and resilience.
  • Implementing and maintaining infrastructure as code using Terraform.
  • Setting up and refining monitoring, logging, and alerting systems.
  • Driving incident management and root cause analysis to improve system reliability.
  • Championing SRE principles and mentoring junior team members.
Qualifications:
  • Collaborating with development teams to enhance CI/CD pipelines.
  • Experience scaling resource intensive systems, be it storage, networking, or compute.
  • Knowledge and experience with Kubernetes or other container scaling solutions
  • Background in software development or automation scripting.
  • Knowledge and experience with Grafana, ELK stack, or similar tools.
  • Cloud security experience.

Equal Employment Opportunity:

We are an equal opportunity employer and do not discriminate on the basis of race, religion, national origin, gender, sexual orientation, age, veteran status, disability or other legally protected statuses.


Top Skills

AWS
Elk Stack
Grafana
Kubernetes
Terraform

Similar Jobs

5 Hours Ago
Easy Apply
Remote
US
Easy Apply
151K-266K Annually
Senior level
151K-266K Annually
Senior level
Cloud • Security • Software • Cybersecurity • Automation
The Senior Site Reliability Engineer will automate operations, maintain systems, develop monitoring tools, respond to incidents, enhance security, and collaborate with engineering teams for optimal system performance.
Top Skills: AnsibleAWSElkGCPGitlabGoKubernetesPrometheusRubyTerraform
3 Days Ago
Easy Apply
Remote or Hybrid
2 Locations
Easy Apply
147K-215K Annually
Senior level
147K-215K Annually
Senior level
Hardware • Information Technology • Security • Software • Cybersecurity • Conversational AI
The role involves developing and managing scalable cloud infrastructure, automating tasks, and leading technical projects in a 24/7 on-call environment.
Top Skills: AnsibleApache AirflowArgoAWSDebianDockerIaasLuigiPythonRubyScalaTerraformUbuntu
5 Days Ago
Easy Apply
Remote
United States
Easy Apply
126K-193K Annually
Senior level
126K-193K Annually
Senior level
Artificial Intelligence • Fintech • Hardware • Information Technology • Sales • Software • Transportation
As a Senior Site Reliability Engineer, you will enhance infrastructure and services, ensuring high availability, developing IaC and CM strategies, and improving monitoring capabilities while collaborating across teams.
Top Skills: AWSBashGoHelmKubernetesPythonRubyTerraform

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account