Renesas Electronics Logo

Renesas Electronics

Staff Site Reliability Engineer

Posted 9 Days Ago
Be an Early Applicant
Hybrid
La Jolla, CA, USA
170K-210K Annually
Senior level
Hybrid
La Jolla, CA, USA
170K-210K Annually
Senior level
Ensure reliability, availability, and performance of large-scale Altium cloud platforms. Responsibilities include improving observability, developing resilience frameworks, automating operations, managing infrastructure and capacity, supporting incident response, promoting Infrastructure as Code, collaborating with engineering teams, and participating in on-call rotations.
The summary above was generated by AI
Job Description

Senior Site Reliability Engineer ensures the reliability, availability, and performance of large-scale software systems through a blend of software engineering and systems administration. Key responsibilities involve automating operational tasks, improving observability, and contributing to incident management, while also collaborating with development and technology teams to build more reliable and scalable applications.  

Join Altium as a Senior Site Reliability Engineer to ensure the reliability and performance of the Altium Cloud Platforms. 

Key Responsibilities:

  • Understanding how an Altium Cloud Platform works 
  • Pioneer improvements in observability, including logging, monitoring, and application performance management (APM), ensuring system reliability and proactive issue detection. 
  • Develop and implement reliability frameworks and patterns that standardize and elevate the resilience of our SaaS products across multiple regions and environments. 
  • Cultivate a shared responsibility model where the SRE team collaborates with and educates engineering teams on reliability best practices. 
  • Contribute to incident response and management, ensuring rapid resolution, clear stakeholder communication, and post-incident analysis for continuous improvement. 
  • Participate in system design consulting, platform management, infrastructure upgrades and capacity planning. 
  • Partner closely with engineering and development teams to enhance product stability, observability, and manageability through best practices in reliability engineering. 
  • Partner closely with DevOps/Operations, drive automation initiatives, promote Infrastructure as Code (IaC), and streamline deployment processes to improve operational efficiency and scalability. 
  • Champion Service-Oriented Organization (SOO) principles to ensure accountability and clarity in service ownership.
  • Participate in on-call rotation and drive operational improvements after incidents.

Qualifications

  • 6+ years in SRE, DevOps or related role in a large-scale environment
  • 3+ years professional experience in software development
  • Software development experience (ideally working with and as a .NET developer)
  • Strong understanding of SDLC, microservice and HA architecture
  • Observability - NewRelic, ELK, Grafana, PagerDuty, OTEL or similar
  • Experience with Kubernetes clusters in production setting, AWS, IOC
  • Experience with operational tasks
  • Knowledge of CI-CD tooling Jenkins, Gitlab, GitHub, ArgoCD or similar
  • Knowledge of IaaC Terraform, Ansible
  • Basic knowledge of networking fundamentals
  • Experience with relational databases (mysql, postgres) as a plus 
  • MUST be a US permanent resident (Citizen or Green Card)

Additional Information

Altium Limited, a part of the Renesas Group and headquartered in San Diego, California, is a global software company accelerating the pace of electronics innovation. We are redefining electronic product creation in a software-defined world with our industry-first cloud-based platform that unites every stakeholder and phase of electronics development.

From startups to world’s technology giants, our digital platforms give more power to PCB designers, supply chain, and manufacturing, letting them collaborate as never before. At Altium, our teams are empowered to innovate, collaborate globally, and help create the future of electronics development.

We believe in rewarding our employees with a competitive benefits package alongside their salary. More information will be provided during the hiring process.

Similar Jobs

9 Days Ago
Hybrid
150K-262K Annually
Senior level
150K-262K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Maintain and improve the reliability, scalability, performance, and operability of ServiceNow’s federal cloud infrastructure. Resolve infrastructure issues, automate repetitive work, develop monitoring solutions, reduce incidents and MTTR, and lead reliability initiatives across hardware, networking, systems, applications, and cloud technologies. This third-shift role supports government customers on a Sunday-Wednesday, 10-hour schedule with no on-call rotation.
Top Skills: AgileAutomationAWSAzureCi/CdDevOpsIp AddressingJavaScriptLinuxMonitoringMySQLNetworkingObservabilityPostgresPythonRoutingRubyScriptingServicenow Platform
21 Days Ago
Hybrid
Los Angeles, CA, USA
211K-263K Annually
Expert/Leader
211K-263K Annually
Expert/Leader
Digital Media • eCommerce • Gaming • Mobile • News + Entertainment
Lead reliability, scalability, observability, automation, incident management, capacity planning, disaster recovery, and security engineering for Crunchyroll’s cloud-native data platforms. Design and operate Kubernetes and GCP infrastructure, implement Infrastructure as Code, improve SLOs and operational excellence, remediate vulnerabilities, strengthen cloud and container security, and mentor engineers across cross-functional teams.
Top Skills: Ci/CdDatadogDockerGCPGoGrafanaIamInfrastructure As CodeJavaKubernetesLinuxOpentelemetryOwasp Top 10PrometheusPythonSecrets ManagementShellSsdlcTerraformZero Trust
21 Days Ago
Hybrid
233K-292K Annually
Senior level
233K-292K Annually
Senior level
Digital Media • eCommerce • Gaming • Mobile • News + Entertainment
Lead reliability, scalability, observability, automation, infrastructure, disaster recovery, and security initiatives for Crunchyroll’s cloud-native data platforms. Design and operate Kubernetes and GCP systems, establish SRE practices including SLIs, SLOs, and error budgets, manage incidents and vulnerabilities, strengthen cloud security, and mentor engineers across teams.
Top Skills: Ci/CdDatadogGCPGoGrafanaIdentity And Access ManagementInfrastructure As CodeJavaKubernetesLinuxOpentelemetryOwasp Top 10PrometheusPythonSecrets ManagementSecure Software Development LifecycleShellTerraform

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account