Group 1001 Logo

Group 1001

Site Reliability Engineer

Posted 6 Days Ago
In-Office or Remote
2 Locations
180K-200K Annually
Expert/Leader
In-Office or Remote
2 Locations
180K-200K Annually
Expert/Leader
As a Site Reliability Engineer, you will ensure system reliability and performance, collaborate with various teams, and implement scalable infrastructure solutions.
The summary above was generated by AI

Group 1001 is a consumer-centric, technology-driven family of insurance companies on a mission to deliver outstanding value and operational performance by combining financial strength and stability with deep insurance expertise and a can-do culture. Group1001’s culture emphasizes the importance of collaboration, communication, core business focus, risk management, and striving for outcomes. This goal extends to how we hire and onboard our most valuable assets – our employees.

Why This Role Matters:

As a Site Reliability Engineer (SRE) at GROUP1001 you will be responsible for ensuring the reliability, availability, and performance of our systems and applications. You will work closely with our development, operations, and security teams to design, implement, and maintain robust and scalable infrastructure solutions. The ideal candidate is passionate about automation, continuous improvement, and delivering exceptional user experiences.

How You'll Contribute:

  • Design, implement, and maintain highly available and scalable infrastructure solutions on cloud platforms (e.g., AWS, Azure, GCP).
  • Implement and manage DevSecOps practices for multi-Cloud, multli-region project lifecycle, enhancing collaboration and efficiency.
  • Experience with monitoring and observability tools (Grafana preferably) for real-time system monitoring and troubleshooting.
  • Strong Git skills, comfort in trunk-based workflows with semver release tagging
  • Design and implement Infra CI/CD pipelines for automated geospatial software deployment and infrastructure management.
  • Conduct regular system audits to identify and address potential issues before they impact project delivery.
  • Ensure compliance with data governance and security policies throughout the geospatial project lifecycle.
  • Provide technical guidance and mentorship to junior team members, fostering a culture of learning and growth.
  • Work on tasks such as preventing incidents with setting up alerts for symptoms.
  • Coordinating with multiple teams such as Data Platforms, NOC/SOC and IT security teams.
  • Building effective monitoring systems with proactive and reactive alerts.
  • Build system health dashboards.
  • Build end user monitoring dashboards.
  • Work with Delivery teams to provide insights into monitoring data.
  • Manage deployments and incidents.
  • Integrate alerts with notifications engine.

What We're Looking For:

  • 10-14 years.
  • Git, GitLab, Infra CI/CD Pipelines
  • Terraform and/or Pulumi
  • Hands on experience as SRE
  • Experience with AWS, Azure
  • Experience with APM tooling
  • Experience automating Operational actions with CI/CD pipelines
  • Experience with Operational Excellence, generating runbooks and working handoffs to L1/L2 teams

 

Preferred Skills: 

  • Worked as SRE similar to an environment such as AWS, Azure, Angular, REST/GraphQL, Neo4j, Event hubs
  • Proven experience in Service Meshes
  • Proven experience with Backups and Patching
  • Proven experience with Policy-as-Code (Rego, OPA) 
  • Proven experience with ZTNA Policies

Compensation:  

 

Our compensation reflects the cost of labor across several U.S. geographic markets. The base pay for this position ranges from $180,000/year in our lowest geographic market up to $200,000/year in our highest geographic market.  Pay is based on a number of factors including market location and may vary depending on job-related knowledge, skills, and experience.

Benefits Highlights:  

Employees who meet benefit eligibility guidelines and work 30 hours or more weekly, have the ability to enroll in Group 1001’s benefits package. Employees (and their families) are eligible to participate in the Company’s comprehensive health, dental, and vision insurance plan options.  Employees are also eligible for Basic and Supplemental Life Insurance, Short and Long-Term Disability. All employees (regardless of hours worked) have immediate access to the Company’s Employee Assistance Program and wellness programs—no enrollment is required.  Employees may also participate in the Company’s 401K plan, with matching contributions by the Company.

 

Group 1001, and its affiliated companies, is strongly committed to providing a supportive work environment where employee differences are valued. Diversity is an essential ingredient in making Group 1001 a welcoming place to work and is fundamental in building a high-performance team. Diversity embodies all the differences that make us unique individuals.  All employees share the responsibility for maintaining a workplace culture of dignity, respect, understanding and appreciation of individual and group differences.

#LI-AS1 #LI-REMOTE

Top Skills

Apm
AWS
Azure
Ci/Cd
GCP
Git
Gitlab
Grafana
Pulumi
Terraform

Similar Jobs

Yesterday
In-Office or Remote
52 Locations
125K-135K Annually
Senior level
125K-135K Annually
Senior level
Artificial Intelligence • Computer Vision • HR Tech • Machine Learning • Software
The Site Reliability Engineer II will manage and ensure the reliability and efficiency of SaaS application platforms, leveraging tools for automation, monitoring, and incident response while collaborating with various teams.
Top Skills: AnsibleArgocdAWSAzureCisDockerElasticsearchFips 140-2Fips 140-3GCPGoGrafanaHelmIptablesJavaJenkinsKubernetesLinuxMongoDBMssqlMySQLPostgresPrometheusPythonSelinuxSolrStigTerraform
2 Days Ago
Easy Apply
Remote
United States
Easy Apply
200K-275K
Senior level
200K-275K
Senior level
Big Data • Fintech • Mobile • Payments • Financial Services
This role involves setting technical strategies, collaborating across teams, managing operations and availability, and fostering a culture of quality and ownership within the Site Reliability Engineering team.
Top Skills: AWSKotlinKubernetesMySQLPythonSpark
2 Days Ago
Remote
United States
161K-180K Annually
Mid level
161K-180K Annually
Mid level
Consumer Web • Digital Media • Information Technology • News + Entertainment • Social Media
The Site Reliability Engineer will enhance infrastructure resilience, optimize system performance, and improve automation within the cloud-based infrastructure while collaborating across engineering teams.
Top Skills: AnsibleArgocdBashCC#C++DockerGoHelmJavaKubernetesLinuxPythonRustTerraform

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account