As a Senior Site Reliability Engineer, you will ensure the reliability and performance of our cloud-native platform by designing robust systems, managing incident responses, and collaborating with development teams.
Sr. Site Reliability Engineer
Location: eMed HQ, Miami, FL 33132
Department: Engineering
Reports to: Sr. Director - Cloud, SRE, Infosec & IT
About eMed
eMed is a pioneering healthcare technology platform company revolutionizing at-home and virtual diagnostics with its innovative 24/7 “Test to Treat” solutions and AI-based technologies. Our primary mission is to provide large employers, state/federal governments, unions, and payers with unique healthcare solutions aimed at reducing obesity, improving employee health, and lowering company healthcare costs. Our integrated GLP-1 medication weight management program utilizes state-of-the-art at-home blood collection kits and connected clinical telehealth services to screen, onboard, and manage qualified candidates, ensuring medication adherence and effective management of side effects through continuous telehealth support.
Position Summary
As a Senior Site Reliability Engineer (SRE) at eMed, you will play a vital role in ensuring the reliability, performance, scalability, and availability of our cloud-native platform. You will partner with engineering, cloud infrastructure, and security teams to design and maintain robust systems that support our growing suite of digital health services. This hands-on technical role combines software engineering, DevOps, and incident response disciplines to enhance operational excellence and customer trust.
Key Responsibilities
Location: eMed HQ, Miami, FL 33132
Department: Engineering
Reports to: Sr. Director - Cloud, SRE, Infosec & IT
About eMed
eMed is a pioneering healthcare technology platform company revolutionizing at-home and virtual diagnostics with its innovative 24/7 “Test to Treat” solutions and AI-based technologies. Our primary mission is to provide large employers, state/federal governments, unions, and payers with unique healthcare solutions aimed at reducing obesity, improving employee health, and lowering company healthcare costs. Our integrated GLP-1 medication weight management program utilizes state-of-the-art at-home blood collection kits and connected clinical telehealth services to screen, onboard, and manage qualified candidates, ensuring medication adherence and effective management of side effects through continuous telehealth support.
Position Summary
As a Senior Site Reliability Engineer (SRE) at eMed, you will play a vital role in ensuring the reliability, performance, scalability, and availability of our cloud-native platform. You will partner with engineering, cloud infrastructure, and security teams to design and maintain robust systems that support our growing suite of digital health services. This hands-on technical role combines software engineering, DevOps, and incident response disciplines to enhance operational excellence and customer trust.
Key Responsibilities
- Design and implement scalable, fault-tolerant, and secure systems in a primarily AWS environment.
- Develop and enforce SLAs, SLOs, and error budgets across services in collaboration with engineering and product teams.
- Own incident response lifecycle—including on-call rotation, root cause analysis, and postmortem documentation.
- Create and maintain infrastructure-as-code (IaC) using tools such as Terraform and Helm.
- Implement monitoring, logging, and alerting systems using Datadog, CloudWatch, Prometheus, etc.
- Build and maintain CI/CD pipelines to improve release velocity and reliability.
- Collaborate with InfoSec to integrate security best practices into all SRE workflows (DevSecOps).
- Proactively address system bottlenecks, capacity challenges, and performance degradation.
- Mentor junior engineers and promote a blameless, learning-driven culture around operations and incidents.
- Bachelor’s degree in Computer Science, Engineering, or a related field (or equivalent practical experience).
- 7+ years of experience in DevOps, SRE, or cloud infrastructure roles in production environments.
- Expertise with AWS services (EC2, RDS, EKS, S3, Lambda, etc.).
- Proficient in Kubernetes, container orchestration, and infrastructure monitoring tools.
- Hands-on experience with Terraform, GitHub Actions, Jenkins, or similar CI/CD tools.
- Solid understanding of TCP/IP networking, DNS, CDN, and load balancing.
- Strong programming or scripting skills (Python, Go, Bash).
- Experience implementing observability and monitoring best practices in a distributed system.
- Familiarity with compliance requirements such as HIPAA, SOC 2, and NIST frameworks.
- Background in healthcare, healthtech, or regulated industries.
- Experience supporting telehealth platforms or mission-critical patient-facing systems.
- Familiarity with zero trust architecture, secrets management, and identity platforms (e.g., Okta).
- Be part of a fast-growing, mission-driven organization that’s transforming healthcare access and outcomes in a collaborative team environment.
- Work on cutting-edge technologies in cloud infrastructure and digital health.
- Collaborate with world-class clinicians, engineers, and innovators.
- Competitive compensation and comprehensive benefits.
Top Skills
AWS
Bash
Cloudwatch
Datadog
Github Actions
Go
Jenkins
Kubernetes
Prometheus
Python
Terraform
Similar Jobs
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
Develop software applications focused on driverless technology, enhance observability systems, provide technical guidance, and manage cloud solutions.
Top Skills:
AWSAzureGCPGoGrafanaKubernetesLinuxPrometheusPythonTsdbs
Artificial Intelligence • Blockchain • Fintech • Financial Services • Cryptocurrency • NFT • Web3
The Senior Site Reliability Engineer will enhance system reliability, optimize cloud deployments, build automation, and mentor engineering teams, ensuring high observational standards and debugging complex issues.
Top Skills:
AWSAzureDatadogDockerEc2GCPGoKibanaKubernetesRubyTerraform
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
As a Senior Site Reliability Engineer, you will maintain cloud infrastructure reliability, automate tasks, and drive technical resolutions across the technology stack, focusing on improving system design and operations.
Top Skills:
AWSAzureJavaScriptLinuxMariadbMySQLPostgresPython
What you need to know about the Los Angeles Tech Scene
Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.
Key Facts About Los Angeles Tech
- Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
- Key Industries: Artificial intelligence, adtech, media, software, game development
- Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
- Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering