Akamai Technologies Logo

Akamai Technologies

Site Reliability Engineer II

Posted Yesterday
Be an Early Applicant
In-Office or Remote
Hiring Remotely in United States
111K-171K Annually
Junior
In-Office or Remote
Hiring Remotely in United States
111K-171K Annually
Junior
Analyzes and resolves availability and performance issues in large-scale production and lab environments. Develops automation, monitoring, alerting, log analysis, debugging tools, and systems programming solutions. Collaborates with software development and engineering teams on CI/CD, platform architecture, incident resolution, and operational best practices within an agile SDLC.
The summary above was generated by AI

Job Title: Site Reliability Engineer II
Work Location: 145 Broadway, Cambridge, MA 02142
Job Description:
Akamai Technologies, Inc. is hiring for the following role in Cambridge, MA (multiple openings): Site Reliability Engineer II Analysis and resolution of availability and performance issues affecting Control Center users and internal stakeholders; Address availability and performance issues reported to the Portal DevOps team in cooperation with Software Development teams and other internal stakeholders; Solve complex problems in a timely and accurate manner through active troubleshooting, automation, and systems programming; Develop and implement proactive monitoring, alerting, log analysis, and trend identification systems in both production and lab environments; Leverage automation and systems programming to solve complex operational problems associated with running large scale, multi-tenant production and lab environments; Integrate Control Center components which will be utilized by Akamai’s engineering teams and customers by working with the Engineering teams to incorporate the DevOps model of Continuous Integration and Continuous Delivery; Contribute to the overall design and architecture of platform by providing insight on best operational practices, as part of the product life cycle and operate in an agile SDLC model (SCRUM) in collaboration with QA, Product Operations and Other Engineering teams; Recommend and advocate best practices and tools in the areas of systems operation, availability and performance throughout the engineering community; Participate in incident resolution processes driving restoration and repair of service-impacting issues; and Design and develop internal debugging and monitoring tools and solutions to support and streamline everyday operations of the Portal DevOps team. Telecommuting permitted from anywhere in the United States. AG0469
#LI-DNI
Required Qualification:
Master’s degree or foreign equivalent in Computer Science, Software Engineering, Computer Engineering, Electrical & Computer Engineering, Electrical/Electronics Engineering, Information Systems, IT, or Telecommunications Engineering, plus one (1) year of experience in the job offered or closely related engineering role.  
Must have: One (1) year of experience with the following: automation scripts, setup codes and bug fixes, optimize performance in devOps, monitor database tasks for large-scale production deployments, troubleshooting in various environments, ansible, bash, powershell, and linux.
About us
At Akamai, we make life better for billions of people, trillions of times a day. Whether you're streaming live events, scrolling social media, watching your favorite series, or managing your savings, we're the engine behind the scenes. We provide the world's most distributed platform from Cloud to Edge to help the giants of the digital world work faster and stay more secure, making the internet a better experience for everyone.

Our focus is simple:
Cloud and Edge: Running apps closer to users for instant performance.
Security: Neutralizing threats before they ever reach your data.
Content Delivery: Scaling the world's biggest moments without a glitch.
AI: Enabling our customers to build, secure, and scale AI apps on the world's most distributed cloud platform.

At Akamai, we don't just support the internet; we power and protect it, because behind every great digital experience is a massive hidden challenge. And we're the ones who solve it. When millions of people hit play or pay, Akamai ensures it just works.

Benefits at Akamai: We support your health, well-being, finances, and life beyond work. See our benefits.

FlexBase adapts to your job's needs

Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. It's not about telling employees where to work; it's about supporting employees to do their best work.

We trust our incredible employees to work in ways that suit them best: at home, in an office, or a combination of both.

Connect with us on social and see what life at Akamai is like!
Compensation

Akamai is committed to fair and equitable compensation practices. For US based candidates only - the base salary for this position ranges from $111,280-$171,000 /year; a candidate’s salary is determined by various factors including, but not limited to, relevant work experience, skills, certifications and location. Compensation for candidates outside the US will vary. The compensation package may also include incentive compensation opportunities in the form of annual bonus or incentives, equity awards and an Employee Stock Purchase Plan (ESPP). Akamai provides industry-leading benefits including healthcare, 401K savings plan, company holidays, vacation (in the form of PTO), sick time, family friendly benefits including parental leave and an employee assistance program including a focus on mental and financial wellness; Eligibility requirements apply.

Similar Jobs

Yesterday
In-Office or Remote
United States
138K-171K Annually
Mid level
138K-171K Annually
Mid level
Cloud • Security • Software • Cybersecurity
Performs site reliability engineering for large-scale cloud infrastructure, focusing on application and network performance, reliability, security, scalability, and capacity. Deploys highly available systems, automates cloud service deployments, monitors and troubleshoots services, analyzes logs and events, maintains SLAs, and resolves infrastructure issues. The role requires expertise in microservices, container orchestration, cloud migration, deployment automation, and multiple DevOps technologies.
Top Skills: AzureChefCloud ComputingContainer OrchestrationGitIntellij IdeaJenkinsKubernetesLinuxMicroservicesMySQLPostgresPythonTerraformVMware
24 Days Ago
In-Office or Remote
United States
146K-264K Annually
Senior level
146K-264K Annually
Senior level
Cloud • Security • Software • Cybersecurity
Architect, develop, test, and distribute software, services, and infrastructure supporting Akamai’s cloud hypervisor platforms. Improve observability, automate infrastructure processes, troubleshoot complex distributed-system issues, mentor engineers, and participate in on-call service restoration. The role requires deep Linux, kernel, virtualization, ARM hardware, large-scale infrastructure, DevOps, and configuration-management expertise.
Top Skills: AnsibleArmDevOpsDistributed SystemsKvm/QemuLinuxLinux KernelNested VirtualizationNvidia GraceObservability InfrastructureSaltstack
One Month Ago
In-Office or Remote
United States
95K-171K Annually
Junior
95K-171K Annually
Junior
Cloud • Security • Software • Cybersecurity
The Site Reliability Engineer II ensures the reliability, availability, performance, and security of critical cloud systems and services. Responsibilities include developing automation for provisioning and configuration management, maintaining monitoring and alerting, optimizing infrastructure performance, supporting high availability, and enabling continuous integration and delivery. The role collaborates with security teams and drives operational improvements across cloud and network infrastructure.
Top Skills: AnsibleAWSAzureChefContinuous DeliveryContinuous IntegrationDnsElk StackGCPGoGrafanaHTTPKubernetesLinuxPrometheusPuppetPythonShellTcp/IpUnix

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account