Akamai Technologies Logo

Akamai Technologies

Senior Site Reliability Engineer Lead

Posted 5 Days Ago
In-Office or Remote
Hiring Remotely in United States
121K-219K Annually
Senior level
In-Office or Remote
Hiring Remotely in United States
121K-219K Annually
Senior level
Lead site reliability engineering for Akamai’s compute infrastructure and services. Develop automation, define reliability requirements and standards, establish SLOs and KPIs, troubleshoot complex distributed-system and hardware issues, manage incidents and postmortems, and participate in on-call rotations. Coordinate engineering teams, support strategic initiatives, mentor engineers, and guide restoration of service-impacting issues.
The summary above was generated by AI

Do you have a passion for cutting edge technologies and tackling system problems?

Are you a self-starting professional who thrives in a dynamic environment?

Join our highly skilled Site Reliability team!

Our team designs, develops, and manages applications and infrastructure that support Akamai's Compute products and services and compute hardware. You will drive automation, operational excellence, and support our customer facing applications and infrastructure. We do this while maintaining Akamai's mission at the forefront of what we do: make life better for billions of people, billions of times a day.

Partner with the best

In this role, you'll focus on developing and implementing automation and efficiency of the Akamai compute infrastructure and services. This requires creative thinking combined with deep domain expertise on Linux, hardware, and performance tuning.

As a Senior Lead Site Reliability Engineer, you will be responsible for:

  • Defining requirements as part of the product lifecycle to influence the new designs and standards
  • Providing support, leadership and motivation for project teams
  • Creating and maintaining SLOs and KPIs, collaborating with Engineering, Product, and Support teams
  • Engaging with our support, operations, hardware, and engineering teams to investigate and troubleshoot complex problems, including incident management and post-mortem reviews
  • Participating in on-call rotations, guiding restoration and repair of service-impacting issues

Do what you love

To be successful in this role you will:

  • Have 5 years of relevant experience and a Bachelor's degree in Computer Engineering, Computer Science or equivalent
  • Have experience with architecting hardware, software, and infrastructure at scale
  • Possess expert-level experience in a Development or SysAdmin role, working with large scale distributed systems
  • Have the ability to align with strategic business initiatives and coordinate the work of other engineers
  • Demonstrate excellent communication, mentorship, and interpersonal skills
  • Possess informed opinions on technology and software design to bring to the table

About us

At Akamai, we make life better for billions of people, trillions of times a day.
Whether you're streaming live events, scrolling social media, watching your favorite series, or managing your savings, we're the engine behind the scenes. We provide the world's most distributed platform from Cloud to Edge to help the giants of the digital world work faster and stay more secure, making the internet a better experience for everyone.
Our focus is simple:
Cloud and Edge: Running apps closer to users for instant performance.
Security: Neutralizing threats before they ever reach your data.
Content Delivery: Scaling the world's biggest moments without a glitch.
AI: Enabling our customers to build, secure, and scale AI apps on the world's most distributed cloud platform.
At Akamai, we don't just support the internet; we power and protect it, because behind every great digital experience is a massive hidden challenge. And we're the ones who solve it. When millions of people hit play or pay, Akamai ensures it just works.

Benefits at Akamai: We support your health, well-being, finances, and life beyond work. See our benefits.

FlexBase adapts to your job's needs

Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. It's not about telling employees where to work; it's about supporting employees to do their best work.
We trust our incredible employees to work in ways that suit them best: at home, in an office, or a combination of both.

Connect with us on social and see what life at Akamai is like!

Compensation

Akamai is committed to fair and equitable compensation practices. For US based candidates only - the base salary for this position ranges from $121,400 - $218,600/year; a candidate’s salary is determined by various factors including, but not limited to, relevant work experience, skills, certifications and location. Compensation for candidates outside the US will vary. The compensation package may also include incentive compensation opportunities in the form of annual bonus or incentives, equity awards and an Employee Stock Purchase Plan (ESPP). Akamai provides industry-leading benefits including healthcare, 401K savings plan, company holidays, vacation (in the form of PTO), sick time, family friendly benefits including parental leave and an employee assistance program including a focus on mental and financial wellness; Eligibility requirements apply.

Akamai Technologies Denver, Colorado, USA Office

Denver, United States

Similar Jobs

12 Days Ago
In-Office or Remote
United States
146K-264K Annually
Senior level
146K-264K Annually
Senior level
Cloud • Security • Software • Cybersecurity
Architect, develop, test, and distribute software, services, and infrastructure supporting Akamai’s cloud hypervisor platforms. Improve observability, automate infrastructure processes, troubleshoot complex distributed-system issues, mentor engineers, and participate in on-call service restoration. The role requires deep Linux, kernel, virtualization, ARM hardware, large-scale infrastructure, DevOps, and configuration-management expertise.
Top Skills: AnsibleArmDevOpsDistributed SystemsKvm/QemuLinuxLinux KernelNested VirtualizationNvidia GraceObservability InfrastructureSaltstack
18 Days Ago
Remote or Hybrid
United States
168K-210K Annually
Senior level
168K-210K Annually
Senior level
Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Lead reliability, scalability, and operational excellence of large-scale database platforms across cloud and on-prem. Build automation-first database infrastructure (Kubernetes operators, IaC, GitOps), drive monitoring/SLOs, incident leadership, performance and cost optimization, and partner with application teams on safe schema/migration practices. Mentor engineers and evaluate AI-assisted workflows to improve productivity and reliability.
Top Skills: AerospikeArgocdAuroraClaudeCloud SqlCursorDatabase OperatorsEksFluxcdGithub CopilotGitopsGkeGoKubernetesMcpMongoDBMySQLPersistent VolumesPostgresPulumiPythonRedisScylladbStatefulsetsTerraform
One Month Ago
Remote
United States
150K-195K Annually
Senior level
150K-195K Annually
Senior level
Artificial Intelligence • Information Technology • Software • Database
As a Site Reliability Engineer, you will design, implement, and maintain scalable infrastructure, ensure system reliability, automate processes, and collaborate with engineering teams.
Top Skills: DockerElk StackGoGrafanaJavaKubernetesNode.jsPrometheusPulumiPythonRubyTerraform

What you need to know about the Colorado Tech Scene

With a business-friendly climate and research universities like CU Boulder and Colorado State, Colorado has made a name for itself as a startup ecosystem. The state boasts a skilled workforce and high quality of life thanks to its affordable housing, vibrant cultural scene and unparalleled opportunities for outdoor recreation. Colorado is also home to the National Renewable Energy Laboratory, helping cement its status as a hub for renewable energy innovation.

Key Facts About Colorado Tech

  • Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
  • Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
  • Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
  • Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account