Akamai Technologies Logo

Akamai Technologies

Senior Site Reliability Engineer (Guardicore AI Platform) - Remote

Posted 4 Days Ago
Be an Early Applicant
In-Office or Remote
Hiring Remotely in Spain
Senior level
In-Office or Remote
Hiring Remotely in Spain
Senior level
Own reliability, availability, performance, and operational readiness of a cloud-native data and AI platform. Operate Kubernetes infrastructure, improve observability, security, performance and cost efficiency, lead production investigations, use AI/LLMs for automation, collaborate across DevOps, Software, Data and Security teams, and participate in on-call rotations.
The summary above was generated by AI

Are you passionate about building reliable and scalable cloud platforms?

Would you enjoy supporting cybersecurity products through advanced data and AI capabilities?

Join our Akamai Guardicore Data and AI Platform Team!

The team develops and manages a cloud-native Data and AI Platform enabling analytics, insights, and AI-driven features for Akamai Guardicore Segmentation. It processes extensive security and contextual data, supporting customers in understanding environments, minimizing risks, and preventing threat proliferation.

Partner with the best

As a Senior Site Reliability Engineer, you will take technical ownership of the reliability, availability, performance, and operational readiness of the Guardicore Data and AI Platform. 

You will be responsible for:

  • Operating secure, highly available Kubernetes infrastructure for core microservices, data pipelines, observability, and internal tooling.
  • Enhancing platform reliability, observability, security, performance, and cost efficiency.
  • Providing guidance to engineers and developers to increase confidence that their services are performing as expected.
  • Leading complex production investigations and driving long-term improvements.
  • Leverage LLMs and AI-driven automation to auto-remediate incidents and streamline operations.
  • Partner across DevOps, Software, Data, AI and Security engineering Teams to investigate and troubleshoot complex problems.
  • Participating in on-call rotations, guiding restoration and repair of service-impacting issues.

Do what you love

To be successful in this role you will:

  • 5+ years of experience in SRE, DevOps, or Platform Engineering, with a proven track record of mastering and troubleshooting complex system architectures."
  • Demonstrate ability to design and implement a comprehensive monitoring and observability strategy using tools like Prometheus and Grafana.
  • Have extensive production expertise with Kubernetes, Docker, Helm, and third-party clouds (GCP, Azure, Linode, AWS) on Linux-based infrastructure.
  • Have exceptional troubleshooting and problem-solving skills across network, system, applications, and database layers.
  • Have experience with GitOps, CI/CD, and Infrastructure as Code.
  • Have scripting and programming proficiency in Python, Go, and Bash.
  • Leverage AI tools in daily operational tasks and actively propose initiatives to improve platform automation.
  • Demonstrate technical leadership and ownership in driving cross-team initiatives, defining tools, and building foundational frameworks.

About us

At Akamai, we make life better for billions of people, trillions of times a day.
Whether you're streaming live events, scrolling social media, watching your favorite series, or managing your savings, we're the engine behind the scenes. We provide the world's most distributed platform from Cloud to Edge to help the giants of the digital world work faster and stay more secure, making the internet a better experience for everyone.
Our focus is simple:
Cloud and Edge: Running apps closer to users for instant performance.
Security: Neutralizing threats before they ever reach your data.
Content Delivery: Scaling the world's biggest moments without a glitch.
AI: Enabling our customers to build, secure, and scale AI apps on the world's most distributed cloud platform.
At Akamai, we don't just support the internet; we power and protect it, because behind every great digital experience is a massive hidden challenge. And we're the ones who solve it. When millions of people hit play or pay, Akamai ensures it just works.

Benefits at Akamai: We support your health, well-being, finances, and life beyond work. See our benefits.

FlexBase adapts to your job's needs

Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. It's not about telling employees where to work; it's about supporting employees to do their best work.
We trust our incredible employees to work in ways that suit them best: at home, in an office, or a combination of both.

 

Connect with us on social and see what life at Akamai is like!

     

Akamai Technologies London, England Office

London, United Kingdom

Akamai Technologies Runnymede, England Office

Runnymede, United Kingdom

Similar Jobs

2 Hours Ago
Remote or Hybrid
Mid level
Mid level
Fintech • Legal Tech • Software • Financial Services • Cybersecurity • Data Privacy
Responsible for end-to-end accounting for investment funds and portfolio companies: prepare periodic financial reports, CNMV and regulatory filings, annual accounts (individual and consolidated), treasury management, assist auditors, and collaborate with tax. Supervise teams to deliver client services, ensuring accuracy, deadlines, and strong client communication in Spanish and English.
14 Hours Ago
Remote
UK
Senior level
Senior level
Information Technology
As a Senior Backend Engineer at DuckDuckGo, you'll lead backend projects, mentor engineers, and develop AI-enhanced features for the company's privacy-centric product line.
Top Skills: Ai ToolingGoNode.jsPerlRag Pipelines
14 Hours Ago
Easy Apply
Remote
Easy Apply
Senior level
Senior level
Big Data • Fintech • Mobile • Payments • Financial Services
Design, build, and operate scalable test platform capabilities (synthetic identities, test data seeding, mocking, deterministic validation, load testing). Improve reliability, observability, and automation of CI/test workflows, lead medium-to-large projects, own operational readiness, mentor engineers, and produce async technical artifacts to enable faster, safer developer validation loops.
Top Skills: AlertsCi/Cd PlatformsContract TestingDashboardsJavaKotlinLoad TestingMockingObservabilityPythonRunbooksSlosSynthetic IdentitiesTest Automation FrameworksTest Data Seeding

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account