Cisco Logo

Cisco

Site Reliability Engineer

Reposted 10 Days Ago
Be an Early Applicant
In-Office
London, Greater London, England, GBR
Mid level
In-Office
London, Greater London, England, GBR
Mid level
Develop and maintain Kubernetes-based microservices configurations across the software lifecycle. Build reusable Helm configurations, manage GitOps deployments with Argo CD, implement canary releases, autoscaling, resiliency, secrets management, observability, policy governance, and disaster recovery. Contribute to a self-service platform using Backstage while ensuring secure, compliant, reproducible deployments.
The summary above was generated by AI
Meet the Team

Cisco's Webex Engineering Group is redefining the future of collaboration. We're building a world where people connect effortlessly to enjoy modern, uncompromised collaboration across every room, desk, pocket, and application.

Our technology powers industry-leading products like Webex Meetings, Webex Calling, and Webex Contact Center. As a member of our team, your work will ensure every user gets a truly immersive experience regardless of device or location, empowering the hybrid working patterns of the future.

Our team is based in one of our central London offices, near Moorgate Street station. We operate a hybrid working model, meeting to collaborate in the office three times a week.

Your Impact
  • Full Lifecycle Development: Own the design, codification, testing, and delivery of Kubernetes-based microservices configurations (Deployments, Services, Ingress, ConfigMaps) using Kubernetes, Helm. Build reusable base configs with environment overlays and enable safe promotion across dev, staging, and production.

  • DevOps & Microservices: Drive operational perfection in a GitOps model using Argo CD. Implement progressive delivery (canary releases, validation gates), codify autoscaling and resiliency, and manage secure secrets via HashiCorp Vault and Kubernetes Secrets.

  • Embrace AI technology: Use AI-assisted tools to accelerate development, validation, and optimization of configuration-as-code, improving delivery speed and reducing deployment risk.

  • Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance.

  • Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full reproducibility, including backup, restore, and disaster recovery.

  • Support & Growth: Advance a self-service platform using Backstage and Argo CD, enabling teams to deploy via declarative configs while promoting standard processes and continuous improvement.

Minimum Qualifications:
  • Analytical Degree or equivalent experience: Bachelor's or higher degree in an analytical subject (e.g. Computer Science, Electronic/Software Engineering).

  • Work experience: 3+ Years demonstrated experience in software development.

  • Coding Aptitude: Extensive architecture and programming experience and the willingness to learn new technologies quickly.

  • Strong Kubernetes + microservices experience

  • Experience with GitOps (CI/CD)

Preferred Qualifications:

Any of the following skills would be particularly advantageous, but we encourage applications without these:

  • Hands-on with Helm or Kustomize

  • Experience with GitOps (e.g., Argo CD)

  • Knowledge of secrets management (e.g., HashiCorp Vault)

  • Experience with observability (metrics/logs/tracing)

CollabHiringWhy Cisco? 

At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations in the AI era – and beyond. We’ve been innovating fearlessly for 40 years to create solutions that power how humans and technology work together across the physical and digital worlds. These solutions provide customers with unparalleled security, visibility, and insights across the entire digital footprint.

Fueled by the depth and breadth of our technology, we experiment and create meaningful solutions. Add to that our worldwide network of doers and experts, and you’ll see that the opportunities to grow and build are limitless. We work as a team, collaborating with empathy to make really big things happen on a global scale. Because our solutions are everywhere, our impact is everywhere. 

We are Cisco, and our power starts with you. 

Similar Jobs

2 Days Ago
Hybrid
London, Greater London, England, GBR
Senior level
Senior level
Financial Services
Leads site reliability engineering for market risk technology services, defining reliability requirements, availability targets, observability strategies, and service-level objectives. Designs resilient systems, debugs complex platform dependencies, automates operational improvements, and promotes safe AI-assisted reliability workflows. Mentors engineers, champions SRE adoption, contributes to engineering communities, and guides production readiness, testing, monitoring, security, auditability, and operational decision-making.
Top Skills: AWSGCPGrafanaJavaOpentelemetryPrometheusPython
2 Days Ago
Remote or Hybrid
London, Greater London, England, GBR
Mid level
Mid level
Cloud • Software
Designs, deploys, and operates highly available, secure, multi-region cloud platforms. Responsibilities include managing AWS and Kubernetes services, automating production operations, improving reliability and scalability, participating in 24x7 incident response, and developing tooling for deployment, testing, failure recovery, and platform security. The role collaborates with application engineering teams and requires Linux administration, Python or Go development, container security, and CI/CD automation experience.
Top Skills: ArgocdAWSCi/CdCncfDastDnsGoHTTPKubernetesLinuxOpentelemetryPrometheusPythonSastService MeshTcp/IpUnix
9 Days Ago
Hybrid
London, Greater London, England, GBR
Senior level
Senior level
Cloud • Information Technology • Security • Software • Cybersecurity
Lead Cloudflare’s Edge SRE team in London, owning production reliability, scalability, observability, incident management, and platform tooling. Mentor and develop engineers, prioritize the SRE roadmap, drive cross-team reliability initiatives, and provide technical leadership for distributed infrastructure systems. The role includes design participation, root-cause analysis, production debugging, availability and performance monitoring, and collaboration with engineering, infrastructure, and product teams.
Top Skills: APIsClickhouseDatabasesDnsElkGrafanaJaegerLinuxOpentracingPrometheusProxiesThanos

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account