Cisco Logo

Cisco

Software Engineer

Posted 9 Days Ago
Be an Early Applicant
In-Office
London, Greater London, England, GBR
Senior level
In-Office
London, Greater London, England, GBR
Senior level
Designs, builds, and deploys AI-powered production intelligence software for global SaaS infrastructure. Responsibilities include developing resilient microservices, AI agents, MCP integrations, observability pipelines, anomaly detection, incident response, self-healing automation, and SRE practices such as SLOs, error budgets, and post-incident reviews. The role also mentors engineers, conducts code reviews, establishes engineering standards, and coordinates cross-team delivery.
The summary above was generated by AI
Meet the Team

The Collaboration Technology Group is redefining the future of teamwork, building services that connect people effortlessly across devices, locations and time zones.

Our team builds, runs and continuously improves the platform services behind Cisco’s collaboration products, operating at global scale across numerous datacentres. We’re a passionate, collaborative team focused on reliability, innovation and engineering excellence.


Your Impact

As a Software Engineer, you will design, build, and deploy the core capabilities of our next-generation AI-powered Production Intelligence platform. You will combine Site Reliability Engineering practices with modern agentic AI (reusable Skills, Model Context Protocol (MCP), and LLM tooling) to transform how engineering and leadership teams monitor, diagnose, and auto-remediate global SaaS infrastructure.


What You'll Do
  • Technical Design & Architecture: Design and implement high-resilience software systems for AI-assisted observability, automated incident response, and self-healing cloud infrastructure.
  • Agentic Workflows & Tooling: Design, build, and maintain production-grade AI agents, MCP tool integrations, and deterministic evaluation pipelines for automated operational decision support.
  • Telemetry & Insights: Implement ingestion and correlation pipelines across distributed logs, metrics, OpenTelemetry traces, change events, and runbooks to accelerate Mean Time to Detection (MTTD) and Resolution (MTTR).
  • Safe Production Automation: Develop proactive anomaly detection and Human-in-the-Loop (HITL) remediation workflows with rigorous safety, security, and quality guardrails.
  • Reliability & Scalability Engineering: Partner with application and infrastructure teams to define SLIs/SLOs, handle error budgets, and lead deep-dive post-incident reviews (PIRs).
  • Mentorship & Best Practices: Mentor mid-level and junior engineers, conduct thorough code reviews, establish engineering guidelines, and drive operational excellence across global development and operations teams.
  • Cross-Team Delivery: Manage priorities and deadlines, communicate progress clearly, and work across teams to turn production needs into reliable software and AI-assisted capabilities.

Minimum Qualifications
  • Bachelor’s degree + 8 years of related experience, Master’s + 6 years, or PhD + 3 years in Computer Science, Software Engineering, or a related technical field.
  • Experience as a Senior / Lead SRE or Software Engineer delivering distributed, high-availability SaaS platforms at scale.
  • Strong proficiency in Python, Go, Java, or C++ with experience designing microservices, APIs, and production automation.
  • Deep experience with Kubernetes, Docker, and container orchestration in large-scale multi-cluster environments.
  • Proven background in SRE practices: SLI/SLO design, observability platforms (metrics/logs/traces), incident management, and automated root cause analysis (RCA).

Preferred Qualifications
  • AI & Agentic Systems: Experience building LLM pipelines, AI Agents, Model Context Protocol (MCP) servers/clients, RAG architectures, and evaluation frameworks.
  • Observability & Telemetry: Experience with OpenTelemetry (OTel), Prometheus, Grafana, Splunk, ThousandEyes, or distributed tracing systems.
  • Cloud & Infrastructure: Expertise in public cloud providers (AWS, GCP, Azure), Terraform/IaC, and GitOps/CI/CD pipelines (Jenkins, GitHub Actions, Harness).
  • Safe Automation & Guardrails: Experience implementing responsible AI guardrails, deterministic fallback logic, and policy-driven remediation engines.
  • Data & Messaging: Experience with streaming and data platforms (Kafka, Redis, PostgreSQL, Elasticsearch/Vector DBs).
  • CollabHiring
Why Cisco? 

At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations in the AI era – and beyond. We’ve been innovating fearlessly for 40 years to create solutions that power how humans and technology work together across the physical and digital worlds. These solutions provide customers with unparalleled security, visibility, and insights across the entire digital footprint.

Fueled by the depth and breadth of our technology, we experiment and create meaningful solutions. Add to that our worldwide network of doers and experts, and you’ll see that the opportunities to grow and build are limitless. We work as a team, collaborating with empathy to make really big things happen on a global scale. Because our solutions are everywhere, our impact is everywhere. 

We are Cisco, and our power starts with you. 

Similar Jobs

Yesterday
Hybrid
London, Greater London, England, GBR
Entry level
Entry level
Cloud • Information Technology • Security • Software • Cybersecurity
Build scalable software for agent-web interactions, including edge services, data pipelines, protocol integrations, developer tools, content controls, and analytics. Develop support for MCP and related emerging standards, ensuring secure, reliable, high-performance systems. Collaborate with engineering, research, product, design, and other stakeholders to take concepts into production while evaluating AI-agent interactions with websites and applications.
Top Skills: APIsCachingCloudflare WorkersHTTPJavaScriptMcpMcp AppsMppProxiesRagRustTypescriptUcpWeb PerformanceWeb SecurityWebmcp
2 Days Ago
Remote or Hybrid
United Kingdom
Internship
Internship
Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Software engineering intern supporting the Platform Engines group and Policy Engine for SailPoint’s cloud identity security platform. Responsibilities include developing software, procedures, and test solutions; collaborating with software and QA engineers, technical writers, and product managers; and working in Agile teams. The intern will gain exposure to cloud and software technologies including Golang, Angular, AWS, Docker, Kafka, Kubernetes, REST, SQL, and Temporal.
Top Skills: AIAngularAWSCloudDockerGoJenkinsKafkaKubernetesMachine LearningPlaywrightRestSaaSSnowflakeSQLTemporalTypescript
2 Days Ago
Hybrid
London, Greater London, England, GBR
Entry level
Entry level
Financial Services
Build, deploy, and maintain full-stack AI applications, services, and data pipelines for enterprise use. Collaborate with data scientists, machine learning engineers, research scientists, and technical teams to productionize AI products. Develop secure, scalable, resilient software; maintain APIs and backend services; conduct code reviews; troubleshoot complex issues; apply CI/CD and software security practices; and support LLM applications using retrieval-augmented generation and prompt design.
Top Skills: APIsBackend ServicesContinuous DeliveryContinuous IntegrationData PipelinesGoLarge Language ModelsPrompt DesignPythonRetrieval-Augmented GenerationTypescript

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account