AgileEngine Logo

AgileEngine

Senior AI Engineer ID80138

Posted Yesterday
Be an Early Applicant
In-Office
Porto
Senior level
In-Office
Porto
Senior level
Design and build agentic AI systems that analyze ServiceNow Discovery logs, metrics, failures, and infrastructure health at enterprise scale. Develop agents for failure classification, root-cause analysis, corrective-action recommendations, observability, and reusable troubleshooting workflows. Integrate OpenAI and Anthropic APIs with RAG and vector databases, implement testing and guardrails, and deliver production-grade Python software with REST APIs and CI/CD. Apply infrastructure, networking, and operational data expertise to improve Discovery reliability.
The summary above was generated by AI
AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

WHY JOIN US
If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!

ABOUT THE ROLE
We are looking for an Agentic AI Engineer to design and build AI agents that continuously analyze ServiceNow Discovery logs, classify failures, perform root cause analysis, and recommend or generate corrective actions at enterprise scale across millions of infrastructure targets. You will develop agentic observability capabilities for MID Servers and Discovery schedules, integrate LLM APIs including OpenAI and Anthropic using RAG and vector databases, and convert operational runbooks into reusable agentic workflows using advanced Python engineering. Strong infrastructure troubleshooting and network protocol knowledge is required alongside production-grade Python and CI/CD experience.

WHAT YOU WILL DO
- Design and build AI agents that continuously analyze Discovery logs, execution results, errors, and failed discoveries at scale.
- Develop agents capable of classifying and correlating failures, performing troubleshooting, identifying probable root causes, and recommending corrective actions.
- Build capabilities for agents to analyze and recommend or generate corrections to ServiceNow Discovery probes and patterns, with appropriate testing, guardrails, and human approval before production changes.
- Convert operational knowledge, runbooks, and recurring troubleshooting procedures into reusable agentic workflows, progressively reducing manual investigation and improving Discovery reliability.
- Design and build AI agents that continuously assess the health of the ServiceNow Discovery ecosystem, including MID Servers, discovery schedules/playbooks, target execution, throughput, latency, failures, retries, and discovery coverage.
- Develop agents that regularly collect, aggregate, and interpret metrics and telemetry from ServiceNow Discovery, MID Servers, logs, monitoring platforms, and other relevant data sources.

MUST HAVES
- Advanced Python development experience of at least 4+ years, and REST API design capabilities with experience building APIs, data-processing pipelines, integrations, automation, and production-grade engineering solutions.
- Strong experience with modern software development and SDLC practices and toolchains, including Git/GitHub, Jenkins or equivalent CI/CD platforms, artifact repositories, automated testing, code quality/security scanning, release management, and deployment automation.
- Practical experience integrating Gen AI APIs into software applications, e.g. OpenAI, Anthropic, and a working understanding of developer-level concepts like Retrieval-Augmented Generation (RAG) and Vector databases. (Note: We need software builders, not Machine Learning researchers.)
- Practical knowledge of Infrastructure Engineering with Linux, Windows Server, compute, virtualization/cloud infrastructure, authentication, processes/services, and infrastructure troubleshooting.
- Working knowledge of TCP/IP, DNS, routing, firewalls, ports, SSH, WMI/WinRM, SNMP, HTTP/S, and common network and infrastructure troubleshooting techniques.
- Experience with high-volume logs, metrics, errors, and operational datasets to identify patterns, anomalies, trends, and root causes and turn them into actionable engineering improvements.
- Upper-Intermediate English level.

NICE TO HAVES
- Hands-on experience with ServiceNow Discovery, including MID Servers, Discovery Patterns/Probes, credentials, schedules, Discovery Status, ECC Queue, and troubleshooting Discovery failures.
- Experience developing or modifying ServiceNow Discovery Patterns and Probes and understanding how infrastructure attributes and relationships are discovered.
- Experience with ServiceNow CMDB/CSDM, Configuration Items (CIs), identification/reconciliation, relationships, and CMDB data quality.
- Experience with observability technologies such as Prometheus, OpenTelemetry, alongside Grafana.

PERKS AND BENEFITS
- Growth without limits: build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget
- Competitive compensation: get recognition that reflects your skills and impact, with regular performance and compensation reviews
- Flexibility: work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm
- Meaningful, modern projects: build impactful products using modern technologies alongside global teams and leading brands
- Collaborative culture: join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized
- Well-being & support: access local well-being programs and people-focused support tailored to your location

Meet Our Recruitment Process
Application → Coding Challenge → Video Interview → Technical Interview or Hiring Manager Interview
Each step helps us understand your skills and overall fit.
If it’s a match, you’ll receive an offer.

Similar Jobs

Yesterday
Hybrid
London, Greater London, England, GBR
Entry level
Entry level
Fintech • Payments • Financial Services
Develop and ship secure cross-platform mobile applications and payment SDKs using Kotlin Multiplatform for Android and iOS. Build payment, banking, ePOS, and device abstraction solutions with offline support, secure storage, APIs, testing, observability, and CI/CD. Collaborate with product and design teams, contribute to architecture and Agile delivery, and mentor teammates while ensuring scalable, maintainable, localized, user-friendly applications.
Top Skills: AndroidCanvaCi/CdExpect/Actual ApisFigmaiOSKotlinKotlin CoroutinesKotlin MultiplatformNfcRestful ApisSdk Development
Yesterday
In-Office or Remote
Senior level
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Leads Pfizer’s enterprise applied AI strategy, architecture, governance, operating model, and transformation efforts. Drives portfolio-level technical decisions, reusable capabilities, standards, and measurable business outcomes. Provides hands-on architecture and code review using modern AI and agent frameworks while leading engineers, contractors, and vendors. Oversees quality, security, performance, spending, and technical direction, mentors technical leaders, and evaluates emerging technologies and industry trends.
Top Skills: Agentic SystemsAi GovernanceAi GuardrailsArtificial IntelligenceData ArchitectureGenerative AiLarge Language ModelsMachine LearningModern Ai And Agent FrameworksMonitoringNeural NetworksObservabilityRetrieval ArchitectureTracing
Yesterday
Hybrid
Senior level
Senior level
Fintech • Payments • Financial Services
The Backend Engineer designs, develops, and maintains scalable backend systems while collaborating with cross-functional teams and mentoring developers, ensuring high-quality software delivery.
Top Skills: .NetAWSGoHelmJavaKotlinKubernetesLaravelReactorSpring

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account