aion Logo

aion

Software Engineer, General

Reposted 17 Days Ago
Be an Early Applicant
In-Office
London, Greater London, England, GBR
Mid level
In-Office
London, Greater London, England, GBR
Mid level
Build and maintain backend platform services for multi-cloud compute and inference, implementing orchestration, resource scheduling, deployment pipelines, autoscaling, APIs, and event-driven systems. Design schemas and service APIs, work with messaging, storage, and databases, ensure quality via testing and observability, participate in design reviews, debug production issues, and document operational runbooks.
The summary above was generated by AI
About aion

Aion is the enterprise AI platform, a full-stack solution for building, fine-tuning, and deploying AI at scale. Whether an organization is modernizing internal operations, launching AI-powered products, or transforming customer experiences, Aion takes them from concept to production on a single, unified platform.

We work differently than most AI companies: our teams deploy alongside our customers, turning production-ready AI into real business outcomes in weeks, not quarters.

We’re a fast-growing, VC-backed startup led by founders with a track record of successful exits. With teams across the US, UK, and India, we’re building the next generation of enterprise AI and we’re looking for exceptional people to help us scale.


Who You Are

You're a solid engineer with 2-4 years of experience building backend systems and platform infrastructure. You write clean, well-abstracted code with proper design patterns and comprehensive test coverage. You're comfortable working on both the Compute Platform (multi-cloud orchestration, resource management) and Inference Platform (model serving, autoscaling) under the guidance of senior engineers and platform leads.

You have strong proficiency in Golang and understand how to build maintainable, production-grade distributed systems. You take pride in code quality, enjoy collaborating on low-level designs, and are eager to learn from experienced engineers while contributing meaningfully to critical infrastructure components.

You're product-minded, you understand how your technical decisions impact developers using aion's platform and think about the end-to-end user experience. You're a team player comfortable wearing multiple hats one day you're building product features, the next you're joining customer calls to understand their deployment challenges, and the day after you're helping with UI/UX, customer success, documentation and product ops.  

What You'll Do

Platform Development & Implementation

  • Build and maintain platform services across aion's Compute and Inference platforms, working closely with senior engineers and platform leads
  • Implement features for multi-cloud orchestration, resource scheduling, model deployment pipelines, and autoscaling systems
  • Write well-maintained, production-grade code with proper abstractions, design patterns, and comprehensive test coverage
  • Contribute to low-level design (LLD) including service APIs, database schema design, data models, and component interactions
  • Collaborate with senior engineers on high-level design discussions, providing implementation perspectives and feasibility inputs

Backend Systems & Distributed Infrastructure

  • Develop RESTful APIs and gRPC services for platform control planes, resource management, and inference serving
  • Design and implement database schemas for storing platform state, resource metadata, billing data, and observability metrics
  • Work with distributed storage systems, message queues (Kafka, RabbitMQ), and databases (PostgreSQL, Redis) to build reliable platform components
  • Build event-driven architectures for asynchronous processing, job scheduling, and platform automation
  • Implement monitoring, logging, and alerting for platform services to ensure production reliability

Code Quality & Engineering Excellence

  • Write comprehensive unit tests, integration tests, and end-to-end tests to ensure code reliability
  • Participate in code reviews, providing constructive feedback and learning from senior engineers' perspectives
  • Refactor existing code to improve maintainability, performance, and scalability
  • Document design decisions, API specifications, and operational runbooks for platform services
  • Debug production issues and contribute to incident response and post-mortems  

RequirementsTechnical Skills & Experience
  • 2-4 years of experience in backend engineering, platform development, or distributed systems
  • Strong proficiency in Golang you write idiomatic Go code with proper error handling, concurrency patterns, and testing
  • Solid understanding of backend systems fundamentals: RESTful APIs, microservices architecture, and API design principles
  • Hands-on experience with databases (PostgreSQL, MySQL) including schema design, query optimization, and transactions
  • Familiarity with storage systems (object storage like S3, block storage, distributed file systems) and their use cases
  • Experience working with message queues (Kafka, RabbitMQ, NATS) and event-driven architectures
  • Understanding of distributed systems concepts: consensus, eventual consistency, fault tolerance, and retry mechanisms
  • Experience with containerization (Docker) and basic Kubernetes concepts
  • Knowledge of testing frameworks and practices (unit tests, integration tests, mocking)
  • Familiarity with Git, CI/CD pipelines, and modern development workflows
  • Exposure to cloud platforms (AWS/GCP/Azure) and their core services is a plus
  • Experience with infrastructure-as-code (Terraform) or observability tools (Prometheus, Grafana) is beneficial
Bonus/ Good to Have
  • HPC & Cluster Management: Experience handling large-scale HPC clusters using Kubernetes and Slurm for job scheduling, resource allocation, and workload orchestration
  • Data Engineering: Expertise with data pipelines, ETL systems, and large-scale data processing frameworks
  • Systems-Level Programming: Experience with low-level systems programming such as storage systems, Kubernetes operators, OS-level software development, or daemon services (llm-d, system agents)
  • ML Platform Engineering: Experience productionizing ML pipelines, batch job orchestration, model fine-tuning workflows, and Jupyter notebook orchestration systems
  • Enterprise Deployment: Experience platformizing and packaging software for on-premises deployments or customer VPC installations with emphasis on security, compliance, and operational simplicity

BenefitsPreferred Attributes:
  • High ownership, self driven and biased for action.
  • Strong strategic thinking and ability to connect technical decisions to business impact.
  • Excellent communication and mentoring skills.
  • Thrives in ambiguity, fast-paced environments, and early-stage startup culture.
Why Join aion?
  • Work directly with high-pedigree founders shaping technical and product strategy.
  • Build infrastructure powering the future of AI computers globally.
  • Significant ownership and impact with equity reflective of your contributions.
  • Competitive compensation, flexible work options, and wellness benefits

Similar Jobs

An Hour Ago
In-Office or Remote
London, Greater London, England, GBR
Mid level
Mid level
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Manage the full customer lifecycle for mid-market SaaS accounts, including implementation, adoption, success planning, renewals, forecasting, and expansion. Partner with engineering leaders and executive stakeholders to align DX solutions with business goals, drive high-value use cases, and integrate platform insights into workflows. Track account metrics, identify renewal risks, lead strategic customer discussions, and collaborate cross-functionally to improve customer outcomes.
Top Skills: AtlassianDx PlatformSaaS
An Hour Ago
In-Office or Remote
London, Greater London, England, GBR
Mid level
Mid level
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Sell DX's developer productivity SaaS to mid-market accounts in the UK. Manage full sales cycle from prospecting and demos to POCs, negotiation, and close. Build and forecast pipeline, champion buyers through complex evaluations, and relay customer feedback to improve product and GTM.
An Hour Ago
In-Office or Remote
London, Greater London, England, GBR
Senior level
Senior level
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Lead technical engagement for EMEA New Logo enterprise accounts: discover customer needs, map solutions to Atlassian products, deliver tailored value-based demos, guide technical decisions in sales cycles, capture competitive intelligence, and partner with sales and partners to expand cross-product opportunities.
Top Skills: AtlassianBitbucketConfluenceJIRAJira Service ManagementTrello

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account