Senior Forward Deployed ML Engineer, Agents
Listed on 2026-09-13
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer
About aion
Aion is the enterprise AI platform, a full-stack solution for building, fine-tuning, and deploying AI ther an organization is modernizing internal operations, launching AI-powered products, or transforming customer experiences, Aion takes them from concept to production on a single, unified platform.
We work differently than most AI companies: our teams deploy alongside our customers, turning production-ready AI into real business outcomes in weeks, not quarters.
We’re a fast-growing, VC-backed startup led by founders with a track record of successful exits. With teams across the US, UK, and India, we’re building the next generation of enterprise AI and we’re looking for exceptional people to help us scale.
Who You AreYou're a hands-on AI engineer with 3-5+ years of experience building production-grade multimodal AI systems and LLM applications. Your responsibilities mirror those of a hands-on AI startup CTO you work in small teams to own delivery of high-stakes customer projects, embedding directly at client sites to architect, build, and deploy intelligent agent solutions.
You're equally comfortable writing production code, presenting technical solutions to C-level executives, and debugging complex AI systems on factory floors or in customer data centers. You've shipped voice agents, video processing systems, or conversational AI to production. You thrive translating ambiguous business requirements into concrete technical solutions that create measurable impact.
You're comfortable working across the full AI deployment lifecycle from use case discovery and solution architecture to multimodal agent development, MLOps pipeline implementation, and production optimization. You understand what makes agents perform well in production and how to systematically improve quality through observability and evaluation. Experience with voice AI platforms, RAG systems, and LLM orchestration frameworks is highly desirable. You bring exceptional communication skills, customer empathy, and the drive to build AI solutions that transform enterprise operations globally.
WhatYou'll Do
Customer Engagement & Multimodal Agent Development
- Work directly at customer sites from factory floors to executive offices conducting discovery workshops and technical assessments to identify high-impact AI opportunities
- Design and architect end-to-end multimodal agent systems (voice + video + text) that leverage aion's distributed GPU infrastructure and managed services
- Build production-grade voice AI systems using STT, TTS APIs, and LLMs deployed on aion's platform
- Develop vision-enabled agents processing real-time video streams using computer vision pipelines on aion's infrastructure
- Implement sophisticated multi-agent orchestration with(or similar) frameworks like Lang Chain or Llama Index—enabling tool use, memory management, and autonomous task completion
- Rapidly prototype POCs in 2-4 weeks, coding alongside client teams to validate concepts and iterate based on feedback
- Optimize for sub-500ms latency, natural conversation flow, turn detection, and interruption handling in real-time systems
- Integrate agents directly into customer codebases via REST/GraphQL/Web Socket APIs and custom SDKs (Python, Type Script)
- Act as trusted technical advisor to customers, shaping AI strategy and guiding roadmap decisions from concept to production
Data Strategy & MLOps Infrastructure
- Design data architectures with efficient processing pipelines and ingestion workflows for training and inference on aion's platform
- Implement RAG systems with vector databases optimizing embedding strategies, chunk sizes, and retrieval methods
- Prepare and validate datasets for fine-tuning, evaluation, and synthetic data generation
- Work with other MLEs, MLOps, SREs to carry out model deployment and productionization
Observability, Evaluation & Production Operations
- Implement LLM and agents observability and monitoring tracking token usage, latency, costs, and quality metrics across deployments on aion's infrastructure
- Instrument applications to trace LLM calls, retrieval operations, agent actions, and data flows
- Build evaluation frameworks with offline benchmarks (accuracy, relevance, safety metrics) and online monitoring (user feedback, drift detection)
- 6-8+ years of hands-on experience building production AI/ML systems, with 3-4+ years deploying LLM applications to production
- Multimodal AI expertise practical experience building voice agents, vision systems, or…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).