×
Register Here to Apply for Jobs or Post Jobs. X

Senior Software Engineer - LLM Inference

Remote / Online - Candidates ideally in
Vancouver, BC, Canada
Listing for: Nutanix
Part Time, Remote/Work from Home position
Listed on 2026-07-31
Job specializations:
  • Software Development
    AI Engineer (Applied/Software), Cloud Engineer - Software
Job Description & How to Apply Below

The Opportunity

At Nutanix, we're simplifying the future of AI with the Nutanix Cloud Platform for AI, enabling organizations to easily build, fine-tune, and run Generative AI, Large Language Models (LLMs), and next-generation Agentic AI applications without the complexity of managing AI infrastructure themselves. Our high-performance, full-stack machine learning cloud platform delivers AI-ready capabilities out of the box ("GPT-in-a-Box") through a software-defined, full-stack infrastructure solution that simplifies AI deployment across on-premises data centers, edge locations, and public clouds.

The Enterprise AI team is at the forefront of this innovation, driving strategic products such as LLM Inference, the AI Gateway, and the Agentic AI Platform, recently showcased at NVIDIA GTC and NEXT 2026. Join the team responsible for the foundational AI technologies powering the next wave of intelligent applications at Nutanix.

About the Team:
We are a fast-paced, globally distributed team building the foundational layers of the enterprise AI stack. By joining our team, you'll help shape the next generation of enterprise AI platforms, working at the intersection of large-scale distributed systems and machine learning infrastructure. This is an opportunity to solve complex technical challenges, influence the direction of key AI technologies, and build systems that power AI workloads at enterprise scale.

You will report to a seasoned Technical Manager who will provide mentorship and guidance as you navigate through your responsibilities. The work setup at Nutanix AI is a hybrid model, offering a blend of in-office collaboration and remote work flexibility. As a new hire, you will be expected to be in the office for 3 days a week, ensuring that you have the opportunity to engage with your team and foster strong working relationships.

Your Role:

  • Architect, design, and develop horizontally scalable, containerized, fault-tolerant services on Kubernetes.
  • Improve the performance of systems to deliver for low-latency and high-throughput use cases.
  • Optimize any part of the stack, including low-level systems.
  • Leverage and contribute to relevant open-source cloud native projects.
  • Develop scalable, efficient, and fault-tolerant observability architectures for collecting, analyzing, and reporting metrics for various platform services.
  • Collaborate closely with globally located product management and backend development teams to deliver high-quality products in a fast-paced environment.
  • Contribute to all stages of the product development cycle: technical design, development, test, experimentation, analysis, and launch.
  • Be a team player by reviewing code and design docs, giving feedback on product specs and mocks, and documentation.
  • Participate in an ongoing process definition and technology selection to ensure our technology stack is current with relevant trends.
  • Continuously learn and improve your technical and non-technical abilities.
  • What You Will Bring

  • 8+ years of experience developing maintainable, modular, resilient, fail-safe, and long-lasting code from a Product Development company.
  • Have strong programming fundamentals, data structure, and algorithms.
  • Strong experience in Docker, Kubernetes, and Cloud native technologies
  • Experience building applications with Go and Python
  • Experience building and managing CI/CD pipelines
  • Strong understanding of datacenter design, including computing, storage, and networking.
  • Familiarity with on-prem, cloud, and hybrid software deployment architectures
  • Good experience in designing and tuning high-performance system software
  • Strong understanding of distributed computing and storage architectures
  • Strong knowledge of OS internals, virtualization, application performance monitoring, compute storage, and networking management
  • Familiarity with machine learning concepts and popular frameworks (like Tensor Flow, PyTorch, etc) is a strong plus
  • Experience with hardware accelerators, such as GPUs, is a strong plus.
  • Experience working with large codebases or contributing to open source is a strong plus.
  • Experience in building multi-tenant services on a virtualized infrastructure is a solid plus.
  • Detail-oriented…
  • Position Requirements
    10+ Years work experience
    Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
    To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
     
     
     
    Search for further Jobs Here:
    (Try combinations for better Results! Or enter less keywords for broader Results)
    Location
    Increase/decrease your Search Radius (miles)
    0
    200
    Filters
    Education Level
    Experience Level (years)
    Posted in last:
    Salary