Overview
Artificial intelligence is advancing from passive assistance to autonomous, always‑on agentic workflows. Our mission is to make this transition flawless, high‑performing and secure for millions of users worldwide, running natively on the GPUs already installed in their PCs.
We are looking for a Senior Software Engineer to build and optimize local runtimes and agent frameworks that bring autonomous AI to Windows and NVIDIA GeForce RTX GPUs. In this deeply technical, code‑first role you will be the go‑to individual contributor, responsible for making open‑source AI agents such as Nemo Claw and Open Claw run locally, safely and efficiently on consumer PCs.
Responsibilities- Local Inference Optimization: profile and optimize local LLMs (Nemotron and others) on GeForce RTX hardware, minimizing latency and memory footprint with TensorRT and CUDA.
- Agent Runtime Engineering: build and optimize agentic harnesses (Nemo Claw, Open Claw) to run natively and reliably on Windows, implementing orchestration logic for multi‑agent systems.
- Sandboxing & Security: implement policy‑based privacy and security frameworks for autonomous agents, handling file system access and network egress within sandboxed execution environments.
- Hardware/Software Integration: integrate agent and inference stacks with NVIDIA’s driver and middleware layers to extract maximum performance from RTX GPUs.
- Cross‑Team
Collaboration:
partner with AI research teams, driver teams, and the open‑source Open Claw community to ensure the consumer hardware platform is optimal for local agents. - Code Quality: write reliable, production‑ready code, contribute to engineering best practices, and raise the technical bar through code review and design input.
- 12+ years of professional software engineering experience with a track record of shipping performance‑critical systems.
- BS, MS, or PhD in Computer Science, Computer Engineering, or a related field (or equivalent experience).
- Hands‑on experience with LLM inference pipelines (Ollama, llama.cpp, vLLM) and GPU‑accelerated computing (CUDA, TensorRT) on consumer‑grade hardware.
- Practical experience with modern agentic frameworks (Open Claw, Lang Chain, AutoGPT) and an understanding of multi‑agent orchestration.
- Strong understanding of Windows OS internals, process isolation, sandboxing technologies, and system‑level security.
- Proficiency in C++, Python, and Type Script for performance‑critical systems, AI logic, and agent plugins.
- Excellent communication skills and ability to translate complex technical decisions into clear documentation.
- Open‑source contributions to AI agent platforms or inference/orchestration tools, especially Open Claw or llama.cpp.
- Deep knowledge of NVIDIA GeForce RTX architecture and its constraints for edge AI.
- Experience building virtualization, containerization, or sandboxing tools natively for Windows.
- Active technical community presence through blogs, talks, or whitepapers at the intersection of AI, security, and local compute.
Base salary ranges from 224,000 USD to 356,500 USD for Level 5 and 272,000 USD to 431,250 USD for Level 6, adjusted for location and experience. Candidates will also be eligible for equity and a generous benefits package.
Final date to receive applicationsApplications will be accepted until July 10, 2026.
Equal Opportunity StatementWe are committed to fostering an inclusive work environment and pride ourselves as an equal opportunity employer. We do not discriminate on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status, or any other characteristic protected by law.
#J-18808-LjbffrTo Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search: