IT Manager
Listed on 2026-07-18
-
IT/Tech
Cloud Computing: Infrastructure & Operations, IT Infrastructure, SRE/Site Reliability, Systems Administrator
IT Manager
Tensormesh
About the companyTensormesh is building the next generation of AI inference infrastructure. Our mission is to make large language models faster, cheaper, and easier to deploy across any environment — cloud, on‑prem, or hybrid. We help enterprises and AI teams optimize GPU utilization and scale inference workloads with up to 10× better performance.
About the roleWe are seeking an IT Manager to manage our development infrastructure, cloud resources, SaaS tools, and internal systems. This hands‑on role is responsible for keeping our engineering environment secure, reliable, and efficient while supporting a growing team in a fast‑paced startup.
What you will do SaaS Administration & SecurityManaging, provisioning, and de‑provisioning of our developer SaaS stack, such as Git Hub Enterprise, Claude Code Enterprise, and Grafana.
Manage identity and access control (IAM) to ensure least‑privilege access across all cloud resources and SaaS applications.
Monitor and optimize SaaS licensing costs, ensuring we scale our subscriptions efficiently as the team grows.
Networking & Remote AccessDesign, deploy, and maintain secure remote access solutions (e.g., modern VPNs, Tailscale, or Zero Trust Network Access) so engineers can securely connect to dev servers from anywhere.
Manage network configurations, firewalls, and VPCs within GCP and local infrastructure to isolate development environments from unauthorized access.
Dev Server & Infrastructure ManagementMaintain and monitor GPU development servers and various CPU instances hosted on cloud providers.
Implement and manage observability tools to monitor server health, resource utilization (CPU/GPU/Memory), and uptime.
Ideal candidate credentials- 3+ years of experience in systems administration, IT operations, Dev Ops, or a related infrastructure role, preferably in a startup or fast‑moving engineering environment.
- Strong experience managing Linux servers (Ubuntu/Debian), including system troubleshooting, performance tuning, and security best practices.
- Hands‑on experience supporting NVIDIA GPU infrastructure, including CUDA drivers, GPU monitoring, and containerized workloads.
- Experience managing cloud infrastructure in Google Cloud Platform (GCP), including Compute Engine, networking, and access controls.
- Familiarity with infrastructure and operational tools such as Git Hub Enterprise, Grafana, and modern networking or remote‑access solutions.
- Ability to automate operational tasks using scripting languages such as Bash or Python.
- Strong problem‑solving skills and a willingness to take ownership of a wide range of IT and infrastructure challenges.
- Comfortable working independently and adapting to changing priorities in a startup environment.
- Excellent communication skills and the ability to support both engineering and non‑technical team members.
- Competitive base salary
- Performance‑based bonus
- Equity options
- Medical, dental, and vision insurance
- 401(k) retirement plan
- Paid time off
- Build infrastructure powering next‑generation AI applications
- Work alongside a highly technical and experienced engineering team
- Make a direct impact in a fast‑growing startup environment
- Take ownership of challenging technical problems at scale
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).