Member of Technical Staff - SRE
Listed on 2026-08-05
-
Software Development
Site Reliability Engineer at Basis
As a Site Reliability Engineer at Basis, you will ensure the reliability, scalability, and performance of our AI-powered accounting platform. You'll join a high-leverage infrastructure team that sits at the intersection of product and platform, owning the systems that keep Basis fast, secure, and always on as we scale.
This is a hands-on role for an engineer who loves building robust systems, reducing complexity through automation, and taking real ownership. While your primary focus will be infrastructure and reliability, strong software engineering skills are essential. As a growing startup, we value engineers who are comfortable wearing multiple hats and jumping in wherever they can have the biggest impact.
If you're excited about building resilient systems from the ground up and shaping reliability practices at a fast-growing company, we'd love to work with you.
What You'll Do- Architect, build, and operate reliable, scalable, and secure infrastructure for our production systems.
- Own cloud infrastructure across compute, storage, and networking, optimizing for availability, performance, and cost efficiency.
- Design and maintain CI/CD pipelines, infrastructure-as-code, and automation to improve developer velocity and system reliability.
- Lead incident response efforts, including on-call rotations, incident coordination, postmortems, and root cause analyses.
- Partner closely with product and engineering teams to balance reliability, performance, cost, and speed of development.
- Automate operational workflows such as capacity planning, safe rollouts, graceful degradation, and data access controls.
- Provide technical leadership and mentorship, helping shape the culture and standards of the infrastructure team.
- 5+ years of experience building and operating production infrastructure at scale.
- Strong software engineering fundamentals and proficiency in at least one programming language.
- Deep understanding of cloud infrastructure, networking, databases, and security principles.
- Experience with CI/CD systems, containerization, and modern infrastructure automation.
- Comfort operating in ambiguous environments and taking ownership end-to-end.
Experience with similar stack (or ability to learn unfamiliar technologies):
- Infrastructure-as-Code tools such as Terraform or Cloud Formation.
- Observability and incident management tools (e.g., Open Telemetry, Prometheus/Grafana, Better Stack, Pager Duty, SLOs/error budgets).
- Data and systems powering analytics and customer-facing workloads, especially in serverless environments. (e.g., Neon, Modal)
We believe that as agents become more capable, two themes are emerging:
The builder stack collapses
- Coding agents provide the ability to move across domains in ways that weren't possible before
- You can offload more decision making to the agents inside your product meaning the lines between ML work and eng work start to disappear.
- Approx 20% of product engineering at Basis is teaching agents to tackle non-deterministic workflows. We see that being 70% by end of year.
The distinctions between frontend vs. backend, infra vs. product, ml vs. eng are already melting away.
Engineering >
Coding Anything that can be clearly defined and verified will be delegated to an agent. What remains is deciding what to build, breaking down problems from first principles, and designing systems that compound.
We hire individuals, not resumes. Given how fast things are changing, it's more important to hire the best people than to hire the person with specific experience.
First principles and systems thinking. Many of the problems we solve are net new. There's no playbook. We want people who can decompose a novel problem, reason from fundamentals rather than pattern match, and design things that compound.
Thinking beyond the engineering. The engineers who do best here understand the domain deeply: the data systems, the workflows, the work our agents are performing. That understanding is what lets you design better systems, train better agents, and make better product decisions.
Ownership…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).