×
Register Here to Apply for Jobs or Post Jobs. X

Software Engineer; Infrastructure

Job in San Francisco, San Francisco County, California, 94199, USA
Listing for: Rad AI
Full Time position
Listed on 2026-09-04
Job specializations:
  • Software Development
    Cloud Engineer - Software, DevOps, Software Architect
Salary/Wage Range or Industry Benchmark: 180000 - 230000 USD Yearly USD 180000.00 230000.00 YEAR
Job Description & How to Apply Below
Position: Software Engineer (Infrastructure)
  • The Platform Engineering organization at Rad AI builds the foundations that power all of our products—Reporting, Impressions, and Continuity—and enables product teams to ship reliably, safely, and at scale
  • Within Platform, the Infrastructure team owns our core cloud infrastructure, platforms, and reliability practices. We’re hiring a multiple Infrastructure Engineers to help us design and operate robust, scalable systems
  • In this role, you’ll contribute to infrastructure architecture, reliability practices, and thoughtful improvements to our workflows
  • Influence the technical direction for infrastructure and platform capabilities that support our rapidly growing AI product suite
  • Architect and evolve our cloud infrastructure (primarily on AWS) across container orchestration (Kubernetes, Elastic Container Service), serverless (e.g., Lambda), virtual machines (e.g., EC2), and data stores to support current and future products
  • Work closely with Platform leadership, product engineering, data, and ML teams to design systems that are robust, observable, and compliant in a healthcare environment
  • Define and drive infrastructure strategy for the Platform org—partnering with engineering leadership to align roadmaps, set standards, and sequence work for maximum business impact
  • Secure networking, identity, and access patterns across environments
  • Improve reliability and operational excellence by defining SLOs, SLIs, and error budgets for core platform services
  • Leading and participating in blameless post-incident reviews and translating learnings into systemic improvements
  • Own observability and monitoring strategy across logging, metrics, and tracing, ensuring we can detect, debug, and prevent issues efficiently
  • Mentor and level up engineers across Platform and product teams—reviewing design docs, guiding architecture decisions, and modeling high standards for reliability, security, and maintainability
  • Partner with security and compliance stakeholders to ensure our infrastructure and operational practices meet HIPAA and other healthcare requirements
  • Advocate for and implement developer experience improvements, such as better CI/CD workflows, faster feedback loops, and tooling that reduces cognitive load for product teams
  • This posting includes multiple open headcount, spanning from Senior to Principal level
Benefits
  • 100% health, dental, & vision
  • Flexible PTO
  • WFH stipend
  • 401k plan
  • Stock options
  • Socials & off-sites

Communicate clearly and empathetically with both technical and non-technical partners, and enjoy mentoring engineers at multiple levels

Have demonstrable experience leading complex, cross-team initiatives from design through rollout—communicating tradeoffs, aligning stakeholders, de-risking launches, and measuring impact

Extensive experience building tooling and automation for other engineers

Bring 4+ years of hands-on infrastructure / platform development experience (or equivalent practical experience) in modern, cloud-native environments, with a track record of owning critical systems in production

Have deep expertise with AWS (preferred) and/or GCP, including core networking, compute, storage, and managed services

If you’re passionate about building resilient platforms and enjoy collaborating across functions, we’d love to hear from you Possess solid Linux fundamentals and are comfortable debugging issues at the OS, networking, and application layers

Are highly proficient in at least one programming/scripting language used for infrastructure work (Python preferred)
Are comfortable with Infrastructure as Code (Terraform preferred, Pulumi, or similar) and Git-based workflows

Have strong experience with Kubernetes, containers (Docker), and container orchestration, and understand how to operate these systems reliably at scale

Take a data-informed, pragmatic approach to decision-making—balancing ideal architecture with business needs, delivery timelines, and team capacity

Familiarity with observability stacks (Cloud Watch, New Relic, Grafana, Open Telemetry, etc.)If you’re passionate about driving innovation and delivering impactful healthcare solutions, we’d love to hear from you!

Prior experience at a fast-growing startup where you’ve helped scale infrastructure, processes, and teams

Background in platform or security engineering, especially around access control, encryption, auditability, and compliance

Experience designing or operating internal developer platforms, SDKs, or reusable frameworks that standardize how services are built and deployed

Experience in regulated environments (e.g., HIPAA) or prior work in healthcare or health tech Experience working closely with ML / data teams or with ML platforms (e.g., Airflow, Ray, ML pipelines, model serving stacks)

#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary