Director of DevOps & SRE
Listed on 2026-08-29
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, IT Project Manager
Investor Flow is the only company of its kind to deliver industry specialized CRM, built on Salesforce, and digital portals to help alternative asset firms find opportunities, create and manage relationships, and turn relationship insights into action with increased productivity and transparency.
We are looking for a Director of Dev Ops & SRE to lead the team responsible for the reliability, security, and scalability of theInvestor
Flowplatform. This is a strategic leadership role. You will set technical direction, own the Dev Ops/SRE roadmap, and lead a team of technical leads, senior and mid-level engineers across both disciplines,while partnering closely with Engineering and Product on technology and product roadmap planning, resourcing, and delivery sequencing.
You will stay technical enough to lead architectural review, challenge design decisions, and get hands-on where the problem warrants it, but your primary leverage will come through strategy, standards, and the people you build. We take an automation-first approach to everything we do;if it is repeatable, it should be codified,and you will set that standard. You will also help shape how AI-assisted engineering and operations tooling is adopted across the organisation, an area where we are investing heavily.
YouWill:
- Strategy and roadmap. The Dev Ops/SRE strategy for the Investor Flow platform, aligned to measurable outcomes across reliability, delivery throughput, and cost.
- The team. Leading, hiring, and developing technical leads, senior and mid-level engineers across Dev Ops and SRE, and setting priorities across infrastructure, CI/CD, security, and production support.
- Cloud platform ownership. Reliability, performance, and cost-efficiency across DEV, QA, STG, and PRD
-compute, data, secrets management, caching, and networking/edge security. Azure is our primary platform, with a supporting AWS and Snowflake footprint. - CI/CD and release engineering. Git Hub Actions end to end, multi-region deployment automation, advanced deployment strategies (blue/green, canary, rolling) and rollback, plus security scanning, compliance checks, and vulnerability management embedded in the pipeline.
- Infrastructure as code. Terraform practice and standards that standardise provisioning and reduce manual, ticket-driven infrastructure work.
- Automation-first operating standard. Setting and holding the expectation that repeatable work is codified rather than performed, across provisioning, releases, access, remediation, and reporting, and driving developer enablement through self-service tooling that increases engineering capacity while maintaining controls.
- Identity and access. Enterprise directory services, role-based and privileged access controls, Auth0 machine-to-machine credentials, and least-privilege access across engineering and QA.
- Incident response and resilience. Escalation paths, root-cause analysis and postmortems, production readiness reviews, automated failover, and our disaster recovery and RTO/RPO commitments. You are the escalation point for high-severity incidents and major client environment changes, ensuring appropriate change management and CAB governance, with occasional off-hours support for the teams.
- Observability and service levels. SLO/SLA and monitoring strategy across Grafana and our wider observability tooling, building alert coverage that is meaningful rather than noisy.
- AI in engineering operations. The operational rollout of AI-assisted engineering and support tooling: automated triage, agent-based workflows, and developer-facing AI capability.
- Architectural review and technical governance. Reviewing significant infrastructure and platform change with a security- and compliance-first lens, ensuring designs meet our obligations by default rather than by exception, and owning the shared-responsibility operating model between Dev Ops/SRE and product engineering.
- Vendor and cost management. Relationships across cloud, security, CI/CD, and observability platforms, including spend and renewal negotiation.
- 8+ years in Dev Ops, SRE, or infrastructure engineering, including 3+ years leading and developing engineering teams.
- Deep hands-on Microsoft Azure experience at production scale: compute, data, secrets, identity services, networking and WAF, and cost/capacity management. Azure depth is essential for this role.
- Strong CI/CD and release engineering background (Git Hub Actions and/or Azure Dev Ops Pipelines) and infrastructure as code with Terraform, including module development and state management.
- Production support and incident response experience for a multi-tenant SaaS platform.
- Containerisation and orchestration in production (Docker, Kubernetes).
- Solid grounding in identity and access management (role-based and privileged access, SSO/SAML, OAuth/Auth0) and secrets management practices.
- A genuine automation-first instinct, backed by strong scripting (Python, Bash, Power Shell or similar) and a track record of replacing manual process with…
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search: