Cloud Engineering & Architecture - Senior Platform Engineer AI - Vice President
Listed on 2026-07-26
-
IT/Tech
Cloud Computing: Infrastructure & Operations, AI Engineer (Applied/Software), Systems Engineer
What We Do
At the company, our Engineers don’t just make things – we make things possible. Change the world by connecting people and capital with ideas. Solve the most challenging and pressing engineering problems for our clients. Join our engineering teams that build massively scalable software and systems, architect low latency infrastructure solutions, proactively guard against cyber threats, and leverage machine learning alongside financial engineering to continuously turn data into action.
Create new businesses, transform finance, and explore a world of opportunity at the speed of markets. the company Engineers are innovators and problem-solvers, building solutions in Artificial Intelligence, risk management, big data, mobile and more.
As part of Core Engineering at the company, the CE&A team is responsible for enabling the use of public cloud services across the firm. You will be working as part of a multi-disciplinary team responsible for researching, architecting and building a cutting‑edge platform that enable the company Engineering teams to deploy and manage services in public cloud safely and securely.
The organization is seeking highly collaborative, creative, and intellectually curious engineers who are passionate about developing and implementing cutting‑edge cloud computing and AI solutions. The ideal candidate will thrive in a Dev Ops culture and contribute to customer‑centric product development. They will work closely with cross‑functional teams, and will be creative collaborators who evolve, adapt to change and thrive in a fast‑paced global environment.
ResponsibilitiesAnd
Qualifications:
We are looking for a senior technical leader to join our Cloud Engineering & Architecture team and play a pivotal role in enabling the firm to maximize its use of cloud infrastructure. This is a hands‑on leadership position requiring deep technical expertise, strategic thinking, and the ability to drive large‑scale platform initiatives from conception to delivery.
Key Responsibilities:- Design, develop, and operationalize enterprise‑grade cloud platform capabilities.
- Architect scalable, resilient, and secure infrastructure solutions on AWS.
- Architect and operationalize autonomous AI based, self‑healing infrastructure
- Define technical standards, best practices, and reference architectures for cloud adoption across the firm.
- Partner with engineering teams to enable seamless migration and modernization of workloads to the cloud.
- Drive automation and infrastructure‑as‑code practices to improve operational efficiency.
- Drive AI‑powered Fin Ops and predictive resource optimization
- Mentor and guide engineers across teams, raising the overall technical bar.
- Collaborate with security, networking, and compliance teams to ensure platform meets regulatory and governance requirements.
- Evaluate emerging technologies and make recommendations for platform evolution.
- Participate in architecture design reviews and provide technical leadership on complex initiatives.
Skills:
- LLM Orchestration & Agentic Workflows: Experience in designing, building, and deploying Large Language Model (LLM) orchestration frameworks (e.g., Lang Chain, Temporal, or custom agentic loops) to coordinate multi‑step diagnostic and remediation tasks.
- AIOps & Intelligent Observability: Ability to integrate traditional observability stacks (e.g., Datadog, Prometheus, Open Telemetry) with AI/ML models to automate root‑cause analysis, anomaly detection, and semantic log clustering.
- Self‑Healing Infrastructure Engineering: Experience designing closed‑loop, self‑healing systems that autonomously execute recovery actions (e.g., traffic shifting, automated rollbacks, or service restarts) with built‑in verification and safety guardrails.
- AI‑Driven Fin Ops & Resource Optimization: Deep understanding of applying machine learning and predictive analytics to dynamically right‑size cloud resources, manage spot instances, and optimize data platform workloads.
- Predictive Capacity Planning: Ability to design algorithms that forecast workload demands and proactively scale infrastructure to prevent…
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search: