Senior AI & Data Consultant
Listed on 2026-09-13
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Cybersecurity, SRE/Site Reliability
Need Help? If you have a disability and need assistance with the application, you can request a reasonable accommodation. Send an email to Accessibility (accommodation requests only; other inquiries won't receive a response).
Regular or Temporary: Regular
Language Fluency: English (Required)
Work Shift: 1st shift (United States of America)
Please review the following job descriptionThe Senior AI and Data Consultant is a senior, hands‑on technical leader within the AI and Data Support Operations organization. This teammate is accountable for elevating the reliability, resiliency, and operational excellence of critical enterprise platforms across hybrid cloud and on prem environments. Acting as both a hands on cloud expert and a cross domain influencer, the Data Consultant drives systemic improvements in observability, automation, AIOps adoption, fault tolerance, and incident management within AWS.
The role partners closely with Application Development, Infrastructure, Production Support, Platform Delivery, Architecture, Cybersecurity, Risk, and Business technology teams to uplift operational practices and deliver stable, predictable, and scalable services. The AIDC delivers measurable impact through deep expertise in cloud technologies, modern operational tooling and enterprise‑scale incident/problem management.
Following is a summary of the essential functions for this job. Other duties may be performed, both major and minor, which are not mentioned below. Specific activities may change from time to time.
Cloud Operational Support and Modernization- Experience with Infrastructure as Code (IaC) products (Terraform, Cloud formation) for rapid and consistent deployment of virtual infrastructure.
- Proficiency with AWS Identity and Access management suite and how it will integrate with Truist Active Directory.
- Expert level knowledge of AWS virtual OS images and management of core components such as EC2 instances and AMIs.
- Familiarity with cloud networking concepts, namely VPCs, load balances and security groups for access to virtual assets.
- Experience with meta tags used for grouping of assets into subcategories.
- Fundamental understanding of monitoring, alerting and observability tools in AWS.
- Lead major and high‑severity incident response efforts, focusing on diagnosing technical root causes therein, and driving multi‑team technical resolution.
- Drive problem management to closure, ensuring systemic fixes replace recurring operational risks.
- Establish and maintain standardized incident playbooks, escalation paths, and communication frameworks.
- Architect and deliver automation solutions that eliminate toil, reduce MTTR, and increase service resilience.
- Implement intelligent alerting, anomaly detection, and event correlation leveraging AI and AIOps tools.
- Guide and enforce SLO/SLI adoption across product teams, ensuring metrics inform decision‑making and prioritization.
- Infrastructure as Code (IaC) creation and updates for deploying assets in Amazon Web Services tenant.
- Deploy patches and device hardening configurations to improve security posture on AI and Data servers.
- Enhance telemetry coverage across logs, metrics, traces, and events using platforms such as Dynatrace and Splunk.
- Define and standardize enterprise observability practices, dashboards, and KPIs.
- Ensure operational readiness of applications and platforms through resiliency testing, chaos engineering, and failure‑mode validation.
- Partner with Delivery, Architecture, Security, and Risk teams to embed reliability and resilience into design and execution.
- Act as a…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).