More jobs:
Technology
Job in
City of Westminster, Central London, Greater London, England, UK
Listed on 2026-09-20
Listing for:
JPMorgan Chase & Co.
Full Time
position Listed on 2026-09-20
Job specializations:
-
IT/Tech
AI Engineer (Applied/Software)
Job Description & How to Apply Below
As a Technology Support I team member in Commercial and Investment bank, you will ensure the operational stability, availability, and performance of our production application flows. Be part of the team responsible for troubleshooting, maintaining, identifying, escalating, and resolving production service interruptions for all internally and externally developed systems, ensuring a seamless user experience.,
- Troubleshoot and monitor production application flows to ensure end-to-end application or infrastructure service delivery to support business operations, addressing anomalies using standard observability tools.
- Uses enterprise-authorized AI capabilities within the work environment to speed up incident triage and initial problem analysis (e.g., summarizing logs/symptoms into hypotheses), validating outputs and handling operational data according to sensitivity and security requirements.
- Assist in the improvement of operational stability and availability through participation in problem management.
- Identify and document basic issues and potential solutions and support the management of incidents, problems, and changes in technology applications or infrastructure, escalating in compliance with firm policy and processes.
- Applies reuse-first, AI-assisted practices within operational stability routines to identify recurring interruption patterns and support validated remediation actions aligned to resiliency and security expectations.
- Execute creative LLM-assisted software solutions; design, develop, and troubleshoot LLM-powered applications and services (e.g., retrieval augmented generation, agent workflows, structured extraction, classification) with a willingness to think beyond routine approaches to break down technical problems and deliver measurable outcomes and think in the novel Agentic AI way.
- Provide Level 3 (L3) support for LLM-assisted production systems, own complex incidents, model and prompt rollouts/rollbacks, dependency issues (vector stores, embeddings, feature stores), and ensure high availability, reliability, and adherence to SLAs including latency and cost budgets.
- Develop data quality rules and controls using LLM; define and enforce guardrails for prompts, retrieved context, model inputs/outputs, and post-processing, including PII redaction, toxicity/safety filters, hallucination mitigation, output schema validation, and policy compliance.
- Create secure, high-quality production code: implement LLM-assisted microservices, synchronous and asynchronous inference pipelines (streaming where appropriate), deterministic fallbacks, circuit breakers, and observability for reliability in production.
- Produce architecture and design artifacts, deliver model cards, system/data lineage, RAG/agent reference architectures, prompt libraries and versioning strategies, evaluation plans, and control evidence ensuring design constraints and regulatory expectations are met during development.
- Drive LLMOps best practices; integrate models, prompts, and evaluation into CI/CD, enforce approvals, segregation of duties, and reproducibility, automate regression and guardrail tests, and manage lifecycle across environments while ensuring LLM-driven systems meet enterprise reliability and resilience expectations (disaster recovery, fallback behaviors, regional resiliency, and performance SLOs).
Formal training or certification on troubleshooting, resolving, and maintaining information technology services concepts and advanced applied experience - Working knowledge of using enterprise-authorized AI capabilities within the work environment to support production support workflows with strong validation habits and awareness of data sensitivity.
- Ability to review and validate AI-assisted incident recommendations before action, escalating when uncertain and following operational and security expectations.
- Familiarity with applications or infrastructure in a large-scale technology environment on-premises or in the public cloud.
- Formal training or certification in software engineering concepts, with practical experience applying them to LLM-enabled systems in regulated environments.
- Strong coding skills in Java/Python and SQL, applied to building LLM-enabled microservices, retrieval pipelines, evaluators, and data tooling; solid understanding of data structures, algorithms, and object-oriented programming as applied to LLM latency, caching, and throughput.
- Hands-on experience with AWS and cloud data management (e.g., Redshift, DynamoDB, Aurora,…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×