AI Foundation Model Engineer Jersey , NJ; Hybrid – Onsite
Listed on 2026-07-26
-
Software Development
AI Engineer (Applied/Software), Machine Learning/ ML Engineer
AI Foundation Model Engineer
Location:
Jersey City, NJ (Hybrid – 4 Days Onsite)
Key Responsibilities
The AI Foundation Model Engineer will design, develop, deploy, and optimize enterprise-grade Generative AI applications powered by Foundation Models, Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), Agentic AI, and modern AI engineering frameworks. This role is responsible for transforming AI concepts into secure, scalable, observable, resilient, and production-ready enterprise solutions on the organization's AI Research & Innovation Platform (AIRP), which is currently hosted on AWS while following a cloud-agnostic architecture.
The organization is building a reusable enterprise AI platform that supports multiple business domains, and this role plays a critical part in engineering production AI capabilities. The engineer will develop AI solutions supporting business use cases such as KYC, credit underwriting, Banker 360, Customer 360, pitch book generation, deal intelligence, financial crime detection, sanctions screening, enterprise knowledge management, and intelligent workflow automation.
Strong hands-on experience with AWS AI services, cloud-native architectures, Terraform, Infrastructure-as-Code (IaC), Kubernetes, Docker, Dev Ops, CI/CD, MLOps, and LLMOps is essential for delivering secure, scalable, and maintainable AI applications.
The successful candidate will own the end-to-end lifecycle of LLM-powered applications, RAG pipelines, AI services, model-serving platforms, APIs, observability, evaluation frameworks, deployment automation, monitoring, rollback strategies, and continuous optimization. Working closely with AI Researchers, Platform Engineers, Cloud Engineering, Product, Security, Risk, Compliance, and Dev Ops teams, the AI Foundation Model Engineer will build reusable AI services that accelerate enterprise AI adoption while maintaining security, governance, Responsible AI, and operational excellence.
Key Responsibilities- Design, develop, and deploy LLM-powered enterprise applications, including knowledge assistants, document intelligence solutions, conversational AI, workflow agents, AI copilots, summarization platforms, enterprise search, decision-support systems, and intelligent automation solutions.
- Build and optimize Retrieval-Augmented Generation (RAG) pipelines using embeddings, semantic search, vector databases, document chunking strategies, reranking, retrieval optimization, response grounding, citation frameworks, and context management.
- Develop scalable AI services integrating Foundation Models, LLM APIs, model gateways, AI orchestration frameworks, AWS AI services, enterprise authentication, data services, and cloud-native platform components.
- Collaborate with Cloud Engineering and Dev Ops teams to implement Terraform Infrastructure-as-Code (IaC), reusable cloud modules, Kubernetes deployments, Docker containers, CI/CD pipelines, environment promotion, release management, deployment automation, rollback procedures, and production operations.
- Adapt and optimize foundation models using techniques such as LoRA, PEFT, instruction tuning, fine-tuning, transfer learning, model distillation, quantization, prompt engineering, domain adaptation, and inference optimization.
- Optimize production AI workloads for latency, throughput, scalability, token efficiency, inference cost, GPU utilization, reliability, resiliency, and user experience.
- Implement comprehensive LLMOps and MLOps practices, including model evaluation, prompt evaluation, retrieval quality assessment, monitoring, observability, automated testing, versioning, deployment governance, rollback strategies, and continuous improvement.
- Build enterprise observability capabilities capturing prompt logs, retrieval performance, hallucination indicators, groundedness, latency, token consumption, model drift, feedback loops, operational metrics, service health, and cost telemetry.
- Embed AI security, Responsible AI, privacy, model risk management, secure data handling, access controls, compliance, and governance into AI application architecture, development, deployment, and production operations.
- Develop…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).