×
Register Here to Apply for Jobs or Post Jobs. X

AI Foundation Model Engineer Jersey , NJ; Hybrid – Onsite

Job in Jersey City, Hudson County, New Jersey, 07310, USA
Listing for: United Software Group
Full Time position
Listed on 2026-07-26
Job specializations:
  • Software Development
    AI Engineer (Applied/Software), Machine Learning/ ML Engineer
Job Description & How to Apply Below
Position: AI Foundation Model Engineer || Jersey City, NJ (Hybrid – 4 Days Onsite)

AI Foundation Model Engineer

Location:

Jersey City, NJ (Hybrid – 4 Days Onsite)

Role Purpose &

Key Responsibilities

The AI Foundation Model Engineer will design, develop, deploy, and optimize enterprise-grade Generative AI applications powered by Foundation Models, Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), Agentic AI, and modern AI engineering frameworks. This role is responsible for transforming AI concepts into secure, scalable, observable, resilient, and production-ready enterprise solutions on the organization's AI Research & Innovation Platform (AIRP), which is currently hosted on AWS while following a cloud-agnostic architecture.

The organization is building a reusable enterprise AI platform that supports multiple business domains, and this role plays a critical part in engineering production AI capabilities. The engineer will develop AI solutions supporting business use cases such as KYC, credit underwriting, Banker 360, Customer 360, pitch book generation, deal intelligence, financial crime detection, sanctions screening, enterprise knowledge management, and intelligent workflow automation.

Strong hands-on experience with AWS AI services, cloud-native architectures, Terraform, Infrastructure-as-Code (IaC), Kubernetes, Docker, Dev Ops, CI/CD, MLOps, and LLMOps is essential for delivering secure, scalable, and maintainable AI applications.

The successful candidate will own the end-to-end lifecycle of LLM-powered applications, RAG pipelines, AI services, model-serving platforms, APIs, observability, evaluation frameworks, deployment automation, monitoring, rollback strategies, and continuous optimization. Working closely with AI Researchers, Platform Engineers, Cloud Engineering, Product, Security, Risk, Compliance, and Dev Ops teams, the AI Foundation Model Engineer will build reusable AI services that accelerate enterprise AI adoption while maintaining security, governance, Responsible AI, and operational excellence.

Key Responsibilities
  • Design, develop, and deploy LLM-powered enterprise applications, including knowledge assistants, document intelligence solutions, conversational AI, workflow agents, AI copilots, summarization platforms, enterprise search, decision-support systems, and intelligent automation solutions.
  • Build and optimize Retrieval-Augmented Generation (RAG) pipelines using embeddings, semantic search, vector databases, document chunking strategies, reranking, retrieval optimization, response grounding, citation frameworks, and context management.
  • Develop scalable AI services integrating Foundation Models, LLM APIs, model gateways, AI orchestration frameworks, AWS AI services, enterprise authentication, data services, and cloud-native platform components.
  • Collaborate with Cloud Engineering and Dev Ops teams to implement Terraform Infrastructure-as-Code (IaC), reusable cloud modules, Kubernetes deployments, Docker containers, CI/CD pipelines, environment promotion, release management, deployment automation, rollback procedures, and production operations.
  • Adapt and optimize foundation models using techniques such as LoRA, PEFT, instruction tuning, fine-tuning, transfer learning, model distillation, quantization, prompt engineering, domain adaptation, and inference optimization.
  • Optimize production AI workloads for latency, throughput, scalability, token efficiency, inference cost, GPU utilization, reliability, resiliency, and user experience.
  • Implement comprehensive LLMOps and MLOps practices, including model evaluation, prompt evaluation, retrieval quality assessment, monitoring, observability, automated testing, versioning, deployment governance, rollback strategies, and continuous improvement.
  • Build enterprise observability capabilities capturing prompt logs, retrieval performance, hallucination indicators, groundedness, latency, token consumption, model drift, feedback loops, operational metrics, service health, and cost telemetry.
  • Embed AI security, Responsible AI, privacy, model risk management, secure data handling, access controls, compliance, and governance into AI application architecture, development, deployment, and production operations.
  • Develop…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary