×
Register Here to Apply for Jobs or Post Jobs. X

LLM Engineer@Louisville - KY

Job in Louisville, Jefferson County, Kentucky, 40201, USA
Listing for: Diverse Lynx
Full Time position
Listed on 2026-09-04
Job specializations:
  • Software Development
    AI Engineer (Applied/Software), Machine Learning/ ML Engineer
Job Description & How to Apply Below

LLM Engineer

Location:

REMOTE, Louisville - KY Project Duration: 1 year+ Top skills required:
1. Experience deploying open-source models such as Llama, Mistral, Mixtral, Phi, Gemma, Qwen, Deep Seek, Granite, Falcon, or domain-specific models.
2. Experience hosting models on GPU infrastructure such as NVIDIA H100, H200, B200, B300, A100, L40S, GH200, or AMD MI300X. 3. Design reusable LLM patterns, services, APIs, and accelerators for Agent Factory adoption.

Role & Responsibilities

The LLM Engineer will design, build, optimize, deploy, and operate Large Language Model and Small Language Model capabilities that power the enterprise Agent Factory. This role is responsible for transforming foundation models into secure, reliable, reusable, and enterprise-ready AI capabilities across agentic workflows, AI for SDLC, knowledge retrieval, model evaluation, private AI hosting, and Agent Ops.

LLM and SLM Model Engineering

Evaluate, build, fine-tune, deploy, and optimize LLMs and SLMs for enterprise use cases. Support domain-specific model development using internal and approved datasets. Build supervised fine-tuning and model adaptation pipelines. Apply model optimization techniques such as LoRA, QLoRA, distillation, quantization, and model compression. Evaluate commercial, open-source, and internally hosted models for suitability, quality, cost, and operational fit. Support model selection strategies based on use case sensitivity, latency, accuracy, cost, and data residency requirements.

Private

AI and On-Prem Model Hosting

Build and support private AI capabilities for hosting SLMs and LLMs in enterprise-controlled environments. Deploy models on on-prem, hybrid, and private cloud infrastructure. Support GPU-enabled model hosting using enterprise AI infrastructure. Optimize model serving for latency, throughput, concurrency, resiliency, and GPU utilization. Build secure inference endpoints for internal agent and application consumption. Support air-gapped or restricted AI environments where required by security or compliance needs.

Partner with infrastructure and platform teams to operationalize private model hosting patterns.

Model Serving and Inference Optimization

Implement scalable model serving using modern inference frameworks. Build high-availability inference patterns for production workloads. Optimize inference performance, token throughput, response latency, and cost efficiency. Implement model routing, load balancing, caching, and fallback strategies. Support batch inference and real-time inference use cases. Develop reusable deployment templates for multiple model families and serving patterns.

LLMOps, Model Ops, and Agent Ops

Build operational practices for managing models and agents across the lifecycle. Implement observability for prompts, retrieval, model responses, latency, cost, and failures. Develop evaluation pipelines for regression testing and continuous quality improvement. Monitor model drift, response quality, hallucination indicators, and safety risks. Support CI/CD and release management for prompts, models, agents, and retrieval pipelines. Build dashboards and metrics for AI quality, reliability, adoption, and operational readiness.

AI

Evaluation and Benchmarking

Define and implement LLM evaluation frameworks. Measure accuracy, groundedness, relevance, hallucination rate, toxicity risk, safety compliance, task completion, and user satisfaction. Build automated test suites for prompts, agents, tools, and RAG pipelines. Benchmark models across enterprise use cases. Compare cloud-hosted, open-source, and on-prem models based on performance, cost, quality, and risk. Support go/no-go quality gates for production AI releases.

7+ Years of Experience

Diverse Lynx LLC is an Equal Employment Opportunity employer. All qualified applicants will receive due consideration for employment without any discrimination. All applicants will be evaluated solely on the basis of their ability, competence and their proven capability to perform the functions outlined in the corresponding role. We promote and support a diverse workforce across all levels in the company.

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary