Application Engineer Senior
Listed on 2026-09-06
-
Software Development
AI Engineer (Applied/Software), AI QA / Validation Engineer
AI Engineer (Generative AI & Agent Quality)
Pay Rate: $62.00 Per Hour Duration:
Contract to 12/31/2026
Location:
Remote Business Hours
Standard business hours are 8:00 AM to 5:00 PM (local time), with some flexibility.
Potential for Extension or Conversion
Yes. There is potential for both extension and conversion. However, candidates must reside near an Osaic hub location to be considered for conversion opportunities.
Top Skills Required
Interview Process
- Two rounds of Zoom interviews (webcam required)
We are seeking an AI Engineer to help design, test, optimize, and support AI-powered agent experiences across customer and internal business workflows. This role combines prompt engineering, agent orchestration, quality assurance, evaluation framework development, and AI observability.
The ideal candidate is curious, analytical, and hands-on, with experience working with Large Language Models (LLMs), prompt design, testing methodologies, and modern software development practices. You will partner closely with product managers, software engineers, architects, and business stakeholders to improve agent behavior, reliability, accuracy, and user experience.
Key Responsibilities:Prompt Engineering & Agent Design
- Design, develop, and iterate on AI agent logic for business use cases.
- Create, test, and optimize prompts, system instructions, tools, and workflows to improve agent performance.
- Conduct prompt experiments and A/B testing to measure effectiveness.
- Analyze agent responses and refine prompting strategies for accuracy, consistency, groundedness, and user satisfaction.
- Document prompt patterns, reasoning frameworks, reusable templates, and agent design decisions.
- Collaborate with engineering teams to deploy prompt and agent updates through established SDLC processes.
AI Quality Assurance & Testing
- Design and execute comprehensive test plans covering AI agents, web applications, APIs, and backend services.
- Develop evaluation frameworks and scoring methodologies for AI output quality.
- Create test cases covering:
- Accuracy
- Hallucination detection
- Safety
- Tone and style adherence
- Instruction following
- Context retention
- Response consistency
- Perform regression testing for prompt and model changes.
- Track, prioritize, validate, and retest defects using Azure Dev Ops.
AI Evaluation & Observability
- Utilize Arize or similar observability platforms to monitor model and agent performance.
- Analyze user interactions and evaluation results to identify trends and improvement opportunities.
- Develop quality metrics and dashboards for:
- Accuracy
- Relevance
- Latency
- User satisfaction
- Defect rates
- Prompt performance
- Investigate production issues and recommend prompt, workflow, or architectural improvements.
Software Engineering & Integration
- Support integration of AI capabilities into enterprise applications and workflows.
- Collaborate with developers to validate API integrations, retrieval pipelines, and agent workflows.
- Assist in CI/CD deployment and release validation activities using Azure Dev Ops.
- Create and maintain automated testing assets where appropriate.
Documentation & Collaboration
- Maintain documentation for prompts, evaluation methodologies, testing procedures, and agent configurations.
- Participate in sprint planning, backlog refinement, and Agile ceremonies.
- Communicate findings and recommendations to technical and non-technical stakeholders.
- Share best practices and lessons learned across the team.
- Bachelor's degree in Computer Science, Engineering, Information Systems, or related field.
- 5-10 years of experience in software engineering, QA engineering, automation testing, business systems analysis, or related technical roles.
- 1+ years of hands-on experience with Generative AI, LLMs, prompt engineering, or conversational AI solutions.
- Experience working with Azure Dev Ops or similar Agile delivery platforms.
- Experience designing and executing structured test plans.
- Strong analytical and troubleshooting skills.
- Experience working with APIs and JSON data structures.
- Excellent written and verbal communication skills.
- Experience with Arize AI or comparable AI observability tools.
- Experience with Azure OpenAI, OpenAI, Anthropic, Gemini, or other LLM platforms.
- Knowledge of Retrieval-Augmented Generation (RAG) architectures.
- Experience evaluating AI systems using human and automated scoring methods.
- Familiarity with vector databases and semantic search.
- Experience creating automated test frameworks.
- Understanding of responsible AI principles and AI governance.
- Experience with Python, SQL, JavaScript, or C#.
- Knowledge of prompt versioning and experimentation practices.
Skills:
AI & GenAI
- Prompt Engineering
- Agent Design
- AI Evaluation Frameworks
- RAG Concepts
- LLM Testing
- AI Quality Assurance
- AI Observability
- Prompt Optimization
Tools
- Arize AI
- Azure Dev Ops
- Azure OpenAI
- Git
- Postman
- Jira (optional)
- Power BI (optional)
Development
- Python
- SQL
- REST APIs
- JSON
- Test Automation…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).