×
Register Here to Apply for Jobs or Post Jobs. X

Mid-Level AI Software Test Engineer

Job in Ann Arbor, Washtenaw County, Michigan, 48103, USA
Listing for: Apex Systems
Full Time position
Listed on 2026-07-09
Job specializations:
  • Software Development
    AI QA / Validation Engineer, AI Engineer (Applied/Software), Software Testing
Job Description & How to Apply Below

Mid-Level AI Software Test Engineer

We are seeking a Mid-Level AI Software Test Engineer to lead the quality assurance, validation, and automation efforts for AI-powered applications, machine learning systems, AI agents, copilots, and generative AI solutions. This role combines traditional software quality engineering practices with emerging AI testing methodologies to ensure AI systems are accurate, reliable, secure, scalable, and production-ready.

The ideal candidate has a strong foundation in software testing and automation, along with experience or exposure to AI technologies such as Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), AI agents, Model Context Protocol (MCP) integrations, and machine learning platforms. This role plays a critical part in establishing AI quality standards, evaluation frameworks, governance controls, and automated testing capabilities across the AI development lifecycle.

Key Responsibilities

  • Develop and execute comprehensive testing strategies for AI applications, platforms, and services.
  • Validate AI-generated outputs for accuracy, consistency, relevance, reliability, and safety.
  • Design and perform functional, integration, end-to-end, regression, and performance testing for AI-powered solutions.
  • Create and maintain test cases for prompt-driven applications, AI agents, RAG systems, and workflow orchestration platforms.
  • Validate AI guardrails, business rules, permissions models, compliance requirements, and governance controls.
  • Conduct adversarial, negative, and edge-case testing to identify hallucinations, unsafe behavior, model drift, and failure scenarios.
  • Establish quality benchmarks and acceptance criteria for AI solutions.

Test Automation & Evaluation

  • Design, build, and maintain automated test frameworks for AI applications and services.
  • Develop automated evaluation pipelines to assess AI responses, workflows, and model behavior.
  • Integrate AI testing processes into CI/CD pipelines.
  • Implement automated quality scoring, benchmarking, and regression detection capabilities.
  • Create reusable test datasets, simulators, mocks, and validation frameworks to support scalable testing.

Platform & Integration Testing

  • Test AI agents, copilots, APIs, workflow engines, MCP integrations, and tool-calling capabilities.
  • Validate integrations with enterprise systems, external APIs, databases, and knowledge repositories.
  • Verify performance, reliability, scalability, resiliency, and availability of AI workloads.
  • Execute load, stress, and performance testing for AI applications and services.
  • Identify, document, and troubleshoot defects across application, infrastructure, model, and integration layers.

Collaboration & Continuous Improvement

  • Partner closely with software engineers, AI engineers, solution architects, product owners, and security teams.
  • Participate in solution design reviews and provide quality-related recommendations early in the development lifecycle.
  • Contribute to testing standards, methodologies, best practices, and AI quality frameworks.
  • Support production readiness reviews, defect triage, root cause analysis, and continuous improvement initiatives.
  • Promote responsible AI practices and help ensure alignment with organizational governance, privacy, risk, and compliance requirements.

Required Qualifications

  • Bachelor's degree in Computer Science, Software Engineering, Information Systems, or a related technical field.
  • 3–6 years of experience in software testing, quality assurance, quality engineering, or test automation.
  • Experience developing automated testing solutions using one or more of the following:
    Python, Java, JavaScript / Type Script, C#.
  • Experience with API testing and automation frameworks.
  • Strong understanding of: test automation methodologies, Software Development Lifecycle (SDLC), Agile development practices, CI/CD pipelines and Dev Ops principles.
  • Experience testing distributed systems, web applications, APIs, and enterprise platforms.
  • Strong analytical, troubleshooting, and problem-solving skills.
  • Excellent verbal and written communication skills.

Preferred Qualifications

  • Experience testing:
    Generative AI applications, LLM-based systems, AI agents and autonomous workflows, Retrieval-Augmented Generation (RAG) solutions, MCP-based integrations.
  • Familiarity with leading AI platforms and models, including:
    OpenAI, Azure OpenAI, Anthropic Claude, Google Gemini.
  • Experience developing AI evaluation, benchmarking, and validation frameworks.
  • Experience testing cloud-native applications on:
    Microsoft Azure, Amazon Web Services (AWS), Google Cloud Platform (GCP).
  • Knowledge of:
    Responsible AI principles, AI governance frameworks, AI risk management and compliance practices, Privacy and security considerations for AI systems.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary