×
Register Here to Apply for Jobs or Post Jobs. X

Sr. Machine Learning Engineer

Job in Phoenix, Maricopa County, Arizona, 85003, USA
Listing for: Prosum
Full Time position
Listed on 2026-08-22
Job specializations:
  • Software Development
    Machine Learning/ ML Engineer, AI Engineer (Applied/Software)
Salary/Wage Range or Industry Benchmark: 140000 - 190000 USD Yearly USD 140000.00 190000.00 YEAR
Job Description & How to Apply Below

Our client is seeking a Sr. Machine Learning Engineering for a direct hire role to sit in North Phoenix, AZ or Hillsboro, OR. This role will be onsite 4 days a week and 1 day remote.

JOB SUMMARY

The role of Senior Machine Learning Engineer will architect and optimize real-time, high-throughput, and ultra-low latency image pipelines for next-generation Mask Inspection Tools. Responsibilities include eliminating hardware bottlenecks through CUDA kernel tuning and GPU parallel computing, ensuring deep learning models and CV algorithms seamlessly processing massive, high-bandwidth streaming data at production scale.

ESSENTIAL DUTIES AND RESPONSIBILITIES High-Performance Computing Pipeline Architecture
  • Design, implement, and optimize high-throughput, low-latency image processing pipelines for real-time optical inspection and machine vision systems.
  • Develop scalable architectures capable of processing large volumes of imaging data while meeting stringent latency and reliability requirements.
  • Profile and optimize system performance across CPU, GPU, memory, and I/O subsystems.
GPU Acceleration
  • Design, develop, and optimize CUDA kernels to accelerate deep learning inference and classical computer vision algorithms.
  • Maximize GPU utilization through efficient memory management, kernel optimization, and parallel programming techniques.
  • Evaluate and implement performance improvements using NVIDIA GPU technologies and profiling tools.
Model Deployment & Optimization
  • Optimize, quantize, and deploy machine learning models using TensorRT, ONNX Runtime, or similar inference frameworks.
  • Integrate AI models into production-grade C++ and Python applications.
  • Improve inference throughput, latency, and resource utilization while maintaining model accuracy.
  • Develop automated deployment and validation pipelines for machine learning models.
Concurrency & Systems Optimization
  • Architect and implement multi-threaded, high-concurrency software components for data acquisition, buffering, streaming, and real-time processing.
  • Design robust synchronization and communication mechanisms between hardware interfaces and AI processing pipelines.
  • Optimize end-to-end system performance for deterministic, real-time execution.
Cross-Functional Collaboration
  • Partner with machine learning scientists, computer vision engineers, hardware engineers, and software developers to deliver integrated AI solutions.
#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary