Machine Learning Engineering Evaluator
Washington, District of Columbia, 20022, USA
Listed on 2026-09-29
-
Software Development
Machine Learning/ ML Engineer, AI Engineer (Applied/Software)
About Open Train
Open Train is the #1 platform for finding and building careers in AI training and data labeling. Create a free profile, discover projects that match your expertise, and build a lasting portfolio of work that demonstrates your contribution to cutting-edge AI.
- Build a credible AI training and evaluation portfolio
- Find flexible remote opportunities aligned with your technical skills
AI training is the human side of building modern artificial intelligence. Specialists develop, review, and evaluate examples, code, model outputs, and technical systems so AI models can become more accurate, reliable, efficient, and useful.
- Work directly on advanced machine learning and software evaluation
- Help assess whether AI-generated technical solutions meet objective standards
- Contribute to a fast-growing field at the forefront of technology
Open Train is recruiting a Machine Learning Engineering Evaluator to create, solve, review, and validate demanding machine learning engineering tasks for an AI training project. The work spans model development, training and inference systems, numerical computing, performance optimization, Python workflows, and technical evaluation.
You will assess whether implementations satisfy requirements for correctness, reproducibility, efficiency, and performance. You will also document technical decisions, trade-offs, limitations, and failure modes clearly.
This is a global, fully remote contractor role requiring approximately 15 hours per week. Scheduling is flexible, including the option to work weekends. The listed compensation is $100-$150 per hour, and the role is available to candidates in the listed countries.
- Employment type:
Part-time contractor - Time commitment:
Approximately 15 hours per week - Schedule:
Flexible, with weekend work available - Compensation: $100-$150 per hour
- Language:
English - Work arrangement:
Fully remote
You will develop and validate machine learning models, training pipelines, inference systems, and supporting infrastructure. The work combines hands-on implementation with rigorous technical review and objective evaluation.
- Implement model components, data pipelines, evaluation systems, and numerical methods
- Build reproducible workflows using Python and command-line tools
- Work with tensor operations, automatic differentiation, model architectures, tokenization, batching, and generation
- Optimize latency, throughput, memory usage, and hardware utilization
- Diagnose numerical instability, tensor errors, memory bottlenecks, distributed-system failures, and performance regressions
- Review AI-generated code and technical solutions
- Design objective tests, benchmarks, and verification criteria
- Document implementation choices, trade-offs, limitations, and failure modes
This role requires advanced technical training and meaningful professional or research experience in machine learning. The listed experience level is entry level, but the stated educational and technical requirements apply.
- Master's degree or PhD in computer science, machine learning, artificial intelligence, applied mathematics, statistics, engineering, or a closely related quantitative discipline
- Strong professional or research experience in machine learning
- Practical proficiency with Python
- Experience building reproducible technical workflows
- Meaningful experience with at least two relevant machine learning frameworks, libraries, or inference tools
- Strong understanding of model training, evaluation, numerical computation, or inference
- Ability to debug machine learning systems beyond surface-level API usage
- Ability to explain implementation decisions, performance trade-offs, and failure modes clearly
Experience may include the tools listed below. Equivalent tools and substantial open-source or academic experience may also qualify.
- Py Torch
- JAX
- Num Py
- Sci Py
- SGLang
- vLLM
- llama.cpp
- Hugging Face Transformers
- Hugging Face Tokenizers
Open Train helps people start and grow careers in AI training and data labeling, including specialized technical evaluation. A stronger profile lets you showcase credible experience, discover projects that fit your skills, and develop a long-term portfolio in a rapidly expanding industry.
- Work remotely with a flexible schedule
- Use your machine learning engineering expertise on AI development tasks
- Build a portfolio around model, code, and systems evaluation
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).