Senior Remote Python Engineer - AI/LLM Evaluation Lead
Fort Worth, Tarrant County, Texas, 76102, USA
Listed on 2026-10-09
-
Software Development
Python, AI Engineer (Applied/Software), AI QA / Validation Engineer
Turing is seeking experienced software engineers for a remote contractor role to help build LLM evaluation datasets and assess real code behavior. You will triage, set up environments with Docker, and evaluate unit test coverage across public repositories.
The role involves hands-on coding, environment automation, and collaboration with researchers to pick challenging repositories for LLM evaluation. 3+ years of experience and Python skills are required.
Consider building your career as a Senior Remote Python Engineer - AI/LLM Evaluation Lead at turing.
Step into the Senior Remote Python Engineer - AI/LLM Evaluation Lead role at turing in United States and grow with us.
Please review the full job details above before applying.
If your experience matches this role, we encourage you to apply.
All applications are reviewed carefully by our team.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).