×
Register Here to Apply for Jobs or Post Jobs. X

Software Engineer, AI Engineer (Applied​/Software), Machine Learning​/ ML Engineer

Job in Somerville, Middlesex County, Massachusetts, 02145, USA
Listing for: Babel Street
Full Time position
Listed on 2026-09-09
Job specializations:
  • IT/Tech
    AI Engineer (Applied/Software), Machine Learning/ ML Engineer
Salary/Wage Range or Industry Benchmark: 85000 - 100000 USD Yearly USD 85000.00 100000.00 YEAR
Job Description & How to Apply Below

Babel Street is the trusted technology partner for the world’s most advanced identity intelligence and risk operations. We deliver advanced AI and data analytics solutions providing unmatched, analysis-ready data regardless of language, proactive risk identification, 360-degree insights, high-speed automation, and seamless integration into existing systems. Babel Street empowers government and commercial organizations to transform high-stakes identity and risk operations into a strategic advantage.

The actionable insights we deliver safeguard lives and protect critical assets around the world. Babel Street is headquartered in Reston, Virginia, with regional offices in Boston, MA and Cleveland, OH, and international offices in Australia, Canada, Israel, Japan, and the U.K. For more information, visit

ROLE

SUMMARY:

As anearly-career

Engineer on the Image & Computer Vision AI team, you will support the developmentanddeployment of computer vision capabilities thatpower

Babel Street’s intelligence applications. You willhelpbuild systems that extract, analyze, and reason over visual data, including image search, object and scene understanding,facial matching workflows,geolocationsupport, and multimodal intelligencefeatures.

This role iswellsuited forsomeonewithfoundational experience in computer vision, image processing, machine learning, or applied AIwhois readytogrow in a hands-on engineering environment.

You will work with senior engineers and cross-functional partners to implement, test, and improvereliable vision capabilities, including integration with multimodal LLM systems that allow users to search and reason over images using natural language.

This is a hybrid roleto be basedout of either our Reston, VA/Washington DCofficeor our Somerville MA office.

ROLE FOCUS;

This role spans three practical execution areas:

Computer Vision & Image Analytics

You willhelpimplement andmaintainimage analytics pipelines that support facial matching, object detection, scene understanding and image similarity. This includessupportingimage preprocessing, feature extraction, model inference, evaluation, and performance improvements under the guidance of more senior team members.

Geospatial & Location Inference from Imagery

You willassistwithcapabilities that infer location, context, or environmental attributes from imageryby usingvisual cues, metadata, and learned representations. Thismay include supporting image-based geolocation, landmark recognition, and contextual scene analysis used in intelligence workflows.

Multi-Modal AI & Image Search

You will contribute tomultimodal AI systems that combine vision models with LLMs, embeddings, and retrieval pipelines tosupportnatural-language search and reasoning over images and image collections. You will help integrate visual understanding into broader intelligence applications and workflows, including supporting entity and event extraction from image-based intelligence data.

KEY RESPONSIBILITIES:
  • Assistin buildingandmaintainingcomputer vision pipelines for image ingestion, preprocessing, inference, and evaluation.
  • Support facial matching and identity-related vision workflows in accordance with accuracy, safety, and compliance requirements.
  • Help develop, test, and improve object detection, image similarity, or scene understanding models.
  • Contribute to image-based geolocation and location inference capabilities using visual features and contextual signals.
  • Support multimodal AI workflows that combine image embeddings with LLM-based search and reasoning.
  • Support multimodal workflows that connect visual understanding with entity and event extraction, helping identify people, places, objects, activities, and relevant contextual signals from image-based intelligence data.
  • Write clean, maintainable Python code and contribute to production services, APIs, and internal tools.
  • Assist with model evaluation, bias testing, and accuracy monitoring and documentation for vision systems.
  • Help optimize inference pipelines for performance, scalability, and cost efficiency, including GPU usage, batching, and model selection.
  • Collaborate with Product, AI, and Engineering teams to integrate vision capabilities into…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary