Research Engineer - Vision Language Models/Multimodal AI/Computer Vision
Listed on 2026-09-14
-
Software Development
AI Engineer (Applied/Software), Robotics
I'm working with a stealth AI robotics startup building AI-powered observability for critical infrastructure
Think: drones and quadrupeds autonomously monitoring solar farms, data centres, refineries, and other massive-scale environments in real time
They're hiring a Research Engineer to build the multimodal agentic layer that connects language, vision, and robotics
You'll be working on:
Imagine enabling customers to ask:
"Inspect this fault"
"Analyze this live stream"
"Recommend the best next action"
And having AI coordinate robots, analyze visual data, and surface the right decisions in seconds
Seed funded with 2+ years runway
Founders with multiple successful exits ($115M & $350M)
Backed by top operators from YC, Dropbox, Nest & South Park Commons
Live deployments already operating in the field
Ideal background:
If you're excited by the intersection of multimodal AI, agentic systems, and real-world robotics, let's chat!
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).