×
Hier anmelden um sich kostenlos auf Stellen zu bewerben oder Stellenanzeigen aufzugeben. X

Head of AI Engineering (f​/m​/x

in Frankfurt, 60306, Frankfurt am Main, Hessen, Deutschland
Unternehmen: Neoshare
Vollzeit position
Verfasst am 2026-08-18
Berufliche Spezialisierung:
  • Software Entwicklung
    Künstliche Intelligenz Ingenieur, DevOps Ingenieur, Backend Entwicklung, Cloud-Ingenieur - Software
Gehalts-/Lohnspanne oder Branchenbenchmark: 140000 - 190000 EUR pro Jahr EUR 140000.00 190000.00 YEAR
Stellenbeschreibung
Stellenbezeichnung: Head of AI Engineering (f/m/x)
Location: Frankfurt

About neoshare

We’re a Munich-based AI-first fintech scale-up (founded 2019) with offices in Munich, Frankfurt, Berlin and Sofia. Our SaaS platform brings banks, investors, and advisors together to collaborate on complex financial deals making due diligence faster, smarter, and more transparent. Our AI features are already live with leading banks. Now we’re scaling.

The Role

Own and evolve our AI engineering function — transforming a 15–20 person ML team from research-heavy to a high-throughput, production-grade organization. You’ll partner with the CTO on strategy, build the platform that unifies LLM access, RAG, and backend services, and ship reliable, scalable AI features that change how banks work.

Key responsibilities
  • Team leadership and org build
    • Hire, mentor, and develop a high-performing team; set the technical bar, operating rhythms, and code/research review practices
    • Organize sub-teams (e.g., Core Modeling, AI Platform/Infra, Integrations) with clear ownership, SLOs, and on-call
    • Manage roadmap, capacity planning, and delivery across parallel initiatives
  • Architecture and platform
    • Own the LLM gateway: unified APIs and proxy layers for multi-provider routing (OpenAI, Gemini, Bedrock), with rate limits, fallbacks, and cost tracking
    • Build high-performance RAG pipelines (ingestion, embeddings, vector stores, caching) with robust observability and safety guardrails
    • Partner with Java/ NestJS teams to define clean async contracts, schemas, and eventing patterns; drive low-latency, scalable inference
  • Model lifecycle and operations
    • Lead end-to-end model and prompt lifecycle: data curation, training/fine-tuning, evaluation, deployment, rollback
    • Establish LLMOps / MLOps : model/prompt registries, CI/CD, canary/A/B tests, offline/online evals, drift and cost monitoring
    • Optimize inference throughput and cost (autoscaling, batching, quantization/distillation, caching)
  • Strategy and collaboration
    • Translate company goals into an AI/ML roadmap with measurable outcomes; balance exploration with reliability and cost
    • Own build-vs-buy/vendor strategy for models, infrastructure, and data services; manage budgets and SLAs
  • Governance and security
    • Implement data privacy, security, and compliance practices (RBAC, secrets, auditability); track prompt/model lineage and reproducibility
    • Define incident response, runbooks, and postmortems for AI features
Your profile
  • 5+ years as a backend engineer and 4+ years leading AI/ML engineering in production (10+ years total experience ideal)
  • Deep architecture expertise in Java (JVM) and/or Node.js ( NestJS ), distributed systems, APIs, microservices, and messaging/streaming
  • Hands-on with LLM stacks: orchestration (e.g., Lang Chain / Llama Index or custom), vector DBs (Pinecone, Qdrant , FAISS), cloud AI (e.g., AWS Bedrock)
  • Proven operation of systems at scale (millions of daily API calls) with strong SLOs, observability, and incident management
  • MLOps foundations: model registries, experiment tracking, CI/CD, Kubernetes, IaC (e.g., Terraform), security best practices
  • Excellent communication and stakeholder management; strong product sense focused on shipping user-facing feature
  • Fluent German and English for daily team collaboration, stakeholder management, and technical documentation
  • 5+ years as a backend engineer and 4+ years leading AI/ML engineering in production (10+ years total experience ideal)
  • Deep architecture expertise in Java (JVM) and/or Node.js ( NestJS ), distributed systems, APIs, microservices, and messaging/streaming
  • Hands-on with LLM stacks: orchestration (e.g., Lang Chain / Llama Index or custom), vector DBs (Pinecone, Qdrant , FAISS), cloud AI (e.g., AWS Bedrock)
  • Proven operation of systems at scale (millions of daily API calls) with strong SLOs, observability, and incident management
  • MLOps foundations: model registries, experiment tracking, CI/CD, Kubernetes, IaC (e.g., Terraform), security best practices
  • Excellent communication and stakeholder management; strong product sense focused on shipping user-facing feature
  • Fluent German and English for daily team collaboration, stakeholder management, and technical documentation
Nice to have
  • Experience with GPU/accelerator serving and…
Bitte beachten Sie, dass derzeit keine Bewerbungen aus Ihrem Zuständigkeitsbereich für diese Stelle über diese Jobseite akzeptiert werden. Die Präferenzen der Kandidaten liegen im Ermessen des Arbeitgebers oder des Personalvermittlers und werden ausschließlich von diesen bestimmt.
Um nach Stellen zu suchen, sie anzusehen und sich zu bewerben, die Bewerbungen aus Ihrem Standort oder Land akzeptieren, klicken Sie hier, um eine Suche zu starten:
 
 
 
Suchen Sie hier nach weiteren Stellen:
(nach Beruf, Fähigkeit)
Standort
Suchradius erweitern (Meilen)
0
200
Filter
Mindest-Bildungsgrad für die Stelle
Mindest-Berufserfahrung für die Stelle
Veröffentlicht in den letzten:
Gehalt