×
Hier anmelden um sich kostenlos auf Stellen zu bewerben oder Stellenanzeigen aufzugeben. X

AI Engineer (LLM​/RAG) (m​/w​/d

in 50667, Köln, Nordrhein-Westfalen, Deutschland
Unternehmen: Nejo
Vollzeit position
Verfasst am 2026-08-23
Berufliche Spezialisierung:
  • Software Entwicklung
    Künstliche Intelligenz Ingenieur, Backend Entwicklung, Maschinelles Lernen, Cloud-Ingenieur - Software
Gehalts-/Lohnspanne oder Branchenbenchmark: 60000 - 90000 EUR pro Jahr EUR 60000.00 90000.00 YEAR
Stellenbeschreibung
Stellenbezeichnung: AI Engineer (LLM/RAG) (m/w/d)

For a company in the fast-growing AI implementation market we are looking for an experienced AI Engineer, starting immediately. The company operates LLM-based systems in production: content generation pipelines, retrieval-augmented generation (RAG) over internal documents, and automated workflows deeply integrated with their business systems. The stack is Type Script and Next.js end to end.

These systems are already live. As AI Engineer you take ownership of them, improve their reliability and quality, and extend them to new use cases. The role combines applied LLM engineering with solid backend engineering in Type Script. It does not involve training or fine-tuning foundation models.

Stack:
Type Script, Next.js, Node.js, Postgres with pgvector, Docker, Azure, Anthropic and OpenAI APIs, Vercel AI SDK.

Tasks
  • Take over and maintain the existing LLM pipelines: assess the current architecture, identify failure modes, prioritise fixes, and refactor and extend without disrupting production
  • Own the RAG systems end to end: document ingestion and parsing, chunking, indexing, hybrid retrieval (BM25 and vector), query rewriting, reranking, grounded generation with citations
  • Implement and maintain chunk-level access control, index freshness and tenant isolation across retrieval systems
  • Develop content generation pipelines that deliver consistent quality at volume, including human review steps
  • Build and operate automated workflows against internal and third-party business systems (ERP, CRM, email, internal APIs), with durable and idempotent execution, retry and dead-letter handling, and approval steps for irreversible actions
  • Establish an evaluation framework for systems currently running without one: golden datasets derived from observed production failures, retrieval metrics and more
  • Implement observability across the full request path
  • Optimise cost and latency through prompt caching, batching, model routing and use of smaller models where appropriate
  • Assess where deterministic logic is the better solution and implement it accordingly
  • Work directly with non-technical colleagues to specify and validate automated processes
Requirements
  • Professional experience with at least one LLM-based system in production use, including responsibility for its operation and incident handling
  • Type Script and Node.js at an advanced level: strict typing of non-deterministic model output, async and concurrency patterns, streaming responses, structured error handling
  • Next.js in production:
    App Router, route handlers, server actions, streaming to the client
  • Practical retrieval expertise: hybrid search, embedding model selection, cross-encoder reranking, metadata filtering, permission-aware retrieval, and structured diagnosis of poor retrieval quality
  • Experience processing real-world documents: PDFs with tables, scanned material, DOCX, HTML, including layout-aware parsing, OCR and evidence-based chunking
  • Structured outputs and tool calling as part of your everyday work: JSON Schema, Zod or comparable runtime validation, function calling, handling of malformed or partial output, context window management
  • Designed and run LLM evaluations
  • Experience with LLM tracing and evaluation tooling in a Type Script codebase (e.g. Braintrust, Langfuse, Promptfoo, Open Telemetry or Arize Phoenix)
  • Familiar with Postgres including vector search (pgvector or a comparable vector store), Docker, Git, CI/CD and one major cloud platform
  • Working experience with the Anthropic and/or OpenAI Type Script SDKs
  • Confident communication in English, German is a plus

If you have experience with any of the following, that’s a plus:

  • durable workflow execution for long-running, unattended processes (Temporal, Inngest or comparable)
  • agent orchestration in production, tool calling, recovery, multi-step workflows (Vercel AI SDK, Lang Graph, Mastra, Claude Agent SDK, MCP Type Script SDK)
  • integration experience with enterprise systems such as ERP or CRM platforms like SAP
  • security and data protection in LLM systems, prompt injection and data exfiltration defences, PII handling, GDPR‑compliant design, EU‑hosted or self‑hosted inference
  • structured or graph‑based retrieval for entity‑heavy data
  • experience…
Bitte beachten Sie, dass derzeit keine Bewerbungen aus Ihrem Zuständigkeitsbereich für diese Stelle über diese Jobseite akzeptiert werden. Die Präferenzen der Kandidaten liegen im Ermessen des Arbeitgebers oder des Personalvermittlers und werden ausschließlich von diesen bestimmt.
Um nach Stellen zu suchen, sie anzusehen und sich zu bewerben, die Bewerbungen aus Ihrem Standort oder Land akzeptieren, klicken Sie hier, um eine Suche zu starten:
 
 
 
Suchen Sie hier nach weiteren Stellen:
(nach Beruf, Fähigkeit)
Standort
Suchradius erweitern (Meilen)
0
200
Filter
Mindest-Bildungsgrad für die Stelle
Mindest-Berufserfahrung für die Stelle
Veröffentlicht in den letzten:
Gehalt