As the Head of AI Engineering (f/m/x) at neoshare in Berlin, you will step into a senior-level IT role to own and evolve our AI engineering function. You will transform our 15 to 20 person machine learning team from research-heavy into a high-throughput, production-grade organization that builds scalable AI features for the banking sector.
Key Responsibilities and Architecture
- Hire, mentor, and organize sub-teams like Core Modeling and AI Platform into clear ownership areas with defined SLOs and on-call rotations.
- Own the LLM gateway, building unified APIs and proxy layers for multi-provider routing across OpenAI, Gemini, and Bedrock with cost tracking and fallbacks.
- Construct high-performance RAG pipelines with vector stores, caching, observability, and safety guardrails, partnering with Java and NestJS teams for low-latency inference.
- Lead the end-to-end model and prompt lifecycle, establishing MLOps and LLMOps practices including CI/CD, canary testing, drift monitoring, and cost optimization.
- Drive governance, data security, compliance, and vendor strategy while managing budgets and translating company goals into measurable roadmaps.
Candidate Profile
- 5+ years of experience as a backend engineer and 4+ years leading AI/ML engineering in production, with 10+ years of total experience ideal and a Bachelor's degree required.
- Deep architectural expertise in Java or Node.js (NestJS), distributed systems, microservices, and messaging or streaming.
- Hands-on experience with LLM orchestration stacks, vector databases like Pinecone, Qdrant, or FAISS, and cloud AI platforms such as AWS Bedrock.
- Proven ability to operate large-scale systems handling millions of daily API calls with strong observability and incident management.
- Fluent in both German and English for daily collaboration and documentation.
Salary for this position is market competitive. If you value ownership and modern AI technology, you can apply directly through the official neoshare channels.