Head of AI Engineering (f/m/x)
Your mission
About neoshare
We’re a Munich-based AI-first fintech scale-up (founded 2019) with offices in Munich, Frankfurt, Berlin and Sofia. Our SaaS platform brings banks, investors, and advisors together to collaborate on complex financial deals making due diligence faster, smarter, and more transparent. Our AI features are already live with leading banks. Now we’re scaling.
The Role
Own and evolve our AI engineering function — transforming a 15–20 person ML team from research-heavy to a high-throughput, production-grade organization. You’ll partner with the CTO on strategy, build the platform that unifies LLM access, RAG, and backend services, and ship reliable, scalable AI features that change how banks work.
Key responsibilities
- Team leadership and org build
- Hire, mentor, and develop a high-performing team; set the technical bar, operating rhythms, and code/research review practices
- Organize sub-teams (e.g., Core Modeling, AI Platform/Infra, Integrations) with clear ownership, SLOs, and on-call
- Manage roadmap, capacity planning, and delivery across parallel initiatives
- Architecture and platform
- Own the LLM gateway: unified APIs and proxy layers for multi-provider routing (OpenAI, Gemini, Bedrock), with rate limits, fallbacks, and cost tracking
- Build high-performance RAG pipelines (ingestion, embeddings, vector stores, caching) with robust observability and safety guardrails
- Partner with Java/ NestJS teams to define clean async contracts, schemas, and eventing patterns; drive low-latency, scalable inference
- Model lifecycle and operations
- Lead end-to-end model and prompt lifecycle: data curation, training/fine-tuning, evaluation, deployment, rollback
- Establish LLMOps / MLOps : model/prompt registries, CI/CD, canary/A/B tests, offline/online evals, drift and cost monitoring
- Optimize inference throughput and cost (autoscaling, batching, quantization/distillation, caching)
- Strategy and collaboration
- Translate company goals into an AI/ML roadmap with measurable outcomes; balance exploration with reliability and cost
- Own build-vs-buy/vendor strategy for models, infrastructure, and data services; manage budgets and SLAs
- Governance and security
- Implement data privacy, security, and compliance practices (RBAC, secrets, auditability); track prompt/model lineage and reproducibility
- Define incident response, runbooks, and postmortems for AI features
Your profile
- 5+ years as a backend engineer and 4+ years leading AI/ML engineering in production (10+ years total experience ideal)
- Deep architecture expertise in Java (JVM) and/or Node.js ( NestJS ), distributed systems, APIs, microservices, and messaging/streaming
- Hands-on with LLM stacks: orchestration (e.g., LangChain / LlamaIndex or custom), vector DBs (Pinecone, Qdrant , FAISS), cloud AI (e.g., AWS Bedrock)
- Proven operation of systems at scale (millions of daily API calls) with strong SLOs, observability, and incident management
- MLOps foundations: model registries, experiment tracking, CI/CD, Kubernetes, IaC (e.g., Terraform), security best practices
- Excellent communication and stakeholder management; strong product sense focused on shipping user-facing feature
- Fluent German and English for daily team collaboration, stakeholder management, and technical documentation
- Experience with GPU/accelerator serving and optimization ( vLLM , TGI, Triton, ONNX Runtime)
- Cost optimization for LLM workloads (token budgets, dynamic routing, caching)
- Evaluation and safety/red-teaming for generative systems; startup/high-growth experience
- Platform: adoption of a unified LLM gateway; standardized observability and cost reporting
- Delivery: 2–3 user-facing AI features shipped with clear SLOs and measurable impact
- Reliability/cost: reduced average latency and cost per request; autoscaling and caching in place
- Org: sub-team structure established ; improved code quality and on-time delivery; targeted hiring completed
- Backend: Java (JVM), Node.js ( NestJS ); event-driven microservices; API gateways/proxies
- AI platform: Python, PyTorch , LLM orchestration, prompt pipelines/registry; vector DBs (Pinecone, Qdrant ); RAG services
- Infra/DevOps: AWS (incl. Bedrock), Kubernetes, Terraform, CI/CD, Observability ( OpenTelemetry , Prometheus/Grafana)
Why us?
International & Inclusive Team: Collaboration with diverse teams at our locations in Munich, Frankfurt, Berlin, and Sofia.
Modern & Dog-friendly Offices: Ergonomic, green, and inspiring for collaboration and productivity.
Flexibility: 30 vacation days, flexible working hours.
Special Time Off: Additional half-day off on Christmas Eve and New Year's Eve.
Workation: Work remotely for a limited period each year from selected destinations.
Wellbeing & Mobility Benefits: Support for well-being and sustainable lifestyle:
- Urban Sports/EGYM Club subsidy: Monthly support for your membership.
- Jobticket: 50% monthly subsidy for the Deutschlandticket.
- JobRad: Leasing of bicycles or e-bikes at attractive conditions.
Empfohlene Jobs
Data Engineering Lead, Agentic Infrastructure
Redefine the future of live entertainment tech Welcome to vivenu, the global leader in event ticketing tech and one of the world’s fastest-growing live entertainment tech firms. We are transformin…
SAP Supply Chain Planung (Senior) Berater/in (m/w/d)
Werde Teil unseres Teams bei SCPLAN GmbH, einem führenden Anbieter ganzheitlicher Beratungsdienstleistungen für Supply-Chain-Management. Unsere Mission ist es, Unternehmen durch die digitale Transfor…
Lagerhelfer/in (m/w/d) in Frankfurt am Main
Über uns TIME to Care – bei TimePartner setzen wir uns seit über 30 Jahren dafür ein, Karrieren voranzubringen. Unsere Mission ist es, dich bestmöglich zu unterstützen, damit du deinen beruflichen Z…
Praktikum Strategieberatung - Energie & Dekarbonisierung (w/m/d)
Du willst Veränderungsprozesse und transaktionsbezogene strategische Projekte in der Energiewirtschaft und für energieintensive Unternehmen begleiten? Bewirb Dich jetzt für ein Praktikum bei uns und …
(Senior) IT Management Consultant / Enterprise IT Advisor (m/w/d)
Das kannst Du bei uns tun ~ Strategische Beratung: Analyse bestehender IT-Strukturen und Entwicklung zukunftsfähiger, ganzheitlicher IT-Strategien (Enterprise Architecture / IT-Roadmaps) für unsere…
Selbstständig als Führungskräfte-Trainer (m/w/d)
Selbstständig als Führungskräfte-Trainer (m/w/d)02.08.2026 Crestcom über ABD Media GmbH deutschlandweit Weitere passende Anzeigen: Jobmailer Ihre Merkliste / Mit Klick auf einen Stern in d…
Ausbildung Zugverkehrssteuerer 2027 (w/m/d)
Die Zugverkehrssteuerung ist ein unverzichtbarer Teil im Eisenbahnverkehr. Hier sorgst du dafür, dass unsere Fahrgäste sicher und pünktlich an ihr Ziel gelangen. Zum 1. September 2027 suchen wir d…
Werkstudentin (m/w/d) Assistenz- und/oder Kursleiterin langfristig (!) Werkstudentin (m/w/d) Assistenz- und/oder Kursleiterin langfristig (!)
Ein gemeinnütziger Verein zur Förderung der MINT-Kompetenzen bei den Kindern ab ca. 5 Jahren. Unsere Kinder lernen auf eine “magische” Weise mit Mathematik umzugehen. Unsere Lehrmethoden sind sehr er…
Junior Logistik Projekt Manager (m/w/d)
Ihre Aufgaben: Unterstützung des Logistics Project Managers bei der Durchführung logistischer Aktivitäten in internationalen Projekten Koordination der Lieferkette durch Abstimmung und Kommun…
Mitarbeiter (m/w) Vertriebsinnendienst
Sie haben bereits erste Erfahrungen im Vertrieb sammeln können und der Umgang mit Kunden gehört zu Ihren Stärken? Dann nutzen Sie Ihre Chance und bewerben Sie sich bei uns! Die DIS AG vermittelt mi…