AI Engineer
Sabio · Madrid · Spain · Hybrid
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
At Sabio Group, we're building the next generation of AI-powered customer experience for some of the world's most demanding enterprise brands — the kind of frontier customers who set the bar for what "great CX" looks like. We pair frontier-class large language models with deep contact centre expertise to deliver agentic experiences that materially improve containment, CSAT and operational efficiency.
We're hiring an AI Engineer to join our customer service bot / AI agent team in Madrid. You'll work end-to-end across the lifecycle — discovery, design, build, evaluate, deploy, operate — engineering robust LLM/NLU solutions, knowledge pipelines and integrations across voice and digital channels, primarily in Spanish.
This is a hands-on engineering role for someone who is curious about where AI is going, comfortable shipping agentic systems into production, and thoughtful about doing so safely and cost-effectively.
Discovery & Design
Translate CX and business needs into AI solution designs spanning NLU/LLM, RAG, ASR/TTS and back-end integrations.
Define functional and non-functional requirements: latency, resiliency, observability, security, cost-per-conversation.
Agent & Model Implementation
Design and build agentic AI solutions — single-agent, multi-agent, tool-using and orchestrated patterns — using modern frameworks and frontier models.
Build and optimise NLU/LLM components: classification, NER, summarisation, RAG over enterprise knowledge bases.
Engineer prompts, tool schemas, context windows and token budgets for accuracy, latency and cost.
Develop data and knowledge pipelines for ingestion, cleansing, PII redaction and evaluation datasets.
Evaluation & Quality
Build evaluation harnesses (offline eval sets, LLM-as-judge, regression suites, A/B testing) and treat eval as a first-class deliverable, not an afterthought.
Drive continuous improvement via conversation analytics, error triage and failure-mode analysis.
LLMOps & Operations
Implement CI/CD for models and prompts, with feature flags, canary releases and rollback.
Track experiments, lineage, and model/prompt versions.
Operate services in production against SLOs with monitoring, tracing, alerting and incident response.
Responsible AI
Apply practical guardrails for safety, bias, hallucination, prompt-injection and jailbreak resistance.
Enforce GDPR / LOPDGDD: consent, minimisation, retention, access control, auditability.
Collaboration
Work alongside conversation designers, software engineers, data scientists, QA and Ops.
Produce clear design docs, runbooks and stakeholder updates.
Required
Hands-on experience delivering conversational AI — voicebots, chatbots or virtual assistants — including conversational flows, open-ended interactions, NLU and generative AI.
Production experience building agents with generative AI: agentic architectures (single-agent, multi-agent, tool/function-calling, planner–executor patterns), RAG, and orchestration.
Working…