Senior AI Engieer - Voice & Agentic Systems
- AI
- AWS
- Azure
- GCP
- Databricks
- Snowflake
- LangGraph
- LangSmith
- CI/CD
- Machine Learning
- Python
- FastAPI
- Design Systems
- Claude Code
- Cursor
- WebRTC
- Twilio
- SIP
- OpenAI
- Gemini
- MCP
- TypeScript
- React.js
Xebia is a global AI-first, digital transformation, and engineering partner.ย With over 25 years of experience and a team of 5,000 professionals across 16 countries, we help organizations design and build scalable products, platforms, and data-driven solutions.ย
We specialize in Artificial Intelligence, Data and Cloud, Intelligent Automation, and Digital Products, combining deep technical expertise with a strong focus on engineering excellence and a people-first culture.ย ย
In the CEE region, weโre a team of nearly 1,000 experts delivering modern applications, data platforms, and AI solutions for clients such as Millennium, ING, Play, Edenred, Arabian Drilling, FedEx, Leroy Merlin, Truecaller, Volotea, Schmitz Cargobull, and many, many more. We work with leading technologies including AWS, Azure, GCP, Databricks, and Snowflake, and combine strong engineering culture with a consulting mindset and a continuous focus on growth and knowledge sharing.ย
About project:
We are looking for a Senior AI Engineer to join a product team building a voice-first AI agent that allows users to interact through natural spoken conversations. The agent reasons, uses tools, and takes actions in real time, with the product currently in active development and its architecture and product direction still evolving.
This is a hands-on engineering role covering the full voice-agent stack, from real-time communication and speech processing to agent orchestration, tool use, memory, and backend integrations. You will have significant influence over both the technical architecture and product decisions, working closely with product, design, and a small engineering team.
You will be:
Agent engineering
- designing and building agent workflows using LangGraph, including state management, tool and function calling, multi-step reasoning, memory, error recovery, and human handover,
- defining how the agent makes decisions during live conversations where latency and interruptions are critical,
- building and maintaining the backend integrations that allow the agent to take actions,
- designing reliable agent workflows that can operate in real-world production environments.
Voice pipeline
- building and tuning real-time voice pipelines using technologies such as Pipecat, LiveKit Agents, or similar,
- working with streaming STT, TTS, voice activity detection, turn-taking, and interruption handling,
- improving conversation quality through end-of-turn detection, barge-in handling, filler and back-channeling, and graceful error recovery,
- measuring and reducing end-to-end latency across the voice and agent pipeline,
- ensuring reliable voice interactions under real-world network conditions.
Architecture and quality
- shaping the overall system architecture and making build-vs-buy decisions for models, speech technologies, and other components,
- building evaluation into the development workflow through conversation-level test sets, regression suites, LLM-as-judge approaches, and latency and quality metrics,
- implementing observability and tracing using tools such as Langfuse, LangSmith, or similar solutions,
- diagnosing production issues through traces, metrics, and real user sessions,
- applying strong engineering practices including code reviews, automated testing, CI/CD, documentation, and maintainable design.
Product
- translating user needs and product goals into practical technical proposals,
- challenging assumptions and suggesting alternative approaches when they can improve the user experience,
- prototyping ideas quickly and hardening successful approaches for production,
- communicating technical trade-offs clearly to both technical and non-technical stakeholders,
- contributing to regular in-person team sessions for planning, design, and technical reviews.
Your profile:
- proven track record of shippingGenAI and agentic systems to production that are used by real users,
- hands-on experience withLangGraph, including designing and running production workflows,
- strongPython skills, including asynchronous programming and building production services such as FastAPI applications,
- strong software engineering and system architecture skills, with the ability to design systems end to end and make informed technical trade-offs,
- senior or lead-level engineering experience with ownership of significant technical decisions,
- strong product sense and evidence of influencing what gets built, not only how it is implemented,
- strong understanding of production-quality software development, testing, observability, and reliability,
- fluent English with the ability to communicate clearly with both technical and non-technical stakeholders,
- willingness to participate in regular in-person team sessions.
Practical experience using AI-powered assistants (e.g. Claude Code, GitHub Copilot, Cursor) to improve productivity, quality, or decision-making in software delivery.
Work from the European Union region and a work permit are required.
Strongly preferred:
- hands-on experience buildingvoice agents or real-time voice pipelines,
- experience withPipecat, LiveKit Agents, or similar voice-agent frameworks,
- practical experience with streaming STT/TTS technologies such as Deepgram, Soniox, ElevenLabs, or Cartesia,
- experience with voice activity detection, turn-taking, interruption handling, and voice latency optimization,
- production experience withWebRTC and/or LiveKit,
- experience with telephony technologies such as Twilio, Telnyx, or SIP.
Nice to have:
- experience with LLM evaluation and observability tooling such as Langfuse, LangSmith, or custom evaluation frameworks,
- experience working with multiple model providers such as OpenAI, Anthropic, and Gemini,
- experience routing between models based on cost, latency, and capability,
- experience with MCP or similar tool-integration patterns,
- enough TypeScript and React experience to contribute to the client-side voice interface,
- experience working in a startup or 0-to-1 product environment.
Experience applying GenAI in a more structured way within the SDLC, including defined workflows, prompt patterns, or tool integrations embedded into daily work.
- Interest in and familiarity with emerging AI-driven practices (e.g. agent-based workflows, automation patterns, AI-augmented development), with a willingness to explore and experiment beyond standard approaches.
Work mode:
- working remotely from Europe, ideally within 1โ2 hours of CET,
- participating in in-person team meetings every 4โ6 weeks,
- being open to collocating in Berlin, with Madrid or Barcelona also possible alternatives.
ย
Recruitment Process:
CV review โ HR call โInterview โClientInterview โDecision
ย
Senior AI Engieer - Voice & Agentic Systems ยท Poland and Eastern Europe