About this role
Job title: Sr AI Engineer- Voice & Agentic AI
About the Role The Senior AI Engineer will design, build, and ship production-grade AI systems focused on real-time, agentic, and voice applications where reliability and latency are critical.
What You'll Do
- Design and build LLM-powered agents and multi-step workflows using orchestration frameworks (LangGraph or equivalent), with explicit, auditable state.
- Build retrieval (RAG) systems that are accurate and fast to meet real-time latency budgets; balance recall and response time.
- Own AI features end to end from prototype to production deployment, monitoring, and iteration.
- Refactor and improve live production systems without disrupting active customers.
- Build and maintain evaluation and regression harnesses so changes are validated against scenarios, not demos; catch bugs before production.
- Create scalable Python APIs and services exposing AI functionality to web, voice, and mobile clients.
- Implement observability, tracing, and performance monitoring across deployed AI features—tracking latency, quality, and cost.
- Work across a multi-LLM stack (OpenAI, Anthropic, Google) behind a provider-abstraction layer; model choice is a configuration decision.
- Collaborate daily in a remote, sprint-based team, reviewing PRs through quality gates, and surfacing risks and blockers early.
What We're Looking For
- 4+ years in AI/ML or AI-focused software engineering with production products.
- Hands-on experience building LLM applications and agents (OpenAI, Claude, open-source) using LangGraph/LangChain or equivalent.
- Strong RAG and retrieval experience—embeddings, vector databases (Pinecone, Weaviate, pgvector), tuning for quality and speed.
- Solid Python with production APIs (FastAPI or similar).
- Experience evaluating and testing AI systems (RAGAS, Maxim, Promptfoo, LLM-as-judge, or equivalent) with a “done means tested, not demoed” mindset.
- Production experience on cloud (AWS preferred) with Docker, CI/CD, and infrastructure-as-code (Terraform).
- Strong software fundamentals: automated testing, code review, Git, modular design.
- Demonstrated experience deploying, monitoring, and refactoring AI systems in production.
- AI-native development workflow (Cursor, Claude Code, Copilot) and discipline to ship AI-assisted code safely through review and automated gates.
- Excellent written and spoken communication, strong autonomy, and comfort working in a remote team with at least 4 hours/day overlap with US Eastern time.
Nice to Have
- Real-time or low-latency voice/conversational systems experience (streaming STT/TTS pipelines, turn-taking, barge-in).
- Familiarity with telephony or real-time media infrastructure (Twilio, LiveKit, WebRTC, Deepgram, ElevenLabs).
- Multi-tenant or configuration-driven platform experience.
- Knowledge-base ETL / data ingestion pipeline experience.
- Exposure to AI guardrails, governance, and compliance in regulated domains.
Compensation & Benefits
- Access to Rootstrap University, conferences/certifications, and a mentorship program.
- Learning bonus opportunities to organize cross-functional initiatives and receive 360 feedback.
- Flexible remote work or office-based options in Montevideo, Buenos Aires, and Medellín, with flexible time schedule and Workation program.
- Gym benefits, psychological counseling, and weekly lunch reimbursements with special foods in the offices.
- Onboarding kit and access to cutting-edge technologies and tools.
- Prizes and gifts to celebrate achievements; People Care support for well-being and growth.
- We value well-being and happiness and support autonomy, creativity, and leadership.