AivarinnovationsActive opening

Senior Voice AI Engineer — Convogent Delivery

Location flexibleOn-siteSeniorFound today
Apply to this job

Free credits included. Sign up to start applying with Jobfinder.

pythonvoice-aipipecatlivekittelephonytwiliokubernetesdocker

AIVAR INNOVATIONS

Senior Voice AI Engineer — Convogent Delivery

3-5 Years | Coimbatore/ Bangalore |Full-Time

About Aivar Innovations

Aivar Innovations is an AI-native services and software company and an AWS Preferred Partner, backed by Bessemer Venture Partners and Sorin Investments. We ship production-grade AI into enterprise environments across fintech, healthcare, and technology. Our work is anchored by four accelerator platforms: Convogent (voice and agent AI automation), Velogent (governed agentic process automation for regulated industries), Kubogent (Kubernetes-native AIOps), and Datagent (data and analytics). We measure ourselves by what runs in production, not by slideware.

Role Overview

Aivar Innovations is hiring a hands-on Senior Voice AI Engineer for its Convogent Delivery team. This role owns end-to-end production voice agent delivery for enterprise customers, including real-time voice pipelines, STT/LLM/TTS orchestration, telephony integration, deployment, scaling, and production reliability. This is not a chatbot, conversation design, managed-platform configuration, or pure ML research role. The ideal candidate has shipped live voice AI systems used by real customers, writes Python inside the voice pipeline, and can deliver production-grade systems on frameworks or infrastructure such as Pipecat, LiveKit Agents, Kubernetes, Docker, AWS/GCP, or coded telephony integrations.

What You'll Do

  • Own customer delivery outcomes. Take voice AI use cases from discovery through production go-live within committed timelines, working inside Rocketlane-tracked delivery plans alongside your Delivery Manager.
  • Customize and configure Convogent. Adapt Aivar's voice AI accelerator (built on Pipecat, LiveKit) to each customer's telephony stack (Twilio, Exotel, Cisco, Vonage, etc.), API/MCP integrations with customer CRM tools, knowledge base (KB) integrations, compliance constraints, and use-case requirements.
  • Engineer real-time audio pipelines for production. Optimize STT and TTS flows for the latency, concurrency, and call quality bar each customer engagement demands.
  • Troubleshoot customer production systems. When a live call does not sound right, quickly narrow down where in the pipeline it broke — capture, STT, orchestration, or TTS — by asking the right questions, reading the right signals, and driving the fix through to resolution.
  • Build and adapt the orchestration layer. Connect STT, TTS, and LLM agents into resilient, observable workflows with guardrails suited to the customer's risk and compliance profile.
  • Translate customer requirements into deployable solutions. Work directly with customer technical stakeholders alongside Aivar's SA and DM to scope, size, and iterate use cases against measurable outcomes — success rate, latency, adoption.
  • Report and govern delivery. Contribute to sprint tracking, status reporting, and stakeholder governance for the use cases you own.

What You'll Bring

  • Hands-on experience building and deploying voice agents with Pipecat, LiveKit, or similar frameworks — ideally across multiple distinct production deployments, not a single system.
  • Deep expertise in STT (Deepgram, Whisper, AssemblyAI, Google STT, Sarvam) and TTS (ElevenLabs, Cartesia, Amazon Polly).
  • Working knowledge of Speech-to-Speech (S2S) services, including Amazon Nova Sonic, and how they change pipeline and orchestration design compared to cascaded STT/LLM/TTS.
  • Experience with multi-speaker diarization and conversation segmentation.
  • Production fluency in Python plus at least one of Node.js, TypeScript, or Go.
  • Practical use of AI/ML libraries (TensorFlow, PyTorch, Scikit-learn) and conversational AI design — prompt engineering and integration with LLMs (GPT-4o, Claude, Bedrock models).
  • Cloud deployment experience on AWS (Lambda, EC2, S3, EKS, Polly, Transcribe, Bedrock) with Docker, Kubernetes, and CI/CD.
  • Demonstrated ability to work under SOW/contract constraints — scoping, prioritizing, and shipping within a defined timeline and budget across live customer engagements.
  • Comfort working across multiple concurrent customer engagements with differing requirements, tech stacks, and stakeholders.

Preferred Experience

  • Prior experience in a customer delivery, consulting, or systems integration role, rather than purely product/platform engineering.
  • Exposure to regulated industries (fintech, healthcare, BFSI) and their compliance and security constraints.
  • Experience with enterprise stakeholder management — running status calls, translating technical tradeoffs into business language.
  • Familiarity with low-code voice platforms alongside custom infrastructure.
  • Voice analytics, transcription tooling, and Indian-language STT/TTS models.
  • AWS or speech technology certifications.

Why This Role, Why Aivar

  • Work that ships to real customers. Every use case you build runs live on enterprise calls, measured by customer outcomes, not demos.
  • Own delivery end-to-end. From discovery call to production go-live, you are the engineering owner of the customer's voice AI outcome.
  • Build on accelerators that work. Convogent has real production traction — you configure and extend a working platform across diverse customer environments.
  • Direct customer exposure. Work face-to-face with enterprise technical stakeholders, not just internal teams.
  • Backed and proven. AI-native, AWS Preferred Partner, BVP- and Sorin-backed, with 100+ enterprise customers in the first year.
AivarinnovationsSenior Voice AI Engineer — Convogent Delivery
Apply to this job