You'll be the person who builds the machinery behind our AI calling agents. At JoyzAI, that means the real-time voice pipeline itself — audio in, speech understood, LLM reasoning, tools called, speech out — and everything around it: the telephony that gets the call connected, the dashboard a client uses to configure their agent, and the webhooks and CRM plumbing that turn a call into a lead.
This is not a feature-ticket role. It's a build-and-own role. You'll own the voice stack end to end, from the socket that streams audio to the screen a client uses to change their agent's greeting. When a call stutters, drops, or talks over the customer, you're the one who finds out why — and fixes it.
What you'll do
Voice pipeline & real-time systems
Build and harden the real-time calling pipeline: telephony ↔ audio streaming ↔ speech/real-time models ↔ tool calls ↔ speech out, with latency you can't feel.
Own the hard voice problems: turn-taking and barge-in, silence and voicemail detection, noise and echo, language switching mid-call (English ↔ Hinglish ↔ regional),
and graceful recovery when a model or carrier hiccups.
Integrate and evaluate voice models and providers across STT/TTS, real-time LLMs, VAD, and telephony carriers — and help decide what we build on next.
Implement non-blocking tool calling inside live calls — lookups, bookings, CRM writes — so the agent never goes silent or loses the thread.
Platform & full-stack product
Build the product around the calls: agent configuration (tasks, behaviour, guardrails, AI config), call logs and transcripts, analytics, and bulk-calling campaign controls.
Design and ship the Node.js/Express services and APIs — call orchestration, queueing, webhooks, retries, recording and transcript storage — that hold up as call volume grows.
Wire calls into the rest of the platform: CRM contacts and tickets, WhatsApp follow-ups after a call, calendar and booking integrations, Google Sheets and Meta lead-form ingestion.
Ship clean, usable
📌 Full Stack Engineer (Noida)
🏢 JoyzAI
📍 Noida