Senior Back End Developer - Real-Time Audio & Voice Systems (Hyderabad)

Senior Back End Developer - Real-Time Audio & Voice Systems (Hyderabad)

20 Sep
|
Hire3 Labs
|
Hyderabad

20 Sep

Hire3 Labs

Hyderabad

Senior Back End Developer - Real-Time Audio & Voice Systems

Hyderabad, Telangana

Role Overview

We are seeking a Senior Backend Engineer - Real-Time Audio & Voice Systems to design, build, and operate low-latency, highly reliable voice and audio pipelines that power conversational Al platform. This role requires deep hands-on experience with real-time audio streaming, speech-to-text (STT), text-to-speech (TTS), and telephony integrations, using Python and Node.js in production environments.

You will play a critical role in building scalable, fault-tolerant systems for high-volume voice interactions.

Key Responsibilities

· Design, develop, and operate real-time audio processing pipelines for conversational voice agents.

· Build and maintain backend services using Python (FastAPI/async frameworks) and Node.js for low-latency workflows.

· Implement bi-directional audio streaming using WebSockets, WebRTC,or similar protocols.

· Integrate and optimize Speech-to-Text (STT) and Text-to-Speech (TTS) engines (cloud-based or self-hosted).

· Build and maintain telephony integrations (inbound/outbound calling, call routing, call recording, DTMF, conferencing).

· Optimize audio latency, jitter, and reliability across distributed systems.

· Handle high-concurrency workloads and real-time session state management.

· Implement monitoring, logging, and alerting for voice pipelines and call flows.

· Collaborate with AI/LLM teams to integrate speech pipelines with conversational logic.

· Own production deployments, on-call rotations, and incident response for voice systems.

Required Qualifications

· 4+ years of backend engineering experience, with a strong focus on real-time systems.





· Strong hands-on experience with Python and Node.js in production environments.

· Proven experience building real-time audio or voice applications.

· Deep understanding of STT and TTS pipelines, audio codecs, and streaming concepts.

· Solid experience integrating telephony platforms (inbound/outbound calls, SIP, PSTN).

· Experience with WebSockets, WebRTC, or RTP-based streaming.

· Solid understanding of asynchronous programming and event-driven architectures.

· Experience with Docker and CI/CD pipelines.

· Ability to debug and optimize distributed, latency-sensitive systems.

Nice-to-Have Skills

· Experience with LLM-powered voice assistants and conversational Al systems.

· Familiarity with audio codecs (Opus, PCM, u-law) and sampling strategies.

· Experience with cloud infrastructure (AWS/GCP) and autoscaling real-time services.

· Exposure to message queues, Redis, or real-time state stores.

· Prior experience in healthcare, contact centers, or telecom domains.

Why Join Interactly.ai?

· High-impact role: Own mission-critical voice infrastructure at scale.

· Challenging engineering problems: Build low-latency, real-time Al systems.

· Growth-stage startup: Influence architecture and technical direction from early stages.

· Global collaboration: Work closely with India- and US-based teams.

· Career progression: Opportunity to grow into Staff or Principal Engineering roles





Requirements added by the job poster

· 3+ years of work experience with Node.js

· 3+ years of work experience with Python (Programming Language)

Pay: ₹200,000.00 - ₹4,000,000.00 per year

Application Question(s)

- 1. How many years of Total experience do you have?

- 1. How many years of Python (FastAPI, asynchronous frameworks) experience do you have?

- 1. How many years of Node.js experience do you have?

- 1. How many years of Real-Time & Audio Streaming Protocols experience do you have?

WebSockets, WebRTC, RTP-based streaming, Bi-directional audio streaming, Real-time audio processing pipelines, Latency and jitter optimization across distributed systems,
- 1. How many years of Speech & Audio Technologies experience do you have?

Speech-to-Text (STT) engines (cloud-based or self-hosted), Text-to-Speech (TTS) engines (cloud-based or self-hosted), Audio codecs (Opus, PCM, μ-law), Audio sampling strategies,
- 1. How many years of Telephony & Communications Infrastructure experience do you have?

Telephony platforms integration (SIP, PSTN), Inbound and outbound calling systems, Call routing, call recording, DTMF, and conferencing,
- 1. How many years of Backend & Systems Architecture experience do you have?

Asynchronous programming, Event-driven architectures , High-concurrency workload management, Real-time session state management, Debugging and optimizing distributed, latency-sensitive systems, Message queues,
- 1. How many years of AI & Conversational Systems experience do you have?

- 1. How many years of Cloud, DevOps & Operations experience do you have?

- 1. What is your current notice period?

Work Location: In person

📌 Senior Back End Developer - Real-Time Audio & Voice Systems (Hyderabad)
🏢 Hire3 Labs
📍 Hyderabad

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior back end developer - real-time audio & voice systems (hyderabad) / hyderabad

Subscribe to this job alert:

Get the latest job offers by email for: senior back end developer - real-time audio & voice systems (hyderabad) / hyderabad