Senior Voice AI Engineer Needed
Budget: ₹12,500 – ₹37,500 INR
Senior Voice AI Engineer
Location: Remote
Type: Full-time
Experience: 2-3 years
About Us
NovaMind Tech is an innovative AI startup revolutionizing customer engagement with intelligent automation. Our flagship product, Oriv AI, enables seamless communication across voice, WhatsApp, and SMS by leveraging real-time AI, speech recognition, and automation technologies. We are committed to building cutting-edge solutions that enhance business interactions and drive efficiency.
Role Overview
We are looking for a highly skilled and motivated Senior Voice AI Engineer to enhance our voice assistant platform, Oriv AI. This role requires expertise in real-time audio processing, speech recognition, and natural language understanding. The ideal candidate will be responsible for ensuring carrier-grade reliability, optimizing AI-driven communication systems, and integrating seamlessly with telephony providers.
Key Responsibilities
- Develop and optimize real-time voice processing pipelines using WebSocket streams.
- Enhance speech-to-text (STT) and text-to-speech (TTS) capabilities with AI-driven improvements.
- Improve multi-turn dialogue management and context-aware conversational AI.
- Integrate Oriv AI with leading telephony providers to enhance voice interactions.
- Implement advanced voice quality monitoring and analytics solutions.
- Ensure high availability, scalability, and security of AI-powered voice systems.
Required Skills
- Proficiency in Python, with expertise in FastAPI and asynchronous programming.
- Hands-on experience with WebSocket/WebRTC for real-time audio streaming.
- In-depth knowledge of speech processing technologies, including STT, TTS, and voice activity detection.
- Experience with cloud speech services (Google Cloud Speech, Deepgram).
- Strong understanding of SQL databases (PostgreSQL) and asynchronous ORM frameworks.
- Knowledge of telephony protocols, SIP, and AI-driven voice automation.
Preferred Qualifications
- Experience with Large Language Models (LLMs) and conversational AI.
- Familiarity with digital signal processing (DSP) and AI-powered audio enhancement.
- Hands-on experience with Docker and cloud-based deployments.
- Expertise in monitoring tools, observability, and performance optimization.
- Multi-language voice processing and AI-driven translation experience.
Technical Environment
- Backend: Python/FastAPI
- Speech Processing: Google Cloud Speech, Deepgram
- Real-time Audio: WebSocket streaming
- Database: PostgreSQL with SQLAlchemy
- Deployment: Docker/Cloud
What We Offer
- A remote-first, flexible work environment.
- An opportunity to work on cutting-edge AI-driven voice technologies.
- A dynamic, innovation-driven team focused on AI advancements.
- Career growth opportunities in a rapidly evolving tech landscape.
- Competitive compensation and performance-based incentives.
How to Apply
To apply for this role, please submit the following:
- Your updated resume.
- A brief description of a voice or audio processing system you have developed.
- Details of your experience with real-time audio streaming.
- GitHub profile or relevant portfolio (if available).
We are seeking a forward-thinking engineer who is passionate about AI-driven voice technology and eager to contribute to the future of intelligent customer interactions.
Location: Remote
Type: Full-time
Experience: 2-3 years
About Us
NovaMind Tech is an innovative AI startup revolutionizing customer engagement with intelligent automation. Our flagship product, Oriv AI, enables seamless communication across voice, WhatsApp, and SMS by leveraging real-time AI, speech recognition, and automation technologies. We are committed to building cutting-edge solutions that enhance business interactions and drive efficiency.
Role Overview
We are looking for a highly skilled and motivated Senior Voice AI Engineer to enhance our voice assistant platform, Oriv AI. This role requires expertise in real-time audio processing, speech recognition, and natural language understanding. The ideal candidate will be responsible for ensuring carrier-grade reliability, optimizing AI-driven communication systems, and integrating seamlessly with telephony providers.
Key Responsibilities
- Develop and optimize real-time voice processing pipelines using WebSocket streams.
- Enhance speech-to-text (STT) and text-to-speech (TTS) capabilities with AI-driven improvements.
- Improve multi-turn dialogue management and context-aware conversational AI.
- Integrate Oriv AI with leading telephony providers to enhance voice interactions.
- Implement advanced voice quality monitoring and analytics solutions.
- Ensure high availability, scalability, and security of AI-powered voice systems.
Required Skills
- Proficiency in Python, with expertise in FastAPI and asynchronous programming.
- Hands-on experience with WebSocket/WebRTC for real-time audio streaming.
- In-depth knowledge of speech processing technologies, including STT, TTS, and voice activity detection.
- Experience with cloud speech services (Google Cloud Speech, Deepgram).
- Strong understanding of SQL databases (PostgreSQL) and asynchronous ORM frameworks.
- Knowledge of telephony protocols, SIP, and AI-driven voice automation.
Preferred Qualifications
- Experience with Large Language Models (LLMs) and conversational AI.
- Familiarity with digital signal processing (DSP) and AI-powered audio enhancement.
- Hands-on experience with Docker and cloud-based deployments.
- Expertise in monitoring tools, observability, and performance optimization.
- Multi-language voice processing and AI-driven translation experience.
Technical Environment
- Backend: Python/FastAPI
- Speech Processing: Google Cloud Speech, Deepgram
- Real-time Audio: WebSocket streaming
- Database: PostgreSQL with SQLAlchemy
- Deployment: Docker/Cloud
What We Offer
- A remote-first, flexible work environment.
- An opportunity to work on cutting-edge AI-driven voice technologies.
- A dynamic, innovation-driven team focused on AI advancements.
- Career growth opportunities in a rapidly evolving tech landscape.
- Competitive compensation and performance-based incentives.
How to Apply
To apply for this role, please submit the following:
- Your updated resume.
- A brief description of a voice or audio processing system you have developed.
- Details of your experience with real-time audio streaming.
- GitHub profile or relevant portfolio (if available).
We are seeking a forward-thinking engineer who is passionate about AI-driven voice technology and eager to contribute to the future of intelligent customer interactions.