Continuous Streaming Speech-to-Text Application Development

Job ID: 38798427

Budget: ₹1,500 – ₹12,500 INR

I am looking for a skilled developer to help me create a continuous streaming speech-to-text desktop application using OpenAI’s Whisper model. The application should take audio input in real time and provide text output continuously and synchronously. While the Whisper model is already set up to handle audio in chunks (WAV format) and provide responses quickly, the main challenge is integrating it into a seamless, continuous streaming system.

Project Requirements:
1. Model Integration: Utilize the existing Whisper model for speech-to-text conversion.
2. Real-Time Streaming: Ensure the system can handle continuous real-time audio input and produce text output synchronously.
3. Integration with PHP: The application should support integration using PHP for backend functionality.
4. Desktop Application: Build a user-friendly desktop application (a basic UI is already available) and improve upon it if needed.

Deliverables:
- A fully functional desktop application that performs continuous streaming and provides synchronous text output.
- Integration of the PHP backend with the application.
- Clear documentation for deployment and usage.

I would also appreciate any recommendations or improvements to optimize the application.

Please let me know your charges for this work and the estimated timeline for completion. I am looking forward to collaborating with someone who has experience with real-time systems and audio processing. Let me know if you need further details.