Looking for Speech Dataset Creation Team (English, Hindi & Hinglish)

Job ID: 40570347

Budget: ₹150,000 – ₹250,000 INR

We are looking for an experienced freelancer or data collection company to create a high-quality speech dataset.

### Project Requirements:
1. **Languages**: English, Hindi, Hinglish
2. **Duration Per Audio Clip**: Each sentence should be around 10–25 seconds.
3. **Audio Specifications**:
- Studio-quality recordings (48 kHz WAV)
- Balanced male/female speakers across different age groups
4. **Emotion and Expression Tags During Annotation**:
- <laugh>, <chuckle>, <sigh>, <gasp>, <nervous>, <frustrated>, <whispers>, <annoyed>, <sad>, <thoughtful>, <pause>, <slow>, <rushed>, <stammers>, <shouting>, <happy>,<excited>,<sarcastic>,<angry>
5. JSONL file deliverables with metadata including:
```
{
"audio_filepath":"/path/to/audio.wav",
"duration": 4.0,
"text": "Sample sentence with emotion tag",
"speaker": "speaker-name"
}
```
6. All data must be human-evaluated to ensure quality and annotation accuracy.
7. Scripts: The freelancer will need to generate custom conversational scripts covering diverse domains and scenarios. We will not provide scripts.

### Deliverables:
1. WAV audio files meeting the above specifications
2. JSONL metadata as specified in the requirements
3. Human-verified quality checks

### Required Experience:
Applicants should have experience in at least one of the following fields:
- **Speech Dataset Creation**
- **Voice Data Collection**
- **Audio Annotation**
- **Speech Corpus Development**

Please include in your proposal:
a) Details of similar speech/audio dataset projects you have completed.
b) Sample audio files along with corresponding JSON/JSONL metadata.
c) Recording setup and quality assurance process.
d) Team size involved.
e) Estimated pricing and delivery timeline.
f) Confirmation that all scripts, recordings, and metadata comply with the outlined specifications.