Multilingual Conversational Audio Data Collection
Budget: ₹12,500 – ₹37,500 INR
Need Large-Scale Conversational Audio Recordings (BPO/Call-Center, 2-Person Dialogues)
Description:
I am looking to procure large-scale conversational audio datasets in the following languages:
All Indian regional Languages
Urdu
Romanian
Czech
Malay
Croatian
Hungarian
Greek
Bulgarian
Hebrew (Modern)
Slovak
Dataset Requirements:
Type: Natural, real two-person conversations
Source Preference: BPO or Call-Center environments (customer–agent or agent–agent conversations)
Format:
High-quality audio recordings
Along with metadata such as duration, speaker gender (if available), conversation type, etc.
Content Guidelines:
Non-sensitive conversations preferred (general inquiries, support calls, etc.)
Must have proper rights, licenses, and legal clearance for data resale/use
No private/personal data should be included unless anonymized
Freelancer Requirements:
Must have a verified network or access to BPO/call-center data providers
Prior experience in audio dataset procurement or linguistic data collection
Ability to provide samples for verification before full procurement
Understanding of data licensing and compliance
Deliverables:
Full audio dataset per language
Metadata files
Legal clearance documentation
Budget:
Open to discussion depending on volume, language, and quality.
How to Apply:
Please include the following:
Languages you can supply
Estimated number of hours available
Sample data (if possible)
Pricing per hour or per dataset
Delivery timeline
Description:
I am looking to procure large-scale conversational audio datasets in the following languages:
All Indian regional Languages
Urdu
Romanian
Czech
Malay
Croatian
Hungarian
Greek
Bulgarian
Hebrew (Modern)
Slovak
Dataset Requirements:
Type: Natural, real two-person conversations
Source Preference: BPO or Call-Center environments (customer–agent or agent–agent conversations)
Format:
High-quality audio recordings
Along with metadata such as duration, speaker gender (if available), conversation type, etc.
Content Guidelines:
Non-sensitive conversations preferred (general inquiries, support calls, etc.)
Must have proper rights, licenses, and legal clearance for data resale/use
No private/personal data should be included unless anonymized
Freelancer Requirements:
Must have a verified network or access to BPO/call-center data providers
Prior experience in audio dataset procurement or linguistic data collection
Ability to provide samples for verification before full procurement
Understanding of data licensing and compliance
Deliverables:
Full audio dataset per language
Metadata files
Legal clearance documentation
Budget:
Open to discussion depending on volume, language, and quality.
How to Apply:
Please include the following:
Languages you can supply
Estimated number of hours available
Sample data (if possible)
Pricing per hour or per dataset
Delivery timeline