Multilingual Conversational Audio Data Collection

Job ID: 39996985

Budget: ₹12,500 – ₹37,500 INR

Need Large-Scale Conversational Audio Recordings (BPO/Call-Center, 2-Person Dialogues)

Description:
I am looking to procure large-scale conversational audio datasets in the following languages:

All Indian regional Languages

Urdu

Romanian

Czech

Malay

Croatian

Hungarian

Greek

Bulgarian

Hebrew (Modern)

Slovak

Dataset Requirements:

Type: Natural, real two-person conversations

Source Preference: BPO or Call-Center environments (customer–agent or agent–agent conversations)


Format:

High-quality audio recordings

Along with metadata such as duration, speaker gender (if available), conversation type, etc.

Content Guidelines:

Non-sensitive conversations preferred (general inquiries, support calls, etc.)

Must have proper rights, licenses, and legal clearance for data resale/use

No private/personal data should be included unless anonymized

Freelancer Requirements:

Must have a verified network or access to BPO/call-center data providers

Prior experience in audio dataset procurement or linguistic data collection

Ability to provide samples for verification before full procurement

Understanding of data licensing and compliance

Deliverables:

Full audio dataset per language

Metadata files

Legal clearance documentation

Budget:

Open to discussion depending on volume, language, and quality.

How to Apply:

Please include the following:

Languages you can supply

Estimated number of hours available

Sample data (if possible)

Pricing per hour or per dataset

Delivery timeline