Serbian Text Annotation Dataset

Job ID: 40338296

Budget: $8 – $15 USD

I need help expanding GienTech’s Serbian language resources by adding high-quality annotations to a corpus that will be used for large-language-model training.

The raw text is already collected and will be delivered in segmented excel files. Your task is to apply the agreed-upon labels, follow the style guide precisely, and return the data in the same file structure so it can be ingested by our pipeline.

Project Background:
1. Overview: Perform multi-dimensional linguistic evaluation and annotation tasks, most of the tasks will be conducted online via the our company’s platform, a few will be completed offline.
2. Service Type: Text Annotation
3. Field: AI Data & Natural Language Processing
4. Language Pair: English (US) →Serbian (Cyrillic, Bosnia)
5. Project Start: After you passed the test, you will start the official task in the next week
6. Entry Requirement: Free test passed

Hiring Process: CV review >> Rate Negotiation >> Test >> Test passed & Onboarding >> Sign Contract >> Join in the Project

Candidates Requirements:
1. Translation and linguistic evaluation skills
2. Prior experience in QA evaluation, annotation, or similar AI/NLP projects is highly preferred
3. Strong command of the target language (Serbian Cyrillic/ Bosnia); native speakers preferred, but non-native candidates with strong proficiency are also welcome
4. Ability to quickly understand guidelines and follow multiple evaluation criteria accurately