Document Conversion & Structuring for Legal Process
Budget: £250 – £750 GBP
Job Title
Document Processing (PDF, Metadata Extraction, Chronological Organisation)
Job Description
I need a freelancer to process a large number of documents for legal preparation.
This is a strict data handling task — no interpretation or summarising.
Scope of Work
1. Convert files to PDF
Emails, images, scans, etc.
Ensure all pages are clear and correctly oriented
2. Remove duplicates
Identify duplicate documents
Keep best quality version only
3. Extract metadata (for each document)
Place at the top of the first page of each document:
Date (DD/MM/YYYY if possible)
Sender
Recipient
Subject (if available)
If missing → write: “not available”
4. Organise documents
Sort ALL documents in strict chronological order (by date)
Oldest → newest
5. Assign document numbers (P-numbers)
Label each document:
P0001, P0002, P0003…
Numbering must follow date order
6. Page tracking (VERY IMPORTANT)
Each document (P-number) may contain multiple pages.
You MUST:
Track the exact page range of each P-number in the final PDF
Example:
P0001 = pages 1–3
P0002 = pages 4–4
P0003 = pages 5–9
7. Create a P-number index (REQUIRED)
Provide a separate file (Excel or PDF) containing:
P-number
Date
Page range in master PDF
Example format:
P0001 | 12/03/2015 | pages 1–3
P0002 | 14/03/2015 | page 4
P0003 | 20/03/2015 | pages 5–9
8. Create outputs
A. Master PDF
All documents combined
In strict date order
Each document clearly labelled with P-number
B. Text extraction file
Full extracted text from every page
Must include:
P-number
Page reference
STRICT RULES (IMPORTANT)
**No summarising
**No interpretation
**No guessing
**No rewriting
**Only extract visible data
**If unreadable → write: “not readable”
Goal
A fully structured legal bundle where:
Documents are in date order
Each has a P-number
Each P-number has a known page range
Every page can be referenced precisely
Requirements
Experience with PDF handling and OCR
Strong attention to detail
Ability to follow strict rules
To Apply
Please confirm:
You understand no interpretation / no guessing
You can track page ranges accurately
Budget
Open
Important Note
Accuracy is critical.
This is a technical structuring task, not a creative task.
Document Processing (PDF, Metadata Extraction, Chronological Organisation)
Job Description
I need a freelancer to process a large number of documents for legal preparation.
This is a strict data handling task — no interpretation or summarising.
Scope of Work
1. Convert files to PDF
Emails, images, scans, etc.
Ensure all pages are clear and correctly oriented
2. Remove duplicates
Identify duplicate documents
Keep best quality version only
3. Extract metadata (for each document)
Place at the top of the first page of each document:
Date (DD/MM/YYYY if possible)
Sender
Recipient
Subject (if available)
If missing → write: “not available”
4. Organise documents
Sort ALL documents in strict chronological order (by date)
Oldest → newest
5. Assign document numbers (P-numbers)
Label each document:
P0001, P0002, P0003…
Numbering must follow date order
6. Page tracking (VERY IMPORTANT)
Each document (P-number) may contain multiple pages.
You MUST:
Track the exact page range of each P-number in the final PDF
Example:
P0001 = pages 1–3
P0002 = pages 4–4
P0003 = pages 5–9
7. Create a P-number index (REQUIRED)
Provide a separate file (Excel or PDF) containing:
P-number
Date
Page range in master PDF
Example format:
P0001 | 12/03/2015 | pages 1–3
P0002 | 14/03/2015 | page 4
P0003 | 20/03/2015 | pages 5–9
8. Create outputs
A. Master PDF
All documents combined
In strict date order
Each document clearly labelled with P-number
B. Text extraction file
Full extracted text from every page
Must include:
P-number
Page reference
STRICT RULES (IMPORTANT)
**No summarising
**No interpretation
**No guessing
**No rewriting
**Only extract visible data
**If unreadable → write: “not readable”
Goal
A fully structured legal bundle where:
Documents are in date order
Each has a P-number
Each P-number has a known page range
Every page can be referenced precisely
Requirements
Experience with PDF handling and OCR
Strong attention to detail
Ability to follow strict rules
To Apply
Please confirm:
You understand no interpretation / no guessing
You can track page ranges accurately
Budget
Open
Important Note
Accuracy is critical.
This is a technical structuring task, not a creative task.
Related categories:
Data Processing
Data Entry
Excel
Web Scraping
Health & Medicine
PDF
Legal Research
OCR
Data Extraction
Data Analysis