Python API to Send Images to ChatGPT Vision and Return Enhanced Image Based on Prompt
Budget: $250 – $750 USD
Project Description:
I’m looking for an experienced Python developer to create an API that sends images directly to ChatGPT (specifically the vision-capable model, e.g. GPT-4 Vision) and receives back an enhanced image based on a specific command prompt.
The goal is to automate a workflow where a user uploads an image (e.g., a product photo), and the API communicates with ChatGPT Vision to process and return a visually improved version of the image, such as:
• Background removal
• Lighting and contrast enhancement
• Color correction
• Optimization for e-commerce, social media, or catalogs
Key Features:
1. Build a REST API (FastAPI or Flask) with an endpoint for image upload.
2. Convert the image to base64 and send it to ChatGPT Vision along with a specific prompt (e.g., “Remove background and enhance lighting for e-commerce”).
3. Receive and return the modified image to the user (if supported by the model), or at least provide a detailed description of the necessary editing steps.
4. Structure the output as JSON and/or include a downloadable image link.
Technical Requirements:
• Python 3.x
• FastAPI or Flask
• Experience with OpenAI API (GPT-4 Vision)
• Image handling and base64 encoding
• Ability to parse and manage the model’s response
Nice to Have:
• Experience with image processing libraries (e.g., PIL, OpenCV)
• Simple frontend for testing uploads and responses
• Clear technical documentation
Timeline: MVP within 1–2 weeks
I’m looking for an experienced Python developer to create an API that sends images directly to ChatGPT (specifically the vision-capable model, e.g. GPT-4 Vision) and receives back an enhanced image based on a specific command prompt.
The goal is to automate a workflow where a user uploads an image (e.g., a product photo), and the API communicates with ChatGPT Vision to process and return a visually improved version of the image, such as:
• Background removal
• Lighting and contrast enhancement
• Color correction
• Optimization for e-commerce, social media, or catalogs
Key Features:
1. Build a REST API (FastAPI or Flask) with an endpoint for image upload.
2. Convert the image to base64 and send it to ChatGPT Vision along with a specific prompt (e.g., “Remove background and enhance lighting for e-commerce”).
3. Receive and return the modified image to the user (if supported by the model), or at least provide a detailed description of the necessary editing steps.
4. Structure the output as JSON and/or include a downloadable image link.
Technical Requirements:
• Python 3.x
• FastAPI or Flask
• Experience with OpenAI API (GPT-4 Vision)
• Image handling and base64 encoding
• Ability to parse and manage the model’s response
Nice to Have:
• Experience with image processing libraries (e.g., PIL, OpenCV)
• Simple frontend for testing uploads and responses
• Clear technical documentation
Timeline: MVP within 1–2 weeks