Multimodal AI Engineer Needed
Budget: $30 – $250 USD
I'm in need of a seasoned AI researcher or engineer who can consolidate various LLaMA-based models into a single, comprehensive multimodal model. This model should be adept at processing vision, audio, and text inputs seamlessly. The primary use-case for this model will be as an interactive AI assistant.
Key Responsibilities:
- Merge multiple LLaMA-based models into one unified model
- Ensure the model can handle vision, audio, and text inputs
- Design the model for use as an interactive AI assistant
Ideal Skills and Experience:
- Extensive knowledge and experience in AI and machine learning
- Proven track record with LLaMA-based models
- Experience designing interactive AI models
- Strong ability to create models with full integration of multiple input modes
Key Responsibilities:
- Merge multiple LLaMA-based models into one unified model
- Ensure the model can handle vision, audio, and text inputs
- Design the model for use as an interactive AI assistant
Ideal Skills and Experience:
- Extensive knowledge and experience in AI and machine learning
- Proven track record with LLaMA-based models
- Experience designing interactive AI models
- Strong ability to create models with full integration of multiple input modes