Aya Vision - Multilingual, Multimodal AI from Cohere

byβ€’
Aya Vision, from Cohere For AI, is the open-weights, multilingual, multimodal models (8B & 32B). Outperforms larger models on multilingual vision tasks. Available on Hugging Face and Kaggle.

Add a comment

Replies

Best

Hi everyone!

Check out Aya Vision, a new set of open-weights models from Cohere For AI, and this is a significant step towards making AI truly global! Most vision-language models are heavily biased towards English. Aya Vision tackles this head-on by supporting 23 languages spoken by over half the world's population.

Here's why it's important:

🌍 Multilingual by Design: Excels at understanding and generating text and processing images/videos across a wide range of languages.
πŸ–ΌοΈ Multimodal: Handles both images/videos and text.
πŸš€ Outperforms Larger Models: Cohere claims Aya Vision (8B and 32B versions) outperforms models many times their size (like Llama 3 90B!) on multilingual multimodal tasks.
πŸ”“ Open Weights: Available on Hugging Face and Kaggle.
πŸ“± Free on WhatsApp: You can even try Aya for free on !

They're also releasing a new benchmark, , specifically for evaluating multilingual multimodal performance. The goal is to build AI that understands the nuances of different cultures and languages, not just add more languages.

I can see this making a huge impact! Great job on the launch. πŸ”₯

Aya Vision makes video meetings feel personal! πŸ‘€ Loving the immersive experience. Like if you’re all about better virtual connections! Wishing you clear and engaging meetings ahead!