# A Deepdive into Aya Vision: Advancing the Frontier of Multilingual Multimodality

DevFeed: [A Deepdive into Aya Vision: Advancing the Frontier of Multilingual Multimodality](<https://devfeed.tech/articles/a-deepdive-into-aya-vision-advancing-the-frontier-of-multilingual-multimodality-7116.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/aya-vision>)

Author: Saurabh Dash; Yiyang Nan; Arash Ahmadian; John Dang

Published: 2025-03-04T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [aya](<https://devfeed.tech/topics/aya.md>), [vlm](<https://devfeed.tech/topics/vlm.md>), [cohere](<https://devfeed.tech/topics/cohere.md>), [Benchmark](<https://devfeed.tech/topics/benchmark.md>), [datasets](<https://devfeed.tech/topics/datasets.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [Large Language Model](<https://devfeed.tech/topics/llm.md>)

Tags: [aya](<https://devfeed.tech/tags/aya.md>), [benchmark](<https://devfeed.tech/tags/benchmark.md>), [cohere](<https://devfeed.tech/tags/cohere.md>), [datasets](<https://devfeed.tech/tags/datasets.md>), [language-models](<https://devfeed.tech/tags/language-models.md>), [llm](<https://devfeed.tech/tags/llm.md>), [multimodal](<https://devfeed.tech/tags/multimodal.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [performance](<https://devfeed.tech/tags/performance.md>), [synthetic](<https://devfeed.tech/tags/synthetic.md>), [technical](<https://devfeed.tech/tags/technical.md>), [techniques](<https://devfeed.tech/tags/techniques.md>), [text-generation](<https://devfeed.tech/tags/text-generation.md>), [vision](<https://devfeed.tech/tags/vision.md>), [vlm](<https://devfeed.tech/tags/vlm.md>)

## AI overview

Cohere For AI introduces Aya Vision, an open-weight multilingual and multimodal model family supporting language and vision understanding across 23 languages. The article describes its training techniques, benchmark results, open-weight releases, and image-processing architecture, including dynamic image tiling and latency improvements.

## Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.