# smolvlm

A family of 2B small vision-language models released by Hugging Face for multimodal image-and-text tasks.

This is one page of public article previews, not the complete archive. Follow Next page to continue. Summaries are not the original full articles.

## Get your VLM running in 3 simple steps on Intel CPUs

DevFeed: [Get your VLM running in 3 simple steps on Intel CPUs](<https://devfeed.tech/articles/get-your-vlm-running-in-3-simple-steps-on-intel-cpus-7431.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/openvino-vlm>)

Author: Ezequiel Lanza; Helena; Nikita; Ella Charlaix; Ilyas Moutawwakil

Published: 2025-10-15T00:00:00Z

Content type: tutorial

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [vlm](<https://devfeed.tech/topics/vlm.md>), [smolvlm](<https://devfeed.tech/topics/smolvlm.md>), [optimum](<https://devfeed.tech/topics/optimum.md>), [quantization](<https://devfeed.tech/topics/quantization.md>), [Inference](<https://devfeed.tech/topics/inference.md>), [intel](<https://devfeed.tech/topics/intel.md>)

Tags: [inference](<https://devfeed.tech/tags/inference.md>), [intel](<https://devfeed.tech/tags/intel.md>), [optimum](<https://devfeed.tech/tags/optimum.md>), [quantization](<https://devfeed.tech/tags/quantization.md>), [smolvlm](<https://devfeed.tech/tags/smolvlm.md>), [vlm](<https://devfeed.tech/tags/vlm.md>)

### AI overview

A tutorial explains how to run SmolVLM locally with Optimum Intel and OpenVINO, then optimize it for lower memory use and faster inference through quantization.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## SmolVLM - small yet mighty Vision Language Model

DevFeed: [SmolVLM - small yet mighty Vision Language Model](<https://devfeed.tech/articles/smolvlm-small-yet-mighty-vision-language-model-7484.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/smolvlm>)

Author: Andres Marafioti; merve; Miquel Farré; Elie Bakouch; Pedro Cuenca

Published: 2024-11-26T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [smolvlm](<https://devfeed.tech/topics/smolvlm.md>), [vlm](<https://devfeed.tech/topics/vlm.md>), [multimodal-ai](<https://devfeed.tech/topics/multimodal-ai.md>), [Transformers](<https://devfeed.tech/topics/transformers.md>), [Fine-tuning](<https://devfeed.tech/topics/fine-tuning.md>), [Inference](<https://devfeed.tech/topics/inference.md>), [Open Source Models & Datasets](<https://devfeed.tech/topics/open-source-models-datasets.md>), [GPU](<https://devfeed.tech/topics/gpu.md>)

Tags: [apache](<https://devfeed.tech/tags/apache.md>), [fine-tuning](<https://devfeed.tech/tags/fine-tuning.md>), [gpu](<https://devfeed.tech/tags/gpu.md>), [inference](<https://devfeed.tech/tags/inference.md>), [llm](<https://devfeed.tech/tags/llm.md>), [multimodal](<https://devfeed.tech/tags/multimodal.md>), [multimodal-ai](<https://devfeed.tech/tags/multimodal-ai.md>), [nlp](<https://devfeed.tech/tags/nlp.md>), [on-device](<https://devfeed.tech/tags/on-device.md>), [open](<https://devfeed.tech/tags/open.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [smolvlm](<https://devfeed.tech/tags/smolvlm.md>), [transformers](<https://devfeed.tech/tags/transformers.md>), [trl](<https://devfeed.tech/tags/trl.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

This article introduces SmolVLM, a family of small, fast, memory-efficient 2B vision-language models released fully open source under the Apache 2.0 license. It describes the model variants, architecture, training resources, Transformers integration, demo, fine-tuning script, and efficient local or on-device deployment.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Introducing Idefics2: A Powerful 8B Vision-Language Model for the community

DevFeed: [Introducing Idefics2: A Powerful 8B Vision-Language Model for the community](<https://devfeed.tech/articles/introducing-idefics2-a-powerful-8b-vision-language-model-for-the-community-7272.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/idefics2>)

Author: Leo Tronchon; Hugo Laurençon; Victor Sanh

Published: 2024-04-15T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [vlm](<https://devfeed.tech/topics/vlm.md>), [multimodal](<https://devfeed.tech/topics/multimodal.md>), [smolvlm](<https://devfeed.tech/topics/smolvlm.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [datasets](<https://devfeed.tech/topics/datasets.md>), [foundation-models](<https://devfeed.tech/topics/foundation-models.md>), [Fine-tuning](<https://devfeed.tech/topics/fine-tuning.md>), [Transformers](<https://devfeed.tech/topics/transformers.md>), [Computer vision](<https://devfeed.tech/topics/computer-vision.md>)

Tags: [computer-vision](<https://devfeed.tech/tags/computer-vision.md>), [cv](<https://devfeed.tech/tags/cv.md>), [datasets](<https://devfeed.tech/tags/datasets.md>), [fine-tuning](<https://devfeed.tech/tags/fine-tuning.md>), [model](<https://devfeed.tech/tags/model.md>), [multimodal](<https://devfeed.tech/tags/multimodal.md>), [nlp](<https://devfeed.tech/tags/nlp.md>), [ocr](<https://devfeed.tech/tags/ocr.md>), [open](<https://devfeed.tech/tags/open.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [performance](<https://devfeed.tech/tags/performance.md>), [research](<https://devfeed.tech/tags/research.md>), [transformers](<https://devfeed.tech/tags/transformers.md>), [vision](<https://devfeed.tech/tags/vision.md>), [vlm](<https://devfeed.tech/tags/vlm.md>)

### AI overview

Idefics2 is an 8B open multimodal vision-language model that accepts text and images and generates text responses. It supports image question answering, visual description, multi-image storytelling, document information extraction, and basic arithmetic, with enhanced OCR and integration with Transformers for fine-tuning.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.