Hugging Face team ships 'Building VLMs' book: hands-on guide to vision-language models

andimarafioti · x · 2026-09-16

Hugging Face multimodal team members Merve Noyan, Miquel Farré, Andrés Marafioti, and Orr Zohar have released a new book, Vision-Language Models — Building VLMs with Hugging Face, available on Amazon and O'Reilly with open-source code and notebooks.

Julien Chaumond amplified the announcement, and Merve Noyan called it her favorite.

Original post →

More from Multimodal

Multimodal channel →