Vision Language Models: Building VLMs with Hugging Face - Marjane Mall - Image 1
222
00DH
444.00 DH
-50%

Vision Language Models: Building VLMs with Hugging Face

Livraison

Détails
Frais de livraison à partir de :
Livraison entre le Samedi 25 juillet 2026 et le Lundi 27 juillet 2026

À propos de cet article :

Marque : GENERIC
Vendu par HEAVENBOOKS.MA

Vision language models (VLMs) combine computer vision and natural language processing to create powerful systems that can interpret, generate, and respond in multimodal contexts. Vision Language Models is a hands-on guide to building real-world VLMs using the most up-to-date stack of machine learning tools from Hugging Face, Meta (PyTorch), NVIDIA ...

1

Mode de paiement

Paiement par carte bancaire
Carte marocaines
Paiement à la livraison
Paiement en espèce à la livraison
Politique de retours
Note de politique de retour

Description produit

Marque
GENERIC
Titre
Vision Language Models: Building VLMs with Hugging Face
Éditeur
O'Reilly Media
Type de produit
paperback
Présentation du livre
paperback
Date de sortie
7/14/2026 12:00:00 AM
Langue d'origine
English
ISBN
4390020161
Dimensions
7 x 2 x 9.19 inches
Nombre de pages
406 pages
Langue
English
Résumé
Vision language models (VLMs) combine computer vision and natural language processing to create powerful systems that can interpret, generate, and respond in multimodal contexts. Vision Language Models is a hands-on guide to building real-world VLMs using the most up-to-date stack of machine learning tools from Hugging Face, Meta (PyTorch), NVIDIA (Cuda), and others, written by leading researchers and practitioners Merve Noyan, Miquel Farré, Andrés Marafioti, and Orr Zohar. From image captioning and document understanding to advanced zero-shot inference and retrieval-augmented generation (RAG), this book covers the full VLM application and development lifecycle.Designed for ML engineers, data scientists, and developers, this guide distills cutting-edge VLM research into practical techniques. Readers will learn how to prepare datasets, select the right architectures, fine-tune and deploy models, and apply them to real-world tasks across a range of industries.Explore core model architectures and alignment techniquesTrain and fine-tune VLMs with Hugging Face, PyTorch, and othersDeploy models for applications like image search and captioningImplement advanced inference strategies, from zero-shot to agentic systemsBuild scalable VLM systems ready for production use Read more
Auteur
Merve Noyan, Andrés Marafioti, Miquel Farré, Orr Zohar
Date de parution
7/14/2026 12:00:00 AM