Shop Vision Language Models by Merve Noyan

Vision Language Models by Merve Noyan

1,750.00

Close
Price Summary
  • 1,750.00
  • 1,750.00
  • 1,750.00
In Stock
Highlights:

BLACK & WHITE Final Release Version
Language ‏ : ‎ English / Size B5
Paperback, 409 Pages, Edition 2026
A+ PDF Printed On Demand Book!
Local Printed Book!
Delivery All Over Pakistan Charges Will Apply.
Due to constant currency fluctuation, prices are subject to change with or without notice.

Compare
Category: Tags: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , ,
Description

Vision Language Models: Building VLMs with Hugging Face

Merve Noyan, Andrés Marafioti, Miquel Farré, Orr Zohar

Vision language models (VLMs) combine computer vision and natural language processing to create powerful systems that can interpret, generate, and respond in multimodal contexts. Vision Language Models is a hands-on guide to building real-world VLMs using the most up-to-date stack of machine learning tools from Hugging Face, Meta (PyTorch), NVIDIA (Cuda), and others, written by leading researchers and practitioners Merve Noyan, Miquel Farré, Andrés Marafioti, and Orr Zohar. From image captioning and document understanding to advanced zero-shot inference and retrieval-augmented generation (RAG), this book covers the full VLM application and development lifecycle.
Designed for ML engineers, data scientists, and developers, this guide distills cutting-edge VLM research into practical techniques. Readers will learn how to prepare datasets, select the right architectures, fine-tune and deploy models, and apply them to real-world tasks across a range of industries.
Explore core model architectures and alignment techniques
Train and fine-tune VLMs with Hugging Face, PyTorch, and others
Deploy models for applications like image search and captioning
Implement advanced inference strategies, from zero-shot to agentic systems
Build scalable VLM systems ready for production use

Reviews (0)
0 ★
0 Ratings
5 ★
0
4 ★
0
3 ★
0
2 ★
0
1 ★
0

There are no reviews yet.

Be the first to review “Vision Language Models by Merve Noyan”

Your email address will not be published. Required fields are marked *

Scroll To Top
Close
Close
Close

My Cart

Shopping cart is empty!

Continue Shopping

Select at least 2 products
to compare