June 29 – July 2, 2026 · San Francisco, CA · imported from ai.engineer's public schedule feed

AI Engineer World's Fair 2026 — unofficial import demo

Unofficial demo. This programme was imported from the AI Engineer World's Fair's own public schedule feed to show vibeboard at real conference scale. Not affiliated with, or endorsed by, the organisers.

All sessions
Vision & OCRSponsor Session

Skill issue: stop deploying vision language models, use them with Skills to build e2e vision apps on edge

Merve Noyan

When
Tuesday, June 3011:40 AM – 12:00 PM · 20 min
Where
Track 2San Francisco, CA · imported from ai.engineer's public schedule feed
Google Calendar

About this session

With the boom of vision language models barrier of entry to build vision apps are much lower so developers tend to use them right away. However, these models are very large and inefficient in production. In this talk, I will go through combining vision language models with Skills to build end-to-end vision apps from training to deployment using HF Skills, on top of showing the state-of-the-art in small computer vision/multimodal models.

Speaker

Merve Noyan
Merve Noyan

MLE, Hugging Face

Works at Hugging Face open-source team, author of the book Vision Language Models with Hugging Face published by O'Reilly.

More in Vision & OCR