June 29 – July 2, 2026 · San Francisco, CA · imported from ai.engineer's public schedule feed

AI Engineer World's Fair 2026 — unofficial import demo

Unofficial demo. This programme was imported from the AI Engineer World's Fair's own public schedule feed to show vibeboard at real conference scale. Not affiliated with, or endorsed by, the organisers.

All sessions
Context EngineeringSession

How long can your skills be before your agent forgets what you told it?

Laurie Voss

When
Wednesday, July 11:30 PM – 1:50 PM · 20 min
Where
Track 8San Francisco, CA · imported from ai.engineer's public schedule feed
Google Calendar

About this session

A year ago, frontier models lost the thread somewhere around 200 simultaneous instructions, so skills files had to stay short and lean on sub-skills and subagents. We re-ran IFScale on the 2026 frontier and found the ceiling has moved by an order of magnitude: closer to 2,000 instructions, up to 5,000 on the strongest models. The more interesting story is how models fail at the new frontier: DeepSeek quietly drops instructions, Opus refuses outright when innocuous words trip a safety classifier, Gemini burns its whole budget on reasoning and emits nothing, and GPT-5.5 stops to tell you your request was unreasonable. The capacity problem is largely solved; verification is wide open. We'll show the data, the failure modes, and what it costs to find out. You’ll come out with hard data on the ceiling for complex instructions to LLMs

Speaker

Laurie Voss
Laurie Voss

Head of Developer Relations, Arize AI

Laurie Voss is Head of Developer Relations at Arize AI, the leading company for AI observability and evaluations. He has been a developer for over 30 years and was co-founder of npm, Inc.. He believes passionately in making the web bigger, better, and more accessible for everyone.

More in Context Engineering