Northbound AI Summit 2026

October 12 – 14, 2026 · Yerba Buena Center for the Arts, San Francisco, California

Full agenda .ics

Tuesday, October 13, 2026

9:00 AM PDT

Featured Keynote

Keynote: The Post-Cloud Developer

9:00 AM – 9:45 AM PDT · Main Hall

The console was never the product — it was the interim UI for infrastructure that could not yet describe itself. This keynote argues that the next platform shift is already visible at the edges: infrastructure declared next to application c…

ID

Inês Duarte

Developer Infrastructure Fellow · Local Cloud Foundation

9:30 AM PDT

Workshop

Hands-on: Observability for LLM Apps

9:30 AM – 11:00 AM PDT · Workshop Room B

You cannot fix what you cannot see, and most LLM apps ship blind. This workshop builds the observability stack for an AI application from first principles: structured traces for every model call and tool invocation, cost and latency attribu…

TN

Tomáš Novák

Observability Architect · Tracegarden

10:00 AM PDT

InnovationTalk

Local-first AI Apps

10:00 AM – 10:30 AM PDT · Room 305

The most reliable AI app is the one that keeps working in airplane mode. Local-first AI stopped being a curiosity when small models crossed the "good enough" line for summarization, classification, and retrieval over personal data — all wor…

KS

Kenji Sato

On-device ML Engineer · Pocket Models

11:00 AM PDT

AI InfrastructurePanel

Panel: The Economics of Inference

11:00 AM – 12:00 PM PDT · Main Hall

Everyone in this industry is spending someone else's margin. This panel brings together people who see the inference market from different seats — a capacity buyer at a scaled AI product, an economist covering compute markets, and an infras…

JB

Jordan Bell

Compute Economist · Capacity Index

IK

Isaac Kim

Capacity Strategy Lead · Tessellate Cloud

11:30 AM PDT

Talk

Caching Strategies for LLM APIs

11:30 AM – 12:00 PM PDT · Room A

The fastest and cheapest LLM call is the one you never make. But semantic caching — reusing an answer because the question is "close enough" — is a correctness gamble that has burned every team that treated it as a drop-in. This talk maps…

BL

Benjamin Liu

Principal Engineer · Cacheline AI

1:00 PM PDT

Talk

Structured Output at Scale

1:00 PM – 1:30 PM PDT · Workshop Room B

Parsing model output with regexes is how you end up debugging production at midnight. Constrained decoding and schema-first output turned our flakiest integration surface into the most boring one, and this talk covers how to get there at sc…

NB

Nia Brooks

Applied AI Lead · SchemaWorks

2:00 PM PDT

PracticeTalk

Agents that Ship: Case Studies

2:00 PM – 2:30 PM PDT · Main Hall

Three agents made it to production. One triages support tickets, one migrates legacy code, one runs infrastructure remediations. All three nearly died in month two, each for a different reason. This talk is the post-mortem series: the tria…

Sofía Álvarez

Sofía Álvarez

Staff Product Engineer · RelayWorks

3:00 PM PDT

Panel

Panel: Build vs Buy for AI Platforms

3:00 PM – 4:00 PM PDT · Room 305

Every platform team eventually faces the question: build the AI platform layer or buy it. Both answers are expensive and one of them is wrong for you specifically. This panel stages the argument properly, with a platform lead who built and…

FZ

Fatima Zahra

Head of Platform · Atlas Commerce

LH

Layla Hassan

VP of Engineering · Harbor Finance

4:00 PM PDT

Talk

Multimodal Pipelines in Practice

4:00 PM – 4:30 PM PDT · Room B

The interesting documents were never plain text. Invoices, engineering drawings, medical forms, dashboards — the high-value pipelines are the ones that read pixels and text together, and they fail in ways pure-text systems never prepared us…

ZA

Zara Amin

Document Intelligence Lead · Papertrail Health