AIMeetup*
Cover for GPT-5 Evals Deep Dive — in-person event, in San Francisco, hosted by AIMeetup San Francisco
In Person· Hosted by AIMeetup San Francisco

GPT-5 Evals Deep Dive

How top teams run production evals for frontier models.

20 spots left30 / 50 going
MON, JUN 01 · 2026

17:00 – 20:00 PDT

San Francisco

San Francisco Innovation Hub, downtown

30 of 50 spots

60% filled

About this event

Join us in San Francisco to pull back the curtain on how elite research teams validate frontier models like GPT-5 before they hit production. We move beyond basic benchmarks to explore the rigorous, multi-stage evaluation pipelines currently securing the world’s most powerful AI systems. You will learn how engineers architect automated testing frameworks to catch subtle regressions and prevent model drift in high-stakes environments. Connect with fellow practitioners at our host venue to discuss the mechanics of human-in-the-loop feedback and robust red-teaming strategies. We dive deep into the specific challenges of scaling evals for multimodal reasoning and long-context performance, ensuring you leave with actionable insights for your own deployments.

Agenda

  1. 17:00

    Doors & networking

    Coffee, snacks, name tags.

  2. 17:30

    Opening talk

    GPT-5 Evals Deep Dive — the state of the art.

  3. 18:15

    Hands-on session

    Small groups, real problems.

  4. 19:15

    Lightning talks + Q&A

  5. 19:00

    Mingling & wrap-up

Speakers

PN

Priya Natarajan

Applied AI Lead, San Francisco

Ships production ML at scale.

CW

Chen Wang

Engineer, San Francisco

Building the next wave of AI tooling.

Location

San Francisco

San Francisco Innovation Hub, downtown

More meetups