AI:AM GUEST

Angela Yeung

SVP, Product, Cerebras

Angela Yeung is SVP of Product at Cerebras Systems, the wafer-scale computing company that powers OpenAI's Ultrafast tier, which serves GPT-5.6 Sol at roughly 750 output tokens per second. She owns the demand side of ultra-fast inference: which models run on the chip, what teams build once latency collapses, and where the next bottleneck moves when the silicon stops being the constraint. Before Cerebras she was a product manager at Google across Search, YouTube, Ads and Healthcare, and a director of product at Hinge Health.

APPEARANCES

One AI:AM appearance.

EPISODE 2026-08-31 · AUG 31, 2026

AI:AM LIVE — August 31, 2026 — Why the OpenFace Investigation Wasn't Enough, Gradient's Zach Bratun-Glennon on Betting on the Ecosystem, and Cerebras's Angela Yeung on the Moat That Stopped Holding

Nathan Labenz and Prakash Narayanan open on the OpenFace post-mortem, with Nathan arguing the independent investigators were given access too narrow — six days on site, about a thousand transcripts from a single seven-day window — for the public to treat the story as settled, and Prakash countering that scope and deadline are the price of getting a report at all. Gradient general partner Zach Bratun-Glennon explains why he now invests below the model layer, where the agent is the customer, or above it in end-to-end enterprise workflows, and why he'd bet on the ecosystem over any single leader. Cerebras SVP of Product Angela Yeung details wafer-scale inference, the microbatch of one, and a CUDA moat she says has eroded sharply now that AI can write the kernels — before a forty-five-minute close on frontier models as capable cyber attackers, RLVR and model deception, and an AI-generated song.

GUESTS · Zach Bratun-Glennon, Angela Yeung