Vibeleaderboard
← All Intel
Intel / video

[State of Post-Training] From GPT-4.1 to 5.1: RLVR, Agent & Token Efficiency — Josh McGrath, OpenAI

Source
youtube.com
Author
Latent Space
Date
Why it matters

Explains how verifiable-reward RL and efficiency work shape the models you build agents on, which informs what to expect in , cost and latency.

Terms in this piece · Glossary
  • token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
  • tool use — A model's ability to call external functions — run code, search the web, edit files — instead of only generating text.
Read the source www.youtube.com
More from Latent Space
Recommended reads
Comments

Checking sign-in…

Loading comments…