LLM — A large language model — the neural network behind tools like Claude and ChatGPT, trained on huge amounts of text to predict what comes next.
multimodal — A model that works with more than text — reading images, audio, or video, and sometimes generating them too.
Why it matters
A single, thematically organized reference to 200+ of the most important 2025 LLMA large language model — the neural network behind tools like Claude and ChatGPT, trained on huge amounts of text to predict what comes next.Full definition → papers — grouped by reasoning, RL, and multimodalA model that works with more than text — reading images, audio, or video, and sometimes generating them too.Full definition → themes — saves researchers and practitioners hours of hunting and lets them go deep on the areas moving fastest.
Key quotes
“Also, as LLM research continues to be shared at a rapid pace, I have decided to break the list into bi-yearly updates.”
“This year, my list is very reasoning model-heavy. So, I decided to subdivide it into 3 categories: Training, inference-time scaling, and more general understanding/evaluation.”
“As you may see, much of the recent progress has centered around reinforcement learning (with verifiable rewards), which I covered in more detail in a previous article.”