Deep (Learning) Focus
Subscribe
Sign in
Home
Notes
The Author
Archive
About
Agentic RL: Frameworks and Best Practices
How LLMs are trained to handle long horizon tasks in complex environments...
READ THE LATEST
Most Popular
View all
Decoder-Only Transformers: The Workhorse of Generative LLMs
Mar 4, 2024
•
Cameron R. Wolfe, Ph.D.
177
16
10
Demystifying Reasoning Models
Feb 18, 2025
•
Cameron R. Wolfe, Ph.D.
292
10
30
Agent Evaluation: A Detailed Guide
May 18
•
Cameron R. Wolfe, Ph.D.
304
21
50
Understanding and Using Supervised Fine-Tuning (SFT) for Language Models
Sep 11, 2023
•
Cameron R. Wolfe, Ph.D.
94
5
8
Latest
Top
Discussions
Agent Evaluation: A Detailed Guide
Best practices and common patterns for effectively evaluating AI agents...
May 18
•
Cameron R. Wolfe, Ph.D.
304
21
50
RL Scaling Laws for LLMs
How scaling laws have evolved from pretraining to reinforcement learning...
Apr 20
•
Cameron R. Wolfe, Ph.D.
132
1
17
The Anatomy of an LLM Benchmark
Common patterns used to create the most effective LLM evaluation datasets...
Mar 30
•
Cameron R. Wolfe, Ph.D.
112
3
17
Applying Statistics to LLM Evaluations
Most LLM evaluations are conducted without a deep consideration of statistics.
Mar 9
•
Cameron R. Wolfe, Ph.D.
142
9
14
Rubric-Based Rewards for RL
Extending the benefits of large-scale RL training to non-verifiable domains...
Feb 16
128
11
17
Continual Learning with RL for LLMs
Exploring the impressive continual learning capabilities of RL training...
Jan 26
•
Cameron R. Wolfe, Ph.D.
153
15
19
GRPO++: Tricks for Making RL Actually Work
How to go from the vanilla GRPO algorithm to functional RL training at scale...
Jan 5
•
Cameron R. Wolfe, Ph.D.
148
10
18
Olmo 3 and the Open LLM Renaissance
Fully-open artifacts with the potential to make LLM research a reality for anyone...
Dec 15, 2025
•
Cameron R. Wolfe, Ph.D.
85
7
14
Group Relative Policy Optimization (GRPO)
How the algorithm that teaches LLMs to reason actually works...
Nov 24, 2025
•
Cameron R. Wolfe, Ph.D.
126
10
14
See all
Deep (Learning) Focus
I contextualize and explain important topics in AI research.
Subscribe
Recommendations
View all 13
💎DiamantAI
Nir Diamant
Ahead of AI
Sebastian Raschka, PhD
LLM Watch
Pascal Biese
AI by Hand ✍️
Prof. Tom Yeh
The Founders Corner®
Ruben Dominguez
Deep (Learning) Focus
Subscribe
About
Archive
Recommendations
Sitemap
This site requires JavaScript to run correctly. Please
turn on JavaScript
or unblock scripts