Stanford CS329A Course on Self-Improving AI Agents Now Free on YouTube

Stanford University's graduate-level course CS329A, titled 'Self-Improving AI Agents,' has been made freely available on YouTube via the Stanford Online channel. The course is co-taught by Azalia Mirhoseini and Aakanksha Chowdhery, both of whom have hands-on industry experience building large-scale AI systems at organizations including Google DeepMind, Anthropic, and Meta. Mirhoseini is known for foundational work on Mixture-of-Experts architecture, AlphaChip, and LLM test-time scaling, while Chowdhery led training of the PaLM 540B model and multiple generations of Gemini pre-training. The curriculum covers topics such as test-time compute, robust verification, and reinforcement learning as tools for building agents that improve continuously through interaction rather than manual updates. The course materials are also referenced on the official website cs329a.stanford.edu, making the full learning path accessible to the public.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in