This is the website for the AI agents class (CS 839, Spring 2026) at UW-Madison.
For submitting assignments, reports, reviews, etc., see canvas.
See course info for class structure, presentation advice, and compute resources.
We begin by providing a class overview and a deep dive into how to build LLM-based AI agents using a variety of techniques and frameworks
Tue: Class overview
Thu: Building agents
We will start by covering "classic" papers that took LLMs from next-word-prediction machines to programs that can reason and act.
Tue:
Thu:
see also Building Effective Agents
This week, we will talk about coding (or software engineering) agents and benchmarks, which has so far been the most successful application of AI agents
Tue:
Thu:
This week, we will cover "deep research" agents, which are agents that generate comprehensive reports on a topic.
Tue:
Thu:
We will now switch attention to agents whose goal is to automate or augment the scientific process.
Tue:
Thu:
AI agents, by virtue of their transformer architecture, mix instructions and data, creating a huge security risk. The following papers study this issue
Tue:
Thu:
Over the next two weeks, we will cover a range of ideas for optimizing and improving agents.
Tue:
Thu:
Tue:
Thu:
We will continue the discussion of agents that search the space of possible solutions.
Tue:
Thu:
We now switch attention to multi-agent systems.
Tue:
Thu:
This week, we will see how agents can use memory to improve their performance and explore some other topics.
Tue:
Thu:
Work on project week
Project presentation week
Project presentation week
18 commits
This is the website for the AI agents class (CS 839, Spring 2026) at UW-Madison.
For submitting assignments, reports, reviews, etc., see canvas.
See course info for class structure, presentation advice, and compute resources.
We begin by providing a class overview and a deep dive into how to build LLM-based AI agents using a variety of techniques and frameworks
Tue: Class overview
Thu: Building agents
We will start by covering "classic" papers that took LLMs from next-word-prediction machines to programs that can reason and act.
Tue:
Thu:
see also Building Effective Agents
This week, we will talk about coding (or software engineering) agents and benchmarks, which has so far been the most successful application of AI agents
Tue:
Thu:
This week, we will cover "deep research" agents, which are agents that generate comprehensive reports on a topic.
Tue:
Thu:
We will now switch attention to agents whose goal is to automate or augment the scientific process.
Tue:
Thu:
AI agents, by virtue of their transformer architecture, mix instructions and data, creating a huge security risk. The following papers study this issue
Tue:
Thu:
Over the next two weeks, we will cover a range of ideas for optimizing and improving agents.
Tue:
Thu:
Tue:
Thu:
We will continue the discussion of agents that search the space of possible solutions.
Tue:
Thu:
We now switch attention to multi-agent systems.
Tue:
Thu:
This week, we will see how agents can use memory to improve their performance and explore some other topics.
Tue:
Thu:
Work on project week
Project presentation week
Project presentation week
18 commits