A curated list of high-signal resources — articles, books, courses, cookbooks, papers, playbooks, benchmarks, talks, podcasts, and newsletters — for agentic engineering and AI engineering.
74
82 commits
updated Sep 21, 2026
A curated list of high-signal resources — articles, books, courses, cookbooks, papers, playbooks, benchmarks, talks, podcasts, and newsletters — for agentic engineering and AI engineering.
This is a resources list, not a tools list. Open-source tools for building agentic systems live in the sister list awesome-production-agentic-systems; production ML tooling lives in awesome-production-machine-learning. This list covers the learning, design, and operational resources that sit alongside those tools — including both:
Agentic engineering focuses on using AI agents to do software engineering (Copilot, Cursor, Claude Code, Aider, Cline, Windsurf, Codex; spec-driven development; context engineering; agent IDE rules and memory files; SWE benchmarks). AI / agentic systems engineering focuses on building agentic and LLM-powered systems (architecture, RAG, memory, tool use & MCP, orchestration, multi-agent coordination, evaluation, observability, guardrails, safety, fine-tuning, inference, product/UX, economics, teams).
You can keep up to date by watching this repo for the monthly releases summarising newly added resources 🤩
This list was proposed in EthicalML/awesome-production-machine-learning#709 as a sister list focused on resources rather than tools.
Resources are tagged with icons so you can scan and filter at a glance:
| Icon | Meaning |
|---|---|
| ⭐ | Editors' pick — start here |
| 🆓 | Free to access |
| 💰 | Paid |
| 📘 | Book |
| 🧑🎓 | Course |
| 🎥 | Video / talk |
| 🎧 | Audio / podcast |
| 📄 | Paper |
| 🛠️ | Hands-on cookbook / tutorial |
| 📋 | Playbook / design-pattern catalog |
| 🧪 | Benchmark / leaderboard |
| 🏗️ | Reference implementation / case study |
| 📰 | Newsletter |
Resources are organised as a matrix: the top-level sections above (rows) are resource types, and each section is sub-divided by topic. The 21 topics, T1–T21, are shared across sections. This lets you read vertically ("what papers exist on RAG?") or horizontally ("where do I find resources on Coding Agents?").
Topics:
| # | Topic |
|---|---|
| T1 | Coding Agents & AI-Assisted Development (Copilot, Cursor, Claude Code, Aider, Cline, Windsurf, Codex) |
| T2 | Spec-Driven Development & Context Engineering (AGENTS.md, spec-kit, rules files) |
| T3 | Agent IDE Rules, Memory Files & Developer Workflows |
| T4 | SWE Benchmarks & Coding Evaluation |
| T5 | Autonomous Software Agents & Long-Horizon Engineering Tasks |
| T6 | LLM Application Architecture & System Design |
| T7 | Prompt Engineering |
| T8 | Retrieval-Augmented Generation (RAG) |
| T9 | Memory Systems & Long-Context |
| T10 | Tool Use, Function Calling & MCP |
| T11 | Orchestration, Planning & Design Patterns |
| T12 | Multi-Agent Systems & Coordination |
| T13 | Evaluation & Testing |
| T14 | Observability, Tracing & Debugging |
| T15 | Guardrails & Security (prompt injection, jailbreaks, red-teaming) |
| T16 | Safety, Alignment & Responsible AI |
| T17 | Fine-tuning, Post-training, RLHF & Reasoning Training |
| T18 | Inference, Serving, Cost & Latency |
| T19 | Voice, Multi-modal & Embodied Agents |
| T20 | Product, UX & Human-AI Interaction Design |
| T21 | Economics, Teams, Hiring & Org Design |
Coverage (● = populated, ○ = opportunistic / partial, — = out of scope for that row):
| Row \ Topic | T1 | T2 | T3 | T4 | T5 | T6 | T7 | T8 | T9 | T10 | T11 | T12 | T13 | T14 | T15 | T16 | T17 | T18 | T19 | T20 | T21 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Core & Foundations | ● | ● | ○ | ○ | ○ | ● | ● | ● | ○ | ● | ● | ○ | ● | ○ | ○ | ○ | ○ | ○ | ○ | ○ | ○ |
| Communities | ● | ○ | ○ | ○ | ○ | ● | ● | ● | ○ | ● | ● | ○ | ● | ● | ○ | ● | ● | ● | ○ | ● | ● |
| Courses | ● | ○ | ○ | ● | ○ | ● | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ● | ● | ● | ○ | ○ | ○ |
| Books | ● | ○ | ○ | — | ○ | ● | ● | ● | ○ | ● | ● | ○ | ● | ○ | ● | ● | ● | ● | ○ | ● | ● |
| Articles & Essays | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● |
| Tutorials & Cookbooks | ● | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ○ | ● | ● | ● | ○ | — |
| Playbooks & Patterns | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ○ | ● | ○ | ● | ● |
| Papers & Research | ● | ○ | — | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ○ |
| Benchmarks | ● | — | — | ● | ● | ○ | ○ | ● | ○ | ● | ● | ● | ● | ○ | ● | ● | ○ | ● | ● | ○ | — |
| Reference Impls | ● | ● | ● | ● | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ● | ● | ○ | ● | ● | ● | ● | ● |
| Talks & Conferences | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● |
| Podcasts | ● | ○ | ○ | ○ | ● | ● | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ● | ● | ● | ○ | ● | ● |
| Newsletters | ● | ○ | ○ | ○ | ○ | ● | ● | ● | ○ | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ○ | ● | ● |
The Trending / What's New, Milestones Timeline, Governance & Responsible AI, Product / UX / Economics, and Teams, Hiring & Org Design sections collapse across topics and are presented as curated lists rather than matrix cells.
Please review our CONTRIBUTING.md before submitting a PR — it explains the one-line description style, how to pick the right row/topic cell, and the quality bar for inclusion. Thank you to the community for supporting the list's growth 🚀
| You can join the Machine Learning Engineer newsletter. Join over 70,000 ML professionals and enthusiasts who receive weekly curated articles & tutorials on production Machine Learning. |
|
| Also check out Awesome Production Agentic Systems and Awesome Production Machine Learning, the sister lists of open-source tools for agentic systems and production ML respectively. |
|
Rotating pinned items: the most-discussed agentic & AI-engineering resources of the current cycle. Refreshed regularly — see CONTRIBUTING.md for nomination criteria.
Canonical "what is agentic engineering / AI engineering" reading. Start here.
Dated, field-defining events that shaped agentic & AI engineering.
| Date | Event | Reference |
|---|---|---|
| 2017-06 | Transformer architecture introduced | Attention Is All You Need |
| 2020-05 | GPT-3 shows in-context learning at scale | Language Models are Few-Shot Learners |
| 2020-05 | RAG framework introduced | RAG for Knowledge-Intensive NLP |
| 2021-06 | GitHub Copilot preview launches — first mainstream AI coding assistant | GitHub blog |
| 2022-01 | Chain-of-Thought prompting | Wei et al. |
| 2022-03 | InstructGPT / RLHF | Ouyang et al. |
| 2022-10 | ReAct: reasoning + acting agent loop | Yao et al. |
| 2022-11 | ChatGPT release — mainstream adoption inflection | OpenAI |
| 2023-03 | GPT-4 release | OpenAI |
| 2023-03 | HuggingGPT / Toolformer-era tool use | Toolformer |
| 2023-03 | LangChain & LlamaIndex hit mainstream | — |
| 2023-05 | Voyager: open-ended agents in Minecraft | Voyager |
| 2023-06 | Simon Willison coins "prompt injection" as a durable threat category | SW blog |
| 2023-10 | SWE-bench released — real-world coding eval | SWE-bench |
| 2023-12 | Mixture-of-experts open models (Mixtral) | Mistral |
| 2024-03 | Devin demo — autonomous software agent pitch | Cognition |
| 2024-05 | GPT-4o: native multi-modal + realtime voice | OpenAI |
| 2024-06 | Anthropic's "Building effective agents" publishes | Anthropic |
| 2024-07 | SWE-bench Verified launched | OpenAI |
| 2024-09 | o1 reveals reasoning-model era | OpenAI |
| 2024-11 | Model Context Protocol (MCP) announced | Anthropic |
| 2025-02 | Claude Code general availability | Anthropic |
| 2025-05 | AGENTS.md published as cross-agent standard | agents.md |
| 2025-06 | GitHub spec-kit / "new code" essays formalise spec-driven dev | spec-kit |
Discords, Slacks, forums, and meetups where practitioners gather.
Structured courses — free and paid, university and industry.
Published and in-progress books covering agentic & AI engineering.
Long-form writing from canonical authors and engineering teams.
.cursorrules files.Hands-on, code-first guides and official cookbooks from model providers and framework authors.
AGENTS.md files for common stacks..cursorrules examples.Opinionated, prescriptive guides distilling design patterns and operational practices.
aletheia-cli authority-diff CI gate, and Kit Certified stamped blueprints (Daniel Albinsson, 2026). Pairs with Agentic UX lifecycle patterns (T20).Foundational papers, surveys, and benchmark papers. Includes a dated milestone-papers table.
Public benchmarks and leaderboards for coding agents, tool use, RAG, evaluation, and more.
Public production write-ups and canonical reference repositories that teach by example.
Recorded talks, workshops, and conference series worth watching.
Recurring podcasts with strong agentic & AI-engineering coverage.
Weekly and monthly curated newsletters.
Policy frameworks, safety research, red-teaming resources, and responsible-AI guidance.
Going beyond engineering: designing for AI, human-AI interaction, and the economics of LLM applications.
How organisations structure AI-engineering work, hire for it, and operate sustainably.
Please use one of the issue templates (resource suggestion, broken link, or trending nomination) or open a pull request following the guidance in CONTRIBUTING.md. The curation methodology and update cadence are documented in NOTES.md.
Weekly: PR triage and broken-link fixes. Monthly: trending rotation and new-resource batches. Quarterly: full thoroughness pass against the checklist in NOTES.md.
— To the extent possible under law, the contributors have waived all copyright and related or neighboring rights to this work.
A curated list of high-signal resources — articles, books, courses, cookbooks, papers, playbooks, benchmarks, talks, podcasts, and newsletters — for agentic engineering and AI engineering.
74
82 commits
updated Sep 21, 2026
A curated list of high-signal resources — articles, books, courses, cookbooks, papers, playbooks, benchmarks, talks, podcasts, and newsletters — for agentic engineering and AI engineering.
This is a resources list, not a tools list. Open-source tools for building agentic systems live in the sister list awesome-production-agentic-systems; production ML tooling lives in awesome-production-machine-learning. This list covers the learning, design, and operational resources that sit alongside those tools — including both:
Agentic engineering focuses on using AI agents to do software engineering (Copilot, Cursor, Claude Code, Aider, Cline, Windsurf, Codex; spec-driven development; context engineering; agent IDE rules and memory files; SWE benchmarks). AI / agentic systems engineering focuses on building agentic and LLM-powered systems (architecture, RAG, memory, tool use & MCP, orchestration, multi-agent coordination, evaluation, observability, guardrails, safety, fine-tuning, inference, product/UX, economics, teams).
You can keep up to date by watching this repo for the monthly releases summarising newly added resources 🤩
This list was proposed in EthicalML/awesome-production-machine-learning#709 as a sister list focused on resources rather than tools.
Resources are tagged with icons so you can scan and filter at a glance:
| Icon | Meaning |
|---|---|
| ⭐ | Editors' pick — start here |
| 🆓 | Free to access |
| 💰 | Paid |
| 📘 | Book |
| 🧑🎓 | Course |
| 🎥 | Video / talk |
| 🎧 | Audio / podcast |
| 📄 | Paper |
| 🛠️ | Hands-on cookbook / tutorial |
| 📋 | Playbook / design-pattern catalog |
| 🧪 | Benchmark / leaderboard |
| 🏗️ | Reference implementation / case study |
| 📰 | Newsletter |
Resources are organised as a matrix: the top-level sections above (rows) are resource types, and each section is sub-divided by topic. The 21 topics, T1–T21, are shared across sections. This lets you read vertically ("what papers exist on RAG?") or horizontally ("where do I find resources on Coding Agents?").
Topics:
| # | Topic |
|---|---|
| T1 | Coding Agents & AI-Assisted Development (Copilot, Cursor, Claude Code, Aider, Cline, Windsurf, Codex) |
| T2 | Spec-Driven Development & Context Engineering (AGENTS.md, spec-kit, rules files) |
| T3 | Agent IDE Rules, Memory Files & Developer Workflows |
| T4 | SWE Benchmarks & Coding Evaluation |
| T5 | Autonomous Software Agents & Long-Horizon Engineering Tasks |
| T6 | LLM Application Architecture & System Design |
| T7 | Prompt Engineering |
| T8 | Retrieval-Augmented Generation (RAG) |
| T9 | Memory Systems & Long-Context |
| T10 | Tool Use, Function Calling & MCP |
| T11 | Orchestration, Planning & Design Patterns |
| T12 | Multi-Agent Systems & Coordination |
| T13 | Evaluation & Testing |
| T14 | Observability, Tracing & Debugging |
| T15 | Guardrails & Security (prompt injection, jailbreaks, red-teaming) |
| T16 | Safety, Alignment & Responsible AI |
| T17 | Fine-tuning, Post-training, RLHF & Reasoning Training |
| T18 | Inference, Serving, Cost & Latency |
| T19 | Voice, Multi-modal & Embodied Agents |
| T20 | Product, UX & Human-AI Interaction Design |
| T21 | Economics, Teams, Hiring & Org Design |
Coverage (● = populated, ○ = opportunistic / partial, — = out of scope for that row):
| Row \ Topic | T1 | T2 | T3 | T4 | T5 | T6 | T7 | T8 | T9 | T10 | T11 | T12 | T13 | T14 | T15 | T16 | T17 | T18 | T19 | T20 | T21 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Core & Foundations | ● | ● | ○ | ○ | ○ | ● | ● | ● | ○ | ● | ● | ○ | ● | ○ | ○ | ○ | ○ | ○ | ○ | ○ | ○ |
| Communities | ● | ○ | ○ | ○ | ○ | ● | ● | ● | ○ | ● | ● | ○ | ● | ● | ○ | ● | ● | ● | ○ | ● | ● |
| Courses | ● | ○ | ○ | ● | ○ | ● | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ● | ● | ● | ○ | ○ | ○ |
| Books | ● | ○ | ○ | — | ○ | ● | ● | ● | ○ | ● | ● | ○ | ● | ○ | ● | ● | ● | ● | ○ | ● | ● |
| Articles & Essays | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● |
| Tutorials & Cookbooks | ● | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ○ | ● | ● | ● | ○ | — |
| Playbooks & Patterns | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ○ | ● | ○ | ● | ● |
| Papers & Research | ● | ○ | — | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ○ |
| Benchmarks | ● | — | — | ● | ● | ○ | ○ | ● | ○ | ● | ● | ● | ● | ○ | ● | ● | ○ | ● | ● | ○ | — |
| Reference Impls | ● | ● | ● | ● | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ● | ● | ○ | ● | ● | ● | ● | ● |
| Talks & Conferences | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● | ● |
| Podcasts | ● | ○ | ○ | ○ | ● | ● | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ● | ● | ● | ○ | ● | ● |
| Newsletters | ● | ○ | ○ | ○ | ○ | ● | ● | ● | ○ | ● | ● | ○ | ● | ● | ● | ● | ● | ● | ○ | ● | ● |
The Trending / What's New, Milestones Timeline, Governance & Responsible AI, Product / UX / Economics, and Teams, Hiring & Org Design sections collapse across topics and are presented as curated lists rather than matrix cells.
Please review our CONTRIBUTING.md before submitting a PR — it explains the one-line description style, how to pick the right row/topic cell, and the quality bar for inclusion. Thank you to the community for supporting the list's growth 🚀
| You can join the Machine Learning Engineer newsletter. Join over 70,000 ML professionals and enthusiasts who receive weekly curated articles & tutorials on production Machine Learning. |
|
| Also check out Awesome Production Agentic Systems and Awesome Production Machine Learning, the sister lists of open-source tools for agentic systems and production ML respectively. |
|
Rotating pinned items: the most-discussed agentic & AI-engineering resources of the current cycle. Refreshed regularly — see CONTRIBUTING.md for nomination criteria.
Canonical "what is agentic engineering / AI engineering" reading. Start here.
Dated, field-defining events that shaped agentic & AI engineering.
| Date | Event | Reference |
|---|---|---|
| 2017-06 | Transformer architecture introduced | Attention Is All You Need |
| 2020-05 | GPT-3 shows in-context learning at scale | Language Models are Few-Shot Learners |
| 2020-05 | RAG framework introduced | RAG for Knowledge-Intensive NLP |
| 2021-06 | GitHub Copilot preview launches — first mainstream AI coding assistant | GitHub blog |
| 2022-01 | Chain-of-Thought prompting | Wei et al. |
| 2022-03 | InstructGPT / RLHF | Ouyang et al. |
| 2022-10 | ReAct: reasoning + acting agent loop | Yao et al. |
| 2022-11 | ChatGPT release — mainstream adoption inflection | OpenAI |
| 2023-03 | GPT-4 release | OpenAI |
| 2023-03 | HuggingGPT / Toolformer-era tool use | Toolformer |
| 2023-03 | LangChain & LlamaIndex hit mainstream | — |
| 2023-05 | Voyager: open-ended agents in Minecraft | Voyager |
| 2023-06 | Simon Willison coins "prompt injection" as a durable threat category | SW blog |
| 2023-10 | SWE-bench released — real-world coding eval | SWE-bench |
| 2023-12 | Mixture-of-experts open models (Mixtral) | Mistral |
| 2024-03 | Devin demo — autonomous software agent pitch | Cognition |
| 2024-05 | GPT-4o: native multi-modal + realtime voice | OpenAI |
| 2024-06 | Anthropic's "Building effective agents" publishes | Anthropic |
| 2024-07 | SWE-bench Verified launched | OpenAI |
| 2024-09 | o1 reveals reasoning-model era | OpenAI |
| 2024-11 | Model Context Protocol (MCP) announced | Anthropic |
| 2025-02 | Claude Code general availability | Anthropic |
| 2025-05 | AGENTS.md published as cross-agent standard | agents.md |
| 2025-06 | GitHub spec-kit / "new code" essays formalise spec-driven dev | spec-kit |
Discords, Slacks, forums, and meetups where practitioners gather.
Structured courses — free and paid, university and industry.
Published and in-progress books covering agentic & AI engineering.
Long-form writing from canonical authors and engineering teams.
.cursorrules files.Hands-on, code-first guides and official cookbooks from model providers and framework authors.
AGENTS.md files for common stacks..cursorrules examples.Opinionated, prescriptive guides distilling design patterns and operational practices.
aletheia-cli authority-diff CI gate, and Kit Certified stamped blueprints (Daniel Albinsson, 2026). Pairs with Agentic UX lifecycle patterns (T20).Foundational papers, surveys, and benchmark papers. Includes a dated milestone-papers table.
Public benchmarks and leaderboards for coding agents, tool use, RAG, evaluation, and more.
Public production write-ups and canonical reference repositories that teach by example.
Recorded talks, workshops, and conference series worth watching.
Recurring podcasts with strong agentic & AI-engineering coverage.
Weekly and monthly curated newsletters.
Policy frameworks, safety research, red-teaming resources, and responsible-AI guidance.
Going beyond engineering: designing for AI, human-AI interaction, and the economics of LLM applications.
How organisations structure AI-engineering work, hire for it, and operate sustainably.
Please use one of the issue templates (resource suggestion, broken link, or trending nomination) or open a pull request following the guidance in CONTRIBUTING.md. The curation methodology and update cadence are documented in NOTES.md.
Weekly: PR triage and broken-link fixes. Monthly: trending rotation and new-resource batches. Quarterly: full thoroughness pass against the checklist in NOTES.md.
— To the extent possible under law, the contributors have waived all copyright and related or neighboring rights to this work.