All drills

Build a Disciplined Coding Agent

What you'll be able to do

Build the habit behind every reliable coding agent — plan before you write, critique before you ship — and prove with real numbers what it costs and buys, not just claim it helps.

The discipline behind Claude Code's plan mode, Cognition's Devin, Cursor, and GitHub Copilot Workspace — all publicly wrestling with the same reliability problem this course builds from scratch.
Start this internship
Create an account to unlock the 10 sections, the workbench, and AskThili.
Begin

Sections

1. Build a Disciplined Coding Agent
🔒 locked
2. Lesson 1 - The Efficient Agentic-Coding Playbook
🔒 locked
3. Lesson 2 - The Undisciplined Baseline
🔒 locked
4. Lesson 3 - The Planning Gate
🔒 locked
5. Lesson 4 - Grading a Plan
🔒 locked
6. Lesson 5 - The Critique Gate
🔒 locked
7. Lesson 6 - Closing the Loop
🔒 locked
8. Lesson 7 - Measuring the Discipline
🔒 locked
9. Lesson 8 - A Real Model, Not a Script
🔒 locked
10. Lesson 9 - Packaging It as a Skill
🔒 locked

Dig deeper

🔗thilidiscipline — the reference implementation you build in this course
code
📄Reflexion: Language Agents with Verbal Reinforcement Learning (Shinn et al., 2023)
paper
📄Self-Refine: Iterative Refinement with Self-Feedback (Madaan et al., 2023)
paper
🔗Ollama — run open LLMs locally
docs