Module 1 · Foundations of Agentic AI · scripted
Exercise: The Agent's Eyes
The goal: let the agent see what you see
Your first agents drove the conversation and you reported back in words. Now upgrade the channel: in this exercise you answer the agent with photographs. The Agent Lab's input bar has a 📷 button — attach an image from your computer or phone (on a phone, it opens the camera), and it is sent into the conversation as part of your message. The agent sees it.
The exercise at a glance
- 1Use a picture to get a task done.
- 2Use a picture to get feedback on your work.
- 3Turn a picture into a simulation.
The exercise playground
The playground is the same Agent Lab you used for your first
agents — everything you built earlier is still there (your saved
agents and conversations included). The one new thing you need: the
input bar's 📷 button attaches an image to your message. The
Prompts view works too: an attached image shows up in the log as
[image attached], one more thing on the growing prompt.
You can also do the exercise in ChatGPT, Claude, or Gemini instead — all of them accept photos. If you're going to use one of those tools, click here for the instructions →.
How the playground works
- 1Open it: /exercises/agent-lab — or scan the QR code on the next card.
- 2Attach photos with the 📷 button in the input bar.On a phone it opens the camera; on a computer it picks a file.
- 3The photo goes into the conversation as part of your message — the agent sees it.
Scan to open the Agent Lab

Step 1 · Use a picture to get a task done
Step 1 · What you'll do
- 1Take a picture of something around you — your desk, a bookshelf, the contents of a bag, a schedule on your screen.
- 2Send it with a task: "plan the cleanup" · "pick my next three reads and an order" · "what am I forgetting for this trip?" · "turn this schedule into a plan for tomorrow".
- 3Watch for the details it uses that you never typed — that's the picture doing work words would have dropped.
The picture is the context; the task tells the agent what the context is for. The test of a good pairing: the agent's answer uses things from the image you would never have thought to mention.
Capture — end of Step 1
Copy into your course document and save:
- 1The photo and the task you gave with it.
- 2One thing the agent used from the image that you had not mentioned — and would not have typed.
Step 2 · Use a picture to get feedback on your work
Step 2 · What you'll do
- 1As a group, draw a plan or an idea — on paper or a whiteboard — or screenshot something you're working on.
- 2Photograph it and send it three times, changing the role:· "Act as a skeptic. Poke holes in this — how does it fail in ways we haven't thought of?" · "What are the gaps and ambiguities in this?" · "What are the critical questions we should be asking?"
- 3Answer its follow-up questions and see how the critique sharpens.
This is the whiteboard move from the lesson, aimed at your own work. The same drawing, prompted three ways, gives you a skeptic, a gap-finder, and a facilitator — and a critique session your group can actually argue with.
Capture — end of Step 2
Copy into your course document and save:
- 1The drawing you photographed.
- 2The strongest single criticism or question the agent produced — the one your group had not thought of.
Step 3 · Turn a picture into a simulation
Step 3 · What you'll do
- 1Draw a process diagram or a user interface sketch and photograph it. (A process from your research works well.)
- 2Send the photo with the simulation prompt (below).
- 3Interact with it: step through the process, or "click" around the interface, and see whether the simulation stays true to your drawing.
The simulation prompt
A drawing of a system is enough for the agent to become the system: a persona-based simulation, generated from a photograph. This is prototyping with a pen — you can test a process or an interface before anything exists.
Before we regroup
Keep one document for this course. Copy each item into it and save — we will compare the best of each, and the best paper-drawn system of the session.
Copy into your course document and save:
- 1The diagram from Step 3 and your opening prompt.
- 2One exchange from the simulation — and one place where it followed your drawing exactly, or departed from it.
Be ready to discuss:
- The photos went into the conversation like any other message. What does that mean for the cost of a conversation full of photos — and for what should happen to old ones?
- Where in your research would an agent's eyes replace a measurement you currently type in by hand?