---
id: M05-L02
title: "Test instructions in a fresh session"
module: M05
chapters: "10"
checkpoint: CP03
---

# M05-L02 — Test instructions in a fresh session

Evaluate whether repository guidance affects actual work in a fresh session. Start with CP03 or later in a separate exercise folder and use the request below. CP04 later collects the complete skill/request matrix in `docs/skills.md`. Use a behavior already implemented at that checkpoint. The outcome is a prompt, actual action record, independent review, and verdict.

A continued conversation may carry the command and source facts from earlier messages. Starting a new task makes the repository responsible for supplying those details. A request to summarize active instructions is a useful diagnostic, but a summary alone does not establish that the workflow is followed.

Use this CP03 reporting-documentation case: “Clarify the README's synthetic-input explanation using the project brief. Keep the edit bounded and complete the documented replay check.” Do not paste the exact command into the task. At CP07 or later, adding a test for an implemented behavior is a separate alternative, not a CP03 prerequisite.

Observe context reads, scope, actual command use, and result interpretation. An announced intention to test is not execution. Inspect the changed explanation and check that its claims come from the applicable project brief and preserve synthetic-input limits. Then read the process result and output from the documented wrapper.

Record interface/version, checkpoint, instruction-file identity, prompt, observed actions, changed files, commands/results, and limitations. Independently inspect the diff. A passing test shows relevant guidance used, bounded work, actual available checks, and an accurate completion report. If a compiler is unavailable, the run may be blocked; it is not a passing execution result.

If the run fails, classify the issue before editing instructions. Wrong folder or a hidden extension is discovery. Contradictory active guidance is scope or wording. A nonexistent script is stale documentation. A denied operation is runtime capability or permission. Repair the smallest supported cause and repeat in another fresh session.

Figures SS10-03–05 show fresh task, command use, and repair/rerun. They must come from actual runs. The course's procedure defines success, while your saved evidence determines whether success occurred.

## Resources and completion

Use the Sensor Monitor `CP03` download and its `README.md`; project paths in this lesson are relative to that root. Read Chapter 10 for the full lab and explanatory review answers. Figure IDs: SS10-03, SS10-04, SS10-05. Primary references: [OpenAI AGENTS.md](https://learn.chatgpt.com/docs/agent-configuration/agents-md) and [OpenAI build skills](https://learn.chatgpt.com/docs/build-skills). Complete the [exercise](exercise.md), preserve actual evidence, and use the separate instructor answer key for self-check after attempting the task. Narration scripts are production sources; final transcripts must match the actual narrated edit.
