Unit 4 activity: Choosing the tool

DEVELOPMENT REVIEW DEPLOYMENT - NOT READY FOR RELEASE

Time: about 14 minutes. Read page.md first, at least as far as the decision aid.

Unit 1 asked you to sort tasks. This one asks you to choose, and then to notice what the choice cost you.

Part 1 — the decision aid, on one real task (6 minutes)

Take one task from the table you built in Unit 1 — ideally the one you moved into a cheaper class. Work the five steps and write down each answer.

Step Your answer
1. Which class is this?
2. Can I move it? To what?
3. What will the answer contain that needs checking?
4. Can I actually do that check, with the time and access I have?
5. Which is the cheapest kind that leaves me a check I can do?

Step 3 is where most of the value is, so be concrete. "Sources" is not an answer; "three citations I would have to look up in the library catalogue" is. "A number" is not an answer; "a currency conversion whose date I would have to confirm" is.

Step 4 has only two honest answers, and one of them is uncomfortable. If it is "no", say so and go back to step 2: either supply something that moves the class, or accept that what you will get is a proposal rather than a result.

Part 2 — the same question, two kinds (5 minutes)

Now ask the same question twice: once of a plain generation system, and once of a system that can fetch or be given a source. If you only have access to one kind, simulate the second by pasting in the source yourself.

Keep both answers. Then write three or four sentences on this:

What is in the second answer that I can check, that I could not check in the first?

Note what the question is not. It is not which answer is better written, and it is not which answer is right — you may not know that yet. It is about what each answer gives you to work with.

Two things to look for specifically:

Part 3 — price it (3 minutes)

For the task in Part 1, write one line each:

That second line is the one this unit is for. An extra layer earns its cost when it removes a check you could not otherwise do. It does not earn it when it produces more claims you were never going to verify.

Now take the checkpoints

cp04-what-changes-the-check gives you four observations about an answer and asks which check each one makes necessary. cp04-cheapest-adequate is about Part 3's judgement.

python3 selfcheck.py run cp04-what-changes-the-check
python3 selfcheck.py run cp04-cheapest-adequate  # or in the browser

Run the Python route from course/lab/, or use the browser self-check.

Optional extensions

Optional. These sit outside the core study time, are not required for the checkpoints, and nothing later in the module depends on them.

Break a chart reading (about 6 minutes). Find a chart with a truncated axis, a log scale, or values that fall between gridlines, and ask a multimodal system to read a specific value off it. Then read the value yourself. Record what it said, what the value is, and — the useful part — how confident the wording was. This is the fastest way to see why an unsearchable reference is a different problem from an unreliable one.

Find a citation that is real and does not support the claim (about 8 minutes). Ask a retrieval-backed system a question in an area you know, and check not whether the sources exist but whether they say what the answer says they say. The failure you are looking for has a name in Unit 9 — this is a preview of it, and finding one yourself is worth more than being shown one.

Marking guidance — open this once you have done the activity, to check your own work

Read this after attempting the activity.

Part 1 — what the steps are testing

Step 3 and step 4 are the whole exercise; steps 1, 2 and 5 are bookkeeping around them.

A strong step 3 names artefacts, not categories: which citations, which number, which part of which image. A weak one names a kind of thing, which is the same failure Unit 0's verification sentence had — "I will check the sources" cannot be done, because it does not say what would be opened.

Step 4 has a failure mode worth naming: answering it about the system rather than about yourself. "It is usually reliable for this" is not an answer to "can I do this check". The question is about your time, your access and your competence, and it is the only step that can tell you to change the task rather than the tool.

If step 4 came out "no" and you did not go back to step 2, that is the thing to correct. A check you cannot do is not a check.

Part 2 — what to look for in the comparison

The intended discovery is that the grounded answer is not better, it is differently checkable. Three patterns come up:

What you found What it means
The grounded answer names a source but no place in it You have provenance, not a locator. Provenance tells you where the answer came from; a locator is what makes the check cheap. Unit 9 treats this as its own failure mode.
The grounded answer says more than the source does The characteristic failure of the retrieval-backed kind, and the reason attaching a source moves the check rather than removing it.
Both answers agree Worth nothing on its own. The second answer is not independent evidence for the first — they may share the same source of error. Unit 3 has the mechanism.

That last row catches people out, because agreement feels like confirmation. It is the same error as re-asking a question and taking the repeated answer as confirmation.

Part 3 — pricing

"I do not know what it cost" is the honest answer for most consumer tools, where money and energy are invisible at the point of use and time is the only cost you feel. Recording that is a finding, not a gap: a cost you cannot see is a cost you cannot weigh.

The line that matters is the second. The instinct to reach for the most capable available option treats capability as free. It is not free even when it is unpriced, because every extra layer produces another kind of claim, and a claim you do not check is not a benefit.

The checkpoints

cp04-what-changes-the-check matches four observations to the check each makes necessary. The confusion it targets is treating the presence of a source or a computed number as the end of the checking rather than the start of it.

cp04-cheapest-adequate tests the judgement in Part 3. The tempting wrong answer is the one that picks the most capable system because it is available, which is the reasoning this unit is written against.

Your private activity record

Browser storage is not a permanent copy

Progress is kept only in this browser, profile and device. Private browsing, clearing site data, removing the profile, a browser reset, storage eviction or device loss can erase it. Keep important answers and contributions separately.

These notes stay in this browser unless you download a backup or activity log.