Mission control sends a short message: "How much should Sky Sorter do on its own?"
Rocket wants it to file every card instantly. "Speed is the whole point," he argues.
Raven shakes her head. "If it files a comet as a star, nobody will notice until much later."
Nova draws a dial on the whiteboard with three settings: Suggest, Decide and Check, and Decide Alone.
"Instead of arguing," she says, "let's test all three settings with real cards and count what happens."
Rocket grabs a stack of index cards. Raven grabs a timer. The lab is open.
You will play Sky Sorter three times, once at each autonomy setting.
A partner plays the person who works with the sorter. You will count mistakes and checks at each setting.
Then you will compare the trade-off between speed and catching mistakes.
| Setting | Who makes the final call? | Time (seconds) | Mistakes left in bins |
|---|---|---|---|
| 1. Suggest | The person | ||
| 2. Decide and Check | The sorter, with some checks | ||
| 3. Decide Alone | The sorter |
Copy this table onto paper and fill in your own times and mistake counts.
Here are the crew's own numbers from the same lab. Use them to answer the questions below.
| Setting | Time (seconds) | Mistakes left in bins |
|---|---|---|
| 1. Suggest | 96 | 0 |
| 2. Decide and Check | 58 | 1 |
| 3. Decide Alone | 31 | 3 |
| What the crew's results show | True or false? |
|---|---|
| More autonomy made the sorting faster. | ? |
| More autonomy let more mistakes slip through. | ? |
| The Suggest setting had no trade-offs at all. | ? |
| A person in the loop can catch mistakes before they matter. | ? |
Great lab work. Tomorrow you will tell AI systems apart from fixed programs.