Files
FabledScribe/plugin/hooks/scribe_turn.awk
T
bvandeusenandClaude Opus 5.5 6071007fe2
CI & Build / Python lint (push) Successful in 3s
CI & Build / Plugin hooks (push) Successful in 13s
CI & Build / TypeScript typecheck (push) Successful in 53s
CI & Build / integration (push) Successful in 1m7s
CI & Build / Python tests (push) Failing after 1m25s
CI & Build / Build & push image (push) Skipped
feat(500): one end-of-turn request - the completion-section check folds into the reply check
The Stop hook sent the finished reply twice: scribe_report_check.sh checked a
task-closing reply for the completion sections in shell and reported to
/report-check, and scribe_reply_check.sh sent the same reply to /reply-rules
for the rule hold. Now the reply goes once. When the turn closed a task the
hook adds the close count and ids, and the server runs the section check
(services/report_check, the same three patterns) beside the reply hold,
folding both into one reason.

The report_check adherence log is still written for every checked reply
(milestone 409's number). A section hold marks the session, so its rewrite is
sent back once with rewrite:true to record how it came out, and is never held.
The block reason carries the completion shape's one line and points at
list_reply_shapes rather than at the skill. #5496 (step 4 of milestone 500).

Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
2026-10-09 15:38:17 -04:00

109 lines
4.5 KiB
Awk

# Scribe plugin — what happened in the LAST TURN of a transcript (#4107).
#
# Reads the flat `IDX<TAB>PATH<TAB>VALUE` stream that scribe_json.awk produces
# in `mode=lines` from a Claude Code transcript, and answers the four questions
# the Stop hooks ask (scribe_reply_check.sh). It replaces a thirty-line jq program; the shape
# of the answer is unchanged, so the hook around it reads the same.
#
# bounded 1 if a user prompt was found in the window, else empty. A window
# with no prompt cannot be cut into a turn, and the hook then says
# nothing rather than guessing — `tail -n 3000` cuts wherever it
# cuts, and a turn that began before the cut is not this hook's.
# closed how many task-closing tool calls SUCCEEDED in that turn.
# task_ids their task ids, comma-joined, for the server to record.
# reply the assistant text after the last action, STILL JSON-ESCAPED and
# on one line. The caller decodes it with scribe_json_unescape —
# decoding here would put newlines into a line-oriented format.
#
# A SIDECHAIN IS NOT THIS SESSION. Subagent records interleave into the same
# file, and a subagent closing a task is not the operator's session closing
# one — counting those made the check fire on turns that closed nothing.
#
# A CLOSE THAT ERRORED IS NOT A CLOSE, which is why the error pass runs first:
# the tool_result carrying `is_error` arrives in a LATER record than the
# tool_use it refutes, so a single forward pass would have already counted it.
# TAB-SEPARATED, and stated rather than assumed. Under awk's default splitting
# a value is cut at its first SPACE, so `$3` of a reply line was the reply's
# first word — the kind of defect that hides completely behind a test whose
# replies are all empty. Found by differential-testing this against the jq
# program it replaces, over real transcript windows (#4107).
BEGIN { FS = "\t" }
{
i = $1 + 0
if (i > maxidx) maxidx = i
p = $2
if (p == ".type") { f[i, "type"] = $3; next }
if (p == ".isSidechain") { f[i, "side"] = $3; next }
if (p == ".isMeta") { f[i, "meta"] = $3; next }
# `.message.content` as a SCALAR is what marks a real user prompt; a tool
# result carries an array at the same path, and counting one as a prompt
# would cut the turn at the wrong place.
if (p == ".message.content") { if ($3 != "null") f[i, "str"] = 1; next }
if (substr(p, 1, 17) != ".message.content[") next
rest = substr(p, 18)
if (rest == "#]") { nb[i] = $3 + 0; next }
c = index(rest, "]")
if (c < 2) next
b[i, substr(rest, 1, c - 1) + 0, substr(rest, c + 1)] = $3
}
function is_mine(i) {
return (f[i, "side"] != "true")
}
function has_tool_use(i, j) {
for (j = 0; j < nb[i]; j++) if (b[i, j, ".type"] == "tool_use") return 1
return 0
}
END {
for (i = 1; i <= maxidx; i++)
if (is_mine(i) && f[i, "type"] == "user" && f[i, "meta"] != "true" && f[i, "str"] == 1)
prompt = i
if (!prompt) { print "bounded\t"; exit 0 }
for (i = prompt + 1; i <= maxidx; i++) {
if (!is_mine(i) || f[i, "type"] != "user") continue
for (j = 0; j < nb[i]; j++)
if (b[i, j, ".type"] == "tool_result" && b[i, j, ".is_error"] == "true")
errored[b[i, j, ".tool_use_id"]] = 1
}
for (i = prompt + 1; i <= maxidx; i++) {
if (!is_mine(i) || f[i, "type"] != "assistant") continue
for (j = 0; j < nb[i]; j++) {
if (b[i, j, ".type"] != "tool_use") continue
if (b[i, j, ".name"] !~ /(^|__)(update|create)_task$/) continue
if (b[i, j, ".input.status"] != "done") continue
if (b[i, j, ".id"] in errored) continue
closed++
t = b[i, j, ".input.task_id"]
if (t != "") ids = ids (ids == "" ? "" : ",") t
}
}
# Where the WORK stopped and the report began. Everything after the last
# action is the reply being checked; text emitted between two tool calls is
# narration mid-work, not a report, and holding it to the report shape would
# block turns that did report properly at the end.
act = prompt
for (i = prompt + 1; i <= maxidx; i++) {
if (!is_mine(i)) continue
if (f[i, "type"] == "user" || (f[i, "type"] == "assistant" && has_tool_use(i))) act = i
}
for (i = act + 1; i <= maxidx; i++) {
if (!is_mine(i) || f[i, "type"] != "assistant") continue
for (j = 0; j < nb[i]; j++)
if (b[i, j, ".type"] == "text")
reply = reply (reply == "" ? "" : "\\n") b[i, j, ".text"]
}
printf "bounded\t1\n"
printf "closed\t%d\n", closed + 0
printf "task_ids\t%s\n", ids
printf "reply\t%s\n", reply
}