You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Idea + plugin: an enforced pain-point check to stop rabbit-holing
中文版见下方 / Chinese version below.
TL;DR
An agent in "solution state" loses meta-cognition and keeps attacking the same problem — three biases inherited from human problem-solving: confirmation bias (designing experiments that support the current hypothesis), sunk cost (refusing to change direction), and narrative closure (wanting to finish the story). The official repeat-tool-reminder is advisory — it nudges an agent that repeats the exact same call. I built the enforced sibling: after two non-converged experiments on the same problem, a gate injects the three questions, denies non-investigative tool calls until the model answers them in its reply text, and blocks same-direction retries.
Per-agent counters on the current problem: failed (errored) calls and consecutive identical calls.
agent/pre-step
Reset on a real user interjection (new problem); while pending, append the three-question block to the next step.
tools/pre-execute
Deny every non-investigative call while pending (allowlist: read, read_image, glob, grep, web_search, ask_user_question, skill, todo_write).
session/event
Detect the three answers (痛点=… 排除=… 更便宜=…, English markers accepted) and lift the gate.
The three questions: is this pain point still the most painful? What did the last negative result actually rule out? Is there a cheaper path? If the model cannot name what the negative result excluded, it has no falsifiable hypothesis — go write one instead of firing another shot.
Tested end-to-end through a real agent loop with a scripted mock adapter (8 tests: arming, denial, allowlist, lifting, partial answers, resets, fail-loud config).
Questions for the team
Is this direction aligned with where the guard family is going? Would you consider an enforced tier next to repeat-tool-reminder?
The weak point is convergence detection — I count failures and identical repeats as proxies. A goal-aware signal (did the user's stated goal advance?) would be much stronger. Any plans in that space?
Any appetite for a first-class "checkpoint" UI (e.g. blocking with a modal asking the user) rather than text-pattern detection?
reacted with thumbs up emoji reacted with thumbs down emoji reacted with laugh emoji reacted with hooray emoji reacted with confused emoji reacted with heart emoji reacted with rocket emoji reacted with eyes emoji
Uh oh!
There was an error while loading. Please reload this page.
Idea + plugin: an enforced pain-point check to stop rabbit-holing
TL;DR
An agent in "solution state" loses meta-cognition and keeps attacking the same problem — three biases inherited from human problem-solving: confirmation bias (designing experiments that support the current hypothesis), sunk cost (refusing to change direction), and narrative closure (wanting to finish the story). The official
repeat-tool-reminderis advisory — it nudges an agent that repeats the exact same call. I built the enforced sibling: after two non-converged experiments on the same problem, a gate injects the three questions, denies non-investigative tool calls until the model answers them in its reply text, and blocks same-direction retries.Plugin: ICCuse/dsh-pain-point-check (
dsh-plugintopic).The mechanism
tools/resultagent/pre-steptools/pre-executeread,read_image,glob,grep,web_search,ask_user_question,skill,todo_write).session/event痛点=… 排除=… 更便宜=…, English markers accepted) and lift the gate.The three questions: is this pain point still the most painful? What did the last negative result actually rule out? Is there a cheaper path? If the model cannot name what the negative result excluded, it has no falsifiable hypothesis — go write one instead of firing another shot.
Tested end-to-end through a real agent loop with a scripted mock adapter (8 tests: arming, denial, allowlist, lifting, partial answers, resets, fail-loud config).
Questions for the team
repeat-tool-reminder?中文版
一句话
进入"解决态"的 agent 会丢掉 meta-cognition,连续猛攻同一个问题——背后是三个继承自人类解题的偏执:确认偏执(只设计支持当前假设的实验)、沉没成本(已投入就舍不得换方向)、叙事偏执(想把故事收尾)。官方
repeat-tool-reminder是建议性的——只对"逐字重复同一条调用"的 agent 递话。我做的是它的强制版:同一问题连续 2 个实验未收敛后,门禁注入三问、否决非调查类工具调用直到模型在回复正文答出三问,并阻止同方向重试。插件:ICCuse/dsh-pain-point-check(
dsh-plugin话题)。机制
tools/resultagent/pre-steptools/pre-executeread、read_image、glob、grep、web_search、ask_user_question、skill、todo_write)。session/event痛点=… 排除=… 更便宜=…,也接受英文标记)并解禁。三问:这个痛点还是最痛的吗?上一轮负结果到底排除了什么?有没有更便宜的路径?若说不出负结果排除了什么 = 没有可证伪假设,先回去写假设,别再开一枪。
已用脚本化 mock 适配器驱动真实 agent loop 端到端测试(8 个用例:触发、拦截、放行、解禁、部分作答、重置、配置校验失败)。
想请教团队
repeat-tool-reminder旁加一个强制层?All reactions