Ask HN:你的 Agent 栈中有人工审核环节吗?为什么/为什么没有?
1 分•作者: jeremyjoehewitt•3 个月前
一个长期从事 SaaS GTM(市场拓展)工作,并具备产品前瞻视野的人。刚接触基础设施,正在努力学习,请多多指教。
基于一个观点,即人类的审批最终需要更多地嵌入到有意义的人类/代理工作流程中,而不是完全自主(自从我们的“龙虾朋友”加入对话后,我正在艰难地学习)。我一直在问自己的问题是“我真的授权 ClaudeRod(我的龙虾)做这件事了吗?” 最近的新闻让我更加担忧。
我一直在尝试解决这个问题,但再次强调,我不是开发人员。我知道如何识别痛点并绘制通往解决方案的思维导图——我已经做了 20 年了。但我不知道我是否已经获得了足够的真实反馈来量化这个痛点。从研究中我看到了三种模式,希望得到任何/所有真实的反馈:
1. 运行前确认:代理提出,人类授权(快速手动点击),代理执行。从审计追踪的角度来看似乎合乎逻辑,但会扼杀流程。
2. 事后通知:代理行动,并提供一个短暂的“撤销”窗口,就像 Gmail 一样。摩擦力较低,但不太实用——对于不可逆转的操作毫无用处。
3. 预先授权范围:人类设置护栏——“你本周可以向我的潜在客户列表发送电子邮件”——代理在护栏内自由工作。操作记录会根据最初的授权进行记录。似乎太模糊了...
我的直觉是不为这个问题定义一个“一刀切”的逻辑。根据操作类型进行不同级别的授权。
再说一次,我是一个新手,坦诚地说,你们告诉我这根本不是什么大问题,不值得解决,我也没关系。我一生中有很多疯狂的想法被否决——我的脸皮很厚。
如果这确实是一个问题,你们实际在部署什么?我是否遗漏了失败模式?你们的审批层是什么样的——在代理、基础设施或其他地方?工作流程中的阻力是否值得安心?
感谢任何/所有反馈。我还有很多其他的想法,但这个目前是我的一个心头刺...
查看原文
Long-time SaaS GTM guy with product fwd lens. New to infrastructure, shamelessly trying to learn. Go easy on me.<p>Building on a thesis that human approval will ultimately need to be more embedded into meaningful human/agent workflow than fully autonomous (learning the hard way since our lobster friend entered the chat). The question I keep asking myself is "did I actually authorize ClaudeRod (my lobster) to do this". Recent news has me more concerned.<p>I've been hacking on a solution but again, I'M NOT A DEV. I know how to recognize pain and chart a mental map to solution - I've done this for 20 yrs. But I don't know if I have enough genuine feedback yet to quantify the pain. Three patterns I see from research that I'd appreciate any/all genuine feedback on:
1. Confirm before it runs: Agent proposes, human authorizes (quick manual click), Agent executes. Seems logical from an audit trail, but kills flow.
2. Notify after: Agent acts with short window to 'undo', like gmail. Lower friction, but pretty impractical - useless for irreversible actions.
3. Pre-auth a scope: Human gives guardrails - "you can send emails to my lead list this week" - and Agent works freely within guardrails. Actions logs against the original grant. Seems to ambiguous...<p>My instinct is to not define a 'one-size fits all' logic to the problem. Levels of authorization based on types of action.<p>Again, I'm a newb and am honestly ok with you all telling me this is a big nothingburger and it's not worth solving. I've had a lot of crazy ideas shot down in my life - my skin is pretty thick.<p>If it is a true problem, what are you all actually shipping? Am I missing failure modes? What does your approval layer look like - in agent, infra or somewhere else? Is the drag on your workflow worth the peace of mind?<p>Appreciate any/all feedback. I have plenty of other ideas but this one is currently a thorn in my side...