跳到主内容
精选75elvis技巧与观点

实测AI员工Viktor:集成层是护城河,人机协作设计

I read agent papers every day.

原文

I read agent papers every day.

我每天阅读智能体论文。

This week I tested an agent that is already working in production, with a real role and a scoreboard.

本周我测试了一个已在生产环境中运行的智能体,它拥有真实角色和记分牌。

I have been paying attention to where agent systems actually break. The breakage shows up in tool access, state, and handoff, the unglamorous layer nobody writes papers about.

我一直在关注智能体系统实际崩溃的地方。问题出现在工具访问、状态和交接这些不引人注目、没人写论文的层面。

So I gave @viktor_com a real job in my own workspace to see how it handles that problem.

所以我在自己的工作区给@viktor_com安排了一份真实工作,看看它如何处理这个问题。

It is an AI employee that lives in Slack. Three things stood out.

它是一个住在Slack里的AI员工。有三点很突出。

The integration layer is the moat. It reaches around 3,000 tools through one connection, scoped OAuth, SOC 2, rather than a graph of brittle per-tool auth.

集成层是护城河。它通过一个连接、范围限定的OAuth、SOC 2,覆盖约3000个工具,而不是脆弱的逐工具认证图。

It carries a standing brief instead of a prompt chain. Role, standards, sign-off rules, persistent across sessions. This is closer to onboarding than to prompting.

它携带一份长期简报,而不是提示链。角色、标准、签署规则,跨会话持久。这更接近入职培训,而不是提示。

Human-in-the-loop is part of the design. Every action is proposed, a person approves, and the trail is there afterwards. That is what makes it trustworthy enough to leave running.

人在回路是设计的一部分。每个动作都被提议,由人批准,之后留有痕迹。这就是让它足够可信、可以持续运行的原因。

One I did not even ask for. It caught a metric being reported on two different windows across two connected sources, and proposed a fix.

还有一个我没想到的。它发现一个指标在两个连接来源的不同窗口中被报告,并提出了修复方案。

Customer side, Antonin Stetina, CEO of KULINA Group: "Mindblowing all-in-one AI which does everything in a single solution." Five campaigns became sixty across 29 markets on $2.5M in spend, no headcount added.

客户方面,KULINA集团CEO Antonin Stetina说:“令人惊叹的一体化AI,在一个解决方案中完成所有事情。”五个活动变成了29个市场的60个,花费250万美元,没有增加人手。

The bigger change is who owns the work. Viktor runs the whole job end-to-end and hands you proposals to approve.

更大的变化是谁拥有工作。Viktor端到端地运行整个工作,并提交提案供你批准。

If you are building agents, look closely at how this handles orchestration and permissions. Happy to go deeper on that if it is useful.

如果你在构建智能体,请仔细看看它如何处理编排和权限。如果有用,我很乐意深入探讨。

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

另一事件,读法相近