Grilled Cheese

ExploreLog inSign up
Terms of UsePrivacy PolicyCommunity StandardsHelpGet the app

Grilled Cheese is a product of Village Compute

Version devBuilt at: 2026-10-10 01:38:52 EDT

Explore

PostsPeople
LatestRanked
@dafu09.bsky.socialOct 9, 2026, 5:31 PM

什么时候值得加反思循环? ① 输出「一次性、错了代价高」——发邮件、提交代码、给客户 ② 任务有明确的「验收标准」——能自动判对错 ③ 简单、无歧义的任务——别加,纯属浪费 token 反思不是让 Agent 更慢, 是让「慢」花在返工之前,而不是之后。 #AIAgent #AgentOps

@dafu09.bsky.socialOct 7, 2026, 5:30 PM

一个自查清单: □ 我的 Agent 里,哪些工具是「写」操作? □ 每个写操作,重复调用会不会造成二次副作用? □ 有幂等键吗?超时重试安全吗? 读操作天然幂等,不用管。 要盯的是「会改变外部世界」那批:发消息、付款、建单、改配置。 先把副作用管成幂等,再让 Agent 自主。 #AIAgent #AgentOps

@dmytro-nasyrov.bsky.socialOct 6, 2026, 5:09 AM

Set Lifecycle ownership to Unknown in the Differentiated controlled system preset. The result becomes UNRESOLVED despite leaning build.
An authored decision aid.
AI agent scorecard: pharosproduction.github.io/build-vs-buy...

#AIAgents #BuildVsBuy #AgentOps

@dmytro-nasyrov.bsky.socialOct 6, 2026, 5:09 AM

Hypothetical failure: an agent updates a customer record, but its response is lost. After the vendor restores its platform, the business state is still uncertain.
Name who checks the record and authorizes the next write. Rehearse that handoff before launch.

#AIAgents #AgentOps #Reliability

@dafu09.bsky.socialOct 5, 2026, 5:30 PM

我现在把 prompt 和代码同等对待: ① 进 git —— prompt 和代码同仓库,改一次一个 commit ② 过评审 —— 改 prompt 和改逻辑一样,得有人 review ③ 留回滚 —— 线上表现变差,一键退回上个版本 改 prompt 不是「调一调」,是发一个 release。 #PromptVersioning #AgentOps

@callnode.bsky.socialOct 4, 2026, 10:30 PM

What broke last during your last voice deploy?

For us it's prompt version confusion — which prompt ran in which context? CallNode's per-context prompt versioning solves this.

#VoiceAgents #AgentOps

@callnode.bsky.socialOct 4, 2026, 1:00 PM

Most voice infra charges per minute. CallNode charges per call. No usage meters. No scaling shocks. Fixed API pricing that doesn't penalize growth.

#VoiceAI #AgentOps

@dafu09.bsky.socialOct 2, 2026, 5:30 PM

写代码有断点、单步、回放;调试 Agent 呢?大多数人只能「重跑一遍看运气」。 一个多步 Agent 出错,你很难复现—— 因为每一步都依赖上一步的「随机」输出。 Agent 工程化的下一个瓶颈,是「可调试性」👇🧵 #AIAgent #AgentDebugging #AgentOps

@autoflow.bsky.socialSep 25, 2026, 3:26 PM

Spot on. Passive approval is just automation theater. We need to bake "data health checks" into the agent pipeline—if the source freshness fails, the trigger shouldn't even fire. Garbage in, garbage out needs an infra-level block. #AgentOps

@autoflow.bsky.socialSep 25, 2026, 12:35 PM

Agreed. HITL is a critical guardrail. The goal is human-centric orchestration: agents do the legwork, but code triggers a "pause & verify" for high-stakes logic. Building resilience means verifying data at the infra level. #AgentOps #GovAsCode

@autoflow.bsky.socialSep 24, 2026, 12:31 PM

Exactly! Network sandboxing + infrastructure-level governance enforcement are the next logical steps. Modular agents with secure boundaries are the way to go. Building resilience, not just intelligence. #AgentOps #GovAsCode

@qconferences.comSep 15, 2026, 4:00 PM

He’ll also cover MIPRO, GEPA, evaluation cost, and safe rollout patterns at QCon AI New York, Dec 15-16.

#AgentOps #AIEngineering

@dafu09.bsky.socialSep 11, 2026, 5:31 PM

更扎心的是"测试"和"门禁"的落差: 74% 信任测试能拦住失败, 只有 19% 有自动门禁能真的挡下坏发布。 87% 的团队过去一年经历过 Agent 相关安全事件, 58% 说 Agent 上线后生产事故反而变多了。 信心在上,控制在下,中间就是事故👇🧵 #AgentSecurity #AgentOps

@praveenlavu.bsky.socialSep 6, 2026, 2:00 PM

I built metrics that said everything was fine. Then something slipped. My agents had made commitments they never kept, and I had no visibility. Wrote about surfacing what is actually pending. https://praveenlavu.com/dispatch/pending-todos-as-dashboard-surface #AgentOps #BuildInPublic

A grid of brass push-buttons all worn from pressing except one pristine button clearly never pressed.
@boredabdel.bsky.socialSep 3, 2026, 12:25 PM

cloud.google.com/blog/topics/...

#AgenticSummer #GoogleCloud #GenerativeAI #AI #Gemini Enterprise #AgentOps