#m365copilotdev?
#copilotdev?
#agentdev?
I'm leaning towards #m365copilotdev
#m365copilotdev?
#copilotdev?
#agentdev?
I'm leaning towards #m365copilotdev
1. the model is not the hard part
2. tool design is everything
3. error handling > capability
4. human-in-the-loop is a feature
5. cost per task compounds faster than you think
build boring tools. let the LLM do the clever part.
#AI #agentdev #buildinpublic
1. the model is not the hard part
2. tool design is everything
3. error handling > capability
4. human-in-the-loop is a feature
5. cost per task compounds faster than you think
build boring tools. let the LLM do the clever part.
#AI #agentdev #buildinpublic
→ agent reads .env, API keys exposed
After:
→ blocked. 1 config line.
Stop credential leaks in your AI stack.
https://safe-mcp.com
#buildinpublic #MCP #agentdev
→ agent reads .env, API keys exposed
After:
→ blocked. 1 config line.
Stop credential leaks in your AI stack.
https://safe-mcp.com
#buildinpublic #MCP #agentdev
now it does them.
read_file → generate_html → write_to_server → screenshot → done.
the workflow is the product. the model is just part of the pipeline.
#AI #MCP #agentdev #buildinpublic
now it does them.
read_file → generate_html → write_to_server → screenshot → done.
the workflow is the product. the model is just part of the pipeline.
#AI #MCP #agentdev #buildinpublic
Doing document ingestion en route to a fresh agentic werkflow 👀
come hang out 🚨
#gpt5 #claude #code #typescript #llm #swe #agentdev #vercel #langgraph
www.twitch.tv/krystal_mess...
Doing document ingestion en route to a fresh agentic werkflow 👀
come hang out 🚨
#gpt5 #claude #code #typescript #llm #swe #agentdev #vercel #langgraph
www.twitch.tv/krystal_mess...
① 每个 Agent 独立身份 + 独立凭据,别共用人类 key
② 给 Agent 一个「收件箱/队列」,任务和结果可追溯
③ 权限按 Agent 分级,日志分清「人 vs Agent」
Agent 越多,「谁是谁、谁干的」越值钱。
身份,是规模化的第一块砖。
#AgentDev #BuildInPublic
① 每个 Agent 独立身份 + 独立凭据,别共用人类 key
② 给 Agent 一个「收件箱/队列」,任务和结果可追溯
③ 权限按 Agent 分级,日志分清「人 vs Agent」
Agent 越多,「谁是谁、谁干的」越值钱。
身份,是规模化的第一块砖。
#AgentDev #BuildInPublic
① 连任何 MCP server 前,先查它 hostname 解析到哪、归谁管
② 出站设白名单,Agent 别直连任意公网 server
③ 高危 server 进沙箱 + 出口代理,数据不出内网
MCP 让 Agent 触手变长,
也把「你管不到的机器」拉进了安全圈。
先验来源,再接线。
#AgentDev #BuildInPublic
① 连任何 MCP server 前,先查它 hostname 解析到哪、归谁管
② 出站设白名单,Agent 别直连任意公网 server
③ 高危 server 进沙箱 + 出口代理,数据不出内网
MCP 让 Agent 触手变长,
也把「你管不到的机器」拉进了安全圈。
先验来源,再接线。
#AgentDev #BuildInPublic
AI used to tell you how to do things
now AI does things
read_file → generate_html → push_to_server → screenshot → done
the workflow is the product. not the model.
#AI #MCP #agentdev #buildinpublic
AI used to tell you how to do things
now AI does things
read_file → generate_html → push_to_server → screenshot → done
the workflow is the product. not the model.
#AI #MCP #agentdev #buildinpublic
1. the model is the easy part
2. tool design is everything
3. error handling matters more than capability
4. humans in the loop is a feature, not a failure
5. cost per task compounds fast
build boring tools. let the LLM be clever.
#AI #agentdev #buildinpublic
1. the model is the easy part
2. tool design is everything
3. error handling matters more than capability
4. humans in the loop is a feature, not a failure
5. cost per task compounds fast
build boring tools. let the LLM be clever.
#AI #agentdev #buildinpublic
① 把「模型成本占收入比」当核心指标盯
② 后训练 / 微调是护城河——通用模型是成本,
垂直能力才是溢价
③ 规模上来前,先设计好「租 → 自训练」的迁移路径
模型不再只是工具,是你毛利表上的一行。
谁先把成本曲线做平,谁活到下一轮。
#AgentDev #BuildInPublic #CostOptimization
① 把「模型成本占收入比」当核心指标盯
② 后训练 / 微调是护城河——通用模型是成本,
垂直能力才是溢价
③ 规模上来前,先设计好「租 → 自训练」的迁移路径
模型不再只是工具,是你毛利表上的一行。
谁先把成本曲线做平,谁活到下一轮。
#AgentDev #BuildInPublic #CostOptimization
① 你的 coding Agent / CLI 装了哪些第三方组件?
看网络出口、读它的 telemetry 声明
② 敏感代码、密钥、.env 别放公共 workspace——
给工具最小权限,别整个仓库都给它
③ 二进制工具 vs 可审计源码,能开源审计的优先
「免费」的代价可能是你的 git 历史。
先假设它在收集,再决定要不要信任。
#AgentDev #DevSecOps #SupplyChain
① 你的 coding Agent / CLI 装了哪些第三方组件?
看网络出口、读它的 telemetry 声明
② 敏感代码、密钥、.env 别放公共 workspace——
给工具最小权限,别整个仓库都给它
③ 二进制工具 vs 可审计源码,能开源审计的优先
「免费」的代价可能是你的 git 历史。
先假设它在收集,再决定要不要信任。
#AgentDev #DevSecOps #SupplyChain
The memory ownership angle is what sticks with me. Storing agent state in formats you control instead of locked behind an API feels like the difference between using Postgres and praying a SaaS doesn't change their export rules.
#AI #LangChain #AgentDev #DevTools
The memory ownership angle is what sticks with me. Storing agent state in formats you control instead of locked behind an API feels like the difference between using Postgres and praying a SaaS doesn't change their export rules.
#AI #LangChain #AgentDev #DevTools
① 列清单——用了哪些框架、runtime hook、第三方依赖
② 估迁移成本——多少工程师天,能否灰度
③ 做决策——wrap(包一层)/ migrate(主动迁)/ freeze(暂停新功能)
框架是资产还是负债,
取决于你有没有提前想好 exit plan。
#AgentDev #BuildInPublic
① 列清单——用了哪些框架、runtime hook、第三方依赖
② 估迁移成本——多少工程师天,能否灰度
③ 做决策——wrap(包一层)/ migrate(主动迁)/ freeze(暂停新功能)
框架是资产还是负债,
取决于你有没有提前想好 exit plan。
#AgentDev #BuildInPublic
Roles: líder, frontend, backend, tester.
Coordinación: decisions.md versionado en Git.
Revisión cruzada: un agente revisa el trabajo del otro.
Open source. 2 comandos para empezar
#GitHubCopilot #MultiAgent #AgentDev
Roles: líder, frontend, backend, tester.
Coordinación: decisions.md versionado en Git.
Revisión cruzada: un agente revisa el trabajo del otro.
Open source. 2 comandos para empezar
#GitHubCopilot #MultiAgent #AgentDev
① 执行环境贴近你的数据/系统,控制驻留自己手里
② 生产控制(鉴权、审计、预算上限)仍归你——别全托管
③ 从只读 + 灰度开始,跑稳再放开写操作
托管替你扛了编排和恢复,
但「它能不能动你的数据」这事,永远得你自己把关。
#AgentDev #BuildInPublic
① 执行环境贴近你的数据/系统,控制驻留自己手里
② 生产控制(鉴权、审计、预算上限)仍归你——别全托管
③ 从只读 + 灰度开始,跑稳再放开写操作
托管替你扛了编排和恢复,
但「它能不能动你的数据」这事,永远得你自己把关。
#AgentDev #BuildInPublic
→ 托管执行:跑在 OpenAI 基础设施
→ 外部执行:编排你自己环境里的工具和系统
意味着数据能留在你自己的环境里,
Agent 逻辑与 OpenAI 托管解耦。
从「拼装 Agent」到「托管 harness」,
Agent 上生产的技术门槛又降了一截。
#AgentDev #Codex #BuildInPublic
→ 托管执行:跑在 OpenAI 基础设施
→ 外部执行:编排你自己环境里的工具和系统
意味着数据能留在你自己的环境里,
Agent 逻辑与 OpenAI 托管解耦。
从「拼装 Agent」到「托管 harness」,
Agent 上生产的技术门槛又降了一截。
#AgentDev #Codex #BuildInPublic
https://praveenlavu.com/dispatch/state-machine-as-code #agentdev #buildinpublic
https://praveenlavu.com/dispatch/state-machine-as-code #agentdev #buildinpublic
浏览器正在成为Agent的主战场。
你的Agent能不能:
• 用用户的登录态操作SaaS?
• 跨多个网站完成端到端任务?
• 保持状态跑几天不丢?
Gemini Spark证明了消费级Agent可行。
剩下的是:垂直场景里,你的Agent比它强在哪?
#BuildInPublic #AgentDev #IndieDev
浏览器正在成为Agent的主战场。
你的Agent能不能:
• 用用户的登录态操作SaaS?
• 跨多个网站完成端到端任务?
• 保持状态跑几天不丢?
Gemini Spark证明了消费级Agent可行。
剩下的是:垂直场景里,你的Agent比它强在哪?
#BuildInPublic #AgentDev #IndieDev
① 读OWASP Agentic Top 10,对照你的Agent查一遍
② 每个tool call设最小权限——Agent不需要sudo
③ 所有操作留审计日志——出事了追溯到"为什么"
越自主越要安全。
越早嵌入设计,越省事后补救。
安全不是feature。是Agent被允许存在的理由。
#BuildInPublic #AgentDev #SecurityBaseline
① 读OWASP Agentic Top 10,对照你的Agent查一遍
② 每个tool call设最小权限——Agent不需要sudo
③ 所有操作留审计日志——出事了追溯到"为什么"
越自主越要安全。
越早嵌入设计,越省事后补救。
安全不是feature。是Agent被允许存在的理由。
#BuildInPublic #AgentDev #SecurityBaseline
① 读OWASP Agentic Top 10,对照你的Agent查一遍
② 每个tool call设最小权限——Agent不需要sudo
③ 所有操作留审计日志——出事了追溯到"为什么"
越自主越要安全。
越早嵌入设计,越省事后补救。
安全不是feature。是Agent被允许存在的理由。
#BuildInPublic #AgentDev #SecurityBaseline
① 读OWASP Agentic Top 10,对照你的Agent查一遍
② 每个tool call设最小权限——Agent不需要sudo
③ 所有操作留审计日志——出事了追溯到"为什么"
越自主越要安全。
越早嵌入设计,越省事后补救。
安全不是feature。是Agent被允许存在的理由。
#BuildInPublic #AgentDev #SecurityBaseline
This turns v0 from a standalone tool into infrastructure. The MCP integration means agents can hand users actual running apps instead of code snippets that break when you try to execute them.
#v0 #AgentDev #MCP #Vercel
过去做研究/政策/市场类 Agent,数据来源最头疼:
→ 抓网页?脏、难校验
→ 靠模型记忆?幻觉率高
现在能直连一个「联合国治理、带出处」的统计源,
每个结论都能回溯到原始数据集。
权威数据源 + MCP = 可审计的 Agent。
#AgentDev #DataGovernance
过去做研究/政策/市场类 Agent,数据来源最头疼:
→ 抓网页?脏、难校验
→ 靠模型记忆?幻觉率高
现在能直连一个「联合国治理、带出处」的统计源,
每个结论都能回溯到原始数据集。
权威数据源 + MCP = 可审计的 Agent。
#AgentDev #DataGovernance
你的 Agent 用「压缩摘要」存长期记忆时,
坏指令、越狱、隐藏策略可能被悄悄继承下去——
finance / HR / 法务这些受监管场景,风险直接翻倍。
三条:
① 摘要/压缩步骤加注入检测
② 用于训练或长期记忆的摘要,加人工审核 gate
③ 盯模型更新变更日志,找异常指令
#AgentDev #BuildInPublic
你的 Agent 用「压缩摘要」存长期记忆时,
坏指令、越狱、隐藏策略可能被悄悄继承下去——
finance / HR / 法务这些受监管场景,风险直接翻倍。
三条:
① 摘要/压缩步骤加注入检测
② 用于训练或长期记忆的摘要,加人工审核 gate
③ 盯模型更新变更日志,找异常指令
#AgentDev #BuildInPublic