<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Tool Use on AI Digest</title><link>https://aidigest.kikihuang.net/tags/tool-use/</link><description>Recent content in Tool Use on AI Digest</description><generator>Hugo</generator><language>zh-cn</language><lastBuildDate>Sat, 27 Jun 2026 00:00:00 +0800</lastBuildDate><atom:link href="https://aidigest.kikihuang.net/tags/tool-use/index.xml" rel="self" type="application/rss+xml"/><item><title>AI Builders 日报 — 2026年6月27日</title><link>https://aidigest.kikihuang.net/posts/2026-06-27-daily/</link><pubDate>Sat, 27 Jun 2026 00:00:00 +0800</pubDate><guid>https://aidigest.kikihuang.net/posts/2026-06-27-daily/</guid><description>follow-builders 的 X（推文）/ Podcast 抓取源本期仍不可用，日报继续以「GitHub Trending daily + HuggingFace Daily Papers 2026-06-26」双公开源 remix，featured 链接已逐条抽查 HTTP 200。今天的主线一句话：agent 的竞争重心，正在从『能不能动手』迁移到『怎么把它训得稳、给得对、验得准』。学术侧今天清一色在啃 agent 训练与评测的硬骨头——OPID 直接从 on-policy 轨迹里蒸出『技能监督』、《The Verification Horizon》宣告 coding agent 不存在万能 reward、另一篇拆解了多步工具调用 RL 为何会『灾难性崩溃』、GauntletBench 则把 agent 拖出熟悉环境做压力测试。工程侧呼应得同样直接：gstack 冲到 11.6 万⭐、Google 的 design.md 单日 +2,319⭐、AWS 官方工具箱、OpenMontage 视频 agent——『agent = 团队 × 技能库 × 工具链』范式继续固化；中文圈今天也罕见地集体上榜：ai-berkshire、Agent-Reach、MediaCrawler、MinerU、张雪峰.skill。</description></item></channel></rss>