<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Interpretability on AI Digest</title><link>https://aidigest.kikihuang.net/tags/interpretability/</link><description>Recent content in Interpretability on AI Digest</description><generator>Hugo</generator><language>zh-cn</language><lastBuildDate>Wed, 08 Jul 2026 00:02:00 +0800</lastBuildDate><atom:link href="https://aidigest.kikihuang.net/tags/interpretability/index.xml" rel="self" type="application/rss+xml"/><item><title>AI Builders 日报 — 2026年7月8日</title><link>https://aidigest.kikihuang.net/posts/2026-07-08-daily/</link><pubDate>Wed, 08 Jul 2026 00:02:00 +0800</pubDate><guid>https://aidigest.kikihuang.net/posts/2026-07-08-daily/</guid><description>follow-builders 的 X / Podcast 源本期正常抓取（16 位 builder、34 条推文、1 集播客、1 篇 blog）。今日主线极其集中——Anthropic 与 @claudeai 官方同步放出《Claude Code 起源史》，Boris Cherny 与 Cat Wu 讲述它如何从『安全研究内部工具』长成产品；同日 Anthropic 发布 J-space 可解释性论文，Swyx 划重点：他们能对模型推理做『脑外科手术』式干预，且模型能『察觉自己被干预了』，逼近 eval awareness。另一条支线是『agent 开始自我进化 / 自我评测』：Replit 宣称已闭环让 agent self-improving，Vercel 的 eve 用 &lt;code&gt;eve eval&lt;/code&gt; 给自己做进化评测。Fable 5 今晚 23:59 PT 下线，Peter Yang 给出最后 5 个值得一试的高价值 prompt。</description></item></channel></rss>