搜索:Goodfire

共命中 4 条(服务端检索)
Goodfire推出模型内部监测器,低成本捕捉异常AI智能体
Goodfire发布基于可解释性的监测器,直接观察AI模型内部状态而非让第二个AI复读输出,以更低成本监控智能体行为,已面向Baseten客户开放。
大模型 TechCrunch · 昨天 阅读 5·访客 5
机械可解释性 2026:打开大模型黑箱,这条路走到了哪一步
从叠加假说、稀疏自编码器到归因图,机械可解释性在 2026 年第一次跑进生产线:Anthropic 能画出 Claude 的思维回路,OpenAI 从 GPT-4 抽出 1600 万特征,Goodfire 本周把内部探针做成低价监测产品。本文梳理核心进展、落地场景与「SAE 是否找到真实特征」的根本性质疑。
研究前沿 精选 · 原创 · 今天 阅读 0·访客 0
Base Labs launches an open-weight AI safety partnership with Hugging Face and Goodfire
Base Labs, the research group Baseten spun up earlier this year, will develop and publish methods for training and monit…
行业动态 TechCrunch · 9-18 阅读 18·访客 18
The fix for rogue AI agents could be more AI
As companies hand off longer and more complex tasks to AI agents, they are running into an oversight problem: Agents can…
智能体 TechCrunch · 9-18 阅读 26·访客 25