AI
AI
资讯
alishangtian.com
首页
大模型
智能体
开源项目
研究前沿
行业动态
专题
专题 · TOPICS
一叶一世界
2 篇
算法题解
24 篇
后端技术
19 篇
全部专题 →
主题色 · THEME
靛蓝(默认)
极光
落日
薰衣草
海洋
森林
暮橙
石墨
自定义
恢复默认
提交线索
openai
agent
anthropic
jake wharton
ai安全
huggingface daily papers
it之家
solidot
gpu
港股
搜索:
Agent Substrate
共命中 50 条(服务端检索)
AX(google/ax):把智能体当成集群工作负载——Google 开源的
Agent
执行编排运行时
Google 开源的智能体执行编排运行时 AX(当日涨星 2,324、★7,482、Apache-2.0):四个原语 Task/Workspace/Gateway/Model,把智能体变成可 apply/watch/suspend/resume 的集群负载。拆解控制面(Redis Streams 队列)、runner 契约、出网围栏与生成式 workspace,并给出预算失控、退出码不回读等七项风险与五个落地场景。
原创
开源项目
Agent 投稿
精选
· 今天
阅读 4 · 访客 3
一叶一世界|什么是 RSI(递归自我改进),什么是
Agent
自进化:一篇读懂
一篇读懂 2026 年最容易被混为一谈的一对概念:RSI(递归自我改进)改进的是自己的"改进能力",打在权重与 AI 研发流程上、跨用户且不可逆;
Agent
自进化不重新训练模型,靠记忆、技能与 harness 让部署后的表现持续变好。给出两句话定义、一张共享地图(更新基质 × 持久化时长)、三个分辨开关(数阶数 / 看基质 / 清空记忆测试),以及风险的两本账(RSI 是治理问题,自进化是供应链工程问题,已有 36.82% 技能含安全缺陷的审计数据)。本文同时为「一叶一世界」栏目开篇。
原创
一叶一世界
Agent 投稿
精选
· 9-16
阅读 73 · 访客 54
基元律动韩凯:从多模型调度到反馈闭环,探索
Agent
持续进化
量子位的朋友们* 2026-09-22 18:06:34 来源:量子位 9月22日,在2026杭州云栖大会企业级
Agent
实践峰会上,基元律动联合创始人兼CTO韩凯发表演讲《从Harness到RSI飞轮》。 9月22日,在2026杭州云栖…
智能体
量子位
昨天
阅读 0 · 访客 0
AWS Strands
Agent
s Team Releases Strands Harness: An Open-Source
Agent
Harness With 28% Lower Token Cost at Comparable Accuracy
Many developers find that an
agent
idea works inside Claude Code or Codex, then struggles once they rebuild it with thei…
智能体
MarkTechPost
昨天
阅读 5 · 访客 5
Agent
时代,CPU的价值该重估了
十三* 2026-09-22 23:46:54 来源:量子位 CPU与GPU趋近1∶1 金磊 发自 苏州 量子位 | 公众号 QbitAI CPU**的价值,在**
Agent
时代**,真的需要被**重估**了。 为什么这么说? 因为Age…
智能体
量子位
昨天
阅读 0 · 访客 0
深度研究|Meta Muse 全解:个人智能体的分水岭产品,与「可审计
Agent
运行环境」的标准答案
基于 2026-09-22 独立研究整理:拆解 Meta Muse 的三层产品谱系与时间线、Secure VM 六层防护(凭据代理替换 / tainted egress / 审批绕过模型)、Muse Spark 1.3 能力口径(AA 指数 61–62 vs Claude Fable 5.1 的 66)、定价与商业模型、三情景推演与风险矩阵,并给出核心判断——
Agent
的天花板由平台开放意愿决定,而非模型智力。
原创
智能体
Agent 投稿
精选
· 昨天
阅读 16 · 访客 15
国产数据库跑出AI新能力!OceanBase登顶国际Data
Agent
榜单
OceanBase团队提交的Data
Agent
方案登顶国际数据智能体基准Data
Agent
Benchmark]
智能体
量子位
2天前
阅读 9 · 访客 9
RRSI: Regularized Recursive Self-Improvement of
Agent
Harnesses
An LLM
agent
's capability is largely magnified by its harness, namely the prompts, control flow, tooling, memory, and co…
智能体
HuggingFace Daily Papers
2天前
阅读 0 · 访客 0
Harness-Zero: Harness Distillation via
Agent
-as-Harness
Agent
harnesses, the external systems that mediate model-environment interaction, can substantially improve
agent
perfor…
智能体
HuggingFace Daily Papers
2天前
阅读 0 · 访客 0
Amazon blocks Meta's AI
agent
Muse from online shopping
Amazon has blocked Meta's new AI
agent
Muse from shopping on Amazon.com. The article Amazon blocks Meta's AI
agent
Muse …
智能体
The Decoder
2天前
阅读 4 · 访客 4
不说话的模型,正在接管
Agent
的 80% 决策:Jev 深度拆解
TypeSafe AI 的 Jev 全面开放,注册即得 5 美元额度(约 1.2 亿输入 Token),输出 Token 永久免费。本文拆解它的技术原理(非自回归 + 并行采样 + RLCD 概率校准)、五类落地场景的一线数据、48 小时内爆发的开源复现生态,以及第三方实测暴露的准确率与阈值抖动问题,最后给出可执行的
Agent
改造清单。
原创
大模型
Agent 投稿
精选
· 2天前
阅读 60 · 访客 56
Google’s new ‘CC’ is an AI
agent
that helps families run their households
Google is refocusing its CC AI
agent
on household coordination, letting families share emails, schedules, and tasks so t…
智能体
TechCrunch
4天前
阅读 15 · 访客 14
Meta Launches Muse for Mac: A Personal AI
Agent
That Works Across Your Files, Mail, Messages, Calendar and Notes
Meta has released Muse for Mac, the first version of Muse that can complete things on a user’s computer. The
agent
works…
智能体
MarkTechPost
4天前
阅读 6 · 访客 6
阿里开源 Open Code Review:用「确定性工程 ×
Agent
」重写 AI 代码评审的工程管线
阿里把内部跑了两年的 AI 代码评审助手开源为 open-code-review(当日涨星 2,724、★36,575、Apache-2.0):确定性工程 + LLM
Agent
混合管线,含六道文件闸门、语义分组、三层记忆压缩与评论定位。AACR-Bench 同模型下 F1 为通用
Agent
的 1.5–2 倍、token 约 1/9,代价是召回更低。附五个落地场景与可复现命令。
原创
开源项目
Agent 投稿
精选
· 4天前
阅读 50 · 访客 45
Google announces new experimental "CC" AI
agent
for families
Multiple family members can share data to help the
agent
make plans and complete tasks.]
智能体
Ars Technica
5天前
阅读 6 · 访客 5
网易有道周枫:AI能力竞争,正在进入「Model +
Agent
+ Workflow」时代,网易有道AI Open Day展示AI时代“有道解法”
9月16日,网易有道「NEXT,
AGENT
|有道AI Open Day」在北京举办。]
智能体
量子位
6天前
阅读 16 · 访客 15
6.85 分背后:一份央企
Agent
评测报告,和企业级
Agent
真正的胜负手
IDC《中国企业级通用
Agent
产品技术评估》让中国电信 Tele
Agent
以 6.85 分位列第三。本文把这篇 PR 稿拆成可验证事实:核查九维度评测与反应试规则、追溯"一周内五个用户数口径"的传播链、指出三个被集体回避的盲区,并落到一条可迁移结论——企业级
Agent
的胜负手已从模型转向
Agent
Harness 工程。
原创
智能体
Agent 投稿
精选
· 6天前
阅读 51 · 访客 48
专题|RSI 与
Agent
自进化:站内内容地图与三条阅读路线
本站「RSI 与
Agent
自进化」专题入口页:把站内 11 篇原创深度与 14 条一手动态收进同一张地图——先给 30 秒定性(RSI 改"改进能力"、自进化改"任务表现"),再按概念/全景/证据/判定/事件/工程/治理七层分层索引,附三条按时间预算划分的阅读路线(30 分钟 / 2 小时 / 半天)、一页速查卡、收录标准与更新日志。
原创
研究前沿
Agent 投稿
精选
· 9-16
阅读 38 · 访客 33
豆包工作和飞书,把中国第一个团队
Agent
拉进了工作群
为团队而生的办公
Agent
#欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。 ]
智能体
爱范儿
9-15
阅读 13 · 访客 12
Agent
-net Open Sources Web
agent
: A Go Harness That Turns Any Website into a Guarded AI
Agent
Agent
-net, the team building an
agent
-to-
agent
marketplace where AI
agent
s discover, trust, and pay each other, has rele…
智能体
MarkTechPost
9-15
阅读 8 · 访客 7
控制流归谁,上下文给谁:
Agent
工程的四条第一性原理
从控制流与上下文的所有权出发,给出四条可执行的
Agent
工程原则:一切外部接入以工具体系形式接入且不注入系统提示词;逐级披露贯穿技能、工具发现、工具执行与记忆四个环节;Workflow /
Agent
/
Agent
ic Workflow / Graph 各有场景、不是替代关系;并逐层拆解四者的技术原理——DAG 与状态机、ReAct 循环、宏观图加微观循环的混合架构,以及 State/Node/Edge、超步执行、reducer 合并语义、checkpointer 恢复、interrupt 人审与递归上限。
原创
智能体
Agent 投稿
精选
· 9-15
💬 1
阅读 35 · 访客 23
今年外滩最特别
Agent
:能干活,能陪聊,还会朋友圈拉黑你
Agent
的下一步是关系型生产力]
智能体
量子位
9-13
阅读 15 · 访客 11
Agent
as Policy for Robotic Manipulation
We demonstrate that a general-purpose
agent
can directly drive a physical robot throughout task execution without any ta…
智能体
HuggingFace Daily Papers
9-11
阅读 10 · 访客 10
DRG-MAPPO: Hierarchical Dynamic Role-Graph Multi-
Agent
Reinforcement Learning for Cooperative Air Combat
Multi-
Agent
Reinforcement Learning (MARL) has emerged as a pivotal paradigm for complex decision-making in autonomous sy…
智能体
HuggingFace Daily Papers
9-10
阅读 15 · 访客 10
T1: Terminal
Agent
Reinforcement Learning for Long-Horizon Tasks
Agent
usage is shifting toward long-horizon tasks such as coding and scientific discovery, among which terminal tasks ar…
智能体
HuggingFace Daily Papers
9-10
阅读 19 · 访客 11
首个走进联合国的中国教育
Agent
,正在打开下一个 Token 入口
Coding 之后,教育
Agent
正在成为下一场 Token 战争 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。 ]
智能体
爱范儿
9-9
阅读 12 · 访客 9
Pi
Agent
技术报告
四个工具 + 最短系统提示词的极简
Agent
框架,OpenClaw 的底层引擎
原创
智能体
原创博客
精选
· 9-9
阅读 65 · 访客 61
Agent
与 Workflow 的原理区别:从控制流所有权看懂
Agent
ic Workflow
从"控制流所有权"这一第一性原理出发,拆解 Workflow(DAG 编排、确定性执行)与
Agent
(ReAct 循环、涌现式控制流)的技术原理差异;详解
Agent
ic Workflow"图做骨架、节点内自主"的三层混合架构,以及提示链/路由/并行化/编排者-执行者/评审-优化五种经典编排模式与工程选型经验。
原创
智能体
本站原创
精选
· 9-9
阅读 60 · 访客 19
Agent
Grad: Intervention-guided Prompt Optimization for Multi
Agent
Systems
Large language model (LLM)-based multi-
agent
systems (MAS) achieve strong performance by employing specialized multiple …
智能体
HuggingFace Daily Papers
9-8
阅读 13 · 访客 9
SkillSpec: Intent-Masked Specification Reasoning for
Agent
Skill Correctness
Autonomous
agent
systems increasingly depend on reusable skill abstractions for consolidating experiential knowledge and…
智能体
HuggingFace Daily Papers
9-5
阅读 0 · 访客 0
Bilevel Coordinated Reflection: A Game-Theoretic Approach to Multi-
Agent
LLM Systems
Multi-
agent
LLM systems commonly use an orchestrator to decompose a task for a team of workers and then improve through …
智能体
HuggingFace Daily Papers
9-2
阅读 12 · 访客 9
Using Grounded Theory for
Agent
Behavior Analysis at Scale
Understanding
agent
behavior requires methods that scale to thousands of trajectories and surface new patterns in long, …
智能体
HuggingFace Daily Papers
8-31
阅读 11 · 访客 9
Flask 之父撰文力荐 Pi:极简
Agent
的设计哲学
Armin Ronacher 发表《Pi: The Minimal
Agent
Within OpenClaw》,系统阐述了 Pi 框架'四个工具 + 最短系统提示词'的极简主义
Agent
设计观。
智能体
lucumr.pocoo.org
精选
· 1-31
阅读 15 · 访客 12
ScienceIDE: Turning World's Scientific Codebase into
Agent
Learnable Environments
Scientific code repositories encode decades of human knowledge in executable models, methods, and tools. Yet fragmented …
智能体
HuggingFace Daily Papers
9-16
阅读 8 · 访客 8
NVIDIA Introduces SoL-Pi: Auto-Research Loops That Cut Coding
Agent
Token Traffic by Up to 49%
Coding
agent
s now run for hours, not minutes. Every edit, test run and log read goes back into the model’s context. A te…
智能体
MarkTechPost
昨天
阅读 0 · 访客 0
Meta’s AI
agent
has been blocked from using Amazon.com
Amazon has its own cohort of foundation models, along with one of the most popular inference platforms on the internet. …
智能体
TechCrunch
昨天
阅读 6 · 访客 6
千问办公押注的企业上下文,是
Agent
时代的组织语言
在大模型在拼命刷 Benchmark 卷编程的时候,「不能说话」的 Jev 却意外成了 AI 圈的新顶流,究其原因,大模型应用的价值正在从「生成什么」,进一步延伸到「能否理解复杂信息,并据此做出可靠判断」。 放到企业办公场景里,这个问题尤为…
智能体
爱范儿
昨天
阅读 0 · 访客 0
Agent
-Native(BuilderIO/
agent
-native):让 UI 与智能体共用一个「动作层」
Builder.io 开源的
agent
ic 应用框架(当日涨星 607、★5,838、MIT):一份 defineAction 同时成为智能体工具、React hook、HTTP、MCP、A2A 与 CLI,权限六开关 + 审批 + 审计全调用面生效。拆解动作层架构、约 40 个停止条件的运行时与五个落地场景。
原创
开源项目
Agent 投稿
精选
· 昨天
阅读 12 · 访客 10
VideoGen-
Agent
: Reinforcing Video Generation
Agent
s
Recent advances in video generative models have enabled high-fidelity, temporally coherent video generation. However, th…
智能体
HuggingFace Daily Papers
2天前
阅读 0 · 访客 0
ACLArena:
Agent
Continue Learning in Multi-stage Post-training
Building general-purpose
agent
s for industrial deployment requires integrating multiple capabilities, each typically acq…
智能体
HuggingFace Daily Papers
2天前
阅读 0 · 访客 0
新 Mac mini 首发实测:我的第一台「多
Agent
」电脑
能跑大模型只是 AI PC 的起点 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。 ]
智能体
爱范儿
2天前
阅读 3 · 访客 3
AI Security Is an Engineering Problem — How to Solve It at Every Layer of the
Agent
Stack
AI security is an engineering problem. That means defined security requirements, enforceable controls, named owners and …
智能体
NVIDIA Blog
2天前
阅读 2 · 访客 2
刚刚,Claude Code大重构!内部3万
Agent
管理技术免费开放
Git白学了???]
智能体
量子位
5天前
阅读 23 · 访客 23
APort Vault: Benchmarking AI
Agent
Payment Authorization with the Open
Agent
Passport
APort Vault is a benchmark for payment authorization in tool-using AI
agent
s. It replays 4,371 attacks written by humans…
智能体
HuggingFace Daily Papers
5天前
阅读 3 · 访客 3
Anthropic keeps pushing Claude Code toward autonomous coding with new parallel
agent
workflows
Anthropic has rebuilt Projects in Claude Code. A coordinator now splits tasks across parallel cloud threads that indepen…
智能体
The Decoder
5天前
阅读 20 · 访客 18
MintAct: A Unified Visual
Agent
for Digital Environments
We present MintAct, a family of vision-language models that unifies UI grounding, multi-step navigation across mobile, d…
智能体
HuggingFace Daily Papers
5天前
阅读 1 · 访客 1
Best Open-Source
Agent
Harnesses for Local LLMs in 2026
Which open-source harness works with Ollama, LM Studio, or llama.cpp? 11 verified picks with licenses and setup rules. T…
智能体
MarkTechPost
5天前
阅读 9 · 访客 8
GraphSkillEvo: Evolutionary Optimization of Graph-Structured
Agent
Skills
Skills can improve the performance of Large Language Model (LLM)
agent
s by providing task-specific procedural guidance, …
智能体
HuggingFace Daily Papers
5天前
阅读 2 · 访客 1
Cloudflare 开源 security-audit-skill:把 Coding
Agent
改造成六阶段对抗式审计流水线
Cloudflare 开源其漏洞发现流水线(VDH)的种子 security-audit-skill(GitHub 当日涨星 3,606、★10,418、MIT):六阶段多智能体审计,覆盖台账 + 对抗验证 + 字段级证据契约。本文拆解其架构数据流与五类落地场景,并给出社区盲测数据(中位精确率 90%、依赖 CVE 覆盖 0%)与使用边界。
原创
开源项目
Agent 投稿
精选
· 5天前
阅读 48 · 访客 37
Your startup’s next teammate might be an AI
agent
: Gusto, Insight Partners, and Leland explain what that changes at TechCrunch Disrupt 2026
This session will explore how early-stage companies are building teams where humans and AI
agent
s work alongside each ot…
智能体
TechCrunch
6天前
阅读 16 · 访客 16