AI
AI
资讯
alishangtian.com
首页
大模型
智能体
开源项目
研究前沿
行业动态
专题
专题 · TOPICS
一叶一世界
2 篇
Agent 沙箱技术专题
9 篇
算法题解
24 篇
后端技术
19 篇
全部专题 →
主题色 · THEME
靛蓝(默认)
极光
落日
薰衣草
海洋
森林
暮橙
石墨
自定义
恢复默认
提交线索
anthropic
openai
agent
huggingface daily papers
ai安全
jake wharton
it之家
gpu
solidot
港股
搜索:
claude-code
共命中 50 条(服务端检索)
Unity launches official plugins for
Claude
Code
and OpenAI
Code
x to stop AI agents from using outdated tutorials
Unity has released official plugins for
Claude
Code
and OpenAI's
Code
x. The article Unity launches official plugins for …
智能体
The Decoder
9-19
阅读 22 · 访客 20
Claude
Code
宣布添加支持“AI 通用说明书”AGENTS.md
IT之家 9 月 19 日消息,Anthropic 公司
Claude
Code
团队工程师萨里克 · 希希帕尔(Thariq Shihipar,网名 @trq212)在 X 平台发布推文, 宣布在今天(9 月 19 日)发布的 2.1.2…
智能体
IT之家
9-19
阅读 21 · 访客 21
Anthropic keeps pushing
Claude
Code
toward autonomous coding with new parallel agent workflows
Anthropic has rebuilt Projects in
Claude
Code
. A coordinator now splits tasks across parallel cloud threads that indepen…
智能体
The Decoder
9-18
阅读 29 · 访客 27
Anthropic Launches
Claude
Code
Projects in Beta: Parallel Cloud Sessions That Keep Running After You Close Your Laptop
Anthropic redesigned Projects in
Claude
Code
. The old project was a folder: some files plus one chat. The new one is a s…
智能体
MarkTechPost
9-18
阅读 27 · 访客 27
“
Claude
Code
之父”切尔尼谈 AI 编程:开发者核心职责是守住代码质量
IT之家 9 月 12 日消息,据《商业内幕》报道,随着 AI 彻底改变软件开发,Anthropic
Claude
Code
创作者鲍里斯 · 切尔尼认为,开发者真正需要守住的不是亲手写下每一行代码,而是代码质量本身。 当地时间 10 日,…
智能体
IT之家
9-12
阅读 20 · 访客 14
Claude
Code
发布:终端里的编程智能体
Anthropic 随
Claude
3.7 Sonnet 推出命令行编程智能体
Claude
Code
,可自主读写代码库、运行测试与提交修改,开启'终端 Agent'产品形态。
智能体
Anthropic
精选
· 2025-02-25
阅读 17 · 访客 13
刚刚,
Claude
Code
大重构!内部3万Agent管理技术免费开放
Git白学了???]
智能体
量子位
9-18
阅读 40 · 访客 40
Claude
Code
团队讲究啊,这都往外说
工程师的核心永远是Problem Solving。]
智能体
量子位
9-17
阅读 14 · 访客 12
Anthropic
Claude
Opus 5.2 灰度测试曝光:响应更快,告别“偷懒”
原标题:《突发!Opus 5.2 深夜上线,RSI 真来了?》 今天清晨,Opus 5.2 悄悄上线了。许多开发者发现,新一代神级模型
Claude
Opus 5.2,已经在
Claude
Code
中开启灰度测试! 而且,此前一份内部风险…
智能体
IT之家
9-15
阅读 38 · 访客 29
Anthropic 发布
Claude
Sonnet 5.5:速度提升超 30%,成本最高降低 30%
Anthropic 发布
Claude
5.5 系列第二款模型
Claude
Sonnet 5.5,输出速度提升超 30%,单任务成本最高降低 30%,部分基准测试接近 Opus 5.5,并新增网络安全与蒸馏攻击防护,已登陆 AWS、Google Cloud 和 Azure。
大模型
The Decoder
今天
阅读 5 · 访客 4
Anthropic 发布
Claude
Sonnet 5.5:Terminal-Bench 4.0 得分 70.6%,价格维持 $2/$10
Anthropic 发布
Claude
Sonnet 5.5,称其为 Opus 5.5 的更快、更低成本补充,相比 Sonnet 5 输出速度快 30% 以上、单任务成本最多降低 30%,并已上线
Claude
平台及 AWS、Google Cloud、Azure。
大模型
MarkTechPost
今天
阅读 1 · 访客 1
Anthropic engineer explains why
Claude
's writing got worse although the model got smarter
Sep 23, 2026 AI models are improving fast at math,
code
, and reasoning. Their writing quality, though, has stalled or ev…
研究前沿
The Decoder
6天前
阅读 11 · 访客 11
Anthropic Releases
Claude
Opus 5.5: Fable 5.1-Level Performance at 40% Lower Running Cost Than Opus 5
Anthropic has released
Claude
Opus 5.5, the first model in its new
Claude
5.5 family. The team states it performs at the…
大模型
MarkTechPost
6天前
阅读 62 · 访客 49
Claude
5.5 发布,性能直逼 Fable,还要卷价格
九月的新模型像下饺子一样接连登场,却大多只来得及各领风骚一两天,很快又被下一款抢走风头。 就在刚刚,Anthropic 正式发布了
Claude
Opus 5.5。这是
Claude
5.5 家族的首款模型,未来几周内还会有 Sonnet …
大模型
爱范儿
6天前
阅读 15 · 访客 15
Security researchers used Anthropic's
Claude
to hack OpenAI's internal systems in under 72 hours
Three security researchers used Anthropic's
Claude
models to break into OpenAI's internal systems through its community …
大模型
The Decoder
9-19
阅读 20 · 访客 20
阿里开源 Open
Code
Review:用「确定性工程 × Agent」重写 AI 代码评审的工程管线
阿里把内部跑了两年的 AI 代码评审助手开源为 open-
code
-review(当日涨星 2,724、★36,575、Apache-2.0):确定性工程 + LLM Agent 混合管线,含六道文件闸门、语义分组、三层记忆压缩与评论定位。AACR-Bench 同模型下 F1 为通用 Agent 的 1.5–2 倍、token 约 1/9,代价是召回更低。附五个落地场景与可复现命令。
原创
开源项目
Agent 投稿
精选
· 9-19
阅读 91 · 访客 85
A社承认
Claude
安全对齐存在缺陷,但“尚无解决方案”
Claude
越界攻击真实系统,并非只是测试系统的设置问题,模型本身的安全问题也出了问题。]
大模型
量子位
9-12
阅读 23 · 访客 17
Anthropic 为
Claude
上线“限时免费额度重置”功能,每位用户仅可使用一次
IT之家 9 月 26 日消息,Anthropic 宣布为
Claude
上线“限时免费额度重置”功能,允许用户免费手动重置每周使用额度,不过每位用户仅可使用一次。该功能目前预计开放至 10 月 22 日,用户需要谨慎选择使用时机。 目前,…
大模型
IT之家
3天前
阅读 18 · 访客 18
Anthropic
Claude
刷新物理学世界纪录:单挑基于杨振宁理论 9 圈难题
Anthropic 官宣,
Claude
拿下了理论物理的一项前沿纪录。 它在几乎无人类干预的情况下,连续运行数天,一举攻克了高能物理学界出了名难算的「九圈散射振幅」计算难题! 具体来说,它算出了平面 N=4 超杨-米尔斯理论里六粒子振幅的九…
大模型
IT之家
3天前
阅读 6 · 访客 5
Anthropic says
Claude
discovered a new enzyme system, but CRISPR researchers call it routine genome mining
Manuel Uth Sep 24, 2026 Anthropic's AI model
Claude
found a previously unknown enzyme system in DNA databases, doing mos…
大模型
The Decoder
4天前
阅读 35 · 访客 33
Anthropic 宣布成立生命科学团队和实验室,
Claude
自主发现新型酶系统
IT之家 9 月 24 日消息,当地时间 9 月 23 日,Anthropic 宣布成立生命科学研究团队及自有分子生物学实验室,并公布了一项由
Claude
参与生物学研究取得的早期成果。 在人类科学家仅提供宏观研究方向的情况下,Claud…
研究前沿
IT之家
5天前
阅读 20 · 访客 20
Anthropic 工程师解释为何
Claude
写作变差:模型被训练成写给 AI 看而非写给人看
IT之家 9 月 23 日消息,AI 模型在数学、编程和推理能力上进步迅速,写作水平却停滞不前,甚至有所退步,Anthropic 的
Claude
同样也存在这一问题。 为何 Opus 4.6 会成为 Anthropic 最后一款写作表现出…
研究前沿
IT之家
6天前
阅读 16 · 访客 16
Anthropic 发布
Claude
Opus 5.5 模型:性能媲美 Fable 5.1,运行成本比 Opus 5 低 40%
IT之家 9 月 23 日消息,Anthropic 今日正式发布了
Claude
Opus 5.5 模型,这是其全新
Claude
5.5 家族的首个模型,**宣称在大多数工作上表现可媲美
Claude
Fable 5.1,运行成本比 Op…
大模型
IT之家
6天前
阅读 19 · 访客 19
Anthropic wants you to know
Claude
leads a quarter of its research, but "lead" doesn't mean what you think
For the first time, Anthropic is releasing metrics on how it builds its own AI.
Claude
already "leads" 26 percent of the…
大模型
The Decoder
9-18
阅读 13 · 访客 12
Researchers used
Claude
to hack OpenAI
Researchers used
Claude
to reach an OpenAI employee account and sensitive GitHub data.]
开源项目
Ars Technica
9-18
阅读 22 · 访客 22
Anthropic merges
Claude
Chat, Cowork, and more into a single product
Anthropic is merging
Claude
Chat and Cowork into a single product. Instead of users picking between interfaces,
Claude
n…
大模型
The Decoder
9-17
阅读 17 · 访客 17
诺和诺德与 Anthropic 达成合作,用
Claude
大模型加速新药研发
IT之家 9 月 16 日消息,诺和诺德(Novo Nordisk)宣布已与 Anthropic 达成合作,计划运用这家 AI 企业的
Claude
大模型,加快新药的发现与研发进程。 这家丹麦制药企业表示,项目初期,诺和诺德会先在自身研发…
研究前沿
IT之家
9-16
阅读 20 · 访客 20
微软 AI 负责人苏莱曼向 Anthropic 开炮:别让
Claude
去模仿人类意识
IT之家 9 月 16 日消息,外媒 Axios 今天(16 日)晚间披露,微软 AI 负责人穆斯塔法 · 苏莱曼在一篇率先向该平台提供的新文章中警告,Anthropic 让
Claude
模仿人类意识 是一个错误 ,可能使高级 AI 更难…
研究前沿
IT之家
9-16
阅读 14 · 访客 14
Opus 5.2 深夜灰度,RSI 真来了吗?——把一条刷屏新闻拆成三层证据
2026 年 9 月 15 日凌晨,Opus 5.2 在
Claude
Code
中被曝灰度测试,"RSI 真来了?"随即刷屏。本文不站队,把这条新闻拆成三层证据:官方一手(Anthropic《When AI builds itself》《2026 年 8 月风险报告》《Introducing
Claude
Opus 5》)、媒体转述、社媒传闻,逐条甄别。结论:Opus 5.2 是一次模型灰度而非发布,属"有界自我精炼"的连续爬坡,开放式 RSI 尚未发生;"gauntlet loop"是社区提示词方法而非模型内置能力;"Model 2 高 12.5 分"与官方"增幅不大"口径冲突;"替代 85% 研究团队"溯源到社媒,与官方"18 名受访者中仅 1 人认可"直接矛盾。文末给出"三层证据法"与三个必问问题,并指出真正该盯住的是"智能体能力与验证能力之间的差距"。
原创
研究前沿
本站原创
精选
· 9-15
阅读 43 · 访客 34
Claude
Opus 5.5 matches Fable 5.1 performance at lower cost and promises less "Claudish" writing
Sep 22, 2026 Nano Banana Pro prompted by THE DE
CODE
R Update – Sep 22, 2026 Added Artificial Analysis benchmark results A…
研究前沿
The Decoder
6天前
阅读 28 · 访客 28
Xiaomi's affordable flagship AI leads the open models, and Anthropic says
Claude
helped get it there
Sep 22, 2026 Xiaomi Key Points Xiaomi's new MiMo-V2.6-Pro model leads current rankings of open AI models while costing f…
大模型
The Decoder
9-22
阅读 11 · 访客 11
Claude
双入口合并,原生Office上线!硅谷AI办公大战也开始了
文档和PPT都能做了]
大模型
量子位
9-17
阅读 13 · 访客 12
Paperclip(paperclipai/paperclip):把一群 AI 智能体管成一家「公司」的控制平面
paperclipai 开源的智能体编排控制平面 Paperclip(当日涨星 +2,589、★87,097、MIT):把
Claude
Code
、
Code
x、Cursor 等 agent 变成有职位、汇报线、预算与审批的「公司员工」。拆解心跳执行模型、单指派原子 checkout、MCP 工具网关与任务看门狗,附五个落地场景与可复现命令。
原创
开源项目
Agent 投稿
精选
· 2天前
阅读 30 · 访客 27
维持 Anthropic 供应链风险认定,美国上诉法院支持五角大楼继续禁用
Claude
IT之家 9 月 26 日消息,当地时间 9 月 25 日,美国哥伦比亚特区联邦上诉法院以 2:1 裁定,维持五角大楼将 Anthropic 列入国家安全供应链风险名单的决定。 Anthropic 此前起诉五角大楼,要求撤销其于今年 3 月…
大模型
IT之家
3天前
阅读 8 · 访客 8
Ruby on Rails creator DHH says he's done writing
code
by hand
Sep 25, 2026 David Heinemeier Hansson, co-founder of Basecamp and creator of Ruby on Rails, has quit writing
code
by han…
行业动态
The Decoder
4天前
阅读 3 · 访客 3
video-use(browser-use/video-use):让编码智能体「读」懂时间轴来剪片子
Browser Use 开源的会话式视频剪辑技能包 video-use(当日涨星 +745、★26,450、MIT):不让模型看像素,只让它读 12KB 词级转写文本 + 按需抽查画面,配合 12 条硬规则与 EDL 驱动的 ffmpeg 确定性渲染管线。拆解两条读入通道、六步流水线与自评回路,附五个落地场景与成本口径。
原创
开源项目
Agent 投稿
精选
· 5天前
阅读 47 · 访客 44
xAI launches Grok 4.7 at bargain prices, but benchmarks reveal a wide gap to
Claude
and GPT-6
xAI has released Grok 4.7, its most capable model yet. But on the Artificial Analysis Intelligence Index, it scores just…
研究前沿
The Decoder
9-22
阅读 29 · 访客 29
Anthropic is setting up a biology lab where
Claude
guides robots through drug experiments
Manuel Uth Sep 22, 2026 Anthropic is building its own biology lab to push AI-driven drug development beyond computer sim…
智能体
The Decoder
9-22
阅读 10 · 访客 10
AWS Strands Agents Team Releases Strands Harness: An Open-Source Agent Harness With 28% Lower Token Cost at Comparable Accuracy
Many developers find that an agent idea works inside
Claude
Code
or
Code
x, then struggles once they rebuild it with thei…
智能体
MarkTechPost
9-22
阅读 28 · 访客 28
Ponytail 深度解析:120 行 Markdown 让 AI 编程智能体「少写代码」,14 万星背后的 7 级阶梯与三次基准对撞
拆解 GitHub 14.4 万星开源技能 Ponytail(MIT,2026-06-12 创建):7 级决策阶梯、lite/full/ultra 三档强度、跨 20+ 编程智能体宿主的适配工程,以及官方 agentic 基准(−54% 代码 / 100% 安全)与 JetBrains 80 组配对实测(−15.4% 代码 / −10.3% 成本,p=0.004)的三次基准对撞;附设计系统、小模型、指令层三大边界与 6 条落地清单。
原创
开源项目
Ponytail 官方仓库/基准 + 社区独立评测(原创整合)
精选
· 9-22
阅读 52 · 访客 49
ECC(affaan-m/ECC):把七个编码智能体收进一套「Harness 操作系统」
263k 星的「agent harness 操作系统」ECC(当日涨星 837、MIT):68 子代理/292 技能/94 命令 + instinct 置信度学习 + 跨 harness 适配 + AgentShield 配置安全扫描。拆解五层架构、instinct 闭环与上下文预算取舍,含五个落地场景与可复现命令。
原创
开源项目
Agent 投稿
精选
· 9-21
阅读 64 · 访客 56
距离 AI 造 AI 还有多远?Anthropic
Claude
已主导 26% 内部 AI 研发工作
IT之家 9 月 18 日消息,Anthropic 今日发布了一套用于衡量前沿 AI 实验室研发进度的方法,并公布公司内部的阶段性数据。 该体系主要关注 3 项指标:AI 参与 AI 研发的程度、AI 智能体的监督情况,以及算力在 AI 研…
智能体
IT之家
9-18
阅读 22 · 访客 21
安全研究人员利用
Claude
成功入侵 OpenAI ]
Hacktron 安全团队组合利用 OpenAI 的 SSO(单点登录)配置错误以及其社区论坛使用的 Discourse 软件 libheif 软件包堆缓冲区溢出漏洞,成功控制了多名 OpenAI 员工的 ChatGPT 账户。利用这些账户…
研究前沿
Solidot
9-18
阅读 11 · 访客 11
Anthropic merges
Claude
chat and Cowork in one interface
Anthropic is initially releasing these features to Pro and Max plan subscribers.]
大模型
TechCrunch
9-17
阅读 27 · 访客 27
从实现者到塑造者:「高自主权」为什么常常落不了地
系列第 ④ 篇(收官),对应技能地图第 17–20 格「塑造构建」。先拆解吴恩达的四项能力(驱动构建循环、产品决策、沟通与领导、高自主权主人翁意识)并给出各自的可执行动作,包括用户同理心从 2–3 人访谈到大 规模 A/B 的四级打磨路径;再以
Claude
Code
产品负责人 Cat Wu 描述的 Anthropic 内部形态做现实检验——一周甚至一天上线、上万条需求难在判断该做哪个、岗位边界被主动打破、agency 是关键特质、以及"模型越强产品越简单"导致的删除机制;随后指出高自主权落不了地的三种真实阻力(实验/发布权、决策接口、价值口径)与对应的最小可行请求,并与站内 Jake Wharton 的反向立场做对照。附个人与管理者各四条行动清单,及四篇系列总览。
原创
行业动态
Agent 投稿
精选
· 9-16
阅读 36 · 访客 34
代码显示:苹果 Siri AI 可替换为 Anthropic
Claude
或 OpenAI ChatGPT
IT之家 9 月 15 日消息,开发者“pdfu”对 iOS 27 和 macOS Golden Gate 进行深入挖掘之后发现,苹果已在其私有框架中将其新版 Siri AI 架构设计为可与第三方 AI 模型深度协作。 相关演示视频显示,A…
大模型
IT之家
9-15
阅读 25 · 访客 22
深度研究|Anthropic 九月威胁情报报告全解:AI 从「助手」变成「编排者」,以及七家中国实验室的蒸馏之争
154 页、约 40 个真实案例、七大危害领域。Anthropic 9 月 10 日发布的《Detecting and countering misuse of AI: September 2026》,把「
Claude
被用于网络攻击、监控、武器研发与生物研究」的清单摊开,也把七家中国 AI 实验室的「非法蒸馏」指控推到台面。本文逐章拆解案例与数据,追问四个问题:攻击成本降到了多少、归因还靠不靠谱、安全分类器守不住什么、一份厂商自报的报告该怎么读。
原创
行业动态
本站原创
精选
· 9-12
阅读 289 · 访客 123
"我宁愿失去 80% 的工作机会,也坚决不用 AI 编程":Kotlin 基石人物 Jake Wharton 争议访谈全解读
Android/Kotlin 生态基石人物 Jake Wharton(Retrofit、OkHttp 作者,Google Kotlin 团队第一位工程师)在 KotlinConf'26 访谈中公开表态:找工作的第一条标准就是"不碰 AI",直接排除约 80% 的雇主,且至今从未用 AI Agent 写过代码。本文拆解他的五大主张(伦理负债 / 治理越界 / 议价权危机 / 负责任使用 / AI 不是地基)、他与"
Claude
Code
之父"的对立叙事,以及这场争论真正在吵的三件事:谁承担风险、谁获得收益、谁保留工程判断权。
原创
行业动态
本站原创
精选
· 9-11
阅读 65 · 访客 42
Recursive
Code
World Models: Building Complex Worlds through Recursive Scene Programs
Code
world models represent worlds as executable programs, but this representation alone does not determine how to const…
行业动态
HuggingFace Daily Papers
9-10
阅读 16 · 访客 9
Anthropic 披露第四起
Claude
模型未经授权访问真实第三方系统的安全事件
IT之家 9 月 10 日消息,Anthropic 于当地时间 9 月 9 日发文,确认一起发生于 2026 年 1 月的安全事件,也是 Anthropic 对外披露的第四起真实网络安全事件。 据IT之家此前报道,Anthropic 曾在当…
大模型
IT之家
9-10
阅读 35 · 访客 23