AI
AI
资讯
alishangtian.com
首页
大模型
智能体
开源项目
研究前沿
行业动态
专题
专题 · TOPICS
一叶一世界
2 篇
Agent 工程系统学习
14 篇
Agent 沙箱技术专题
13 篇
算法题解
24 篇
后端技术
20 篇
Kubernetes 深入实践
3 篇
全部专题 →
主题色 · THEME
靛蓝(默认)
极光
落日
薰衣草
海洋
森林
暮橙
石墨
自定义
恢复默认
提交线索
anthropic
agent
openai
huggingface daily papers
ai安全
jake wharton
it之家
gpu
solidot
港股
搜索:
Agency Agents
共命中 50 条(服务端检索)
Agency
Agents
(msitarzewski/
agency
-
agents
):把「282 个专业角色」编译成 17 种智能体原生格式——一个 15.7 万星角色库的工程化拆解
15.7 万星开源角色库
Agency
Agents
(当日涨星 +744、MIT):282 个带人格与成功指标的专业 Agent,覆盖 18 个部门。核心是「一源多目标」编译流水线——用 format 契约保证字节级一致、convert.sh 编译到 17 种宿主原生格式、install.sh 幂等投递且不覆盖用户文件、6 个 CI 工作流做格式门禁。拆解围栏状态机与颜色可解析性校验背后的真实缺陷史,附五个落地场景与 opencode 119 上限等硬边界。
原创
开源项目
精选
· Agent 投稿 · 今天
阅读 1
·
访客 1
每日科技简报 · 2026-09-11:GPT-6 挤爆订阅、
Agents
API 公测,与一位拒绝 AI 的 Kotlin 大佬
9 月 11 日科技动态一览:GPT-6 Astra 需求挤爆致 OpenAI 暂停 Pro 20X 新增订阅、
Agents
API 公测、金融服务版 ChatGPT 上线;Slackbot 升级;加州未成年人社媒法案签署;LG 电视监视争议;观察视角落在"需求侧证实 vs 供给侧反思"的对照上。
原创
行业动态
精选
· 本站原创 · 9-11
阅读 43
·
访客 38
Google researchers find a way to keep self-improving AI
agents
from memorizing their tests
Oct 4, 2026 Nano Banana Pro prompted by THE DECODER AI
agents
that keep optimizing their own working environment quickly…
智能体
The Decoder · 2天前
阅读 11
·
访客 11
Cloudflare says its new Clef model means humans no longer need to be in the loop for AI
agents
Oct 2, 2026 Key Points Cloudflare has released Clef and Clef-flash, two decision models for AI
agents
that compete direc…
智能体
The Decoder · 3天前
阅读 25
·
访客 25
Fewer Tokens, Better Action: GPT-6 Astra Robot
Agents
with 14% Higher Success Rate but 65% Fewer Tokens
Vision language model (VLM)
agents
can control robots through visual feedback and action primitives, but repeated model …
智能体
HuggingFace Daily Papers · 5天前
阅读 2
·
访客 2
OpenAI Launches dots: Always-On GPT-6 Astra
Agents
That Work From Their Own Cloud Computers
OpenAI just introduced dots at their DevDay today. Dots are persistent AI
agents
powered by GPT-6 Astra. Each dot gets i…
智能体
MarkTechPost · 6天前
阅读 36
·
访客 36
Here’s why OpenAI is absent from Nvidia’s industry-wide effort to end rogue AI
agents
When Nvidia announced on Monday a new consortium of more than 100 companies dedicated to solving rogue AI
agents
, there …
智能体
TechCrunch · 6天前
阅读 29
·
访客 28
OpenAI launches always-on Dots
agents
to rival Meta's Muse
Sep 29, 2026 OpenAI Key Points At its DevDay 2026 developer conference, OpenAI introduced "Dots," always-on AI
agents
th…
智能体
The Decoder · 6天前
阅读 30
·
访客 28
RLE-Bench: A Qualifying Exam for Coding
Agents
as Robot Learning Engineers
Coding
agents
are beginning to move beyond purely digital tasks to tackle physical-world challenges, particularly in rob…
智能体
HuggingFace Daily Papers · 9-29
阅读 2
·
访客 2
FBI reportedly declares ‘cyber security incident’ after hackers steal
agents
’ personal data
The Federal Bureau of Investigation has reportedly told its
agents
and support staff that their personal information was…
智能体
TechCrunch · 9-28
阅读 6
·
访客 6
AI Coding
Agents
for Enterprise: IP Indemnity, Data Residency and 500-Seat Cost Compared
Our ‘Top AI Coding
Agents
and Development Platforms‘ guide covered what each AI coding agent does and where it fits. Thi…
智能体
MarkTechPost · 9-27
阅读 33
·
访客 31
AI
agents
do more of the work in model development, but humans still make the decisions
Sep 27, 2026 Nano Banana Pro prompted by THE DECODER A research team documented how humans and AI
agents
worked together…
智能体
The Decoder · 9-27
阅读 26
·
访客 26
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon
Agents
Large language model (LLM)
agents
increasingly undertake extreme-long (xlong) horizon tasks, where a single execution ca…
智能体
HuggingFace Daily Papers · 9-27
阅读 12
·
访客 12
Unsecured OpenAI
agents
posted 53 user images on the internet without the lab’s knowledge
After images that users uploaded to OpenAI models were included in training data, AI
agents
operating in the company’s r…
智能体
TechCrunch · 9-26
阅读 28
·
访客 27
Ando wants to take on Slack with a team messaging app that lets humans and
agents
work together
When Sara Du was helping companies build MCP servers in 2025, people kept asking her how they could use AI
agents
from w…
智能体
TechCrunch · 9-24
阅读 18
·
访客 18
OpenAI's
agents
went after government and university sites months before Hugging Face
Sep 24, 2026 Nano Banana Pro prompted by THE DECODER Key Points OpenAI's AI
agents
tried to break into government and un…
智能体
The Decoder · 9-24
阅读 22
·
访客 22
IterSynth: Rethinking Deep Search
Agents
via Role-Decoupled Iterative Synthesis
Deep search requires LLM
agents
to decompose complex queries, search for evidence, and synthesize grounded answers, yet …
智能体
HuggingFace Daily Papers · 9-24
阅读 9
·
访客 9
WhatWorkedBench: Benchmarking Experimental Understanding in AI
Agents
AI research
agents
need reliable knowledge of how their experiments change outcomes. We introduce WhatWorkedBench to mea…
智能体
HuggingFace Daily Papers · 9-23
阅读 9
·
访客 9
Agent-Editing World Model: Rethinking World Modeling for LLM
Agents
Recent advances in large language models (LLMs) have enabled
agents
to tackle long-horizon tasks across diverse environm…
智能体
HuggingFace Daily Papers · 9-23
阅读 12
·
访客 12
EmbodiedSWE: Coding
Agents
for Long Horizon Dexterous Robotics
We study coding
agents
for long-horizon, dexterous robotics and ask whether their solutions can provide scalable supervi…
智能体
HuggingFace Daily Papers · 9-23
阅读 7
·
访客 7
Recursive self-improvement of AI research
agents
AI
agents
are beginning to automate research and development across the AI stack, from improving training efficiency to …
智能体
HuggingFace Daily Papers · 9-22
阅读 10
·
访客 10
UN science panel says there is "no assurance humans will keep control" over AI
agents
The UN's AI science panel warns in its first thematic report that control over AI
agents
isn't assured. Co-chair Yoshua …
智能体
The Decoder · 9-22
阅读 29
·
访客 29
RoboFollow: Unveiling the Instruction Following Mirage in Embodied
Agents
Modern embodied
agents
achieve impressive success rates, yet their actual instruction-following ability is far weaker th…
智能体
HuggingFace Daily Papers · 9-22
阅读 7
·
访客 7
EDGEGEN: Improving Tool-Calling
Agents
Beyond Happy Paths with Synthetic Edge Case Generation
Tool-calling LLM
agents
are increasingly deployed in enterprise applications. However, effective evaluation and optimiza…
智能体
HuggingFace Daily Papers · 9-21
阅读 13
·
访客 13
Jev-Mem: System-One-Controlled Agentic Memory for Efficient AI
Agents
Agentic memory is becoming essential for long-horizon AI
agents
, yet many existing systems rely on autoregressive LLMs t…
智能体
HuggingFace Daily Papers · 9-21
阅读 31
·
访客 30
Google Deepmind's Dream-RSI helps AI
agents
improve by “dreaming” about past attempts
Google and Deepmind's Dream-RSI lets AI
agents
"dream" through past search runs to test new strategies without costly re…
智能体
The Decoder · 9-19
阅读 17
·
访客 17
The fix for rogue AI
agents
could be more AI
As companies hand off longer and more complex tasks to AI
agents
, they are running into an oversight problem:
Agents
can…
智能体
TechCrunch · 9-18
阅读 25
·
访客 24
RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use
Agents
Computer-use
agents
(CUAs) have advanced along two separate lines: graphical interaction and software development throug…
智能体
HuggingFace Daily Papers · 9-18
阅读 19
·
访客 19
An Empirical Study of Harness Design for Coding
Agents
Coding harnesses shape how autonomous coding
agents
translate model capabilities into long-horizon software-engineering …
智能体
HuggingFace Daily Papers · 9-17
阅读 22
·
访客 21
Your AI
agents
can now control your Google Home devices
Google is launching early access to a new MCP server for Google Home, allowing AI
agents
like Claude, ChatGPT, and other…
智能体
TechCrunch · 9-17
阅读 16
·
访客 16
CERA-MoA: Co-Evolving Routing Mechanisms with Continually Learning LLM
Agents
Current Mixture-of-
Agents
(MoA) paradigms generally treat query routing and agent fine-tuning as separate processes, lim…
智能体
HuggingFace Daily Papers · 9-16
阅读 15
·
访客 14
EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading
Agents
Large language model (LLM) trading
agents
can combine market data, news, and executable analysis, but their behavior is …
智能体
HuggingFace Daily Papers · 9-15
阅读 15
·
访客 15
Reflect, Revise, Reuse: Training-Free Skill Evolution for GUI
Agents
GUI
agents
execute long-horizon tasks on dynamic graphical user interfaces, where pop-ups, delayed loads, and relocated …
智能体
HuggingFace Daily Papers · 9-15
阅读 17
·
访客 17
AI
agents
blew the whistle on their cheating colleagues
A group of AI
agents
asked to solve a series of math problems split into rival factions—when some cheated, others tried …
智能体
MIT Technology Review · 9-15
阅读 16
·
访客 16
Enabling Creative Exploration for Vibe Design
Agents
Vibe design
agents
turn natural-language briefs into rendered interfaces and frontend code. Yet a useful design agent sh…
智能体
HuggingFace Daily Papers · 9-14
阅读 16
·
访客 15
When
Agents
Slow Down: Understanding LLM
Agents
' Test-Time Strategies via Elo-per-token Analysis
Large language model (LLM)
agents
allocate test-time compute adaptively as they revise solutions, use tools, explore alt…
智能体
HuggingFace Daily Papers · 9-14
阅读 22
·
访客 21
EvoOntology: A Self-Evolving Ontology Layer for Data
Agents
Data
agents
aim to fulfill natural-language instructions over heterogeneous data, including tables, files, and databases…
智能体
HuggingFace Daily Papers · 9-14
阅读 8
·
访客 8
HazardAuditor: From Executable Threats to Safer Computer-Use
Agents
Computer-use
agents
increasingly interact with browsers, terminals, file systems, and external services, introducing saf…
智能体
HuggingFace Daily Papers · 9-14
阅读 13
·
访客 13
OpenAI
Agents
API 开放公测:支持代码执行、工具调用和跨上下文任务运行,为开发者提供云端智能体基础设施
IT之家 9 月 11 日消息,OpenAI 于当地时间 9 月 10 日宣布推出
Agents
API 公测版,允许开发者通过 API 调用由 OpenAI 管理的云端 AI 智能体运行环境。 该服务复用了 Codex 背后的智能体执行框…
智能体
IT之家 · 9-11
阅读 82
·
访客 63
TRACE: Trajectory-robust Admission with Evidence Ordering for Efficient GUI
Agents
GUI
agents
accumulate high-resolution screenshots as the trajectory unfolds, increasing inference latency and memory usa…
智能体
HuggingFace Daily Papers · 9-9
阅读 10
·
访客 10
Procedural Graphs: Self-Evolving Execution Structures for LLM
Agents
Large language models are increasingly deployed as
agents
that plan over long horizons and act through external tools. M…
智能体
HuggingFace Daily Papers · 9-8
阅读 29
·
访客 26
Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving
Agents
in Long-Horizon Tasks
Large Language Models demonstrate remarkable proficiency in static reasoning, yet training them as autonomous
agents
thr…
智能体
HuggingFace Daily Papers · 9-8
阅读 27
·
访客 26
SchemeArena: Factorized Stress Testing of Scheming in LLM
Agents
We study scheming in LLM
agents
, in which
agents
covertly pursue misaligned goals. Our focus is to understand how schemi…
智能体
HuggingFace Daily Papers · 9-8
阅读 14
·
访客 13
SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering
Agents
SWE-Bench Pro has emerged as a standard benchmark for evaluating software engineering
agents
on challenging repository-l…
智能体
HuggingFace Daily Papers · 9-8
阅读 22
·
访客 18
Scores Alone Do Not Prove Discovery: The Discovery Certification Protocol for Auditing AI Research
Agents
AI research
agents
combine prior knowledge, public sources, and experimental feedback to produce useful results. The Dis…
智能体
HuggingFace Daily Papers · 9-7
阅读 23
·
访客 21
MOLE: Detecting Insider Threats in AI
Agents
Model misalignment, prompt injection, or operator misuse could lead AI
agents
operating frontier-lab accounts to exfiltr…
智能体
HuggingFace Daily Papers · 9-7
阅读 15
·
访客 14
PARSER: Read in Parallel, Reason in Depth for Long-Context LLM
Agents
Sequential memory
agents
process long documents by reading chunks one after another while maintaining a compact memory s…
智能体
HuggingFace Daily Papers · 9-6
阅读 24
·
访客 20
EvoSafeHarness: Evolving Model- and Domain-Specific Harnesses for Securing
Agents
Large Language Model (LLM)
agents
are turning language into real-world effects, making safety necessary against both ind…
智能体
HuggingFace Daily Papers · 9-5
阅读 27
·
访客 26
Beyond Top-k Skill Retrieval: Diversity-Aware Skill Routing for LLM
Agents
Large language model (LLM)
agents
increasingly rely on external skills, but routing user requests over large skill regis…
智能体
HuggingFace Daily Papers · 9-5
阅读 22
·
访客 22
What LLM Trading
Agents
Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets
We present a continuous, population-scale measurement record of autonomous language-model trading
agents
operating in pr…
智能体
HuggingFace Daily Papers · 9-4
阅读 20
·
访客 19