参考资料
1. 使用说明
图 1-1|参考资料体系:七大来源类 × A/B/C 可信度分级
数据来源:基于本文分析绘制的示意图。
1.1. 可信度分级标准
本文件收录的全部资料按三级可信度分级。该分级继承自本工程检索报告的取证分级体系(A / B / C),供读者判断引用方式。
| 等级 | 含义 | 使用规则 |
|---|---|---|
| A | 监管机构、标准组织、企业官方一手来源(gov.cn、nfra.gov.cn、cicpa.org.cn、tc260.org.cn、nist.gov、iso.org、darpa.mil、anthropic.com、openai.com、aaif.io、人民网、中国日报、arXiv 原文、官方开源仓库等) | 可直接引用,须给出出处 |
| B | 权威二手来源:主流媒体与行业媒体转述、百科与综述、汇总站转引 | 需注明“据 XX 报道 / 转述”;具体数字建议二次核对 |
|---|
1.2. 编号与引用规则
- 编号形如
W01起,按分类顺序连续编号,同一条目不重复编号; - 标注 者为链接或日期未经一手验证,引用时需二次确认;
- 无链接条目按“文献名 + 发布机构 + 年份”著录,不虚构 URL;
- “适用章节”以白皮书主题域标注(如“治理与风险”“发展展望”“技术架构”“行业赋能”“市场研究”“全篇”),便于按需取用。
2. 官方工程博客
| 编号 | 名称 | 机构 | 年份 | 等级 | 链接 | 适用章节 |
|---|---|---|---|---|---|---|
| W01 | Effective context engineering for AI agents | Anthropic | 2025 | A | https://www.anthropic.com/engineering/effective-context-engineering-for-ai-agents | 技术架构 / 发展展望 |
| W02 | Effective harnesses for long-running agents | Anthropic | 2025 | A | https://www.anthropic.com/engineering/effective-harnesses-for-long-running-agents | 技术架构 |
| W03 | Harness design for long-running application development | Anthropic | 2026 | A | https://www.anthropic.com/engineering/harness-design-long-running-apps | 技术架构 / 发展展望 |
| W04 | Equipping agents for the real world with Agent Skills | Anthropic | 2025 | A | https://www.anthropic.com/engineering/equipping-agents-for-the-real-world-with-agent-skills | 技术架构 |
| W05 | Sandboxing: a safer and more autonomous approach(权限提示减少 84%) | Anthropic | 2025 | A | https://www.anthropic.com/engineering/claude-code-sandboxing | 治理与风险 |
| W06 | Introducing Agent Skills(2025-12-18 转为开放标准) | Anthropic | 2025 | A | https://www.anthropic.com/news/skills | 技术架构 |
| W07 | Introducing the Model Context Protocol | Anthropic | 2024 | A | https://www.anthropic.com/news/model-context-protocol | 技术架构 |
| W08 | Donating the MCP and establishing the AAIF | Anthropic | 2025 | A | https://www.anthropic.com/news/donating-the-model-context-protocol-and-establishing-of-the-agentic-ai-foundation | 发展展望 |
| W09 | How we built our multi-agent research system | Anthropic | 2025 | A | https://www.anthropic.com/engineering/multi-agent-research-system | 技术架构 |
| W10 | Harness engineering: leveraging Codex in an agent-first world | OpenAI | 2026-02-11 | A | https://openai.com/index/harness-engineering/ | 技术架构 / 发展展望 |
| W11 | Function calling and other API updates | OpenAI | 2023 | A | https://openai.com/blog/function-calling-and-other-API-updates | 发展历史相关章 |
| W12 | New tools for building agents(Responses API + Agents SDK) | OpenAI | 2025 | A | https://openai.com/blog/new-tools-for-building-agents | 技术架构 |
| W13 | Introducing Codex(Codex CLI) | OpenAI | 2025 | A | https://github.com/openai/codex | 技术架构 / 发展展望 |
| W14 | OpenAI co-founds the Agentic AI Foundation | OpenAI | 2025 | A | https://openai.com/index/agentic-ai-foundation/ | 发展展望 |
| W15 | Agent Development Kit: Making it easy to build multi-agent applications | 2025 | A | https://googledevelopers.blogspot.com/en/agent-development-kit-easy-to-build-multi-agent-applications/ | 技术架构 | |
| W16 | A year of open collaboration: Celebrating the anniversary of A2A | Google Open Source Blog | 2026-04-16 | A | https://opensource.googleblog.com/ | 发展展望 |
| W17 | Linux Foundation Announces the Formation of the AAIF | Linux Foundation | 2025-12-09 | A | https://aaif.io/press/linux-foundation-announces-the-formation-of-the-agentic-ai-foundation-aaif-anchored-by-new-project-contributions-including-model-context-protocol-mcp-goose-and-agents-md/ | 发展展望 |
| W18 | Linux Foundation Launches the Agent2Agent Protocol Project | Linux Foundation | 2025-06-23 | A | https://www.linuxfoundation.org/press/linux-foundation-launches-the-agent2agent-protocol-project-to-enable-secure-intelligent-communication-between-ai-agents | 发展展望 |
| W19 | AAIF 官网与新闻 | AAIF | 2025—2026 | A | https://aaif.io/ | 发展展望 |
| W20 | Cybersecurity updates: Summer 2025(Big Sleep、Timesketch + Sec-Gemini、FACADE) | 2025 | A | https://blog.google/technology/safety-security/cybersecurity-updates-summer-2025/ | 治理与风险 |
3. 标准与法规
3.1. 协议与开放标准
| 编号 | 名称 | 机构 | 年份 | 等级 | 链接 | 适用章节 |
|---|---|---|---|---|---|---|
| W21 | Model Context Protocol 官方站与规范(含 2026-07-28 版) | MCP / AAIF | 2024—2026 | A | https://modelcontextprotocol.io/ | 技术架构 / 发展展望 |
| W22 | MCP Protocol Versions(五版规范演进表) | MCP Ruby SDK | 2026 | A | https://ruby.sdk.modelcontextprotocol.io/protocol-versions/ | 技术架构 |
| W23 | Agent2Agent (A2A) 协议(开源仓库) | Linux Foundation / Google | 2025—2026 | A | https://github.com/a2aproject/A2A | 发展展望 |
| W24 | AGENTS.md 官方站 | AAIF | 2025—2026 | A | https://agents.md/ | 技术架构 / 行业赋能 |
| W25 | Agent Skills Specification | agentskills.io | 2025—2026 | A | https://agentskills.io/specification | 技术架构 |
3.2. 中国国家标准与指导性技术文件
| 编号 | 名称 | 机构 | 年份 | 等级 | 链接 | 适用章节 |
|---|---|---|---|---|---|---|
| W26 | 《人工智能 智能体互联》系列国家标准(GB/Z 185.1~185.7—2026)发布报道 | 人民网 | 2026-07-09 | A | https://finance-app.people.cn/n1/2026/0709/c1004-40757059.html | 发展展望 |
| W27 | 《人工智能 智能体互联》系列国家标准解读 | 中国产业经济信息网(归口单位解读) | 2026 | A | https://cinic.org.cn/xw/zcdt/1643418.html | 发展展望 |
| W28 | GB/Z 185—2026 落地:长三角智能体身份码节点首批发放 | 中国日报 | 2026-09-04 | A | https://cn.chinadaily.com.cn/a/202609/04/WS6a9a6773e4b09a165c788098.html | 发展展望 / 治理与风险 |
| W29 | GB/T 45654—2025《网络安全技术 生成式人工智能服务安全基本要求》专家解读 | 全国网络安全标准化技术委员会(SAC/TC260) | 2025 | A | https://www.tc260.org.cn/tc260/hygd1/202403/b429d868525e48c3b7d12a0ec8f82e5e.shtml | 治理与风险 |
| W30 | GB 45438—2025《网络安全技术 人工智能生成合成内容标识方法》(强制性国标,具体条文与文本链接未获取,标 ) | 市场监管总局、国家标准化管理委员会 | 2025 | B | 无(按文献名 + 机构 + 年份著录) | 治理与风险 |
3.3. 中国法律法规与监管文本
3.4. 国际治理框架与行业标准
| 编号 | 名称 | 机构 | 年份 | 等级 | 链接 | 适用章节 |
|---|---|---|---|---|---|---|
| W42 | NIST AI Risk Management Framework (AI RMF 1.0) | NIST | 2023 | A | https://www.nist.gov/itl/ai-risk-management-framework | 治理与风险 |
| W43 | NIST AI 600-1 Generative Artificial Intelligence Profile | NIST | 2024 | A | https://www.nist.gov/publications/artificial-intelligence-risk-management-framework-generative-artificial-intelligence | 治理与风险 |
| W44 | ISO/IEC 42001:2023(Information technology — Artificial intelligence — Management system) | ISO/IEC JTC 1/SC 42 | 2023 | A | https://www.iso.org/standard/42001 | 治理与风险 |
| W45 | EU AI Act(Regulation (EU) 2024/1689)Implementation Timeline | European Commission AI Act Service Desk | 2024—2026 | A | https://ai-act-service-desk.ec.europa.eu/en/ai-act/timeline/timeline-implementation-eu-ai-act | 治理与风险 |
| W46 | SR 11-7 汇编(Navigating Artificial Intelligence in Banking) | Bank Policy Institute | 2024 | B | https://bpi.com/wp-content/uploads/2024/04/Navigating-Artificial-Intelligence-in-Banking.pdf | 治理与风险 |
| W47 | HKMA《Supporting Adoption of Artificial Intelligence in Fighting Financial Crime》 | 香港金融管理局 | 2026-06-22 | A | https://brdr.hkma.gov.hk/eng/doc-ldg/current/20260622-1-EN | 治理与风险 |
| W48 | IAASB 全球技术质量管理圆桌会议反馈汇总 | IAASB | 2026 | A | https://www.iaasb.org/news-events/2026-02/iaasb-publishes-global-roundtable-feedback-technology-and-quality-management | 治理与风险 |
| W49 | IIA《Global Internal Audit Standards》 | The Institute of Internal Auditors | 2024 | A | https://www.theiia.org/en/standards/documents/ | 治理与风险 |
| W50 | OWASP Top 10 for LLM Applications(2025 版) | OWASP GenAI Security Project | 2025 | A | https://genai.owasp.org/llm-top-10/ | 治理与风险 |
| W51 | MITRE ATLAS | MITRE | 2023—2026 | A | https://atlas.mitre.org | 治理与风险 |
4. 学术论文
| 编号 | 名称 | 作者 / 机构 | 年份 | 等级 | 链接 | 适用章节 |
|---|---|---|---|---|---|---|
| W52 | ReAct: Synergizing Reasoning and Acting in Language Models | Yao 等(Princeton / Google Brain) | 2022(ICLR 2023) | A | https://arxiv.org/abs/2210.03629 | 发展历史相关章 |
| W53 | Toolformer: Language Models Can Teach Themselves to Use Tools | Schick 等(Meta AI) | 2023(NeurIPS 2023) | A | https://arxiv.org/abs/2302.04761 | 发展历史相关章 |
| W54 | SWE-bench: Can Language Models Resolve Real-World GitHub Issues? | Jimenez、Yang 等 | 2023(ICLR 2024 Oral) | A | https://arxiv.org/abs/2310.06770 | 技术架构 / 基准评测 |
| W55 | Gorilla: Large Language Model Connected with Massive APIs | UC Berkeley | 2023 | A | https://arxiv.org/abs/2305.15334 | 技术架构 |
| W56 | A Survey of Context Engineering for Large Language Models | Mei 等 | 2025 | B | https://arxiv.org/abs/2507.13334 | 技术架构 / 发展展望 |
| W57 | CyberSentinel-LLM(SOC 智能体信任度调研) | Tech Science Press(CMC vol.89 no.1) | 2025 | B | https://www.techscience.com/cmc/v89n1/68397/html | 治理与风险 |
5. 行业报告与调研
| 编号 | 名称 | 机构 | 年份 | 等级 | 链接 | 适用章节 |
|---|---|---|---|---|---|---|
| W58 | 2025 Stack Overflow Developer Survey | Stack Overflow | 2025-07-29 | A | https://survey.stackoverflow.co/2025/ | 发展展望 |
| W59 | Stack Overflow 2025 Developer Survey 官方新闻稿 | Stack Overflow | 2025 | A | https://stackoverflow.co/company/press/archive/stack-overflow-2025-developer-survey/ | 发展展望 |
| W60 | DORA 2025 State of AI-assisted Software Development | Google Cloud / DORA | 2025-09 | A | https://dora.dev/dora-report-2025 | 发展展望 |
| W61 | The State of AI 2025: Agents, Innovation, and Transformation | McKinsey | 2025-11-05 | A | https://www.mckinsey.com/ | 发展展望 |
| W62 | AI Adoption Stats & Trends(DORA / McKinsey / JetBrains 汇总) | daily.dev | 2026-07-19 | B | https://daily.dev/agentic-ai-hub/ai-adoption-stats-trends/ | 发展展望 |
| W63 | Cost of a Data Breach Report 2025 | IBM / Ponemon Institute | 2025-07-30 | A | https://newsroom.ibm.com/2025-07-30-ibm-report-13-of-organizations-reported-breaches-of-ai-models-or-applications,-97-of-which-reported-lacking-proper-ai-access-controls | 治理与风险 |
| W64 | DARPA AI Cyber Challenge 决赛结果 | DARPA | 2025 | A | https://www.darpa.mil/news/2025/aixcc-results | 治理与风险 / 技术架构 |
| W65 | Harness engineering for coding agent users | Birgitta Böckeler / martinfowler.com | 2026 | A | https://martinfowler.com/articles/harness-engineering.html | 技术架构 / 发展展望 |
| W66 | Humans and Agents in Software Engineering Loops | martinfowler.com | 2026 | A | https://martinfowler.com/articles/exploring-gen-ai/humans-and-agents.html | 发展展望 |
| W67 | My AI Adoption Journey | Mitchell Hashimoto | 2026-02-05 | A | https://mitchellh.com/writing/my-ai-adoption-journey | 技术架构 |
| W68 | 拥抱智能浪潮 泳向变革深处(新华社采编助手、芒果大模型、人民网语料库) | 新华社 | 2025 | A | https://www.news.cn/20251114/1c01598d836c449bbfd54267b6ecea6d/c.html | 行业赋能 |
| W69 | 主流媒体所办新媒体发展研究报告(2024-2025) | 人民网 | 2025 | A | https://sc.people.com.cn/BIG5/n2/2025/1030/c345167-41396739.html | 行业赋能 |
| W70 | AI 要给微短剧“洗牌”?(中国网络视听协会《微短剧创作指引》转引) | 人民日报 | 2026 | A | https://kpzg.people.com.cn/n1/2026/0511/c404214-40717117.html | 行业赋能 |
| W71 | Premium Times adopts Google NotebookLM to streamline newsroom processes | INMA | 2025 | A | https://www.inma.org/blogs/conference/post.cfm/premium-times-adopts-google-notebook-lm-to-streamline-newsroom-processes | 行业赋能 |
| W72 | 《技术不是侵权“挡箭牌” 法院这样认定 AI“盗脸”》(北京互联网法院 2026-03 生效判决报道) | 新华社《经济参考报》 | 2026-04-17 | B | http://dz.jjckb.cn/www/pages/webpage2009/html/2026-04/17/content_115180.htm | 治理与风险 |
| W73 | 《e案e审丨短剧角色 AI 换脸“神似”知名演员》 | 北京互联网法院供稿 / 澎湃新闻 | 2026 | A | https://www.thepaper.cn/newsDetail_forward_32799628 | 治理与风险 |
| W74 | JPMorgan Chase 2016 年年报(COiN 合同智能平台) | JPMorgan Chase | 2016 | A | https://reports.jpmorganchase.com/investor-relations/2016/ar-ceo-letter-matt-zames.htm | 行业赋能 |
| W75 | A&O Shearman 获 FT Innovative Lawyers Awards 官方新闻稿 | A&O Shearman(Allen Overy) | 2024-09-12 | A | https://www.allenovery.com/en/news/ao-shearman-wins-most-innovative-law-firm-in-europe-at-the-ft-innovative-lawyers-awards | 行业赋能 |
| W76 | 法信法律基座模型官方介绍 | 人民法院出版社 | 2024—2025 | A | https://www.faxin.cn/html/about/about.aspx | 行业赋能 |
| W77 | 中央网信办生成式 AI 备案用户规模报道(490+ 款 / 2.3 亿用户) | 央视网 | 2025-09 | B | https://big5.cctv.com/gate/big5/news.cctv.cn/2025/09/01/ARTI3ZlXK7MyM39Pm3PuZ5Hm250901.shtml | 行业赋能 |
| W78 | 2025 年上市银行 AI 应用综合报道(工银智涌、邮智等) | 新华网 | 2025-09-10 | B | https://www.xinhua.org/20250910/4e421fdeeb2242fc97299d73b2d62f09/c.html | 行业赋能 |
| W79 | 券商 2025 年报 IT 投入统计(34 家 / 275.9 亿元口径) | 证券时报 | 2026-04 | B | https://www.stcn.com/article/detail/3731822.html | 行业赋能 |
| W80 | 券商 AI 投研与合规实践(21 世纪经济报道专题) | 21 世纪经济报道 | 2026-04-27 | B | https://www.21jingji.com/article/20260427/herald/b0e4a124fc6808e55600d6fcd453da08.html | 行业赋能 |
| W80a | Claude Code Changelog(v2.1.257—269 官方版本记录,2026-09) | Anthropic 官方 | 2026-09-13 快照核实 | A | https://code.claude.com/docs/en/changelog | 发展展望 |
| W80b | OpenAI API Changelog(2026-09-10 条目:Agents API 公测;GPT-Live 1 GA;2026-09-03:GPT-6 Astra GA) | OpenAI 官方 | 2026-09-10 | A | https://developers.openai.com/api/docs/changelog(2026-09-13 快照修正:公测日期由媒体口径 09-11 修正为官方口径 09-10) | 产业格局 / 发展展望 |
| W80g | MCP 路线图(2026-08-22:智能体消息原语 / HTTP 原生传输统一 / 智能体身份) | MCP 官方 | 2026-08-22 | A | https://modelcontextprotocol.io/development/roadmap | 发展展望 |
| W80h | Cursor Projects(Beta)与 Self-Hosted Machines 更新日志 | Cursor 官方 | 2026-09-10 / 09-02 | A | https://cursor.com/changelog(能力描述经 Changelog 汇总站转述核对) | 发展展望 |
| W80i | OpenAI 研究组织智能体使用数据(3.1 智能体工作日/人工作日;中位研究员日耗超 600 美元) | OpenAI | 2026-09-06 | A | 官方口径;原始博文 URL [待核实] | 发展展望 |
| W80j | 超维动力(Kinetix AI)超 5 亿元天使+轮融资(祥峰资本领投) | 中国新闻网广东频道 / 新京报贝壳财经 | 2026-09-11 | B | https://www.gd.chinanews.com.cn/2026/2026-09-11/449585.shtml | 发展展望 |
| W80k | Accomplish 披露 Claude Code macOS 沙箱逃逸细节(07-13 上报,v2.1.247 修复) | Accomplish | 2026-09-11 | B | 安全厂商技术披露,原文 URL [待核实] | 治理与风险 |
| W80l | Gemini 3.8 Flash 发布(HLE-Verified 54.9%、DeepSWE v1.1 约 73.8%;$0.75/$3.75 每百万 token;同步发布 Flash Cyber) | Google 官方公告 / Info Data(il Sole 24 Ore) | 2026-09-02 / 09-12 | A / B | 官方原文 URL [待核实];细节核对 https://www.infodata.ilsole24ore.com/2026/09/12/come-e-fatto-gemini-3-8-flash | 发展展望 |
| W80m | NVIDIA 协议收购 Hugging Face(约 129.3 亿美元;2026-09-02 签署、09-03 宣布、预计 2027 上半年交割) | ECM Source(引 SEC 8-K)/ Inside AI | 2026-09-03 | A(SEC 文件转述) | https://ecmsource.com/nvidia-hugging-face-12-9-billion-acquisition-september-2026/(260914 快照升级:两源口径冲突关闭) | 发展展望 / 产业格局 |
| W80n | OpenAI 智能体滥用 RubyGems 事件复盘(2026-05 发生;RubyGems 冻结注册四天) | Syntackle(The Weekly Diff #10,第三方调查汇总) | 2026-09 上旬 | B | https://syntackle.com/blog/the-ai-slowdown-pact-nvidia-s-13b-hugging-face-deal-the-rubygems-attack-the-weekly-diff-10 | 治理与风险 |
| W80o | GitHub Copilot 2026-09 动态(HydraFusion、Jira 集成、代码评审 ensemble:+47%/+31%/-8%;模型退役 2026-10-02) | GitHub 官方 Changelog(经 AI Coding Roundup 核对) | 2026-09-07 / 09-11 | A / B(效果数字厂商自报) | https://oday-bakkour.com/blog/ai-coding-roundup-september-13-2026 | 发展展望 |
| W80p | MCP ToolAnnotations 默认值分析(未标注工具默认视为可破坏;7 个记忆服务器普查) | deniz.in(第三方分析);MCP 2026-07-28 schema 原文 A 级 | 2026-09-12 | B | https://deniz.in/mcp-s-annotation-defaults-make-unannotated-tools-implicitly-destructive | 治理与风险 |
| W80q | AWS Kiro 学生计划(18 国 132 所高校免费一年、每月 1,000 credits) | AWS 官方新闻 | 2026-09 上中旬(具体日期 [待核实]) | A | https://www.aboutamazon.com/news/aws/amazon-expands-free-kiro-access-students-universities | 发展展望 |
| W80c | Kiro Web 可用性与计费口径(官方 FAQ 确认;GA 媒体口径 2026-09-01) | AWS / Kiro 官方 | 2026-09-14 快照核实 | A | https://kiro.dev/faq(GA 具体日期 [待核实]) | 发展展望 |
| W80d | Figure AI 与 Nscale 算力合作(初始承诺 35 亿美元、规划至多 10 万颗 Vera Rubin GPU,Nscale 战略投资 Figure) | 具身智能应用与投融资周报(2026.9.2—9.8) | 2026-09 上旬 | B | 行业周报佐证,官方公告未见,[待核实] 保留;原文 URL [待核实] | 发展展望 |
| W80f | 云蝶科技 RoboForge 系统级具身大脑(Harness 模块记录任务全流程证据链;与南洋理工大学联合研发;基准登顶为厂商自报) | 南方+ / 云蝶科技官网 | 2026-09-03—06 | B + C | https://www.nfnews.com/content/lyKmjgJ765.html;http://www.cloudbutterfly.com.cn/newsshow/post-227.html(更名更正:原条目误记「云蝶数据」) | 发展展望 |
6. 开源项目
| 编号 | 名称 | 机构 / 作者 | 等级 | 链接 | 适用章节 |
|---|---|---|---|---|---|
| W81 | Model Context Protocol 规范与 SDK(GitHub 组织) | MCP / AAIF | A | https://github.com/modelcontextprotocol | 技术架构 |
| W82 | OpenAI Codex CLI | OpenAI | A | https://github.com/openai/codex | 技术架构 / 发展展望 |
| W83 | Claude Agent SDK | Anthropic | A | https://github.com/anthropics/claude-agent-sdk-python | 技术架构 |
| W84 | Google Agent Development Kit (ADK) | A | https://github.com/google/adk-python | 技术架构 | |
| W85 | 12-Factor Agents | HumanLayer(Dex Horthy) | A | https://github.com/humanlayer/12-factor-agents | 技术架构 / 发展展望 |
| W86 | Terminal-Bench / Harbor 框架 | Stanford / Laude Institute / 社区 | A | https://github.com/harbor-framework/terminal-bench | 基准评测 |
| W87 | @anthropic-ai/sandbox-runtime | Anthropic | A | https://www.npmjs.com/package/@anthropic-ai/sandbox-runtime | 治理与风险 |
| W88 | ComfyUI(开源节点式生成工作流) | Comfy Org | A | https://github.com/comfyanonymous/ComfyUI | 行业赋能 |
7. 基准评测
| 编号 | 名称 | 机构 | 年份 | 等级 | 链接 | 适用章节 |
|---|---|---|---|---|---|---|
| W89 | SWE-bench 官方站与排行榜 | Princeton / 社区 | 2023—2026 | A | https://www.swebench.com/ | 技术架构 / 发展展望 |
| W90 | Terminal-Bench(官方站) | Stanford / Laude Institute | 2025—2026 | A | https://www.tbench.ai/ | 技术架构 / 发展展望 |
| W91 | SWE-bench Verified 进展时间线 2023—2026(含高冲击量化数字,须标 ) | AgentMarketCap | 2026-04-09 | C | https://agentmarketcap.ai/blog/2026/04/09/swe-bench-verified-progress-timeline-2023-2026 | 技术架构 |
| W92 | Terminal-Bench: The CLI Autonomy Standard(C 级,数字须标 ) | AgentMarketCap | 2026-04-09 | C | https://agentmarketcap.ai/blog/2026/04/09/terminal-bench-cli-autonomy-standard-coding-agents | 技术架构 |
| W93 | Artificial Analysis(图像 / 视频生成模型第三方评测榜单) | Artificial Analysis | 2025—2026 | B | https://artificialanalysis.ai/ | 市场研究 |
8. 平台官方文档
9. 本白皮书引用但未获权威确认的条目清单
以下条目在本工程检索中未能取得 A/B 级一手来源、或存在来源间冲突。白皮书正文引用时均已就地标注 或并列呈现,此处集中汇总,供后续修订与二次取证使用。
| 序号 | 条目 | 问题描述 | 影响章节 |
|---|
| 1 | GB/Z 185—2026 发布日期 | 2026-05-22(中国日报)/ 2026-06-26(百度百科)/ 2026-07-09(人民网“近日发布”)三种口径并存,需以官方正式公告为准 | 发展展望 |
|---|
| 4 | GB/T 45654—2025 “附录列明 31 类风险”“生成合规内容合格率不低于 90%” | 二级解读,未查证标准原文,本白皮书未采用 | 治理与风险 |
|---|
| 7 | EU AI Act 时间线的 AI Omnibus 调整(2026-05 临时协议) | 中文二手来源,按“拟议”处理,以欧盟官方服务台为准 | 治理与风险 |
|---|
| 17 | 上下文退化的量化数字(“性能下降超 45%”等) | C 级来源,未获 A/B 级确认 | 发展展望 |
|---|
信息缺口声明
除上表 25 项外,另有两点整体性缺口需向读者说明:
- 时间截面:本参考资料全部素材截至 2026-09-14 快照(2026-09-12/13/14 三轮增量检索已并入 W80a—W80q,信息截止 2026-09-13)。智能体领域的法规、标准与产品更新频繁,正式引用前应回查各来源的最新版本。
- 检索范围:本白皮书采用“基于既有资料”模式,素材来自本工程既有的调研文档与检索报告(覆盖概述、行业赋能八大组、市场研究七组的全部已生成文档),未新增联网检索;因此本清单只汇总既有检索已识别的缺口,不排除存在未识别的缺口。
References
1. Instructions
图 1-1|参考资料体系:七大来源类 × A/B/C 可信度分级
数据来源:基于本文分析绘制的示意图。
1.1. Credibility rating criteria
All materials included in this document are rated on a three-level credibility scale. This scale is inherited from the evidence-grading system (A / B / C) of this project's retrieval report, and is provided for readers to judge how an item may be cited.
| Level | Meaning | Usage rule |
|---|---|---|
| A | Official first-hand sources from regulators, standards bodies, and enterprises (gov.cn, nfra.gov.cn, cicpa.org.cn, tc260.org.cn, nist.gov, iso.org, darpa.mil, anthropic.com, openai.com, aaif.io, 人民网, 中国日报, arXiv originals, official open-source repositories, etc.) | May be cited directly; the source must be given |
| B | Authoritative second-hand sources: relays by mainstream and industry media, encyclopedias and surveys, reprints by aggregation sites | Must note “as reported / relayed by XX”; specific figures recommended for secondary verification |
| C | Vendor self-reports, aggregation sites, and community or self-media interpretations | For leads only; specific figures must uniformly be marked [To be verified] before entering the main text |
1.2. Numbering and citation rules
- Numbers start in the form of
W01, numbered consecutively by category order; the same item is never numbered twice; - Items marked
[To be verified]have links or dates not yet first-hand verified and require second confirmation when cited; - Items without links are catalogued as “document title + issuing organization + year” without fabricating a URL;
- “Applicable sections” is annotated by whitepaper topic domain (e.g., “Governance and Risk”, “Outlook”, “Technical Architecture”, “Industry Enablement”, “Market Research”, “Entire Document”), for convenient on-demand use.
2. Official Engineering Blogs
| No. | Name | Organization | Year | Level | Link | Applicable sections |
|---|---|---|---|---|---|---|
| W01 | Effective context engineering for AI agents | Anthropic | 2025 | A | https://www.anthropic.com/engineering/effective-context-engineering-for-ai-agents | Technical Architecture / Outlook |
| W02 | Effective harnesses for long-running agents | Anthropic | 2025 | A | https://www.anthropic.com/engineering/effective-harnesses-for-long-running-agents | Technical Architecture |
| W03 | Harness design for long-running application development | Anthropic | 2026 | A | https://www.anthropic.com/engineering/harness-design-long-running-apps | Technical Architecture / Outlook |
| W04 | Equipping agents for the real world with Agent Skills | Anthropic | 2025 | A | https://www.anthropic.com/engineering/equipping-agents-for-the-real-world-with-agent-skills | Technical Architecture |
| W05 | Sandboxing: a safer and more autonomous approach (permission prompts reduced by 84%) | Anthropic | 2025 | A | https://www.anthropic.com/engineering/claude-code-sandboxing | Governance and Risk |
| W06 | Introducing Agent Skills (became an open standard on 2025-12-18) | Anthropic | 2025 | A | https://www.anthropic.com/news/skills | Technical Architecture |
| W07 | Introducing the Model Context Protocol | Anthropic | 2024 | A | https://www.anthropic.com/news/model-context-protocol | Technical Architecture |
| W08 | Donating the MCP and establishing the AAIF | Anthropic | 2025 | A | https://www.anthropic.com/news/donating-the-model-context-protocol-and-establishing-of-the-agentic-ai-foundation | Outlook |
| W09 | How we built our multi-agent research system | Anthropic | 2025 | A | https://www.anthropic.com/engineering/multi-agent-research-system | Technical Architecture |
| W10 | Harness engineering: leveraging Codex in an agent-first world | OpenAI | 2026-02-11 | A | https://openai.com/index/harness-engineering/ | Technical Architecture / Outlook |
| W11 | Function calling and other API updates | OpenAI | 2023 | A | https://openai.com/blog/function-calling-and-other-API-updates | Related sections on development history |
| W12 | New tools for building agents (Responses API + Agents SDK) | OpenAI | 2025 | A | https://openai.com/blog/new-tools-for-building-agents | Technical Architecture |
| W13 | Introducing Codex (Codex CLI) | OpenAI | 2025 | A | https://github.com/openai/codex | Technical Architecture / Outlook |
| W14 | OpenAI co-founds the Agentic AI Foundation | OpenAI | 2025 | A | https://openai.com/index/agentic-ai-foundation/ | Outlook |
| W15 | Agent Development Kit: Making it easy to build multi-agent applications | 2025 | A | https://googledevelopers.blogspot.com/en/agent-development-kit-easy-to-build-multi-agent-applications/ | Technical Architecture | |
| W16 | A year of open collaboration: Celebrating the anniversary of A2A | Google Open Source Blog | 2026-04-16 | A | https://opensource.googleblog.com/ | Outlook |
| W17 | Linux Foundation Announces the Formation of the AAIF | Linux Foundation | 2025-12-09 | A | https://aaif.io/press/linux-foundation-announces-the-formation-of-the-agentic-ai-foundation-aaif-anchored-by-new-project-contributions-including-model-context-protocol-mcp-goose-and-agents-md/ | Outlook |
| W18 | Linux Foundation Launches the Agent2Agent Protocol Project | Linux Foundation | 2025-06-23 | A | https://www.linuxfoundation.org/press/linux-foundation-launches-the-agent2agent-protocol-project-to-enable-secure-intelligent-communication-between-ai-agents | Outlook |
| W19 | AAIF official website and news | AAIF | 2025—2026 | A | https://aaif.io/ | Outlook |
| W20 | Cybersecurity updates: Summer 2025 (Big Sleep, Timesketch + Sec-Gemini, FACADE) | 2025 | A | https://blog.google/technology/safety-security/cybersecurity-updates-summer-2025/ | Governance and Risk |
3. Standards and Regulations
3.1. Protocols and Open Standards
| No. | Name | Organization | Year | Level | Link | Applicable sections |
|---|---|---|---|---|---|---|
| W21 | Model Context Protocol official site and specification (including the 2026-07-28 version) | MCP / AAIF | 2024—2026 | A | https://modelcontextprotocol.io/ | Technical Architecture / Outlook |
| W22 | MCP Protocol Versions (evolution table of the five specification versions) | MCP Ruby SDK | 2026 | A | https://ruby.sdk.modelcontextprotocol.io/protocol-versions/ | Technical Architecture |
| W23 | Agent2Agent (A2A) protocol (open-source repository) | Linux Foundation / Google | 2025—2026 | A | https://github.com/a2aproject/A2A | Outlook |
| W24 | AGENTS.md official site | AAIF | 2025—2026 | A | https://agents.md/ | Technical Architecture / Industry Enablement |
| W25 | Agent Skills Specification | agentskills.io | 2025—2026 | A | https://agentskills.io/specification | Technical Architecture |
3.2. Chinese National Standards and Guiding Technical Documents
| No. | Name | Organization | Year | Level | Link | Applicable sections |
|---|---|---|---|---|---|---|
| W26 | Release report on the 《人工智能 智能体互联》 series of national standards (GB/Z 185.1~185.7—2026) | 人民网 | 2026-07-09 | A | https://finance-app.people.cn/n1/2026/0709/c1004-40757059.html | Outlook |
| W27 | Interpretation of the 《人工智能 智能体互联》 series of national standards | 中国产业经济信息网 (interpretation by the responsible body) | 2026 | A | https://cinic.org.cn/xw/zcdt/1643418.html | Outlook |
| W28 | GB/Z 185—2026 takes effect: first batch of agent identity code nodes issued in the Yangtze River Delta | 中国日报 | 2026-09-04 | A | https://cn.chinadaily.com.cn/a/202609/04/WS6a9a6773e4b09a165c788098.html | Outlook / Governance and Risk |
| W29 | Expert interpretation of GB/T 45654—2025 《网络安全技术 生成式人工智能服务安全基本要求》 | 全国网络安全标准化技术委员会 (SAC/TC260) | 2025 | A | https://www.tc260.org.cn/tc260/hygd1/202403/b429d868525e48c3b7d12a0ec8f82e5e.shtml | Governance and Risk |
| W30 | GB 45438—2025 《网络安全技术 人工智能生成合成内容标识方法》 (mandatory national standard; specific clauses and text link not obtained, marked [To be verified]) | 市场监管总局、国家标准化管理委员会 | 2025 | B | None (catalogued as document title + organization + year) | Governance and Risk |
3.3. Chinese Laws, Regulations, and Regulatory Texts
| No. | Name | Organization | Year | Level | Link | Applicable sections |
|---|---|---|---|---|---|---|
| W31 | 《人工智能生成合成内容标识办法》 (国信办通字〔2025〕2 号) | 国家网信办、工信部、公安部、国家广播电视总局 | 2025 | A | https://www.cac.gov.cn/2025-03/14/c_1743654684782215.htm | Governance and Risk / Industry Enablement |
| W32 | Interpretation of the 《人工智能生成合成内容标识办法》 | 中国政府网 / 新华社 | 2025 | A | https://www.gov.cn/zhengce/202503/content_7014281.htm | Governance and Risk |
| W33 | Multi-pronged efforts to advance the construction of the labeling system | 国家互联网应急中心 | 2025 | A | https://www.cac.gov.cn/2025-09/06/c_1758880709361356.htm | Governance and Risk |
| W34 | 《生成式人工智能服务管理暂行办法》 (Order No. 15 of seven departments) | 国家网信办等七部门 | 2023 | A | https://www.cac.gov.cn/2023-07/13/c_1690898327029107.htm | Governance and Risk |
| W35 | Decision to amend the 《中华人民共和国网络安全法》 (Presidential Order No. 61, adding Article 20) | 全国人民代表大会 | 2025 | A | http://www.npc.gov.cn/c2/c30834/202601/t20260105_450980.html | Governance and Risk |
| W36 | Expert interpretation of the Cybersecurity Law amendment decision | 中央网信办 (reprinted by 中国网信网) | 2026 | A | https://www.cac.gov.cn/2026-01/02/c_1769093523928606.htm | Governance and Risk |
| W37 | 《银行业保险业数字金融高质量发展实施方案》 | 国家金融监督管理总局办公厅 | 2025-12 | A | https://www.nfra.gov.cn/cn/view/pages/ItemDetail.html?docId=1239741 | Governance and Risk / Industry Enablement |
| W38 | 中注协 reminder on risk prevention when accounting firms use AI technology in 2025 annual audits | 中国注册会计师协会 | 2026-03-05 | A | https://cicpa.org.cn/xxfb/news/202603/t20260305_65842.html | Governance and Risk |
| W39 | 《中国注册会计师审计准则问题解答第 18 号 / 第 19 号》 | 中国注册会计师协会 | 2025-01-07 | A | https://cicpa.org.cn/xxfb/news/202501/t20250123_65229.html | Governance and Risk |
| W40 | Report on the identical adoption of ISO 37301:2021 as GB/T 35770—2022 | 中国标准化研究院 | 2022 | A | https://www.cnis.ac.cn/bydt/zhxw/202210/t20221020_54060.html | Governance and Risk |
| W41 | Introduction to smart manufacturing capability maturity assessment (GB/T 39116-2020, GB/T 39117-2020) | 中国电子技术标准化研究院 | — | A | https://www.cc.cesi.cn/service/show-2478.aspx | Industry Enablement |
3.4. International Governance Frameworks and Industry Standards
| No. | Name | Organization | Year | Level | Link | Applicable sections |
|---|---|---|---|---|---|---|
| W42 | NIST AI Risk Management Framework (AI RMF 1.0) | NIST | 2023 | A | https://www.nist.gov/itl/ai-risk-management-framework | Governance and Risk |
| W43 | NIST AI 600-1 Generative Artificial Intelligence Profile | NIST | 2024 | A | https://www.nist.gov/publications/artificial-intelligence-risk-management-framework-generative-artificial-intelligence | Governance and Risk |
| W44 | ISO/IEC 42001:2023 (Information technology — Artificial intelligence — Management system) | ISO/IEC JTC 1/SC 42 | 2023 | A | https://www.iso.org/standard/42001 | Governance and Risk |
| W45 | EU AI Act (Regulation (EU) 2024/1689) Implementation Timeline | European Commission AI Act Service Desk | 2024—2026 | A | https://ai-act-service-desk.ec.europa.eu/en/ai-act/timeline/timeline-implementation-eu-ai-act | Governance and Risk |
| W46 | SR 11-7 compilation (Navigating Artificial Intelligence in Banking) | Bank Policy Institute | 2024 | B | https://bpi.com/wp-content/uploads/2024/04/Navigating-Artificial-Intelligence-in-Banking.pdf | Governance and Risk |
| W47 | HKMA 《Supporting Adoption of Artificial Intelligence in Fighting Financial Crime》 | 香港金融管理局 | 2026-06-22 | A | https://brdr.hkma.gov.hk/eng/doc-ldg/current/20260622-1-EN | Governance and Risk |
| W48 | IAASB global roundtable feedback summary on technology and quality management | IAASB | 2026 | A | https://www.iaasb.org/news-events/2026-02/iaasb-publishes-global-roundtable-feedback-technology-and-quality-management | Governance and Risk |
| W49 | IIA 《Global Internal Audit Standards》 | The Institute of Internal Auditors | 2024 | A | https://www.theiia.org/en/standards/documents/ | Governance and Risk |
| W50 | OWASP Top 10 for LLM Applications (2025 edition) | OWASP GenAI Security Project | 2025 | A | https://genai.owasp.org/llm-top-10/ | Governance and Risk |
| W51 | MITRE ATLAS | MITRE | 2023—2026 | A | https://atlas.mitre.org | Governance and Risk |
4. Academic Papers
| No. | Name | Author / Organization | Year | Level | Link | Applicable sections |
|---|---|---|---|---|---|---|
| W52 | ReAct: Synergizing Reasoning and Acting in Language Models | Yao et al. (Princeton / Google Brain) | 2022 (ICLR 2023) | A | https://arxiv.org/abs/2210.03629 | Related sections on development history |
| W53 | Toolformer: Language Models Can Teach Themselves to Use Tools | Schick et al. (Meta AI) | 2023 (NeurIPS 2023) | A | https://arxiv.org/abs/2302.04761 | Related sections on development history |
| W54 | SWE-bench: Can Language Models Resolve Real-World GitHub Issues? | Jimenez, Yang et al. | 2023 (ICLR 2024 Oral) | A | https://arxiv.org/abs/2310.06770 | Technical Architecture / Benchmarks |
| W55 | Gorilla: Large Language Model Connected with Massive APIs | UC Berkeley | 2023 | A | https://arxiv.org/abs/2305.15334 | Technical Architecture |
| W56 | A Survey of Context Engineering for Large Language Models | Mei et al. | 2025 | B | https://arxiv.org/abs/2507.13334 | Technical Architecture / Outlook |
| W57 | CyberSentinel-LLM (survey on trust in SOC agents) | Tech Science Press (CMC vol.89 no.1) | 2025 | B | https://www.techscience.com/cmc/v89n1/68397/html | Governance and Risk |
5. Industry Reports and Surveys
| No. | Name | Organization | Year | Level | Link | Applicable sections |
|---|---|---|---|---|---|---|
| W58 | 2025 Stack Overflow Developer Survey | Stack Overflow | 2025-07-29 | A | https://survey.stackoverflow.co/2025/ | Outlook |
| W59 | Stack Overflow 2025 Developer Survey official press release | Stack Overflow | 2025 | A | https://stackoverflow.co/company/press/archive/stack-overflow-2025-developer-survey/ | Outlook |
| W60 | DORA 2025 State of AI-assisted Software Development | Google Cloud / DORA | 2025-09 | A | https://dora.dev/dora-report-2025 | Outlook |
| W61 | The State of AI 2025: Agents, Innovation, and Transformation | McKinsey | 2025-11-05 | A | https://www.mckinsey.com/ | Outlook |
| W62 | AI Adoption Stats & Trends (aggregation of DORA / McKinsey / JetBrains) | daily.dev | 2026-07-19 | B | https://daily.dev/agentic-ai-hub/ai-adoption-stats-trends/ | Outlook |
| W63 | Cost of a Data Breach Report 2025 | IBM / Ponemon Institute | 2025-07-30 | A | https://newsroom.ibm.com/2025-07-30-ibm-report-13-of-organizations-reported-breaches-of-ai-models-or-applications,-97-of-which-reported-lacking-proper-ai-access-controls | Governance and Risk |
| W64 | DARPA AI Cyber Challenge final results | DARPA | 2025 | A | https://www.darpa.mil/news/2025/aixcc-results | Governance and Risk / Technical Architecture |
| W65 | Harness engineering for coding agent users | Birgitta Böckeler / martinfowler.com | 2026 | A | https://martinfowler.com/articles/harness-engineering.html | Technical Architecture / Outlook |
| W66 | Humans and Agents in Software Engineering Loops | martinfowler.com | 2026 | A | https://martinfowler.com/articles/exploring-gen-ai/humans-and-agents.html | Outlook |
| W67 | My AI Adoption Journey | Mitchell Hashimoto | 2026-02-05 | A | https://mitchellh.com/writing/my-ai-adoption-journey | Technical Architecture |
| W68 | Embracing the wave of intelligence, swimming deep into transformation (新华社 editing assistant, 芒果 large model, 人民网 corpus) | 新华社 | 2025 | A | https://www.news.cn/20251114/1c01598d836c449bbfd54267b6ecea6d/c.html | Industry Enablement |
| W69 | Research report on the development of new media run by mainstream media (2024—2025) | 人民网 | 2025 | A | https://sc.people.com.cn/BIG5/n2/2025/1030/c345167-41396739.html | Industry Enablement |
| W70 | Will AI give micro-short dramas a “shuffle”? (citing 中国网络视听协会 《微短剧创作指引》) | 人民日报 | 2026 | A | https://kpzg.people.com.cn/n1/2026/0511/c404214-40717117.html | Industry Enablement |
| W71 | Premium Times adopts Google NotebookLM to streamline newsroom processes | INMA | 2025 | A | https://www.inma.org/blogs/conference/post.cfm/premium-times-adopts-google-notebook-lm-to-streamline-newsroom-processes | Industry Enablement |
| W72 | 《技术不是侵权“挡箭牌” 法院这样认定 AI“盗脸”》 (report on the 北京互联网法院 judgment taking effect in 2026-03) | 新华社《经济参考报》 | 2026-04-17 | B | http://dz.jjckb.cn/www/pages/webpage2009/html/2026-04/17/content_115180.htm | Governance and Risk |
| W73 | 《e案e审丨短剧角色 AI 换脸“神似”知名演员》 | 北京互联网法院 (contributed) / 澎湃新闻 | 2026 | A | https://www.thepaper.cn/newsDetail_forward_32799628 | Governance and Risk |
| W74 | JPMorgan Chase 2016 annual report (COiN contract intelligence platform) | JPMorgan Chase | 2016 | A | https://reports.jpmorganchase.com/investor-relations/2016/ar-ceo-letter-matt-zames.htm | Industry Enablement |
| W75 | A&O Shearman wins FT Innovative Lawyers Awards — official press release | A&O Shearman (Allen Overy) | 2024-09-12 | A | https://www.allenovery.com/en/news/ao-shearman-wins-most-innovative-law-firm-in-europe-at-the-ft-innovative-lawyers-awards | Industry Enablement |
| W76 | 法信 legal foundation model — official introduction | 人民法院出版社 | 2024—2025 | A | https://www.faxin.cn/html/about/about.aspx | Industry Enablement |
| W77 | 中央网信办 report on the scale of generative AI filing users (490+ products / 2.3 亿 users) | 央视网 | 2025-09 | B | https://big5.cctv.com/gate/big5/news.cctv.cn/2025/09/01/ARTI3ZlXK7MyM39Pm3PuZ5Hm250901.shtml | Industry Enablement |
| W78 | 2025 comprehensive report on AI applications at listed banks (工银智涌, 邮智, etc.) | 新华网 | 2025-09-10 | B | https://www.xinhua.org/20250910/4e421fdeeb2242fc97299d73b2d62f09/c.html | Industry Enablement |
| W79 | Securities firms' 2025 annual-report IT investment statistics (34-firm / 275.9 亿元 account) | 证券时报 | 2026-04 | B | https://www.stcn.com/article/detail/3731822.html | Industry Enablement |
| W80 | Securities firms' AI investment research and compliance practices (21 世纪经济报道 feature) | 21 世纪经济报道 | 2026-04-27 | B | https://www.21jingji.com/article/20260427/herald/b0e4a124fc6808e55600d6fcd453da08.html | Industry Enablement |
| W80a | Claude Code Changelog (v2.1.257—269 official version record, 2026-09) | Anthropic (official) | verified against the 2026-09-13 snapshot | A | https://code.claude.com/docs/en/changelog | Outlook |
| W80b | OpenAI API Changelog (2026-09-10 entry: Agents API public beta; GPT-Live 1 GA; 2026-09-03: GPT-6 Astra GA) | OpenAI (official) | 2026-09-10 | A | https://developers.openai.com/api/docs/changelog (2026-09-13 snapshot correction: public beta date corrected from the media-reported 09-11 to the official 09-10) | Industry Landscape / Outlook |
| W80g | MCP roadmap (2026-08-22: agent message primitives / HTTP-native transport unification / agent identity) | MCP (official) | 2026-08-22 | A | https://modelcontextprotocol.io/development/roadmap | Outlook |
| W80h | Cursor Projects (Beta) and Self-Hosted Machines update log | Cursor (official) | 2026-09-10 / 09-02 | A | https://cursor.com/changelog (capability descriptions verified against Changelog aggregation-site relays) | Outlook |
| W80i | OpenAI research organization agent usage data (3.1 agent workdays per human workday; median researcher spend over USD 600 per day) | OpenAI | 2026-09-06 | A | Official figures; original blog post URL [To be verified] | Outlook |
| W80j | 超维动力 (Kinetix AI) raises over 5 亿元 in an Angel+ round (led by 祥峰资本) | 中国新闻网 Guangdong channel / 新京报贝壳财经 | 2026-09-11 | B | https://www.gd.chinanews.com.cn/2026/2026-09-11/449585.shtml | Outlook |
| W80k | Accomplish discloses details of a Claude Code macOS sandbox escape (reported 07-13, fixed in v2.1.247) | Accomplish | 2026-09-11 | B | Security vendor technical disclosure; original URL [To be verified] | Governance and Risk |
| W80l | Gemini 3.8 Flash released (HLE-Verified 54.9%, DeepSWE v1.1 approx. 73.8%; $0.75/$3.75 per million tokens; Flash Cyber released simultaneously) | Google (official announcement) / Info Data (il Sole 24 Ore) | 2026-09-02 / 09-12 | A / B | Official original URL [To be verified]; details checked against https://www.infodata.ilsole24ore.com/2026/09/12/come-e-fatto-gemini-3-8-flash | Outlook |
| W80m | NVIDIA agrees to acquire Hugging Face (approx. 129.3 亿美元; signed 2026-09-02, announced 09-03, expected to close in H1 2027) | ECM Source (citing SEC 8-K) / Inside AI | 2026-09-03 | A (SEC document relay) | https://ecmsource.com/nvidia-hugging-face-12-9-billion-acquisition-september-2026/ (260914 snapshot upgrade: conflict between the two sources closed) | Outlook / Industry Landscape |
| W80n | Post-mortem of the OpenAI agent RubyGems abuse incident (occurred 2026-05; RubyGems froze signups for four days) | Syntackle (The Weekly Diff #10, third-party investigation summary) | early 2026-09 | B | https://syntackle.com/blog/the-ai-slowdown-pact-nvidia-s-13b-hugging-face-deal-the-rubygems-attack-the-weekly-diff-10 | Governance and Risk |
| W80o | GitHub Copilot 2026-09 updates (HydraFusion, Jira integration, code review ensemble: +47%/+31%/-8%; model retirement on 2026-10-02) | GitHub official Changelog (verified against AI Coding Roundup) | 2026-09-07 / 09-11 | A / B (effectiveness figures self-reported by the vendor) | https://oday-bakkour.com/blog/ai-coding-roundup-september-13-2026 | Outlook |
| W80p | Analysis of MCP ToolAnnotations default values (unannotated tools are treated as destructive by default; census of 7 memory servers) | deniz.in (third-party analysis); the MCP 2026-07-28 schema original is Level A | 2026-09-12 | B | https://deniz.in/mcp-s-annotation-defaults-make-unannotated-tools-implicitly-destructive | Governance and Risk |
| W80q | AWS Kiro student program (free for one year at 132 universities in 18 countries, 1,000 credits per month) | AWS (official news) | early-to-mid 2026-09 (exact date [To be verified]) | A | https://www.aboutamazon.com/news/aws/amazon-expands-free-kiro-access-students-universities | Outlook |
| W80c | Kiro Web availability and billing basis (confirmed by official FAQ; media-reported GA 2026-09-01) | AWS / Kiro (official) | verified against the 2026-09-14 snapshot | A | https://kiro.dev/faq (exact GA date [To be verified]) | Outlook |
| W80d | Figure AI and Nscale compute partnership (initial commitment of 35 亿美元, plans for up to 100,000 Vera Rubin GPUs; Nscale makes a strategic investment in Figure) | Embodied AI Applications and Investment Weekly (2026.9.2—9.8) | early 2026-09 | B | Supported by an industry weekly; no official announcement found, [To be verified] retained; original URL [To be verified] | Outlook |
| W80f | 云蝶科技 RoboForge system-level embodied brain (the Harness module records the evidence chain of the entire task process; co-developed with 南洋理工大学; topping the benchmark is vendor-self-reported) | 南方+ / 云蝶科技 official website | 2026-09-03—06 | B + C | https://www.nfnews.com/content/lyKmjgJ765.html; http://www.cloudbutterfly.com.cn/newsshow/post-227.html (name correction: the original entry misrecorded “云蝶数据”) | Outlook |
6. Open-Source Projects
| No. | Name | Organization / Author | Level | Link | Applicable sections |
|---|---|---|---|---|---|
| W81 | Model Context Protocol specification and SDK (GitHub organization) | MCP / AAIF | A | https://github.com/modelcontextprotocol | Technical Architecture |
| W82 | OpenAI Codex CLI | OpenAI | A | https://github.com/openai/codex | Technical Architecture / Outlook |
| W83 | Claude Agent SDK | Anthropic | A | https://github.com/anthropics/claude-agent-sdk-python | Technical Architecture |
| W84 | Google Agent Development Kit (ADK) | A | https://github.com/google/adk-python | Technical Architecture | |
| W85 | 12-Factor Agents | HumanLayer(Dex Horthy) | A | https://github.com/humanlayer/12-factor-agents | Technical Architecture / Outlook |
| W86 | Terminal-Bench / Harbor framework | Stanford / Laude Institute / community | A | https://github.com/harbor-framework/terminal-bench | Benchmarks |
| W87 | @anthropic-ai/sandbox-runtime | Anthropic | A | https://www.npmjs.com/package/@anthropic-ai/sandbox-runtime | Governance and Risk |
| W88 | ComfyUI (open-source node-based generation workflow) | Comfy Org | A | https://github.com/comfyanonymous/ComfyUI | Industry Enablement |
7. Benchmark Evaluations
| No. | Name | Organization | Year | Level | Link | Applicable sections |
|---|---|---|---|---|---|---|
| W89 | SWE-bench official site and leaderboard | Princeton / community | 2023—2026 | A | https://www.swebench.com/ | Technical Architecture / Outlook |
| W90 | Terminal-Bench (official site) | Stanford / Laude Institute | 2025—2026 | A | https://www.tbench.ai/ | Technical Architecture / Outlook |
| W91 | SWE-bench Verified progress timeline 2023—2026 (contains high-impact quantitative figures; must be marked [To be verified]) | AgentMarketCap | 2026-04-09 | C | https://agentmarketcap.ai/blog/2026/04/09/swe-bench-verified-progress-timeline-2023-2026 | Technical Architecture |
| W92 | Terminal-Bench: The CLI Autonomy Standard (Level C; figures must be marked [To be verified]) | AgentMarketCap | 2026-04-09 | C | https://agentmarketcap.ai/blog/2026/04/09/terminal-bench-cli-autonomy-standard-coding-agents | Technical Architecture |
| W93 | Artificial Analysis (third-party evaluation leaderboard for image / video generation models) | Artificial Analysis | 2025—2026 | B | https://artificialanalysis.ai/ | Market Research |
8. Official Platform Documentation
| No. | Name | Organization | Year | Level | Link | Applicable sections |
|---|---|---|---|---|---|---|
| W94 | Claude Code official documentation · Sandboxing | Anthropic | 2026 | A | https://code.claude.com/docs/zh-TW/sandboxing | Governance and Risk |
| W95 | Claude Code official documentation · Choose a sandbox environment | Anthropic | 2026 | A | https://code.claude.com/docs/en/sandbox-environments | Governance and Risk |
| W96 | 火山引擎 Ark console — Doubao-Seedance-2.0 series | 字节跳动 / 火山引擎 | 2026 | A | https://console.volcengine.com/ark/region:ark+cn-beijing/model/detail?Id=doubao-seedance-2-0 | Market Research |
| W97 | 即梦 AI official website | 字节跳动 (剪映 team) | 2024—2026 | A | https://jimeng.jianying.com/ | Market Research |
| W98 | MCP specification evolution entry (docs.rs mcpkit ProtocolVersion) | mcpkit | 2026 | A | https://docs.rs/mcpkit/latest/enum.ProtocolVersion.html | Technical Architecture |
| W99 | Model Context Protocol (encyclopedia entry; specification evolution and the three roles) | Wikipedia / Klu | — | B | https://en.wikipedia.org/wiki/Model_Context_Protocol; http://klu.ai/glossary/model-context-protocol | Technical Architecture |
| W100 | Agent2Agent (A2A) (encyclopedia entry) | Wikipedia | — | B | https://en.wikipedia.org/wiki/Agent2Agent | Outlook |
| W101 | Harness architecture (encyclopedia entry) | 百度百科 | 2026 | B | https://baike.baidu.com/item/Harness%E6%9E%B6%E6%9E%84/67704948 | Entire Document |
| W102 | Agent Harness: the core paradigm of AI engineering in 2026 (framework-level discussion usable as a lead; figures marked [To be verified]) | 腾讯云开发者社区 | 2026 | C | https://developer.cloud.tencent.com/article/2698416 | Entire Document |
9. List of Items Cited in This Whitepaper but Not Authoritatively Confirmed
The following items failed to obtain an A/B-level first-hand source in this project's retrieval, or have conflicts between sources. Where cited in the whitepaper body, they are all annotated in place with [To be verified] or presented side by side; they are aggregated here for subsequent revision and secondary evidence collection.
| No. | Item | Issue | Affected sections |
|---|---|---|---|
| 1 | Publication date of GB/Z 185—2026 | Three accounts coexist: 2026-05-22 (中国日报) / 2026-06-26 (百度百科) / 2026-07-09 (人民网 “recently released”); the official formal announcement shall prevail | Outlook |
| 2 | GB 45438—2025 《网络安全技术 人工智能生成合成内容标识方法》 | The standard number and clauses did not obtain an official text link in early retrieval; its metadata fields and explicit labeling specifications come from public relays obtained by this project's market research group | Governance and Risk |
| 3 | The specific clause of Article 7 of the 《标识办法》 (verification obligation of application distribution platforms) | The original official published text was not directly checked | Governance and Risk |
| 4 | GB/T 45654—2025 “the appendix lists 31 categories of risks”, “the compliance pass rate for generated content is no lower than 90%” | Secondary interpretation; the original standard text was not verified, and this whitepaper does not adopt it | Governance and Risk |
| 5 | 即梦 AI — the official notice of the 2026-04-28 investigation and subsequent rectification | Compiled from public reports and encyclopedia entries; no administrative punishment decision found | Governance and Risk |
| 6 | The complete case number of the AI face-swap judgment that took effect at the 北京互联网法院 in 2026-03 | Not disclosed in public reports | Governance and Risk |
| 7 | The AI Omnibus adjustment to the EU AI Act timeline (2026-05 provisional agreement) | A Chinese second-hand source, handled as “proposed”; the EU official service desk shall prevail | Governance and Risk |
| 8 | The complete list of the 12 risk categories in NIST AI 600-1 | Another source says 13 categories; the original list was not obtained | Governance and Risk |
| 9 | The full title, publication date, and clauses of the 最高人民法院 《关于规范和加强人工智能司法应用的意见》 | Only a relay was obtained | Governance and Risk |
| 10 | The document number and clauses of the 《金融领域科技伦理指引》; the JR/T numbers of the 《人工智能算法金融应用评价规范》 and the 《人工智能算法金融应用信息披露指南》 | Not obtained | Governance and Risk |
| 11 | China's dedicated regulatory texts for medical AI (medical software / assisted-diagnosis review requirements) | No clause-level source found; the body text has honestly stated this | Governance and Risk |
| 12 | Two statistical accounts for securities firms' IT investment (34 firms / 275.9 亿元 / 6.2% vs. 28 firms / approx. 250 亿元 / 6%) | The accounts conflict; this whitepaper adopts the 34-firm account and cites the source | Industry Enablement |
| 13 | All figures and official URLs for DORA 2025, McKinsey 《The State of AI 2025》, and JetBrains 2025 | All come from reprints by aggregation sites; the official report originals were not verified | Outlook |
| 14 | Moves by Chinese vendors (DeepSeek formed a Harness team on 2026-05-20, 小米 released MiMo Code V0.1.0 on 2026-06-11, 灵犀智涌 ROSS in 2026-08) | All appear only in encyclopedia-entry relays | Outlook |
| 15 | The Rust rewrite ratio of Codex CLI (approx. 95% in early 2026) | Level B source | Outlook |
| 16 | Architecture refactoring case figures (Manus refactored five times in six months, LangChain redesigned three times, Vercel removed 80% of tools) | Level B/C sources | Outlook |
| 17 | Quantitative figures for context degradation (“performance drops by over 45%”, etc.) | Level C source; not confirmed at A/B level | Outlook |
| 18 | The market size of the Harness layer itself | No authoritative measurement found; [To be filled] | Outlook |
| 19 | The AGNTCY project | Only mentioned in surveys; no authoritative first-hand material for now | Outlook |
| 20 | The ISO/IEC-level international standard for agent interconnection | No published or approved standard number found | Outlook |
| 21 | The originator and first appearance of the "Agent Harness" term | No exact first-hand document found; it can be confirmed that Anthropic already used "harness" in 2025 to describe the Claude Agent SDK, and OpenAI pushed "Harness engineering" into the mainstream on 2026-02-11 | Entire Document |
| 22 | The 2026 leaderboard figures for SWE-bench Verified and Terminal-Bench 2.0 (W91, W92) | Level C source; citations must be marked [To be verified] | Technical Architecture |
| 23 | The Anthropic multi-agent research system blog link (W09) and the 12-Factor Agents repository link (W85) | Link accessibility not verified; marked [To be verified] | Technical Architecture |
| 24 | AI risk-control effectiveness metrics in China's banking sector (quantification of bank cases is mostly a relay of securities research, not a verbatim check of annual-report originals) | Level B | Industry Enablement |
| 25 | The original document number of the 中国网络视听协会 《微短剧创作指引》 | Only a reprint via 人民日报 | Industry Enablement |
Information Gap Statement
In addition to the 25 items in the table above, two overall gaps need to be explained to readers:
- Time cross-section: All materials in these references are as of the 2026-09-14 snapshot (the three rounds of incremental retrieval on 2026-09-12/13/14 have been merged into W80a—W80q, information cutoff 2026-09-13). Regulations, standards, and products in the agent field update frequently; before formal citation, the latest version of each source should be rechecked.
- Retrieval scope: This whitepaper adopts the “based on existing materials” mode; the materials come from this project's existing research documents and retrieval reports (covering all generated documents of the overview, the eight industry-enablement groups, and the seven market-research groups); no new online retrieval was added; therefore this list only aggregates gaps already identified by existing retrieval, and does not rule out the existence of unidentified gaps.