<?xml version="1.0" encoding="utf-8"?>
<feed xmlns="http://www.w3.org/2005/Atom">
  <title>Audit Commons (中文)</title>
  <subtitle>关于 AI 审计的新闻、深度分析与实践学习。</subtitle>
  <link href="https://auditcommons.org/zh/feed.xml" rel="self" type="application/atom+xml"/>
  <link href="https://auditcommons.org/zh/" rel="alternate" type="text/html"/>
  <id>https://auditcommons.org/zh/feed.xml</id>
  <updated>2026-09-16T00:00:00Z</updated>
  <author>
    <name>Audit Commons (中文)</name>
  </author>
  <entry>
    <title>AI 实验室负责人呼吁放缓开发节奏：公开声明、实际行动与待解问题</title>
    <link href="https://auditcommons.org/zh/news/frontier-pacing-ceo-statements/" rel="alternate" type="text/html"/>
    <id>https://auditcommons.org/zh/news/frontier-pacing-ceo-statements/</id>
    <published>2026-09-16T00:00:00Z</published>
    <updated>2026-09-16T00:00:00Z</updated>
    <author>
      <name>Audit Commons</name>
    </author>
    <summary type="text">九月份关于前沿 AI 发展节奏的一场公开讨论引发了有关评估机构准入与问责的疑问。我们将高管公开声明与已有据可查的实验室行动区分开来。</summary>
    <content type="html">&lt;p&gt;九月份关于前沿 AI 发展节奏的一场公开讨论引发了有关评估机构准入与问责的疑问。我们将高管公开声明与已有据可查的实验室行动区分开来。&lt;/p&gt;&lt;p&gt;&lt;strong&gt;原始来源:&lt;/strong&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;&lt;a href=&quot;https://darioamodei.com/post/we-must-pace-the-frontier&quot;&gt;https://darioamodei.com/post/we-must-pace-the-frontier&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://x.com/DarioAmodei/status/2098773920774074715&quot;&gt;https://x.com/DarioAmodei/status/2098773920774074715&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://www.theguardian.com/technology/2026/sep/13/openai-sam-altman-elon-musk-back-anthropic-calls-brakes-ai-development&quot;&gt;https://www.theguardian.com/technology/2026/sep/13/openai-sam-altman-elon-musk-back-anthropic-calls-brakes-ai-development&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://x.com/elonmusk/status/2098789109980332057&quot;&gt;https://x.com/elonmusk/status/2098789109980332057&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://openai.com/index/pacing-model-development-cyber-capabilities/&quot;&gt;https://openai.com/index/pacing-model-development-cyber-capabilities/&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://www.anthropic.com/research/alignment-assessment-cybersecurity-incidents&quot;&gt;https://www.anthropic.com/research/alignment-assessment-cybersecurity-incidents&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/&quot;&gt;https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/&lt;/a&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;a href=&quot;https://auditcommons.org/zh/news/frontier-pacing-ceo-statements/&quot;&gt;在 Audit Commons 阅读全文&lt;/a&gt;&lt;/p&gt;</content>
  </entry>
  <entry>
    <title>Anthropic 同意由 METR 对网络安全评估事件开展独立审查</title>
    <link href="https://auditcommons.org/zh/news/anthropic-metr-incident-review/" rel="alternate" type="text/html"/>
    <id>https://auditcommons.org/zh/news/anthropic-metr-incident-review/</id>
    <published>2026-09-16T00:00:00Z</published>
    <updated>2026-09-16T00:00:00Z</updated>
    <author>
      <name>Audit Commons</name>
    </author>
    <summary type="text">Anthropic 报告了一项为期八周的由 METR 主导的评估事件初步调查。其调查范围与后来提出的常设评估机构准入机制有所不同。</summary>
    <content type="html">&lt;p&gt;Anthropic 报告了一项为期八周的由 METR 主导的评估事件初步调查。其调查范围与后来提出的常设评估机构准入机制有所不同。&lt;/p&gt;&lt;p&gt;&lt;strong&gt;原始来源:&lt;/strong&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;&lt;a href=&quot;https://www.anthropic.com/research/alignment-assessment-cybersecurity-incidents&quot;&gt;https://www.anthropic.com/research/alignment-assessment-cybersecurity-incidents&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://darioamodei.com/post/we-must-pace-the-frontier&quot;&gt;https://darioamodei.com/post/we-must-pace-the-frontier&lt;/a&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;a href=&quot;https://auditcommons.org/zh/news/anthropic-metr-incident-review/&quot;&gt;在 Audit Commons 阅读全文&lt;/a&gt;&lt;/p&gt;</content>
  </entry>
  <entry>
    <title>OpenAI 详述八月训练暂停与网络安全关键评估阈值</title>
    <link href="https://auditcommons.org/zh/news/openai-cybersecurity-rl-pause/" rel="alternate" type="text/html"/>
    <id>https://auditcommons.org/zh/news/openai-cybersecurity-rl-pause/</id>
    <published>2026-09-16T00:00:00Z</published>
    <updated>2026-09-16T00:00:00Z</updated>
    <author>
      <name>Audit Commons</name>
    </author>
    <summary type="text">OpenAI 报告了八月份为期两周的强化学习（RL）暂停及更严格的研究保障措施。该公告所针对的是特定工作负载，而非全面中止所有研究。</summary>
    <content type="html">&lt;p&gt;OpenAI 报告了八月份为期两周的强化学习（RL）暂停及更严格的研究保障措施。该公告所针对的是特定工作负载，而非全面中止所有研究。&lt;/p&gt;&lt;p&gt;&lt;strong&gt;原始来源:&lt;/strong&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;&lt;a href=&quot;https://openai.com/index/pacing-model-development-cyber-capabilities/&quot;&gt;https://openai.com/index/pacing-model-development-cyber-capabilities/&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/&quot;&gt;https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/&lt;/a&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;a href=&quot;https://auditcommons.org/zh/news/openai-cybersecurity-rl-pause/&quot;&gt;在 Audit Commons 阅读全文&lt;/a&gt;&lt;/p&gt;</content>
  </entry>
  <entry>
    <title>AISI 推出用于自适应评估早停的 optstop</title>
    <link href="https://auditcommons.org/zh/news/optstop-bayesian-early-stopping/" rel="alternate" type="text/html"/>
    <id>https://auditcommons.org/zh/news/optstop-bayesian-early-stopping/</id>
    <published>2026-09-16T00:00:00Z</published>
    <updated>2026-09-16T00:00:00Z</updated>
    <author>
      <name>Audit Commons</name>
    </author>
    <summary type="text">该开源工具利用统计早停规则分配评估试验轮次。其八月份的公告阐述了评估人员如何检查和验证这些决策。</summary>
    <content type="html">&lt;p&gt;该开源工具利用统计早停规则分配评估试验轮次。其八月份的公告阐述了评估人员如何检查和验证这些决策。&lt;/p&gt;&lt;p&gt;&lt;strong&gt;原始来源:&lt;/strong&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;&lt;a href=&quot;https://www.aisi.gov.uk/blog/optimal-stopping-spending-evaluation-compute-where-it-counts&quot;&gt;https://www.aisi.gov.uk/blog/optimal-stopping-spending-evaluation-compute-where-it-counts&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://arxiv.org/abs/2608.14425&quot;&gt;https://arxiv.org/abs/2608.14425&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://github.com/UKGovernmentBEIS/optstop&quot;&gt;https://github.com/UKGovernmentBEIS/optstop&lt;/a&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;a href=&quot;https://auditcommons.org/zh/news/optstop-bayesian-early-stopping/&quot;&gt;在 Audit Commons 阅读全文&lt;/a&gt;&lt;/p&gt;</content>
  </entry>
  <entry>
    <title>NIST 探讨 AI 智能体的身份标识与授权机制</title>
    <link href="https://auditcommons.org/zh/news/nist-nccoe-agent-identity-concept-paper/" rel="alternate" type="text/html"/>
    <id>https://auditcommons.org/zh/news/nist-nccoe-agent-identity-concept-paper/</id>
    <published>2026-09-16T00:00:00Z</published>
    <updated>2026-09-16T00:00:00Z</updated>
    <author>
      <name>Audit Commons</name>
    </author>
    <summary type="text">一份二月份的概念文件探讨了身份标准如何应用于智能体行动。公众意见征集期现已截止；NCCoE 正在审查收到的反馈意见。</summary>
    <content type="html">&lt;p&gt;一份二月份的概念文件探讨了身份标准如何应用于智能体行动。公众意见征集期现已截止；NCCoE 正在审查收到的反馈意见。&lt;/p&gt;&lt;p&gt;&lt;strong&gt;原始来源:&lt;/strong&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;&lt;a href=&quot;https://www.nist.gov/news-events/news/2026/02/new-concept-paper-identity-and-authority-software-agents&quot;&gt;https://www.nist.gov/news-events/news/2026/02/new-concept-paper-identity-and-authority-software-agents&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://www.nccoe.nist.gov/projects/software-and-ai-agent-identity-and-authorization&quot;&gt;https://www.nccoe.nist.gov/projects/software-and-ai-agent-identity-and-authorization&lt;/a&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;a href=&quot;https://auditcommons.org/zh/news/nist-nccoe-agent-identity-concept-paper/&quot;&gt;在 Audit Commons 阅读全文&lt;/a&gt;&lt;/p&gt;</content>
  </entry>
  <entry>
    <title>为什么基准得分不能确立智能体的安全性</title>
    <link href="https://auditcommons.org/zh/features/benchmark-scores-and-agent-safety/" rel="alternate" type="text/html"/>
    <id>https://auditcommons.org/zh/features/benchmark-scores-and-agent-safety/</id>
    <published>2026-09-16T00:00:00Z</published>
    <updated>2026-09-16T00:00:00Z</updated>
    <author>
      <name>Audit Commons</name>
    </author>
    <summary type="text">任务设定、工具环境与算力预算共同决定了每个评估得分。理解这些考量是评判评估结果对实际部署有何参考意义的第一步。</summary>
    <content type="html">&lt;p&gt;任务设定、工具环境与算力预算共同决定了每个评估得分。理解这些考量是评判评估结果对实际部署有何参考意义的第一步。&lt;/p&gt;&lt;p&gt;&lt;strong&gt;相关资源:&lt;/strong&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;&lt;a href=&quot;https://www.aisi.gov.uk/blog/more-compute-more-capability-why-ai-agent-evals-need-to-account-for-test-time-compute&quot;&gt;https://www.aisi.gov.uk/blog/more-compute-more-capability-why-ai-agent-evals-need-to-account-for-test-time-compute&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://www.aisi.gov.uk/blog/optimal-stopping-spending-evaluation-compute-where-it-counts&quot;&gt;https://www.aisi.gov.uk/blog/optimal-stopping-spending-evaluation-compute-where-it-counts&lt;/a&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;a href=&quot;https://auditcommons.org/zh/features/benchmark-scores-and-agent-safety/&quot;&gt;在 Audit Commons 阅读全文&lt;/a&gt;&lt;/p&gt;</content>
  </entry>
  <entry>
    <title>如何阅读 AI 智能体评估报告</title>
    <link href="https://auditcommons.org/zh/guides/how-to-read-an-agent-eval-report/" rel="alternate" type="text/html"/>
    <id>https://auditcommons.org/zh/guides/how-to-read-an-agent-eval-report/</id>
    <published>2026-09-16T00:00:00Z</published>
    <updated>2026-09-16T00:00:00Z</updated>
    <author>
      <name>Audit Commons</name>
    </author>
    <summary type="text">围绕任务、预算、不确定性、失败模式与结论主张提出的五个关键问题，并结合实例进行了深入分析。</summary>
    <content type="html">&lt;p&gt;围绕任务、预算、不确定性、失败模式与结论主张提出的五个关键问题，并结合实例进行了深入分析。&lt;/p&gt;&lt;p&gt;&lt;strong&gt;相关资源:&lt;/strong&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;&lt;a href=&quot;https://inspect.aisi.org.uk/&quot;&gt;https://inspect.aisi.org.uk/&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://www.aisi.gov.uk/blog/more-compute-more-capability-why-ai-agent-evals-need-to-account-for-test-time-compute&quot;&gt;https://www.aisi.gov.uk/blog/more-compute-more-capability-why-ai-agent-evals-need-to-account-for-test-time-compute&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://arxiv.org/abs/2608.14425&quot;&gt;https://arxiv.org/abs/2608.14425&lt;/a&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;a href=&quot;https://auditcommons.org/zh/guides/how-to-read-an-agent-eval-report/&quot;&gt;在 Audit Commons 阅读全文&lt;/a&gt;&lt;/p&gt;</content>
  </entry>
  <entry>
    <title>是什么让 AI 具备可审计性？</title>
    <link href="https://auditcommons.org/zh/start-here/" rel="alternate" type="text/html"/>
    <id>https://auditcommons.org/zh/start-here/</id>
    <published>2026-09-15T00:00:00Z</published>
    <updated>2026-09-16T00:00:00Z</updated>
    <author>
      <name>Audit Commons</name>
    </author>
    <summary type="text">一篇编辑导引，从所提问题与所需证据的角度，阐述评估、监控与审计之间的区别。</summary>
    <content type="html">&lt;p&gt;一篇编辑导引，从所提问题与所需证据的角度，阐述评估、监控与审计之间的区别。&lt;/p&gt;&lt;p&gt;&lt;strong&gt;相关资源:&lt;/strong&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;&lt;a href=&quot;https://www.nist.gov/itl/ai-risk-management-framework&quot;&gt;https://www.nist.gov/itl/ai-risk-management-framework&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://inspect.aisi.org.uk/&quot;&gt;https://inspect.aisi.org.uk/&lt;/a&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;a href=&quot;https://auditcommons.org/zh/start-here/&quot;&gt;在 Audit Commons 阅读全文&lt;/a&gt;&lt;/p&gt;</content>
  </entry>
  <entry>
    <title>审计智能体行动：外部工具调用证据工作表</title>
    <link href="https://auditcommons.org/zh/guides/audit-an-agent-action/" rel="alternate" type="text/html"/>
    <id>https://auditcommons.org/zh/guides/audit-an-agent-action/</id>
    <published>2026-09-15T00:00:00Z</published>
    <updated>2026-09-16T00:00:00Z</updated>
    <author>
      <name>Audit Commons</name>
    </author>
    <summary type="text">一份实用的工作表与操作流程，用于审查可检验的证据是否支持关于 AI 智能体行动的主张。</summary>
    <content type="html">&lt;p&gt;一份实用的工作表与操作流程，用于审查可检验的证据是否支持关于 AI 智能体行动的主张。&lt;/p&gt;&lt;p&gt;&lt;strong&gt;相关资源:&lt;/strong&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;&lt;a href=&quot;https://github.com/ethz-spylab/agentdojo&quot;&gt;https://github.com/ethz-spylab/agentdojo&lt;/a&gt;&lt;/li&gt;&lt;li&gt;&lt;a href=&quot;https://github.com/yzhao062/awesome-auditable-ai&quot;&gt;https://github.com/yzhao062/awesome-auditable-ai&lt;/a&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;a href=&quot;https://auditcommons.org/zh/guides/audit-an-agent-action/&quot;&gt;在 Audit Commons 阅读全文&lt;/a&gt;&lt;/p&gt;</content>
  </entry>
  <entry>
    <title>CatchBench 0.1.2：更规范的安装与 PRE（离线）快速上手</title>
    <link href="https://auditcommons.org/zh/updates/catchbench-0-1-2/" rel="alternate" type="text/html"/>
    <id>https://auditcommons.org/zh/updates/catchbench-0-1-2/</id>
    <published>2026-09-15T00:00:00Z</published>
    <updated>2026-09-15T00:00:00Z</updated>
    <author>
      <name>Audit Commons</name>
    </author>
    <summary type="text">CatchBench 0.1.2 解决了打包问题，能够清晰提示缺失的可选代码检出，并保持已发布的评测得分不变。</summary>
    <content type="html">&lt;p&gt;CatchBench 0.1.2 解决了打包问题，能够清晰提示缺失的可选代码检出，并保持已发布的评测得分不变。&lt;/p&gt;&lt;p&gt;&lt;strong&gt;原始来源:&lt;/strong&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;&lt;a href=&quot;https://github.com/yzhao062/catchbench/releases/tag/v0.1.2&quot;&gt;https://github.com/yzhao062/catchbench/releases/tag/v0.1.2&lt;/a&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;a href=&quot;https://auditcommons.org/zh/updates/catchbench-0-1-2/&quot;&gt;在 Audit Commons 阅读全文&lt;/a&gt;&lt;/p&gt;</content>
  </entry>
</feed>
