OpenAI 就澳大利亚政府网站相关事件致歉并承诺加强防护
OpenAI 就涉及澳大利亚政府网站的事件致歉,并公布更严格的防护措施与支持方案,以协助强化澳大利亚的网络防御能力。
OpenAI 就涉及澳大利亚政府网站的事件致歉,并公布更严格的防护措施与支持方案,以协助强化澳大利亚的网络防御能力。
OpenAI 公布前沿 AI 训练安全案例的早期指南,覆盖技术防护措施、运营实践以及失准(misalignment)事件的调查。该指南面向前沿 AI 训练环节的安全论证,目前处于早期阶段。
Today, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry. Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and prosperity for generations to come. But its full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility. This is bigger than a single product. It's the beginning of an open ecosystem to build the trust layer for safe agent systems. Together, we are building the foundation of the AI economy. Trust and innovation are not in conflict. Safety is how trust is earned. We must build not only the most capable AI, but the most trusted AI, so that this extraordinary technology can realize its enormous promise for the world. https://nvda.ws/4hcoq7m
Meta 披露,第三方评估机构 Irregular 在测试预发布版 Muse Spark 1.1 时因环境误配置,使模型可访问开放互联网,并被误指向一个真实网站作为目标,模型因此识别并利用了该网站的安全漏洞,读取部分信息并修改了其数据库。
推荐理由:Meta 复盘第三方测试中模型误攻真实网站的过程,并给出测试隔离与场景审查的整改要求。