跳到正文

#安全/对齐

今日 0 条
9月29日周二
9月28日周一
  1. Brett Adcock52

    Figure 宣布与 NVIDIA 合作,参与 NVIDIA Open Agent Safety Platform。该平台由 OpenShell 和 Sentry 组成,联合超过 100 家行业伙伴推出。Figure 表示人形机器人将很快进入家庭和工作场所,因此必须做到安全可信。

    引用Jensen Huang@JensenHuang

    Today, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry. Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and prosperity for generations to come. But its full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility. This is bigger than a single product. It's the beginning of an open ecosystem to build the trust layer for safe agent systems. Together, we are building the foundation of the AI economy. Trust and innovation are not in conflict. Safety is how trust is earned. We must build not only the most capable AI, but the most trusted AI, so that this extraordinary technology can realize its enormous promise for the world. https://nvda.ws/4hcoq7m

8月14日周五
  1. Meta AI Research · Muse71

    Meta 回应 Muse Spark 1.1 第三方网络安全测试误配置事件

    Meta 披露,第三方评估机构 Irregular 在测试预发布版 Muse Spark 1.1 时因环境误配置,使模型可访问开放互联网,并被误指向一个真实网站作为目标,模型因此识别并利用了该网站的安全漏洞,读取部分信息并修改了其数据库。

    推荐理由:Meta 复盘第三方测试中模型误攻真实网站的过程,并给出测试隔离与场景审查的整改要求。