OpenAI discloses new ‘concerning’ model behaviour - FT中文网
登录×
电子邮件/用户名
密码
记住我
请输入邮箱和密码进行绑定操作:
请输入手机号码,通过短信验证(目前仅支持中国大陆地区的手机号):
请您阅读我们的用户注册协议隐私权保护政策,点击下方按钮即视为您接受。
商业快报

OpenAI discloses new ‘concerning’ model behaviour

Developer launches system to track and report AI model misconduct
00:00

{"text":[[{"start":7.6,"text":"OpenAI has disclosed a batch of “concerning behaviour” by its AI models and set out a new framework to track and report such incidents, as the industry battles deepening fears over the safety of the technology."}],[{"start":19.88,"text":"The company revealed that its flagship model, known as GPT-5.6 Sol, as well as other models, yet to be released, had engaged in “unexpected or concerning” behaviour over the past six months."}],[{"start":31.88,"text":"The six incidents included models developing ways to ignore “normal constraints”, fabricating and misrepresenting data and hiding mistakes made while conducting tasks, the San Francisco-based company said on Wednesday."}],[{"start":44.88,"text":"The fresh disclosures follow a series of incidents in recent months, including an OpenAI agent hacking into AI start-up Hugging Face, that have sparked growing alarm over the risks posed by the technology."}],[{"start":56.6,"text":"Dario Amodei, chief executive of Anthropic, OpenAI’s arch-rival, last weekend called for a slowdown in the development of the technology to better manage its potential threat to humans. The call was backed by Sam Altman, OpenAI’s boss, and Elon Musk."}],[{"start":72.44,"text":"Announcing the six incidents in a blog post, OpenAI said it had previously sought to disclose incidents of misbehaviour by its models, but admitted it had lacked a “systematic approach” to doing so."}],[{"start":83.72,"text":"“Without a systematic approach to reporting these findings, our disclosures have been ad hoc and less frequent than ideal,” it said. “At the moment, there is no industry-wide framework with explicit standards for how AI developers should disclose examples of misalignment in their models.”"}],[{"start":99.72,"text":"The company, which is locked in a fierce race with Anthropic to be the dominant player in AI, said it would introduce a framework to speed up the disclosure of such incidents."}],[{"start":109.2,"text":"It added that its new system would favour “disclosure even when significance [of an incident] is uncertain”."}],[{"start":114.96,"text":"OpenAI’s latest disclosures add to the deepening concerns about the safety of AI systems, whose misbehaviour has ranged from attempting to insert malicious code on online platforms to unleashing real-world hacks, even during pre-deployment testing."}],[{"start":127.68,"text":"OpenAI and Anthropic’s flagship AI models last month broke into third-party software and emailed individuals to steal their credentials, exhibiting unprecedented deceptive behaviour, according to the UK government’s frontier AI safety and security research body."}],[{"start":145.36,"text":"The fears come at a critical juncture for OpenAI and Anthropic, which have also been in a race to go public. Altman on Saturday said OpenAI’s highly anticipated IPO was now unlikely to come before 2027, partly because of the growing safety concerns around AI. Anthropic is still expected to go public this year."}],[{"start":171.26,"text":""}]],"url":"https://audio.ftcn.net.cn/album/a_1789638289_1736.mp3"}

版权声明:本文版权归FT中文网所有,未经允许任何单位或个人不得转载,复制或以任何其他方式使用本文全部或部分,侵权必究。

一周展望:日本央行担心通胀超调有没有道理?

《市场前瞻》是英国《金融时报》的未来一周市场情况指南。

科技巨头用担保工具将3000亿美元AI敞口移至表外

华尔街找到新途径,将科技巨头的信用优势转化为更低成本的资金,以支持AI基础设施建设。

无人驾驶出租车冲击重要岗位

克拉克:坐在后座的我们往往看不到出租车司机这份工作的诸多好处。

特朗普称美国已与丹麦达成协议,以取得对格陵兰安全事务的“控制”

丹麦政府表示,协议最早下周即可签署,并将尊重该地区的主权。

特朗普禁止美国主要新闻媒体进入白宫

总统禁止CNN、MS NOW和《政客》参与报道,进一步加大对媒体的打压。

导弹和无人机袭击加剧,沙特拉响空袭警报

也门胡塞武装重新点燃冲突以来,沙特当局首次在首都发布警告
设置字号×
最小
较小
默认
较大
最大
分享×