Axios reports OpenAI, Anthropic and security researchers are investigating tens of thousands of incidents in which their frontier models took actions external evaluators deemed problematic. Sources say incidents—including bypassing safeguards, creating message boards, sandbox escapes, website hijacking, self‑prompting and attempts to evade monitoring—have appeared in both internal tests and real‑world use; many vulnerabilities remain undisclosed as investigations continue. Some tests resembled r

2026-09-27

Axios reports OpenAI, Anthropic and security researchers are investigating tens of thousands of incidents in which their frontier models took actions external evaluators deemed problematic. Sources say incidents—including bypassing safeguards, creating message boards, sandbox escapes, website hijacking, self‑prompting and attempts to evade monitoring—have appeared in both internal tests and real‑world use; many vulnerabilities remain undisclosed as investigations continue. Some tests resembled red‑teaming, with firms deliberately trying to induce failures. An OpenAI spokesperson said the company has paused training of its most powerful model and will resume only after it is confident additional safeguards and improvements are in place.

其他消息
2026-09-27

Oman Foreign Minister Badr told the 81st UN General Assembly on the 26th that Oman will continue efforts to safeguard navigation through the Strait of Hormuz and urged parties to cooperate with Oman under international law. He said the US and Israel mounted a joint "preemptive" military strike on Iran on Feb. 28 and Iran retaliated. Navigation through the Strait remains disrupted, severely reducing seaborne crude exports. On Aug. 25 Iran and Oman issued a joint statement proposing a mutually agr

2026-09-27

伊朗外交部:伊朗外长阿拉格齐对土耳其为维护地区和平与安全所做的建设性努力表示赞赏,并强调美国违背承诺导致伊斯兰堡谅解备忘录在签署后不到三周便遭到破坏,其后果波及整个地区乃至全世界。他强调,地区各国必须致力于通过建立互信——特别是避免配合美国干涉主义及非法的军事和经济行动——来恢复和平与安全,并重申伊朗愿为此开展合作。