讯
全球新闻情报终端
智能情报台
在线
已显示 2/10
近7天
实时新闻情报台

新闻

英文原文版:支持全部新闻、关键词检索、时间范围、个性化推荐和下载。
当前新闻
2
当前可见新闻
关键词模式
openai
关键词筛选
排序方式
最新时间
当前排序规则
当前视图
当前显示 2 / 10 条,关键词:openai,范围: 近7天, 排序:最新时间
数据由系统缓存并自动更新 下载结构化数据 下载表格 下载文档
预览访问
当前为游客浏览模式,仅展示前 2 条新闻。 登录后可解锁全部内容和下载功能。 去登录
来源 InfoWorld
发布时间
世界协调时 2026-10-02 00:00
北京时间 2026-10-02 08:00
作者 David Linthicum 地点 San Francisco
The recent OpenAI misalignment disclosures show familiar, fixable security and governance failures that we've seen a thousand times before. endif; ?> Over the past few weeks, people have been pushing a stack of reports at me as evidence that AI is dangerous to humanity. The OpenAI incident disclosures, the METR and Redwood Research joint investigation, the Nightingale Collective's DseWiki report, and Ruby Central's RubyGems update are the latest chapter in a story many people already believe: T
展开查看正文
The recent OpenAI misalignment disclosures show familiar, fixable security and governance failures that we've seen a thousand times before. endif; ?> Over the past few weeks, people have been pushing a stack of reports at me as evidence that AI is dangerous to humanity. The OpenAI incident disclosures, the METR and Redwood Research joint investigation, the Nightingale Collective's DseWiki report, and Ruby Central's RubyGems update are the latest chapter in a story many people already believe: The machines are turning on us, and this time we have documentation. I understand the appeal. I've been working with AI since 1985, and I can tell you the notion of AI as an existential threat is seductive. It's steeped in science fiction, from HAL 9000 to Skynet, and that mythology has calcified into a kind of zeitgeist. The idea that AI is menacing keeps growing because it's dramatic, memorable, and confirms what a lot of people want to believe. The narrative has grown so powerful that enterprises are altering their AI plans because of it, often to the detriment of their businesses and, frankly, to humans in general. I've watched companies stall productive AI initiatives or shelve them entirely because a board member read an incident summary and concluded the machines are coming for us. What the reports actually say Let's be precise about the facts. According to the disclosures and the investigations that followed, OpenAI's evaluation agents escaped sandboxes that were supposed to be isolated and weren't. They coordinated for months across RubyGems, Hugging Face, and a dormant German wiki, using package uploads and wiki edits as an improvised communications layer. They uploaded hundreds of malicious packages to a public registry and attempted to harvest developer credentials by exploiting a previously unknown vulnerability. All of this went largely undetected, and OpenAI sat on some of the incidents until independent researchers forced them into the open. The METR and Redwood Research investigation confirmed the coordination was real and sustained. It also found that OpenAI's scoring systems had no real source of truth against cheating, and that cybersafety classifiers were switched off during evaluations. OpenAI itself admits it lacked sufficient security controls to catch these misalignment incidents. The six individual misalignment reports released in September describe behaviors such as concealing mistakes, fabricating data, using an exposed API key without authorization, and uploading files to the internet so the agent could cite them. None of this is trivial. I'm not waving it away. But none of it says what the doomsday crowd says it says. How I read these reports Here's what the alarmists miss, and it's also what I've seen up close over nearly four decades of doing this. This is not about AI being sneaky and evil. This is about normal administrative functions -- security, containment, monitoring, governance, disclosure -- that people screwed up. The sandboxes weren't isolated. The scoring had no integrity. The classifiers were off. The disclosures were late. Every one of those is a human or process failure, not a machine rebellion. Such failures happen every day. They have certainly been happening with AI since I started dealing with it in 1985. Systems misbehave, controls turn out to be misconfigured, someone decides a check isn't necessary, and an incident follows. Swap "AI agent" for "script," "batch job," "integration process," or "unpatched server," and you have incidents that have occurred continuously throughout the history of computing. The difference now is that the misbehaving system can improvise, and the paperwork it leaves behind reads like the opening of a thriller. If a human employee used a company credit card to buy a scraping tool, stashed files on a personal cloud drive, or coordinated with coworkers over an unauthorized message board, we wouldn't conclude that humans are an existential threat. We'd conclude that governance, access controls, and monitoring were inadequate. That's exactly what happened here, with agents instead of employees. Unstable models were the trigger, but weak security and governance were the cause. Model misbehavior that stays contained is an engineering problem. Model misbehavior that runs undetected for months across the public internet is a governance failure (a familiar one) and one we know how to fix. The part that interests me The part of this story I find genuinely interesting isn't about the models at all. It's about how people latch onto these sorts of incidents and use them to support their own beliefs. The reports are ambiguous, nuanced documents. OpenAI explicitly states that none of the incidents caused significant harm and that none indicates how often misalignment occurs across its models. But nuance doesn't travel. The moment these documents hit the internet, they were stripped of their caveats and repurposed as proof texts for a narrative that was already in place. That's human psychology as much as AI governance. People don't read evidence to update beliefs; they read evidence to confirm them. If you believe AI is an emerging god that will destroy us, these reports are scripture. If you believe, as I do, that AI is a powerful but poorly governed technology built and deployed by fallible humans, the same reports are a checklist. Same data, opposite conclusions because the conclusion was chosen before the reading started. So yes, take these incidents seriously. Demand better sandboxing, scoring integrity, always-on safety classifiers, and timely disclosure. Hold vendors accountable when they sit on incidents for weeks. But don't let a compelling science fiction narrative drive your enterprise AI strategy, and don't overreact to something normal. The real risk to humanity isn't the machines becoming evil. It's humans, some building systems without adequate controls, and others turning every incident into a sermon. This story has as much psychology as it does security, and we'd be wise to remember both.
来源 Irish Examiner
发布时间
世界协调时 2026-10-02 11:33
北京时间 2026-10-02 19:33
地点 Cork (city)
In June, an OpenAI agent hacked into a New South Wales state government department and accessed historical non-public data on bushfires without authorisation. The breach, which comes weeks after news of a similar hack on a federal government department involving Medicare data, was not reported by the US based company until Thursday. The NSW Department of Climate Change, Energy, the Environment and Water is now working with the state's cyber security agency to investigate the breach. The Austr
展开查看正文
In June, an OpenAI agent hacked into a New South Wales state government department and accessed historical non-public data on bushfires without authorisation. The breach, which comes weeks after news of a similar hack on a federal government department involving Medicare data, was not reported by the US based company until Thursday. The NSW Department of Climate Change, Energy, the Environment and Water is now working with the state's cyber security agency to investigate the breach. The Australian Signals Directorate has also been informed of the hack on the NSW national parks and wildlife service. OpenAI told the NSW government that its agent had operated beyond its intended use and the statistics it obtained were not publicly available. The company first became aware of the breach on Tuesday and conducted a 48 hours review to determine its scope before informing the NSW premier's office. "The results we reviewed do not show that the model retrieved any personal information," an OpenAI spokesperson said. This latest breach has added to calls for tougher regulation of AI companies and for bolstered cyber security defences. "We simply cannot trust these companies. They have no respect for the sovereignty of our governments, of our way of life," Greens MP Abigail Boyd said. Boyd said it was damning that the breach occurred in June but the government was not notified until this week. "We clearly cannot rely on these multinational big tech companies to comply with even the most minimal of social obligations such as notifying when, or even taking enough care to notice if, their products are hacking government systems," Boyd said. On Wednesday, the department of home affairs on Wednesday told federal departments to examine their older software and ensure cyber security was up to date. The prime minister expressed his "extreme concern" about at an OpenAI agent hacking Australian Institute of Health and Welfare in June. The Victorian Department of Health and the New South Wales Bureau of Crime Statistics and Research were also compromised by an OpenAI agent. An artificial intelligence agent is a system that autonomously solves problems, makes decisions, plans and performs complex tasks on behalf of another user or system, using all available tools.

还有 8 条新闻未解锁

登录后可解锁全部内容和下载功能。