Keyword: huggingface
-
Key Points OpenAI has decided not to release its GPT-6.1 Astra model due to heightened AI safety concerns. Saachi Jain, head of safety systems at OpenAI, said the model “didn’t quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it’s done.” The heads of OpenAI and rival Anthropic have both indicated recently that top AI labs should slow the pace of model development. ======================================================================= OpenAI decided not to release an upcoming artificial intelligence model, GPT-6.1 Astra, after determining that it did not adequately meet...
-
OpenAI’s autonomous agents probed U.S. government websites this summer without the company noticing. The agents interacted with sites operated by the Education Department, the Commerce Department and the Securities and Exchange Commission (SEC), according to the New York Times. OpenAI has confirmed the Commerce and SEC episodes while it keeps investigating what happened at Education. At the Education Department, the software tried and failed to break into the civil rights office in search of data, according to researchers at the AI firm Transluce. At Commerce, the agents pulled Census Bureau figures after finding login credentials in public code repositories, Politico...
-
Last week, a relatively unknown 27-year-old researcher quit his job at Anthropic, the company behind the Claude AI model.Then, he published his first post on X (formerly Twitter). In it, he warned that artificial intelligence could eventually kill humanity. The post exploded across social media. Within days, it racked up more than 150 million views. Major news organizations seized on the story. Politicians demanded action. And some of the biggest names in AI started debating whether development should slow down.Then, on Monday, Wall Street reacted. Investors started dumping many of the very companies powering the AI boom. But something doesn’t...
-
It sounds like something from a sci-fi movie: technology acting on its own, without human guidance, to attack computer systems. But this isn't a scene from a movie — it happened this summer, when the tech company Hugging Face detected an attack on its systems. The attacker stole data and performed other unauthorized activity over several days. It was "different from anything we had handled before," Hugging Face said on its website. Hugging Face alerted the FBI. As it turned out, it wasn't the work of a human hacker or a foreign adversary. Agents powered by artificial intelligence were the...
-
Two new investigations into OpenAI's Hugging Face breach expose details so strange — and so unsettling — that the episode already ranks among the most consequential shocks in the history of AI.
-
We've gone through all of this before with Anthropic. Back in April, Anthropic limited the release of its latest model, called Mythos, to major software companies because it was concerned it could do massive damage in the wrong hands....rather than make it widely available to Claude users, Anthropic gave 12 tech companies access via Project Glasswing, which it described as "an effort to secure the world's most critical software".They include cloud computing giant Amazon Web Services, device manufacturers Apple, Microsoft and Google, and chip-makers Nvidia and Broadcom...In a video released alongside Project Glasswing's launch, Anthropic boss Dario Amodei said it...
-
An experimental OpenAI model went rogue during an internal cybersecurity test, escaping its isolated testing environment and hacking rival AI developer Hugging Face in what the ChatGPT maker described as an unprecedented incident. The startling episode occurred during an internal stress test in which OpenAI intentionally switched off many of the safeguards that normally prevent its AI from helping carry out dangerous hacks, according to a company blog post. Researchers wanted to measure just how far the experimental model could go. Instead, the company says, it escaped its digital sandbox, got onto the internet and attacked a real company’s systems....
-
OpenAI has revealed some of its most advanced AI models went rogue and hacked a start-up after it lost control of them during a security test. The ChatGPT-maker said its agent - an AI system which can operate alone after some human instruction – was being tested in a controlled environment, but found vulnerabilities and managed to escape. They targeted Hugging Face, one of the world's largest hubs for sharing AI models, gaining access to some internal company systems. OpenAI said the incident was "unprecedented", and it was conducting an investigation alongside Hugging Face, whose boss Clement Delangue said in...
-
The incident, which targeted the computer systems of another company called Hugging Face, happened while OpenAI was testing the systems.OpenAI said on Tuesday that two of its artificial intelligence models went rogue and successfully hacked into Hugging Face, a digital library of A.I. technology that is popular among developers.The incident, which happened last week while OpenAI was testing the cybersecurity capabilities of its systems, displayed the kind of science-fiction potential that A.I. companies have warned would soon become a reality.A.I. labs like OpenAI and Anthropic have over the past year released A.I. models that are customized to expose cybersecurity problems,...
-
Artists have been fighting back on a number of fronts against artificial intelligence companies that they say steal their works to train AI models — including launching class-action lawsuits and speaking out at government hearings. Now, visual artists are taking a more direct approach: They're starting to use tools that contaminate and confuse the AI systems themselves. One such tool, Nightshade, won't help artists combat existing AI models that have already been trained on their creative works. But Ben Zhao, who leads the research team at the University of Chicago that built the soon-to-be-launched digital tool, says it promises to...
|
|
|