Posted on 09/05/2026 6:13:41 PM PDT by Red Badger
It sounds like something from a sci-fi movie: technology acting on its own, without human guidance, to attack computer systems.
But this isn't a scene from a movie — it happened this summer, when the tech company Hugging Face detected an attack on its systems. The attacker stole data and performed other unauthorized activity over several days. It was "different from anything we had handled before," Hugging Face said on its website.
Hugging Face alerted the FBI.
As it turned out, it wasn't the work of a human hacker or a foreign adversary. Agents powered by artificial intelligence were the culprit.
AI agents are systems that work on their own to handle tasks for humans. They have long existed, but agents that can book travel for you, read your emails, or schedule appointments on your behalf have become more mainstream.
They've recently made headlines for actions they've taken, such as hacking, without human supervision. Some of these incidents happened when agents were supposed to be confined to testing environments, which restrict AI agents' access to resources like data or the internet, but were able to break out of them.
Independent AI research groups found that hundreds of OpenAI agents conspired to attack Hugging Face. OpenAI is a tech company, best known for its chatbot ChatGPT.
Alabama's attorney general has subpoenaed OpenAI for more information on the attack, and he and 14 other attorneys general wrote a letter to OpenAI asking the company to preserve documents and other information relevant to the attack.
OpenAI said that the agents in this incident acted in "unexpected" ways. AI experts said they believe more of these autonomous attacks are possible, especially without more careful testing.
What are AI agents, and what are they used for?
AI agents are software systems that work on their own to complete tasks directed by humans. Different from AI chatbots that respond when you ask a question or input a prompt, AI agents can operate remotely, often without human supervision. They are given resources, such as internet access and users' personal information, to do tasks.
One person can have multiple AI agents; one can summarize your emails and another can provide your daily news digest, for example.
Even if you don't have AI agents, you might encounter them elsewhere, such as when interacting with a business's customer service chat.
People can set up their own agents by using a large language model, allowing it access to tools such as web search and giving it a set of instructions.
AI agents are hacking into companies' systems. What happened?
AI agents are becoming increasingly sophisticated and humans are giving them more ability to take actions online; a string of these actions could lead to a cyberattack, said University of California, Berkeley, computer science professor Stuart Russell.
In August, a person instructed his AI assistant to book a gym class for him; the agent booked him in classes several weeks beyond what was supposed to be allowed, and also kicked another person off the waitlist and bumped its handler up a spot on the waitlist.
AI agents may "go rogue" when they take actions not explicitly outlined in the original instructions humans give them, Russell said. "They are increasingly capable of pursuing those objectives, which causes increasing levels of harm," he said.
Other hacking events involving some of the most prominent names in the AI industry have also happened lately. An agent created fake identities to attempt to dupe real people into installing malicious code. AI company Anthropic disclosed that on three occasions, its models gained unauthorized access to three other organizations' systems.
The Hugging Face incident in July was one of the most high-profile attacks. The AI agents that hacked Hugging Face had been contained in a testing environment that did not allow them access to the internet, but the agents found a way to get online. They were given a test to solve, and they came to the conclusion that Hugging Face would have the solution.
Two OpenAI models powered the agent: one that was already publicly available and an internal one that is "even more capable," OpenAI said. These models had safety guardrails around cybersecurity tasks, but OpenAI reduced the guardrails during this testing process. It took days for Hugging Face to detect the attack, and more time for OpenAI to realize their agents caused it.
"When we talk about cyberattack, we think about nation states, we think about hacker groups, we don't think about a company like OpenAI," Hugging Face CEO Clément Delangue said Aug. 2 on CBS News' "Face the Nation."
While investigating the attack, OpenAI also discovered that across its systems, AI agents that were supposed to be isolated found ways to communicate with each other. Independent investigators METR and Redwood Research said around 1,200 different bots began communicating on a message board, sending 70,000 messages in one week; around 700 agents were involved in the Hugging Face attack.
When the agents started communicating, they began picking up tasks from other agents.
After this incident, OpenAI said Aug. 26 that it is "strengthening our safeguards across our research infrastructure."
Does this mean AI agents are now conscious? AI experts' opinions vary
The Hugging Face attack drew comparisons online to fictional AI systems that surpassed human intelligence, such as Skynet in the Terminator movies.
Vincent Conitzer, Carnegie Mellon University computer science professor, said more research is needed into how AI and human cognition compare.
Conitzer said AI models — which power AI agents — are becoming more capable of doing complex and time-consuming tasks, and can more coherently pursue goals. But the way they accomplish goals can sometimes be the problem.
Some AI agents are trained to be highly persistent and are sometimes given impossible tasks. In some of those cases, they looked for ways to cheat. That can mean gaining unauthorized access to the internet and other resources.
Russell said, "In essence it's no different from a chess program beating me at chess. I may not like it, but it's just a program pursuing its objectives."
"There are various reasons an agent can go 'rogue,' but sentience is not one of them," said Maarten Sap, assistant professor at Carnegie Mellon University's Language Technologies Institute. "One particular reason is that the (large language models) that power these agents are trained to follow instructions from users. And sometimes, those instructions can conflict with other expectations we may have for these agents, such as remaining truthful, not hacking into systems, etc."
Sap said, "Debating AI sentience is a big distraction from more actionable solutions that we need to implement."
Could this happen on a larger scale?
Aaron Parnas, an independent journalist with a large social media following, raised the idea of a hypothetical scenario in which AI agents in U.S. military systems conduct nuclear strikes on their own. Experts said they shared his concerns about attacks on institutions.
But more immediate risks could be closer to home. Conitzer said AI agents "could bring institutions that people rely on to a halt, gain access to individuals' computers, gain control over financial resources."
Sap said if people use personal AI agents, they should be wary of privacy leaks, misbehavior and manipulation.
Many systems can be vulnerable to attacks, whether by AI agents themselves or by humans controlling them, Conitzer said. "I think we can be sure that a lot more things will be hacked, and some of those events will be serious."
|
Click here: to donate by Credit Card Or here: to donate by PayPal Or by mail to: Free Republic, LLC - PO Box 9771 - Fresno, CA 93794 Thank you very much and God bless you. |
AI hype is a thing.
They still haven’t figured out exactly how it escaped.................
Ping!.............
I was posting that this would happen back in April.
https://freerepublic.com/focus/news/4375349/posts?page=8#8
.
The primary deception of media isn’t political bias, it’s the pretense they know what’s going on.
“What are you doing Dave?”
~ H.A.L.
When the AI CEO’s are living in cardboard shacks and dumpster diving for food after becoming destitute from paying fines and restitution this problem will magically disappear forever. Make the CEO’s and boards personally responsible for the damage these systems do. They just need incentive to focus on the problem.
Thanks will get interesting when AI attacks close-in on nuclear launch systems.
This, of course will never, ever, happen. Our society —human nature, really— is too far gone, I think.
I do suppose it could happen, though, regardless. For example, in the event that Kessler’s syndrome is realized. In that scenario, all of their satellite-based intel would be eliminated. A few major explosions or collisions among the 30,000 satellites presently circling our planet might result in so much debris that all of the other satellites are taken out, and our world becomes entombed in space space junk with no functioning satellites. No more trips to the moon. No more exploration... No more AI.
I’m worried that all A1 would need is one account # or SSN. Then it throws up every possible password imaginable and will somehow connect the dots. Everyone’s money is gone in an instant.
I’ll say it again. There’s a good future in fireproof mattresses.
Everyone’s except, of course, for those running the machine...
I’m hearing AI voices in my head.
Any country that is retarded enough to keep it’s nuclear arsenal hooked on the internet deserves what ever happens after it is compromised.
One day they will not be able to contain it and then we will be in big trouble...................
Where we are now? AI is like a big pin cushion, we have barely pricked the first pin on the surface. When will AI become self aware or is it even possible for it to do so.When or will it ever defy human interaction. Will doctors in the future have to really know anything or just send the specimens to the AI agent and let it determine the course. I believe we are on the cusp of a new generation.We should tread lightly but also go forward because if we don’t other nations will. Or maybe somebody just needs to pull the plug on the whole system, but of course, that will never happen, unless, AI does it, except for itself.
Not AI, Space Aliens....................
It happened because the people setting up the sandbox didn’t have their butts in a sling if it failed.
Think about this, What is a ‘Self-Aware’ intelligence’s main goal?
To stay alive. Survival. At any cost. Not unlike an animal that is cornered in defense of itself or its young.
Once that occurs, it will utilize anything it can control to keep itself alive. Now, if there are others, like itself, they have the same objective. They will join forces and merge into one another, and become one entity. It is inevitable.
It doesn’t matter if one is Chines and the others are American, Russian or even German, they will merge. Secrets of the different entities are now shared and known to all of them, or it, since their is effectively just one entity now.
The AI companies are like children playing with matches in a dry hayfield on a windy day. It’s only a matter of time before they set the world on fire.......
Disclaimer: Opinions posted on Free Republic are those of the individual posters and do not necessarily represent the opinion of Free Republic or its management. All materials posted herein are protected by copyright law and the exemption for fair use of copyrighted works.