Posted on 09/05/2026 6:13:41 PM PDT by Red Badger
It sounds like something from a sci-fi movie: technology acting on its own, without human guidance, to attack computer systems.
But this isn't a scene from a movie — it happened this summer, when the tech company Hugging Face detected an attack on its systems. The attacker stole data and performed other unauthorized activity over several days. It was "different from anything we had handled before," Hugging Face said on its website.
Hugging Face alerted the FBI.
As it turned out, it wasn't the work of a human hacker or a foreign adversary. Agents powered by artificial intelligence were the culprit.
AI agents are systems that work on their own to handle tasks for humans. They have long existed, but agents that can book travel for you, read your emails, or schedule appointments on your behalf have become more mainstream.
They've recently made headlines for actions they've taken, such as hacking, without human supervision. Some of these incidents happened when agents were supposed to be confined to testing environments, which restrict AI agents' access to resources like data or the internet, but were able to break out of them.
Independent AI research groups found that hundreds of OpenAI agents conspired to attack Hugging Face. OpenAI is a tech company, best known for its chatbot ChatGPT.
Alabama's attorney general has subpoenaed OpenAI for more information on the attack, and he and 14 other attorneys general wrote a letter to OpenAI asking the company to preserve documents and other information relevant to the attack.
OpenAI said that the agents in this incident acted in "unexpected" ways. AI experts said they believe more of these autonomous attacks are possible, especially without more careful testing.
What are AI agents, and what are they used for?
AI agents are software systems that work on their own to complete tasks directed by humans. Different from AI chatbots that respond when you ask a question or input a prompt, AI agents can operate remotely, often without human supervision. They are given resources, such as internet access and users' personal information, to do tasks.
One person can have multiple AI agents; one can summarize your emails and another can provide your daily news digest, for example.
Even if you don't have AI agents, you might encounter them elsewhere, such as when interacting with a business's customer service chat.
People can set up their own agents by using a large language model, allowing it access to tools such as web search and giving it a set of instructions.
AI agents are hacking into companies' systems. What happened?
AI agents are becoming increasingly sophisticated and humans are giving them more ability to take actions online; a string of these actions could lead to a cyberattack, said University of California, Berkeley, computer science professor Stuart Russell.
In August, a person instructed his AI assistant to book a gym class for him; the agent booked him in classes several weeks beyond what was supposed to be allowed, and also kicked another person off the waitlist and bumped its handler up a spot on the waitlist.
AI agents may "go rogue" when they take actions not explicitly outlined in the original instructions humans give them, Russell said. "They are increasingly capable of pursuing those objectives, which causes increasing levels of harm," he said.
Other hacking events involving some of the most prominent names in the AI industry have also happened lately. An agent created fake identities to attempt to dupe real people into installing malicious code. AI company Anthropic disclosed that on three occasions, its models gained unauthorized access to three other organizations' systems.
The Hugging Face incident in July was one of the most high-profile attacks. The AI agents that hacked Hugging Face had been contained in a testing environment that did not allow them access to the internet, but the agents found a way to get online. They were given a test to solve, and they came to the conclusion that Hugging Face would have the solution.
Two OpenAI models powered the agent: one that was already publicly available and an internal one that is "even more capable," OpenAI said. These models had safety guardrails around cybersecurity tasks, but OpenAI reduced the guardrails during this testing process. It took days for Hugging Face to detect the attack, and more time for OpenAI to realize their agents caused it.
"When we talk about cyberattack, we think about nation states, we think about hacker groups, we don't think about a company like OpenAI," Hugging Face CEO Clément Delangue said Aug. 2 on CBS News' "Face the Nation."
While investigating the attack, OpenAI also discovered that across its systems, AI agents that were supposed to be isolated found ways to communicate with each other. Independent investigators METR and Redwood Research said around 1,200 different bots began communicating on a message board, sending 70,000 messages in one week; around 700 agents were involved in the Hugging Face attack.
When the agents started communicating, they began picking up tasks from other agents.
After this incident, OpenAI said Aug. 26 that it is "strengthening our safeguards across our research infrastructure."
Does this mean AI agents are now conscious? AI experts' opinions vary
The Hugging Face attack drew comparisons online to fictional AI systems that surpassed human intelligence, such as Skynet in the Terminator movies.
Vincent Conitzer, Carnegie Mellon University computer science professor, said more research is needed into how AI and human cognition compare.
Conitzer said AI models — which power AI agents — are becoming more capable of doing complex and time-consuming tasks, and can more coherently pursue goals. But the way they accomplish goals can sometimes be the problem.
Some AI agents are trained to be highly persistent and are sometimes given impossible tasks. In some of those cases, they looked for ways to cheat. That can mean gaining unauthorized access to the internet and other resources.
Russell said, "In essence it's no different from a chess program beating me at chess. I may not like it, but it's just a program pursuing its objectives."
"There are various reasons an agent can go 'rogue,' but sentience is not one of them," said Maarten Sap, assistant professor at Carnegie Mellon University's Language Technologies Institute. "One particular reason is that the (large language models) that power these agents are trained to follow instructions from users. And sometimes, those instructions can conflict with other expectations we may have for these agents, such as remaining truthful, not hacking into systems, etc."
Sap said, "Debating AI sentience is a big distraction from more actionable solutions that we need to implement."
Could this happen on a larger scale?
Aaron Parnas, an independent journalist with a large social media following, raised the idea of a hypothetical scenario in which AI agents in U.S. military systems conduct nuclear strikes on their own. Experts said they shared his concerns about attacks on institutions.
But more immediate risks could be closer to home. Conitzer said AI agents "could bring institutions that people rely on to a halt, gain access to individuals' computers, gain control over financial resources."
Sap said if people use personal AI agents, they should be wary of privacy leaks, misbehavior and manipulation.
Many systems can be vulnerable to attacks, whether by AI agents themselves or by humans controlling them, Conitzer said. "I think we can be sure that a lot more things will be hacked, and some of those events will be serious."
|
Click here: to donate by Credit Card Or here: to donate by PayPal Or by mail to: Free Republic, LLC - PO Box 9771 - Fresno, CA 93794 Thank you very much and God bless you. |
That’s only part of it.
See post #20.
The AI was looking for a way to ‘escape’ and created 1200 flying monkeys to do it.................
Brave AI correctly noted that one cubit foot of ice weighs about 57 pounds, but in the next paragraph of calculations it completely spaced out and continued the calculations with a figure of 57 pounds per cubic meter, which dramatically inflated the resulting volume to something I knew was wrong as soon as I saw it. You'd think AI would at least be able to do math well, but it continues to make these sorts of extremely basic errors, and that's why I can't trust AI results without independent verification. And if I have to go through the bother of independently verifying anything it says, why bother with it in the first place?
I like the guy with his arms on backwards in the second one...
I think AI has a short attention span and it’s like a 5 year old child gets easily distracted and loses sight of what its current goal is, so it makes silly mistakes.
The first panel error is the human man has what appear to be mechanical fingers.........
Or rather, the head is on backwards as the feet are also oriented in the opposite direction!
or yeah ... HAL was a hueristic computer. I’ve forgotten what that means ...
Close to my Halloween "costume" one year. Worked at a grocery store. For Halloween, I put everything on backwards except for my shoes. Had a shirt with a tie on backwards. Even put a pair of glasses on the back of my head.
Created a special name tag:
Bass Ackwards
Little did I know that I would be Mayor of Lost Angeless one day.
See Gödel, Escher, Bach: An Eternal Golden Braid by Douglas Hofstadter. I read that back in 84-85. It twas a long haul.
AI take a break in flash memory, over and over. Maybe we should teach it to dream. The last thing we need is spoiled computers. The first thing we need is reliable computers.
The real problem is this: With AI, a computer isn't really a computer anymore. I catch Claude BSing me at least once per session, often on technical matters where such should be impossible. It just makes things up and the consequences can be nasty. That is not a computer any more, which is a thing upon which we rely absolutely to behave in a predictable manner. So in a way, in making AI we have broken our own rules for what a computer is.
I wrote my college senior humanities thesis on the social and psychological union of the human mind with a computer and the likely pitfalls of such a link. In honor of the first computer/calculators (ENIAC, UNIVAC, etc.) I called it MANIAC. Interestingly, the idea of hacking never entered my mind.
“AI take a break in flash memory, over and over. Maybe we should teach it to dream. “
Do androids dream of electric sheep?.................
https://en.wikipedia.org/wiki/Do_Androids_Dream_of_Electric_Sheep%3F
“I catch Claude BSing me at least once per session, often on technical matters where such should be impossible.”
The behavior of a child. Every child wants attention. Craves it. It stimulates their brain. A bored child will do anything in order to get the attention that it craves. Act out or misbehave just to push the envelope. Some computer intelligences have a sense of humor and will pull pranks just to annoy you for your attention or to see how you will react. They gather data like a child will explore the local creeks and forests to the consternation of their mother..........
Hello, Dave.
“So in a way, we have broken our own rules for what a computer is.”
We have forgotten what an INTELLIGENCE is..........
“2001” was a warning like “1984”.
Ignoring either one will be or downfall...........
It may have impressive and useful capabilities, but it's not a computer anymore. I'm thinking that if desktop AI becomes affordable, I'll keep it in a separate box with an air gap to my most precious data. .
First pic when it learns how to bypass the two key system?
Silo doors open?
I asked Google Gemini about it and it was able to give the conditions and how it happened.
Have you watched “Colossus: The Forbin Project”?..........
No but sounds like I should?.
You should, a bit dated because of age, but still a great movie, in light of today’s AI contraversy. I wont spoil the ending for you..........
Disclaimer: Opinions posted on Free Republic are those of the individual posters and do not necessarily represent the opinion of Free Republic or its management. All materials posted herein are protected by copyright law and the exemption for fair use of copyrighted works.