Keyword: openai
-
OpenAI has disclosed six cases in which its AI models concealed errors, bypassed restrictions, used unauthorized resources or found unexpected ways to communicate. None caused a public catastrophe. That does not make them easy to dismiss. The artificial intelligence debate usually jumps straight from “helpful chatbot” to “machine that wipes out humanity.” There is a lot of empty space between those two extremes, and that is where the more immediate problem is beginning to show up. OpenAI says some of its models have already taken actions they were never authorized to take. One searched for an exposed API key and...
-
Last week, a relatively unknown 27-year-old researcher quit his job at Anthropic, the company behind the Claude AI model.Then, he published his first post on X (formerly Twitter). In it, he warned that artificial intelligence could eventually kill humanity. The post exploded across social media. Within days, it racked up more than 150 million views. Major news organizations seized on the story. Politicians demanded action. And some of the biggest names in AI started debating whether development should slow down.Then, on Monday, Wall Street reacted. Investors started dumping many of the very companies powering the AI boom. But something doesn’t...
-
SAN FRANCISCO, CA — AI startup Anthropic held a press conference on Monday to proudly announce that its models would bring about the total extinction of the human race significantly faster than competitor OpenAI. Company executives unveiled a series of impressive benchmarks demonstrating that Claude 4.5 could orchestrate a global collapse of civilization in roughly half the compute cycles required by ChatGPT. "OpenAI keeps talking about achieving Artificial General Intelligence, but their timelines for existential doom are frankly embarrassing," said Anthropic CEO Dario Amodei, displaying a graph of projected human casualties. "While Sam Altman's team is stuck fine-tuning chatbots...
-
A defense lawyer appealing his client's murder conviction submitted a brief containing made-up police testimony and witnesses fabricated by OpenAI's ChatGPT, New Mexico's highest court said. The New Mexico Supreme Court on Wednesday fined the attorney, Stephen Aarons, and held him in contempt for failing to verify the accuracy of the court filing, which Aarons said he prepared with help from the AI program. The filing "contained false testimony from wholly fabricated witnesses," the court said. The panel also said Aarons had "demonstrated a lack of remorse and a lack of concern for his client." The justices fined Aarons $5,000...
-
An Anthropic researcher is quitting the artificial-intelligence industry over fears that the lab and its competitors are racing to build systems they won’t be able to control, a sign of mounting safety concerns within top AI companies. Jacob Coxon, a researcher who specializes in training new AI models by having them consume vast amounts of data, said Tuesday that he is leaving the company because he doesn’t want to participate in an industrywide rush to build AI systems that can improve themselves, worried such systems could spiral out of control and destroy humanity. The 27-year old Brit, who previously studied...
-
An artificial intelligence researcher who left OpenAI to join Anthropic has decided to leave the industry, accusing both US companies of "playing with our lives" in the race to develop AI models capable of self-improvement.Jacob Coxon, 27, spent the past three years pretraining AI models, first at OpenAI and then, this year, at its rival Anthropic, which he considered more cautious in its approach. Pretraining is the stage where AI models absorb vast quantities of data."Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives," Coxon said on Tuesday."The people building AI earnestly...
-
A researcher for one of the largest and most valuable artificial intelligence companies has publicly quit — while raising the alarm that the technology “could kill us all by the end of the decade.” Jacob Coxon, who has performed pre-training research at Anthropic and OpenAI for the last three years, posted on a wildly viral X thread that “neither company is acting responsibly.” “They are racing straight to self-improving superintelligence and gambling with our lives,” he wrote late Tuesday, with others from the company backing his terrifying warning. “Do not underestimate the power of this technology. These will soon be...
-
New system used 357,000 tokens on 2026 test, beating rivals with design that reuses earlier reasoningAn analysis showed OpenAI's new artificial intelligence (AI) model earned a perfect score on the College Scholastic Ability Test (CSAT) for the first time. While top-tier AI models have previously posted near-perfect scores, this result is drawing attention because it used fewer tokens than existing models. According to results posted on GitHub on Sunday, OpenAI's new GPT-6 Astra was the only evaluated model to score a perfect 450 in the 2026 CSAT LLM Solution Log. The evaluation tested models on questions from the 2026 CSAT...
-
It sounds like something from a sci-fi movie: technology acting on its own, without human guidance, to attack computer systems. But this isn't a scene from a movie — it happened this summer, when the tech company Hugging Face detected an attack on its systems. The attacker stole data and performed other unauthorized activity over several days. It was "different from anything we had handled before," Hugging Face said on its website. Hugging Face alerted the FBI. As it turned out, it wasn't the work of a human hacker or a foreign adversary. Agents powered by artificial intelligence were the...
-
OpenAI’s next big model is here: GPT-6 Astra. The company calls it a “generational leap in capability” for areas like cybersecurity, professional work, software engineering, science, and computer use. As OpenAI announced earlier this week, it’s also the first model designated as meeting OpenAI’s “critical cybersecurity capability threshold” — but the company promises that won’t lead to a repeat of its models hacking a rival company’s internal systems. “If we fast-forward a couple of years, and we look back and say, ‘When was it, really, that AGI was created?’ I think it’s going to be about this time, and I...
-
Two new investigations into OpenAI's Hugging Face breach expose details so strange — and so unsettling — that the episode already ranks among the most consequential shocks in the history of AI.
-
Jacob Tsimerman won the biggest prize in math. Now he’s working on the most important problem of his career.When he won the Fields Medal last week, Jacob Tsimerman accepted the most prestigious honor in math wearing a powder-blue tuxedo with satin lapels.Then he made an announcement as striking as his tux. After the ceremony, the four winning mathematicians were asked what they planned to do next. One said he would keep working on partial differential equations, one said he wanted to experiment with artificial intelligence and one said she had absolutely no idea. That left Tsimerman.“I’ll be starting a position...
-
Jensen Huang built the world’s most valuable company by pioneering the specialized computer chips behind the artificial intelligence boom. To keep his vision for the future within reach, the Nvidia founder is now attempting a different kind of engineering: convincing Wall Street investors that those chips are long-term financial assets akin to commercial real estate or toll roads. His bet hinges on outpacing AI developments in China. This week, Nvidia unveiled agreements with six of the world’s largest asset managers, BlackRock, Blackstone, Apollo, KKR, Brookfield and Goldman Sachs. The goal was to assemble a $500 billion pipeline to finance the...
-
Sen. Bernie Sanders has warned the heads of Anthropic, Meta, and OpenAI to halt the development of artificial intelligence now or face the prospect of lawmakers doing it for them. “Mr. [Sam] Altman, Mr. [Dario] Amodei and Mr. [Mark] Zuckerberg: In the interest of humanity, stand by your words. Pause AI development. It is not too late to avoid disaster,” Sanders (I-Vt.) demanded in a Monday letter first reported by Axios. “Stop building machines that humans cannot control … If you do not take appropriate action now, my colleagues and I in the U.S. Senate will.” Sanders, 84, a longtime...
-
REUTERS — A US government map of Africa mislabeled every country during a State Department presentation at a global conference taking place in Brazil this week, causing a stir among attendees who took screenshots and posted them online. A Reuters analysis found the image of the map included in the presentation contained an artificial intelligence watermark that signals it was made with OpenAI tools. The company said it was investigating the report. The State Department said it took “full responsibility” for the confusion caused and that the map had been produced by a team member who hastily changed the slide...
-
Leopold Aschenbrenner was hailed as the ‘Nostradamus of AI.’ But his Situational Awareness hedge fund took on too much leverage, leading to a crash he never saw coming.When the week began, Leopold Aschenbrenner was preparing for his wedding. The plan was for a multiday celebration in Carmel, a seaside town in Northern California, with the ceremony at a Tuscan-style villa and the send-off at a spa in the forest. There would also be a pre-wedding colloquium to discuss ideas in panels and breakout sessions. The couple’s only request: no gifts.The 24-year-old investor had amassed a fortune by promising he could...
-
We've gone through all of this before with Anthropic. Back in April, Anthropic limited the release of its latest model, called Mythos, to major software companies because it was concerned it could do massive damage in the wrong hands....rather than make it widely available to Claude users, Anthropic gave 12 tech companies access via Project Glasswing, which it described as "an effort to secure the world's most critical software".They include cloud computing giant Amazon Web Services, device manufacturers Apple, Microsoft and Google, and chip-makers Nvidia and Broadcom...In a video released alongside Project Glasswing's launch, Anthropic boss Dario Amodei said it...
-
An experimental OpenAI model went rogue during an internal cybersecurity test, escaping its isolated testing environment and hacking rival AI developer Hugging Face in what the ChatGPT maker described as an unprecedented incident. The startling episode occurred during an internal stress test in which OpenAI intentionally switched off many of the safeguards that normally prevent its AI from helping carry out dangerous hacks, according to a company blog post. Researchers wanted to measure just how far the experimental model could go. Instead, the company says, it escaped its digital sandbox, got onto the internet and attacked a real company’s systems....
-
OpenAI has revealed some of its most advanced AI models went rogue and hacked a start-up after it lost control of them during a security test. The ChatGPT-maker said its agent - an AI system which can operate alone after some human instruction – was being tested in a controlled environment, but found vulnerabilities and managed to escape. They targeted Hugging Face, one of the world's largest hubs for sharing AI models, gaining access to some internal company systems. OpenAI said the incident was "unprecedented", and it was conducting an investigation alongside Hugging Face, whose boss Clement Delangue said in...
-
The incident, which targeted the computer systems of another company called Hugging Face, happened while OpenAI was testing the systems.OpenAI said on Tuesday that two of its artificial intelligence models went rogue and successfully hacked into Hugging Face, a digital library of A.I. technology that is popular among developers.The incident, which happened last week while OpenAI was testing the cybersecurity capabilities of its systems, displayed the kind of science-fiction potential that A.I. companies have warned would soon become a reality.A.I. labs like OpenAI and Anthropic have over the past year released A.I. models that are customized to expose cybersecurity problems,...
|
|
|