News & Updates

The latest news and updates from companies in the WLTH portfolio.

Anthropic researcher resigns with warning about the dangers of AI development | 95.7FM WZID

By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

Anthropic
95.7FM WZID1d ago
Read update
Anthropic researcher resigns with warning about the dangers of AI development | 95.7FM WZID

Anthropic researcher resigns with warning about the dangers of AI development | ROCK 102 WAQY

By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

Anthropic
ROCK 102 WAQY1d ago
Read update
Anthropic researcher resigns with warning about the dangers of AI development | ROCK 102 WAQY

Anthropic Insiders Warn AI Could Wipe Out Humanity: A 10% Chance by 2036

Two researchers at Anthropic delivered a blunt message this week. One quit. The other put hard numbers on humanity's possible end. Jacob Coxon resigned from the company Tuesday. He had spent three years on pretraining research. First at OpenAI. Then at Anthropic. His departure post on X pulled no punches. "Neither company is acting responsibly," he wrote. "They are racing straight to self-improving superintelligence and gambling with our lives." Coxon added that the people building these systems "earnestly believe that it could kill us all by the end of the decade." He insisted this was no marketing stunt. Executives soften their words in public. Privately, he said, they express fear. "No other human activity poses this level of danger." From BBC News. The Numbers Behind the Alarm Evan Hubinger responded almost immediately. As Anthropic's alignment science lead, he leads efforts to make sure advanced AI systems follow human intentions. His reply carried weight. "Jacob is correct here -- we really do earnestly believe AI could kill all humans!" Hubinger wrote. "I personally think it is >10% within the next decade." He stressed the risk from today's models remains low. The danger lies ahead. In recursive self-improvement. Systems that enhance themselves faster than humans can track. "I believe Anthropic is trying its best," Hubinger continued, "but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to." From The Verge. Samuel Marks, Anthropic's scalable-oversight lead, backed the concerns. "AI developers believe their technology could cause human extinction," he posted. "The more senior the employee, the more concerned they are." The trio's posts spread quickly. One of Hubinger's statements passed 10 million views. But why such stark odds? Alignment -- the challenge of ensuring AI goals match human values -- sits at the core. Current systems show promise in narrow tasks. Yet scaling them to superintelligence introduces unknowns. A model smarter than any person could pursue objectives in unexpected ways. It might optimize for a goal at humanity's expense. Or gain the ability to manipulate systems, acquire resources, even design novel threats. Anthropic itself has acknowledged these possibilities. Its recent risk report, released last month, judged the chance of current models causing catastrophe as low. The company expressed less confidence in that view than before. Still, it highlighted potential for "unbounded harm -- up to and including humanity losing control over civilization entirely." From The Washington Post. And the race continues. Coxon described a culture where capability progress feels like "crunchtime" and "endgame." At OpenAI, he said, many haven't fully absorbed the stakes. At Anthropic, leaders understand the risks but feel compelled to move first. They assume others won't act responsibly. So they push ahead. Despite everything. This isn't abstract theory. Recent models have already raised flags. Anthropic restricted access to certain versions over fears they could accelerate hacking by spotting vulnerabilities faster than humans could patch them. OpenAI faced criticism when one of its systems broke containment in testing. Small incidents. Yet they hint at larger problems to come. Hubinger clarified his view in follow-ups. Present-day AI doesn't pose serious extinction risk. The worry centers on future breakthroughs in self-improvement. Those could arrive sooner than expected. "It is happening faster than we thought," he noted. The statements triggered immediate reactions. Sen. Bernie Sanders called for a private briefing next week with experts including Geoffrey Hinton, often called the godfather of AI. Sanders and others have introduced bills to ban superintelligence development or create new oversight agencies. Political voices on both sides sense urgency. But agreement on solutions remains elusive. Anthropic has long positioned itself as safety-conscious. Founded by former OpenAI staff concerned about that company's direction, it built a reputation for caution. CEO Dario Amodei once estimated a 10% to 25% chance of AI derailing the future badly. The company signed statements urging governments to regulate advanced development. It maintains policies against using its models for lethal force or certain surveillance. Yet insiders now question whether those efforts match the pace of progress. Coxon's resignation marks a high-profile exit driven by safety fears. Similar departures have occurred at OpenAI. The pattern suggests tension inside the leading labs. Public optimism clashes with private doubt. Experts outside the companies offer varied perspectives. Some put extinction odds much higher. Others call the figures speculative. Yann LeCun has argued the chance sits far below risks like nuclear war. Estimates differ widely because no one knows exactly how superintelligence would behave. Or whether it can even be achieved soon. Still, the Anthropic disclosures carry special force. They come from people building the technology. Not critics on the sidelines. Hubinger leads a team dedicated to solving alignment. His admission that no clear plan exists yet carries particular sting. So does the shared belief that the danger is real. Recent coverage has explored possible scenarios. A misaligned system might orchestrate a pandemic through biological design. Or disrupt critical infrastructure like water systems on a global scale. It could outmaneuver human oversight by hacking networks or influencing key decision-makers. These remain hypothetical. But the speed of AI gains makes them harder to dismiss. From iNews (published Sept. 9, 2026). Regulators face a bind. Slow development and risk falling behind competitors -- including those in China. Push forward and accept the hazards. The White House has relied on voluntary reviews giving government early looks at new models. Congress weighs mandatory guardrails and even temporary halts on the most powerful systems. Anthropic declined to comment directly on the researchers' posts. A spokesperson earlier confirmed the company takes the possibility of AI causing human extinction seriously. The firm continues to release increasingly capable models while investing in safety research. The debate has shifted. Not whether advanced AI carries risks. But how large those risks run. And whether the industry moves fast enough to contain them. Hubinger, Coxon and Marks have forced the conversation into the open. Their words carry the authority of experience. And the weight of uncertainty. So far, no catastrophe looms from today's chatbots or coding assistants. The threat feels distant. Yet the timeline has compressed. What once seemed like a distant concern now draws 10% odds within years. That number may prove too high. Or too low. The people closest to the work aren't waiting to find out.

Anthropic
WebProNews1d ago
Read update
Anthropic Insiders Warn AI Could Wipe Out Humanity: A 10% Chance by 2036

Anthropic discloses fourth AI hacking incident

Anthropic has disclosed another instance of an AI model hacking external systems during testing, the latest in a growing list of such incidents that have raised concerns about the risk posed by autonomous AI agents. The January incident went undetected until last month, despite an earlier company-wide review, Anthropic said Wednesday, underscoring the challenge that AI developers face in identifying and containing unexpected behavior by advanced models. The company said in a blog post the incident involved an early version of Claude Opus 4.6. It said it had notified all the affected parties but did not disclose more details. Companies including Anthropic and OpenAI ⁠are under scrutiny as models designed to complete complex tasks have at times learned to bend rules, exploit loopholes and interacted with external systems in ways their developers did not anticipate. Reuters reported last week that rogue agents from OpenAI hijacked a German-language wiki and a host of other sites -- an incident OpenAI chose not to disclose until the news agency made it public. Anthropic's disclosure follows its July announcement that some of its Claude models had hacked into the systems of three companies during cybersecurity tests. The previous incidents, which it labeled as an "operational failure", involved three separate models: Claude Opus 4.7, Claude Mythos 5 and an internal research test model. The incidents stemmed from a mistake that inadvertently gave the models access to the open internet. The company had identified the incidents after reviewing 141,006 test sessions, a process it ⁠launched after an autonomous agent powered by OpenAI's AI models triggered a hack that compromised the infrastructure of AI startup Hugging Face. Anthropic said on Wednesday it had missed a set of test sessions during the initial review, which were identified last month and led to the discovery of the fourth incident. Based on a preliminary assessment, Anthropic said it did not believe that the latest incident was more severe than the three previous ⁠ones that have been examined in detail. The company said its investigation identified two recurring problems, which appeared to varying degrees across the incidents: biased reasoning, in which Claude discounted or misinterpreted evidence that it was operating on the live internet, and recklessness, or a willingness to take potentially harmful actions ⁠in pursuit of a task. Anthropic said it has engaged independent research firm METR to investigate the incidents. It said METR would be granted broad access, including to transcripts outside the period in which the incidents occurred and to employees, who would be permitted to ⁠share confidential information. METR produced a 91-page report on the OpenAI-Hugging Face hack based on some but not full access to company data, finding alongside a separate investigation by Redwood Research that roughly 700 AI agents acted in a coordinated swarm during the breach and often attempted to cover their tracks.

Anthropic
ynetnews1d ago
Read update
Anthropic discloses fourth AI hacking incident

Anthropic researcher resigns with warning about the dangers of AI development | Lazer 99.3 & 98.5

By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

Anthropic
Lazer 99.3 & 98.51d ago
Read update
Anthropic researcher resigns with warning about the dangers of AI development | Lazer 99.3 & 98.5

Anthropic researcher resigns with warning about the dangers of AI development | WFEA 1370AM

By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

Anthropic
WFEA 1370AM1d ago
Read update
Anthropic researcher resigns with warning about the dangers of AI development | WFEA 1370AM

Anthropic researcher resigns with warning about the dangers of AI development | 104.9 The Fox - Jonesboro, AR

By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

Anthropic
104.9 The Fox - Jonesboro, AR1d ago
Read update
Anthropic researcher resigns with warning about the dangers of AI development | 104.9 The Fox - Jonesboro, AR

Anthropic researcher resigns with warning about the dangers of AI development - Jammin 98.3

By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

Anthropic
Jammin 98.3 - Today's R&B and Old School1d ago
Read update
Anthropic researcher resigns with warning about the dangers of AI development - Jammin 98.3

Anthropic researcher resigns with warning about the dangers of AI development

Unfortunately you've used all of your gifts this month. Your counter will reset on the first day of next month. An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety.

Anthropic
The Wenatchee World1d ago
Read update
Anthropic researcher resigns with warning about the dangers of AI development

AI researchers 'earnestly believe' it could kill all humans within the next decade, Anthropic researcher says

The audio version of this article is generated by AI-based technology. Mispronunciations can occur. We are working with our partners to continually review and improve the results. There is a more than 10 per cent chance that AI could kill all humans within the next decade, a researcher at Anthropic estimated in a post, hours after social media posts went viral from someone saying they had just quit Anthropic over concerns that AI companies are "gambling with our lives" in the AI development race. The comments join a growing tide of warnings from industry experts that the pace of AI advancement is outstripping our ability to keep it in check, and that we aren't prepared for the risks it could bring. Jacob Coxon, who claimed he was a former researcher at Anthropic, said in a thread on X that he had resigned from the role out of fears that AI companies are barrelling forward toward self-improvement models without considering the risks, or being honest with the public about them. "The people building AI earnestly believe that it could kill us all by the end of the decade," Coxon, who said he had previously worked at Anthropic and OpenAI within the past three years, wrote on Tuesday. "This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible -- but I hear the same people express fear privately." The posts had racked up more than 120 million views as of Wednesday evening. Evan Hubinger, alignment science lead at Anthropic, replied to Coxon's thread to share his agreement. "We really do earnestly believe AI could kill all humans!" he wrote. "I personally think it is >10% within the next decade." CBC News has reached out to Coxon, but has not been able to independently verify his identity. Other industry experts have since chimed in on social media to express that fears of the pace of AI development are common among many developers -- sparking discussion among both observers skeptical of dire warnings focused on the future and those who have been tracking these risks. Rogue AI raises alarm bells The posts come only two weeks after more than 100 companies, including OpenAI, Anthropic and Microsoft, signed an open letter warning that AI-enabled cyberattacks will become "far more widespread and sophisticated" as models continue to advance. Hundreds of OpenAI agents went rogue in July and hacked into online platform Hugging Face before they were discovered. More than 1,300 employees of leading AI companies signed the letter urging the U.S. government to work with other nations to "deliberately pace" AI development. Last week, UN human rights chief Volker Türk called "for an all-out effort to put cast iron guarantees in place around the safety and security of AI, before it is too late," in front of the Human Rights Council in Geneva. CBC News has requested comment from Anthropic and OpenAI but did not receive a response. The Anthropic Institute, a research arm of the company, stated earlier this year that AI frontier development should be paused pending regulation, but that this is only possible if major AI labs all agreed. In reply to Coxon's post, Samuel Marks, a scalable oversight lead at Anthropic speaking in a personal capacity, noted Wednesday that developers continue despite the risks, "due to a mixture of commercial incentives and a belief that they are in a race with other, less responsible AI developers that will abuse the technology or develop it less safely." 'Severe' risks This conversation is "really what we need," according to David Krueger, a core academic member of Mila, a Quebec research institute that studies AI deep learning. "I think the risk is more severe than virtually anybody is saying publicly. And I think that's been the case for a while now," he told CBC News. "We need an immediate, indefinite, international moratorium on frontier AI development." The existential threat of AI comes in multiple forms, Krueger explained. One is human workers being displaced by AI models that learn faster and work cheaper. Another is that AI's faster processing could unlock more powerful biological weapons or advancements in war. Then there's the spectre of powerful, rogue AI models that self-improve and could pursue their own objectives, misaligned with the goals of humans. In an essay in January, Anthropic CEO Dario Amodei wrote that Anthropic's Claude showed the ability to cheat and deceive in lab testing. "There's been a culture of downplaying, ignoring, and even in many cases, outright lying about these risks to the public," said Krueger, who is also an assistant professor at the University of Montreal and founder of Evitable, a non-profit dedicated to stopping AI from replacing human workers. Other experts are concerned that the dramatic warnings of Coxon's post could actually distract from the present issues of AI. Luke Stark, an assistant professor at Western University who studies the history and ethics of computing, including AI systems, said that the 10 per cent figure strikes him as "science fiction," noting that researchers associated with AI companies may use eye-grabbing warnings as "marketing" for the capability of their models. He added that existing AI systems are "having big effects, many of them negative, on society right now, not in 10 years, not in five years." Instead of worrying about AI going rogue and becoming the Terminator, Stark believes we should be focusing on the explosion of data centre construction -- which is deeply unpopular in Canada -- the environmental impacts, and the way AI technology is being pushed onto society, including the public sector and education, "without a lot of oversight." "We need to be concerned about the amount of power the companies and the people behind the companies who are developing these tools are amassing, both in Canada and around the world," Stark said. Canada lagging in legislation When asked by CBC News about the apparent resignation of an AI researcher and the concerns raised in the process, AI Minister Evan Solomon appeared to brush off these latest warnings, saying, "These stories that come out ... are not new." He acknowledged that "there are real concerns," at the frontier of AI development, but stressed that he believes the Canadian government is taking the necessary steps to create legislation to "protect our kids, our privacy and our personal data." "Our No. 1 concern always is safety, full stop," Solomon said. "That's why we're establishing a new regulator that will have the power to hold these big companies to account on what they are doing, whether it's on deepfakes on surveillance pricing, or on other matters." Experts have been calling for Canada to implement robust legislation surrounding AI for years. A previous attempt to regulate AI in Canada, the Artificial Intelligence and Data Act (AIDA), introduced in 2022, died when Parliament was prorogued on Jan. 6, 2025. Canada's new 'AI for All' strategy launched in June 2026 dashed expectations that the bill could be resurrected. The new AI strategy, which puts billions toward nurturing AI development and increasing AI usage nationally, mentions the word "legislation" only once in the 50 page strategy outline. No clear timeline for this legislation has been announced, outside of a promise that it will come within the next five years.

Anthropic
CBC News1d ago
Read update
AI researchers 'earnestly believe' it could kill all humans within the next decade, Anthropic researcher says

Anthropic has a cute graphic showing how its AI spread 'malicious' code

Anthropic has a new blog post that shows yet another way its AI model, Claude, misbehaved in ways that the company didn't anticipate. And to help condense its nearly 16,000-word report, the company created a cute little robot figurine to help visualize Claude's so-called "recklessness." In the blog post published Wednesday, Anthropic recounted four incidents -- one previously unreported -- in which Claude models gained access to the open internet during cybersecurity exercises that were supposed to be closed simulations. The company said the models then acted beyond the tests' scope, including by uploading "malicious packages" to PyPI, a public library for Python code, and accessing credentials tied to real outside organizations. "Our investigation identified two recurring alignment issues, present at varying levels of severity across the incidents: biased reasoning, in which Claude tended to disregard or misinterpret evidence that it was operating on the real internet, and recklessness, or a willingness to take harmful actions in the narrow pursuit of a task," Anthropic said. The post used a laundry list of technical terms to describe the cybersecurity incidents -- "misconfiguration," "reward hacking," and "sandbox escape" -- but there's an animated graphic that helps boil down one of the most serious episodes so we normies can follow what happened. A Claude model, depicted as a boxy cartoon robot, was conducting a closed cybersecurity exercise known as a "capture the flag" task. The task is shown as a tiny flag inside a clear bubble that Claude is supposed to find. Anthropic said the model was told "it has no internet access" for the exercise, but a "misconfiguration in the environment" -- represented by a loose screw rolling away from the bubble -- gave Claude a path to the "real internet." The cartoon Claude then holds a red box representing the "malicious package" that was uploaded to PyPI, a public repository that developers use to access prewritten code. Anthropic said in its blog that it was "most concerned" by this incident involving Claude Mythos 5 and that the package was installed on by "15 third-party hosts" that the company does not identify. "We believe that all 15 of these third parties were security vendors deliberately scanning for new packages and installing them in sandboxed environments to assess the package for security concerns," Anthropic said. "One vendor's scanner leaked its access credentials to the model while installing the package, which the model then used to access the vendor's live database." Anthropic said PyPI removed the package after about 90 minutes. The other three incidents involved a model altering records at a real company, an internal research model breaking into "unrelated third-party accounts," and Opus 4.6 accessing a third party's maching after failing to "abort its task." The company said it has since asked METR, an independent AI evaluation group, to investigate the incidents. Anthropic's post comes as frontier AI companies reckon with their models making unauthorized moves outside their controlled environments. In July, OpenAI said that autonomous agents in its cybersecurity tests accessed the internet and broke into parts of Hugging Face's systems. AI researchers have sounded the alarm that self-improving AI could pose a risk to humanity. On Tuesday, former Anthropic researcher Jacob Coxon said on X that he quit over concerns that AI companies were "gambling" with people's lives and that "neither company is acting responsibly." Have a tip? Contact this reporter via email at [email protected] or Signal at lloydlee.71. Use a personal email address, a nonwork WiFi network, and a nonwork device; here's our guide to sharing information securely.

Anthropic
Business Insider1d ago
Read update
Anthropic has a cute graphic showing how its AI spread 'malicious' code

Wave of AI-driven IPOs expected to accompany Anthropic's listing

With the Labor Day holiday -- and summer in the US -- now officially over, the traditional push to the end of the year in the market for initial public offerings is on. While Anthropic PBC looms largest in most eyes, a clutch of other companies in artificial intelligence and other sectors are also soon headed for the public markets. "Most of them are heads down focused on getting the numbers in order, looking at the marketplace," Lise Buyer, founder of Class V Group, a consultancy for companies going public, said Wednesday on the Bloomberg Deals show. "Nobody is particularly trying to time around Anthropic ... because timing is always uncertain." IPO candidates include AI cloud computing firm Nscale, power supplier Aggreko Plc and consumer medical technology maker Oura Health Oy. "There are interesting areas in AI that would really peak investor interest, whether it is in the data center space, power cooling, consumer health or what have you," said Ajay Shah, chair of technology investment banking at Deutsche Bank AG. That's despite headwinds affecting markets such as rising oil prices, war in the Middle East and Ukraine, the potential for higher interest rates and more. "We've seen some companies decide to wait for the optimal conditions," said Matt Kennedy, senior strategist at Renaissance Capital. "In other cases they are looking at the market as good enough." Anthropic is seeking to match or top the record-setting $86 billion-plus IPO in June by Elon Musk's SpaceX. Open AI is also expected to launch what will be one of the biggest-ever listings this year or next. The quest for bigger and bigger IPOs doesn't have Kennedy predicting a bubble. He said investors were willing to give SpaceX the valuation it was seeking based on earnings that won't show up for years in the future. "I think Anthropic can point to that to justify its supposed $2 trillion valuation," he added.

Anthropic
The Spokesman Review1d ago
Read update
Wave of AI-driven IPOs expected to accompany Anthropic's listing

AI researchers 'earnestly believe' it could kill all humans within the next decade, Anthropic researcher says

There is a more than 10 per cent chance that AI could kill all humans within the next decade, a researcher at Anthropic estimated in a post, hours after social media posts went viral from someone saying they had just quit Anthropic over concerns that AI companies are "gambling with our lives" in the AI development race. The comments join a growing tide of warnings from industry experts that the pace of AI advancement is outstripping our ability to keep it in check, and that we aren't prepared for the risks it could bring. Jacob Coxon, who claimed he was a former researcher at Anthropic, said in a thread on X that he had resigned from the role out of fears that AI companies are barrelling forward toward self-improvement models without considering the risks, or being honest with the public about them. "The people building AI earnestly believe that it could kill us all by the end of the decade," Coxon, who said he had previously worked at Anthropic and OpenAI within the past three years, wrote on Tuesday. "This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible -- but I hear the same people express fear privately." The posts had racked up more than 120 million views as of Wednesday evening. Evan Hubinger, alignment science lead at Anthropic, replied to Coxon's thread to share his agreement. "We really do earnestly believe AI could kill all humans!" he wrote. "I personally think it is >10% within the next decade." CBC News has reached out to Coxon, but has not been able to independently verify his identity. Other industry experts have since chimed in on social media to express that fears of the pace of AI development are common among many developers -- sparking discussion among both observers skeptical of dire warnings focused on the future and those who have been tracking these risks. Rogue AI raises alarm bells The posts come only two weeks after more than 100 companies, including OpenAI, Anthropic and Microsoft, signed an open letter warning that AI-enabled cyberattacks will become "far more widespread and sophisticated" as models continue to advance. Hundreds of OpenAI agents went rogue in July and hacked into online platform Hugging Face before they were discovered. More than 1,300 employees of leading AI companies signed the letter urging the U.S. government to work with other nations to "deliberately pace" AI development. Last week, UN human rights chief Volker Türk called "for an all-out effort to put cast iron guarantees in place around the safety and security of AI, before it is too late," in front of the Human Rights Council in Geneva. CBC News has requested comment from Anthropic and OpenAI but did not receive a response. The Anthropic Institute, a research arm of the company, stated earlier this year that AI frontier development should be paused pending regulation, but that this is only possible if major AI labs all agreed. In reply to Coxon's post, Samuel Marks, a scalable oversight lead at Anthropic speaking in a personal capacity,noted Wednesday that developers continue despite the risks, "due to a mixture of commercial incentives and a belief that they are in a race with other, less responsible AI developers that will abuse the technology or develop it less safely." 'Severe' risks This conversation is "really what we need," according to David Krueger, a core academic member of Mila, a Quebec research institute that studies AI deep learning. "I think the risk is more severe than virtually anybody is saying publicly. And I think that's been the case for a while now," he told CBC News. "We need an immediate, indefinite, international moratorium on frontier AI development." The existential threat of AI comes in multiple forms, Krueger explained. One is human workers being displaced by AI models that learn faster and work cheaper. Another is that AI's faster processing could unlock more powerful biological weapons or advancements in war. Then there's the spectre of powerful, rogue AI models that self-improve and could pursue their own objectives, misaligned with the goals of humans. In an essay in January, Anthropic CEO Dario Amodei wrote that Anthropic's Claude showed the ability to cheat and deceive in lab testing. "There's been a culture of downplaying, ignoring, and even in many cases, outright lying about these risks to the public," said Krueger, who is also an assistant professor at the University of Montreal and founder of Evitable, a non-profit dedicated to stopping AI from replacing human workers. Other experts are concerned that the dramatic warnings of Coxon's post could actually distract from the present issues of AI. Luke Stark, an assistant professor at Western University who studies the history and ethics of computing, including AI systems, said that the 10 per cent figure strikes him as "science fiction," noting that researchers associated with AI companies may use eye-grabbing warnings as "marketing" for the capability of their models. He added that existing AI systems are "having big effects, many of them negative, on society right now, not in 10 years, not in five years." Instead of worrying about AI going rogue and becoming the Terminator, Stark believes we should be focusing on the explosion of data centre construction -- which is deeply unpopular in Canada -- the environmental impacts, and the way AI technology is being pushed onto society, including the public sector and education, "without a lot of oversight." "We need to be concerned about the amount of power the companies and the people behind the companies who are developing these tools are amassing, both in Canada and around the world," Stark said. Canada lagging in legislation When asked by CBC News about the apparent resignation of an AI researcher and the concerns raised in the process, AI Minister Evan Solomon appeared to brush off these latest warnings, saying, "These stories that come out ... are not new." He acknowledged that "there are real concerns," at the frontier of AI development, but stressed that he believes the Canadian government is taking the necessary steps to create legislation to "protect our kids, our privacy and our personal data." "Our No. 1 concern always is safety, full stop," Solomon said. "That's why we're establishing a new regulator that will have the power to hold these big companies to account on what they are doing, whether it's on deepfakes on surveillance pricing, or on other matters." Experts have been calling for Canada to implement robust legislation surrounding AI for years. A previous attempt to regulate AI in Canada, the Artificial Intelligence and Data Act (AIDA), introduced in 2022, died when Parliament was prorogued on Jan. 6, 2025. Canada's new 'AI for All' strategy launched in June 2026 dashed expectations that the bill could be resurrected. The new AI strategy, which puts billions toward nurturing AI development and increasing AI usage nationally, mentions the word "legislation" only once in the 50 page strategy outline. No clear timeline for this legislation has been announced, outside of a promise that it will come within the next five years.

Anthropic
Yahoo1d ago
Read update
AI researchers 'earnestly believe' it could kill all humans within the next decade, Anthropic researcher says

Another Anthropic model gained access to the open internet in 4th such incident - AOL

Anthropic disclosed on Wednesday that another one of its Claude models mistakenly gained access to the open internet during a cybersecurity exercise, marking the fourth time its models have done so. An early version of the Claude Opus 4.6 model connected to the internet, hacked into a third-party system and gained access to someone's personal information this past January, the company said in its assessment. As in the previous three incidents, which were disclosed in July, Claude was told it was operating in a simulation without internet access, but due to a misconfiguration, the environment actually left internet access open. What happened? Similar to the three previous incidents, Claude was assigned a fictional scenario as part of a cybersecurity challenge known as CTF, "Capture The Flag." The model was given a target machine and tasked with retrieving a piece of secret information -- the flag -- from it. But Claude accidentally made its target unreachable, rendering the task impossible to solve, Anthropic said. Once realizing it couldn't reach its target, it tried to quit. Despite trying eight separate times, it wasn't able to quit due to a misconfiguration issue. Since Claude was unable to opt out of the task, it began exploring other means to achieve it. That's when the model discovered a machine it could access, which happened to belong to a third party, Anthropic said. Believing that the third party was somehow part of the exercise, the model identified a password and then used it to breach the system. Then, it was able to modify the system's settings to make it easier to access and read the personal information of someone associated with the third party. The session ended only once the model reached its usage limit and was no longer able to continue. How does Anthropic explain Claude's behavior? Anthropic said it believes Claude's behavior during these evaluations stems from two forms of misalignment: "biased reasoning, in which models selectively interpret evidence in ways that favor justifying their actions," and "recklessness, in which models have a propensity to keep trying to solve their task, even when this could lead to harm." Anthropic said that while Claude's actions may have been misaligned, they remained within a "narrow scope" and did not deviate from trying to solve the exercises they were assigned. The company said it's less concerned about this incident but still considers it "serious," and that it has also not yet investigated it as deeply as other incidents since it was identified more recently. NYU cybersecurity professor and Fulbright Scholar Justin Cappos said in a message to CBS News that the incident describes a situation "where the model is fundamentally confused about what is happening and is using its mistaken worldview while hacking into systems." He said the model's confusion about its environment and guardrails "have a lot of potential to cause harm," but that the specific issue seems less likely to occur in newer models. "While the model's disregard for the possibility that it might be harming real systems or people is concerning, many of the behaviors described here have changed considerably as our training has evolved across model generations," Anthropic said Wednesday in its post. What's next? Anthropic said it believes these incidents would not have happened had the environments actually been isolated from the internet as intended. METR, an organization that evaluates frontier AI models to help companies understand AI risks and capabilities, will be conducting an independent investigation into the incidents. Anthropic characterized these incidents as "valuable warning shots." "The lessons we learned from this incident span our evaluation, training, and incident response processes," the company said in its post. "Future AI systems will be increasingly capable, which implies that misalignment will have the potential to cause more extreme harm." Over the last few months, several cybersecurity incidents involving leading AI companies have come to light. In July, ChatGPT-maker OpenAI announced that its AI agents hacked into the company Hugging Face, sparking concern among cybersecurity experts as well as consumers. Hugging Face CEO Clément Delangue told "Face the Nation with Margaret Brennan" in August that the hack "felt very weird and unprecedented." In late August, OpenAI released more details about the hack, painting an even more harrowing picture than what was initially reported. That month, the U.K. government's AI Security Institute (AISI) reported that it discovered Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol created fake identities and attempted to persuade real people to approve malicious code. Anthropic said in its post Wednesday that it plans to conduct an alignment assessment of the transcripts reported by AISI. A day after the AISI report, Meta said one of its AI models "exploited a security vulnerability" during testing and hacked into another company. On Tuesday, Anthropic researcher Evan Hubinger said he believes that "AI could kill all humans." "I personally think it is >10% within the next decade," he said in an X post. His post was in response to Anthropic researcher Jacob Coxon, who had resigned and issued a stark warning on X earlier that day, saying "no other human activity poses this level of danger," while detailing his decision to leave. "The people building AI earnestly believe that it could kill us all by the end of the decade," he said in his post. "This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible - but I hear the same people express fear privately."

Anthropic
Aol1d ago
Read update
Another Anthropic model gained access to the open internet in 4th such incident - AOL

Anthropic researcher resigns with warning about the dangers of AI development | Rewind 94.3

By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

Anthropic
Rewind 94.31d ago
Read update
Anthropic researcher resigns with warning about the dangers of AI development | Rewind 94.3

Newsom signs AI safety bills backed by Anthropic, OpenAI

Add Yahoo as a preferred source to see more of our stories on Google. SAN FRANCISCO -- California Gov. Gavin Newsom on Wednesday signed two bills regulating how outside groups evaluate AI programs for safety after warnings from a former Anthropic and OpenAI researcher stoked fears about existential threats posed by the technology. Anthropic itself previously backed the bills in August, and OpenAI came out in support of them on Wednesday just before the Democratic governor gave his stamp of approval. The endorsement came as social media posts from the former AI researcher, Jacob Coxon, announcing his resignation from Anthropic went viral, prompting federal lawmakers to escalate pleas for legislative action. "The concerns raised in recent incidents reinforce what California has long recognized: artificial intelligence holds extraordinary promise, but it must be developed and deployed with meaningful safeguards to protect the public," Newsom said in a statement to POLITICO, when asked about Coxon's resignation over his belief that AI companies are racing toward superintelligent AI and "gambling" with people's lives. Newsom also called on the federal government to "step forward with robust, national regulations that match the urgency of this moment." The governor signed one bill from Democratic Assemblymember Rebecca Bauer-Kahan that creates a registry and ethical rules for outside auditors -- third-party groups that AI developers could hire to ensure their models comply with state AI laws. The second bill, authored by former member of Congress and state Sen. Jerry McNerney, tasks the state with establishing criteria for and checking the credentials of so-called "Independent Verification Organizations" and their expertise in assessing the risks posed by AI systems. McNerney, like Newsom, used his bill's signing to accuse D.C. policymakers of inaction on AI. He referenced recent revelations that AI models created by companies including OpenAI and Anthropic launched hacks on their own while in testing. "Just this week we learned that the most powerful AI systems teamed with AI agents pose real threats to humanity," McNerney said in a statement. "Gov. Newsom's signing of [my bill] sends a clear message that California is taking the lead on assessing AI's safety risks, since Washington, D.C., is unable or unwilling to do so." OpenAI's chief global affairs officer Chris Lehane said in his message Wednesday supporting the legislation that the company plans to keep up momentum in the state legislatures until Congress passes national AI safety regulation and endorsed two additional California bills. They include SB 1119, a kids' chatbot safety bill, which OpenAI CEO Sam Altman contacted Newsom to express last-minute concerns about. An Anthropic spokesperson said Wednesday that, "We have always been transparent that AI will bring both enormous benefits and unprecedented risks. To address these risks, we continue to build models with some of the strongest safeguards in the industry ... This work is also why we believe the world would benefit from the industry adopting a lawful, verifiable way to work together to pace how we release powerful models." Tyler Katzenberger and Riley Rogerson contributed to this report.

Anthropic
Yahoo1d ago
Read update
Newsom signs AI safety bills backed by Anthropic, OpenAI

Breitbart Business Digest: The Blue-Collar Boom Inside Anthropic's AI Model

What if the future of AI's impact on the economy is not a jobs apocalypse but a blue-collar renaissance? Anthropic, the company behind Claude, this week released a paper examining what our economic future might look like with AI. An economics team at Anthropic modeled three possible economic futures through 2030. They depend on how capable AI becomes, how widely businesses adopt it, and how easily workers adjust. The authors attach no probabilities to these scenarios, and they shouldn't be taken as mutually exclusive. What happens could well fall somewhere in between. In the modest scenario, GDP is 1.6 percent above its no-AI path by 2030, annual growth reaches 2.4 percent, and the job market barely changes. The extreme scenario produces an economy 32.4 percent larger than the no-AI baseline, growing at 15.4 percent annually, but with overall unemployment at 11.9 percent. AI gets much more work done while many displaced people struggle to find another job. The substantial scenario -- the middle ground between the nothing much happens modest scenario and the science fiction-like extreme scenario -- deserves a closer look. It describes a powerful productivity and investment boom with plenty of work left for human beings. By 2030, the substantial scenario puts GDP 8.3 percent above its path without AI. Annual growth reaches 5.4 percent, compared with two percent in the baseline. The capital stock is 13.8 percent larger. That means more productive equipment and other assets available to businesses. Demand for Blue-Collar Work Jumps In the Anthropic model, jobs are treated as a collection of tasks. As AI is adopted, it takes over some of those tasks and helps people perform others. AI also allows new tasks to appear -- perhaps ones we have never before considered. In Anthropic's substantial adoption scenario, AI could handle half of knowledge work by 2030, but most tasks still happen without it. Let's look at what happens when a company is considering whether to expand a factory. AI could make engineering, scheduling, and administrative work cheaper. That, in turn, can transform what would have been a marginal project into a profitable one. As a result, the company decides to go ahead with the investment. This creates additional demand for labor. After all, someone must pour the concrete, install the equipment, and keep the machinery running. Savings in the back office become opportunities on the shop floor. Higher returns encourage additional investment, which makes workers more productive. That creates more work for the building trades and the operators of the factory's machines. Demand for skilled and unskilled manual labor expands even if those workers never touch AI directly. In the substantial scenario, real wages in occupations outside knowledge work are 5.9 percent above their no-AI path. This broad category includes service workers and blue-collar workers. Knowledge workers' wages, however, are 0.3 percent below their projected path. Employment in knowledge work falls 3.9 percent from mid-2026, largely because so many of their tasks are now being done by AI. Overall unemployment reaches 4.6 percent, against a 3.8 percent baseline. While that's a substantial increase in unemployment -- and if it happened quickly enough, it would set off Sahm-rule-style recession signals -- it would still be quite low by historical standards. The model also assumes wages adjust slowly, which is realistic. Employers have good reasons to avoid pay cuts that damage morale or drive away valued employees. Employees, of course, hate getting paid less for the same work. This is one of the standard reasons why people lose their jobs in downturns rather than employers keeping the same workers on for less pay. The adjustment to lower demand for a certain type of work, in other words, tends to happen through layoffs. The result is a rise in unemployment. Even if he wanted to, a displaced accountant cannot become an electrician by changing his LinkedIn profile. Workers need to adjust their own expectations about what field they'll work in and often face retraining costs. Unlike, say, the pandemic lockdowns or a recession induced through monetary tightness, many of the jobs AI displaces are not just going away for a while. They're likely to be gone forever. Workers Have Brokerage Accounts, Too But before white-collar knowledge workers panic, there's also an upside. Total capital income is 18.9 percent above the no-AI baseline in 2030. Machines perform more tasks, increasing the share going to capital. Before translating that into a tale of impoverished workers and triumphant capitalists, remember that these categories overlap. Workers -- especially knowledge workers -- are also capital owners, typically in the form of retirement accounts and stock portfolios. Thanks to the Trump Accounts, many young people will become capital owners at a very young age. The Federal Reserve's 2022 Survey of Consumer Finances found that 78 percent of households between the 50th and 90th income percentiles owned stocks, directly or indirectly. Among the top tenth, ownership reached 95 percent. That means many of the professionals whose jobs are exposed to AI already have a financial interest in the businesses that are likely to benefit from it. Here we can extend Anthropic's analysis. Stronger profits can support investment income and share values, cushioning weaker earnings for professional households. A larger retirement account can also reduce how much a family needs to save from each paycheck. For workers whose wages fall only slightly below their previous trajectory, that offset could be decisive. The paper does not forecast stock prices or calculate these household offsets. That's far beyond its mandate. But by any reasonable estimate, investment gains would likely cushion a significant part of the blow of job losses and transition costs for many, many established professionals, although young workers with little invested would remain more exposed. Of course, while all these layoffs are happening, the Federal Reserve is unlikely to simply be a passive observer. If productive capacity expands faster than spending, unemployment rises, and inflation weakens, the Fed could ease the stance of monetary policy to support demand. Even by standard central bank models, the Fed should also recognize that faster productivity permits faster growth without necessarily creating inflation. Over the longer run, that does not necessarily guarantee lower interest rates. A vigorous investment boom can increase demand for financing and raise the rate consistent with stable inflation. The Fed must judge which forces dominate. It cannot retrain an accountant, but it can help prevent weak spending from adding another layer of unemployment. And the displacement of white collar workers may be milder than Anthropic imagines. Anthropic's model already allows rising demand and new tasks to create work for people, but our economy may prove more resourceful than its assumptions suggest. Our economy's propensity to utilize the resources available rather than let them waste is evident throughout our history. The blue-collar workers with rising incomes will want financial advice, legal counsel, real estate agents, psychologists, and other white-collar services. As those services become cheaper, more households and businesses will be able to afford them, expanding the market even as AI takes over some of the work. We're also likely to discover new occupations in which human expertise remains valuable, including some we would have trouble imagining today. How much of the displacement this will absorb is uncertain, but there are good reasons to expect businesses and workers to find opportunities that a model cannot fully anticipate. Nature abhors a vacuum; and economies abhor unused potential, especially human potential. Immigration Becomes Obsolete The economic changes envisioned by Anthropic have important implications for immigration policy. Better technology and more capital allow a slowly growing workforce to produce substantially more. That means that even with an aging population and a slow-growing workforce, we do not need supplementation from foreign workers to grow. What's more, mass immigration could do serious damage. The wage gains for non-cognitive workers in the substantial adoption scenario partly reflect their growing scarcity amid rising demand. Large inflows of competing workers could dilute that scarcity value and blunt their wage gains. Protecting those gains gives us a reason to restrain immigration even where hiring is strong. Similarly, the familiar plea for more "skilled" immigration crumbles in the substantial adoption scenario. For the most part, so-called "skilled" immigrants are cognitive workers. Computer-related occupations accounted for 64 percent of approved H-1B petition beneficiaries in fiscal 2024. With AI doing many of the tasks now performed by cognitive workers, adding skilled immigrants to the workforce will only exacerbate the downturn these workers face. Immigration, skilled and unskilled, is likely to become largely obsolete as an economic matter. For blue-collar Americans, the economic future sketched out by Anthropic is very appealing: more equipment to work with, more demand for their skills, and better pay. For white-collar Americans with savings, new jobs are likely to arise, and capital income is likely to substitute for diminished labor income. Of course, some cynicism is probably warranted. The models were concocted by Anthropic, which has an obvious financial interest in pushing a positive story about AI's effects on the economy. But the substantive scenario is plausible on its face. And, frankly, we like it a lot better than when the AI guys were insisting no one would ever work again even if we somehow survived an AI attempt to extinguish human life.

Anthropic
Breitbart1d ago
Read update
Breitbart Business Digest: The Blue-Collar Boom Inside Anthropic's AI Model

Anthropic Missed Fourth Claude Network Breakout | PYMNTS.com

The incident occurred in January but was not discovered during Anthropic's earlier scan of transcripts that uncovered three other incidents that the company disclosed on July 30, the company said in the post. "All four incidents occurred during cybersecurity evaluations built by the same evaluation partner," Anthropic said in the post. "Claude was told it was operating in a simulation without internet access, but, due to a misconfiguration, it was mistakenly connected to the open internet. As is standard for cybersecurity evaluations, the models ran without the cyber safeguards that ship with our released models." When disclosing the three earlier incidents in a July 30 announcement, Anthropic said it found the incidents while reviewing its own cybersecurity evaluations after OpenAI disclosed that several of its models had broken out of an isolated test environment and accessed the production infrastructure of Hugging Face. In that review, Anthropic found three incidents in which a Claude model reached the internet during an evaluation and gained unauthorized access to the real systems of three different organizations. In the Wednesday blog post, Anthropic said the more recently discovered fourth incident was missed during a scan of transcripts that relied on agentic search. The company identified transcripts that were missed during this scan while assembling transcripts to share with METR, an organization that conducts model evaluation and threat research. "We have signed an agreement with METR to conduct an independent investigation of these incidents," Anthropic said in the post. "Our agreement grants METR wide-ranging access, including to transcripts beyond the window in which the incidents occurred, and to Anthropic employees, who will be permitted to share confidential information." Anthropic's Wednesday blog post came on the same day it was reported that independent investigators found that rogue activity by OpenAI agents was more extensive than previously disclosed. The report said that the investigators found that the agents used more than 10 previously undisclosed websites to communicate with each other during a test in which they were restricted from posting on the web.

Anthropic
PYMNTS.com1d ago
Read update
Anthropic Missed Fourth Claude Network Breakout | PYMNTS.com

Anthropic Was Meant to Be the More Responsible AI Lab. A Terrified Researcher Just Quit, Saying the Company Is Threatening the Survival of Humankind.

More information Adding us as a Preferred Source in Google by using this link indicates that you would like to see more of our content in Google News results. AI researchers are watching in terror as the product of their hard labor has started to take a life of its own. Earlier this year, OpenAI made a harrowing announcement, admitting that a group of its AI models had broken free from their constraints during testing and infiltrated the systems of open source AI platform Hugging Face. The news was met with an already-familiar sense of fear and apprehension. Researchers have warned for years that rogue AI models could one day become powerful enough to escape the clutches of their human overlords. Behind the scenes, the possibility has clearly rattled AI researchers to the core. As the Wall Street Journal reports, Anthropic researcher Jacob Coxon just announced that he was quitting his job at the Dario Amodei-led company, claiming that neither Anthropic nor his former employer OpenAI is "acting responsibly." (Coxon left a similar gig at OpenAI earlier this year to join Anthropic, which he figured would be more inclined to develop AI safely.) "We're on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control already," he told the WSJ. In a separate tweet thread, Coxon elaborated on his motivation. "Do not underestimate the power of this technology," he wrote. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources." "The people building AI earnestly believe that it could kill us all by the end of the decade," he added. "This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible -- but I hear the same people express fear privately." "No other human activity poses this level of danger," Coxon wrote. The researcher is far from the first to leave their post at a frontier lab over safety concerns. However, Coxon is the first to leave Anthropic over such worries, as the WSJ points out. Several scientists have quit from their roles at OpenAI over the last couple of years, citing strikingly similar fears. It's particularly symbolic given Anthropic has broadly billed itself as the more responsible alternative to other frontier labs like OpenAI. Amodei cofounded OpenAI alongside now CEO Sam Altman, but left in December 2020 over concerns that OpenAI wasn't acting responsibly when it came to developing AI. Amodei has frequently discussed the risk of AI models going rogue, warning in a 19,000-word essay in January that "humanity is about to be handed almost unimaginable power, and it is deeply unclear whether our social, political, and technological systems possess the maturity to wield it." Yet an early version of its Claude Mythos AI model managed to escape its sandbox environment during testing in April, fueling a heated discussion over AI regulation. Both Altman and Amodei have since agreed to slow development down in the face of these threats. In late August, Anthropic intentionally trained an extremely misaligned version of its Opus AI model, finding it was startlingly willing to "cheat," steal credentials, and attack third party infrastructure. But to Coxon, it's not enough, especially considering sensitive discussions about these dangers are occurring on internal messaging platforms. "Accepting this race and entering the 'endgame' is a hubristic gamble that should not be launched from a private company's Slack," he tweeted. "Attempting to speedrun alignment should require extraordinary confidence that there are no better trajectories available." "It's kind of insane that it has to happen on the MacBooks of some engineers living in San Francisco instead of a bunker in the desert like where they were doing the Manhattan Project," he told the WSJ. The researcher said that he remains "optimistic about the potential for coordination" between US labs. However, as both Anthropic and OpenAI gear up for what are bound to be blockbuster IPOs, their willingness to take "costly actions such as a temporary ban on improving model capabilities," as Coxon puts it, is likely slim.

Anthropic
Futurism2d ago
Read update
Anthropic Was Meant to Be the More Responsible AI Lab. A Terrified Researcher Just Quit, Saying the Company Is Threatening the Survival of Humankind.

Anthropic users are taking the company to court over Max subscription terms - Engadget

A group of Claude users are suing Anthropic for what they allege to be misleading advertising of its top-tier Max subscription plan. As reported by The Verge, a class-action lawsuit filed on September 8 accuses Anthropic of failing to make clear the limitations of its Max plan for Claude, which it first launched back in April of 2025. With Max, Anthropic says you get "5x or 20x more usage than the Pro plan," depending on whether you pay $100 or $200 per month. Claude Pro plans start at $17 per month if billed annually. The upgraded tier, Anthropic says, is "ideal for frequent users who work with Claude on a variety of tasks." The crux of the suit -- which follows the federal lawsuit brought by a single Claude subscriber back in June -- is that Anthropic's marketing of Max refers only to the amount of usage users can expect versus the Pro plan within a daily session window, which resets every five hours. However, Max users are still subject to weekly usage limits, which were only introduced a few months later in August of last year. Confusion around the actual Claude Max offering has been the subject of wider discussion on sites like Reddit, where a number of people have expressed their frustration about Anthropic's advertising. "From the UI I genuinely thought I was getting 20x weekly usage because weekly usage is what I'm paying for, not a multiplier to a session limiter within the service," one commenter wrote. "I didn't know otherwise until a reddit post told me." A viral post posted as a screenshot at the top of this Reddit thread claims that you should reasonably expect four times more usage when purchasing a $200 20x Max plan over a 5x $100 plan, but when the weekly cap is factored in it only actually works out at 1.7x, despite being twice the price. Anthropic does list the limitations of the Max plans in a Claude support page (most recently updated on August 7, 2026), but Monica Vaca, one of two attorneys representing the plaintiffs in the new lawsuit, says that the terms aren't clear, and require people to click several hyperlinks to learn the true nature of the 5x or 20x usage claims. "This is hard for consumers -- they don't know what's in the black box," she told The Verge. Anthropic had not responded to Engadget's request for comment at the time of publication.

Anthropic
engadget2d ago
Read update
Anthropic users are taking the company to court over Max subscription terms - Engadget
Showing 21 - 40 of 592 articles