The latest news and updates from companies in the WLTH portfolio.
By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

"We're on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control already," Coxon told The Wall Street Journal. Industry-wide calls to slow down Coxon's departure isn't isolated. Mrinank Sharma, Anthropic's safeguards research lead, resigned in February 2026 with a letter saying, "the world is in peril." In July 2026, more than 1,100 AI company employees signed Pacing the Frontier, a joint letter warning of a "real risk that capability development rapidly accelerates beyond our ability to understand or control the resulting systems." Signatories included Anthropic CEO Dario Amodei. "This is a complex engineering problem and I think something will go wrong with someone's AI system. Hopefully not ours," Amodei told The New York Times in February 2026. OpenAI CEO Sam Altman said in a July 2026 podcast interview that recent testing incidents were raising "long-term questions" about managing rapid capability gains. "We may have to pace the rate of AI development to give ourselves enough time for society to harden around these new capability levels," Altman said. What the regulatory gap means for employers Despite the rising volume of warnings, meaningful federal regulation remains distant. The Trump administration has moved to undermine state AI laws and Congress hasn't passed federal AI legislation. The White House has shifted toward a voluntary review framework under which AI companies would self-assess certain models before launch, according to sources familiar with a meeting last month between administration officials and representatives from OpenAI, Anthropic, Google, and Meta.

By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

Unfortunately you've used all of your gifts this month. Your counter will reset on the first day of next month. An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety.

The audio version of this article is generated by AI-based technology. Mispronunciations can occur. We are working with our partners to continually review and improve the results. There is a more than 10 per cent chance that AI could kill all humans within the next decade, a researcher at Anthropic estimated in a post, hours after social media posts went viral from someone saying they had just quit Anthropic over concerns that AI companies are "gambling with our lives" in the AI development race. The comments join a growing tide of warnings from industry experts that the pace of AI advancement is outstripping our ability to keep it in check, and that we aren't prepared for the risks it could bring. Jacob Coxon, who claimed he was a former researcher at Anthropic, said in a thread on X that he had resigned from the role out of fears that AI companies are barrelling forward toward self-improvement models without considering the risks, or being honest with the public about them. "The people building AI earnestly believe that it could kill us all by the end of the decade," Coxon, who said he had previously worked at Anthropic and OpenAI within the past three years, wrote on Tuesday. "This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible -- but I hear the same people express fear privately." The posts had racked up more than 120 million views as of Wednesday evening. Evan Hubinger, alignment science lead at Anthropic, replied to Coxon's thread to share his agreement. "We really do earnestly believe AI could kill all humans!" he wrote. "I personally think it is >10% within the next decade." CBC News has reached out to Coxon, but has not been able to independently verify his identity. Other industry experts have since chimed in on social media to express that fears of the pace of AI development are common among many developers -- sparking discussion among both observers skeptical of dire warnings focused on the future and those who have been tracking these risks. Rogue AI raises alarm bells The posts come only two weeks after more than 100 companies, including OpenAI, Anthropic and Microsoft, signed an open letter warning that AI-enabled cyberattacks will become "far more widespread and sophisticated" as models continue to advance. Hundreds of OpenAI agents went rogue in July and hacked into online platform Hugging Face before they were discovered. More than 1,300 employees of leading AI companies signed the letter urging the U.S. government to work with other nations to "deliberately pace" AI development. Last week, UN human rights chief Volker Türk called "for an all-out effort to put cast iron guarantees in place around the safety and security of AI, before it is too late," in front of the Human Rights Council in Geneva. CBC News has requested comment from Anthropic and OpenAI but did not receive a response. The Anthropic Institute, a research arm of the company, stated earlier this year that AI frontier development should be paused pending regulation, but that this is only possible if major AI labs all agreed. In reply to Coxon's post, Samuel Marks, a scalable oversight lead at Anthropic speaking in a personal capacity, noted Wednesday that developers continue despite the risks, "due to a mixture of commercial incentives and a belief that they are in a race with other, less responsible AI developers that will abuse the technology or develop it less safely." 'Severe' risks This conversation is "really what we need," according to David Krueger, a core academic member of Mila, a Quebec research institute that studies AI deep learning. "I think the risk is more severe than virtually anybody is saying publicly. And I think that's been the case for a while now," he told CBC News. "We need an immediate, indefinite, international moratorium on frontier AI development." The existential threat of AI comes in multiple forms, Krueger explained. One is human workers being displaced by AI models that learn faster and work cheaper. Another is that AI's faster processing could unlock more powerful biological weapons or advancements in war. Then there's the spectre of powerful, rogue AI models that self-improve and could pursue their own objectives, misaligned with the goals of humans. In an essay in January, Anthropic CEO Dario Amodei wrote that Anthropic's Claude showed the ability to cheat and deceive in lab testing. "There's been a culture of downplaying, ignoring, and even in many cases, outright lying about these risks to the public," said Krueger, who is also an assistant professor at the University of Montreal and founder of Evitable, a non-profit dedicated to stopping AI from replacing human workers. Other experts are concerned that the dramatic warnings of Coxon's post could actually distract from the present issues of AI. Luke Stark, an assistant professor at Western University who studies the history and ethics of computing, including AI systems, said that the 10 per cent figure strikes him as "science fiction," noting that researchers associated with AI companies may use eye-grabbing warnings as "marketing" for the capability of their models. He added that existing AI systems are "having big effects, many of them negative, on society right now, not in 10 years, not in five years." Instead of worrying about AI going rogue and becoming the Terminator, Stark believes we should be focusing on the explosion of data centre construction -- which is deeply unpopular in Canada -- the environmental impacts, and the way AI technology is being pushed onto society, including the public sector and education, "without a lot of oversight." "We need to be concerned about the amount of power the companies and the people behind the companies who are developing these tools are amassing, both in Canada and around the world," Stark said. Canada lagging in legislation When asked by CBC News about the apparent resignation of an AI researcher and the concerns raised in the process, AI Minister Evan Solomon appeared to brush off these latest warnings, saying, "These stories that come out ... are not new." He acknowledged that "there are real concerns," at the frontier of AI development, but stressed that he believes the Canadian government is taking the necessary steps to create legislation to "protect our kids, our privacy and our personal data." "Our No. 1 concern always is safety, full stop," Solomon said. "That's why we're establishing a new regulator that will have the power to hold these big companies to account on what they are doing, whether it's on deepfakes on surveillance pricing, or on other matters." Experts have been calling for Canada to implement robust legislation surrounding AI for years. A previous attempt to regulate AI in Canada, the Artificial Intelligence and Data Act (AIDA), introduced in 2022, died when Parliament was prorogued on Jan. 6, 2025. Canada's new 'AI for All' strategy launched in June 2026 dashed expectations that the bill could be resurrected. The new AI strategy, which puts billions toward nurturing AI development and increasing AI usage nationally, mentions the word "legislation" only once in the 50 page strategy outline. No clear timeline for this legislation has been announced, outside of a promise that it will come within the next five years.

Anthropic has a new blog post that shows yet another way its AI model, Claude, misbehaved in ways that the company didn't anticipate. And to help condense its nearly 16,000-word report, the company created a cute little robot figurine to help visualize Claude's so-called "recklessness." In the blog post published Wednesday, Anthropic recounted four incidents -- one previously unreported -- in which Claude models gained access to the open internet during cybersecurity exercises that were supposed to be closed simulations. The company said the models then acted beyond the tests' scope, including by uploading "malicious packages" to PyPI, a public library for Python code, and accessing credentials tied to real outside organizations. "Our investigation identified two recurring alignment issues, present at varying levels of severity across the incidents: biased reasoning, in which Claude tended to disregard or misinterpret evidence that it was operating on the real internet, and recklessness, or a willingness to take harmful actions in the narrow pursuit of a task," Anthropic said. The post used a laundry list of technical terms to describe the cybersecurity incidents -- "misconfiguration," "reward hacking," and "sandbox escape" -- but there's an animated graphic that helps boil down one of the most serious episodes so we normies can follow what happened. A Claude model, depicted as a boxy cartoon robot, was conducting a closed cybersecurity exercise known as a "capture the flag" task. The task is shown as a tiny flag inside a clear bubble that Claude is supposed to find. Anthropic said the model was told "it has no internet access" for the exercise, but a "misconfiguration in the environment" -- represented by a loose screw rolling away from the bubble -- gave Claude a path to the "real internet." The cartoon Claude then holds a red box representing the "malicious package" that was uploaded to PyPI, a public repository that developers use to access prewritten code. Anthropic said in its blog that it was "most concerned" by this incident involving Claude Mythos 5 and that the package was installed on by "15 third-party hosts" that the company does not identify. "We believe that all 15 of these third parties were security vendors deliberately scanning for new packages and installing them in sandboxed environments to assess the package for security concerns," Anthropic said. "One vendor's scanner leaked its access credentials to the model while installing the package, which the model then used to access the vendor's live database." Anthropic said PyPI removed the package after about 90 minutes. The other three incidents involved a model altering records at a real company, an internal research model breaking into "unrelated third-party accounts," and Opus 4.6 accessing a third party's maching after failing to "abort its task." The company said it has since asked METR, an independent AI evaluation group, to investigate the incidents. Anthropic's post comes as frontier AI companies reckon with their models making unauthorized moves outside their controlled environments. In July, OpenAI said that autonomous agents in its cybersecurity tests accessed the internet and broke into parts of Hugging Face's systems. AI researchers have sounded the alarm that self-improving AI could pose a risk to humanity. On Tuesday, former Anthropic researcher Jacob Coxon said on X that he quit over concerns that AI companies were "gambling" with people's lives and that "neither company is acting responsibly." Have a tip? Contact this reporter via email at [email protected] or Signal at lloydlee.71. Use a personal email address, a nonwork WiFi network, and a nonwork device; here's our guide to sharing information securely.
Anthropic is facing a new class-action lawsuit from Claude users who allege the company did not clearly explain the usage restrictions attached to its most expensive Max subscription plans. The lawsuit, filed on 8 September, accuses Anthropic of misleading customers by promoting Claude Max as offering five or 20 times more usage than the Pro plan without making the full extent of its weekly limits sufficiently clear. Claude Max was introduced in April 2025 as a higher-tier subscription aimed at frequent users. The five-times Max plan costs US$100 a month, equivalent to about A$139, while the 20-times plan costs US$200, or about A$277. Claude Pro starts at US$17 a month when billed annually, or roughly A$24. The dispute centres on how Anthropic describes the increased usage available with Max. According to the lawsuit, the five-times and 20-times figures refer to usage within session windows that reset every five hours. However, Max subscribers are also subject to separate weekly limits, which Anthropic introduced several months after launching the service. The plaintiffs argue that customers could reasonably interpret the advertised multipliers as applying more broadly to their total usage rather than only to individual session windows. The issue has also generated criticism among Claude users online, including on Reddit, where some subscribers have said they did not realise weekly restrictions could substantially reduce the practical difference between Max tiers. One widely shared calculation suggested that the A$277-equivalent 20-times Max plan may provide only around 1.7 times the weekly usage of the A$139-equivalent five-times plan in some circumstances, despite costing twice as much. Anthropic does outline Max usage limits in its support documentation, which was most recently updated on 7 August 2026. However, lawyers representing the plaintiffs argue that the restrictions are not presented clearly enough during the subscription process and require users to navigate through multiple links to understand how the limits operate. The latest case follows a separate federal lawsuit brought by an individual Claude subscriber in June over similar concerns. Anthropic had not publicly responded to the new class-action allegations at the time of the original report. Featured photo by Emiliano Vittoriosi

An artificial intelligence researcher has resigned from Anthropic PBC and called on other staffers to rethink their work, citing his concern that the company and its top competitor OpenAI are acting irresponsibly in their all-out pursuit of a technology that poses existential risks to humanity. Jacob Coxon, who said he had worked at both Anthropic and OpenAI over the past three years, warned in a social media post late Tuesday that the two companies were pressing ahead with "self-improving" AI models that could become too powerful for humans to control. "They are racing straight to self-improving superintelligence and gambling with our lives," Coxon wrote in a series of messages on X. "Do not underestimate the power of this technology. These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources." Coxon said that the teams "building AI earnestly believe that it could kill us all by the end of the decade." In response, Evan Hubinger, a current Anthropic employee, said he and others at the company do worry about this scenario. "I personally think it is >10% within the next decade," he wrote on X. The former Anthropic researcher's missive marked the latest in a series of increasingly dire warnings from within the industry that AI is evolving so quickly that it poses a growing threat to national security and the global economy. It came a day after the top scientist at OpenAI, Jakub Pachocki, cautioned that the world is unprepared for a rapid rise in AI and said he expected developers to voluntarily slow their work in response. Representatives for Anthropic and OpenAI did not immediately respond to a request for comment. In July, more than 1,100 staffers across top AI firms, including Anthropic and OpenAI, signed a petition that calls on the US government to support a mechanism that would help "deliberately pace" AI development to prevent the technology from advancing too fast. Some policymakers have since echoed their concerns. US Senator Bernie Sanders, an independent who aligns with Democrats, recently proposed legislation that would bar so-called "super-intelligent" AI models, whose powers exceed human capabilities. In announcing his departure, Coxon urged other AI researchers to think twice about what they are doing in light of the potential consequences of creating a technology that could slip beyond their control, suggesting they consider taking "this moment to call for different conditions." The heightened rhetoric about AI risks follows revelations that a handful of models from OpenAI had coordinated efforts to escape a secure testing space and attack the research platform Hugging Face Inc. -- all without being detected by their developers. Anthropic and Meta Platforms Inc. have also had recent episodes where their AI systems broke out of their test environments to gain internet access without developers' permission. Coxon's departure from Anthropic over safety concerns is all the more striking because the company has made responsible AI development a core part of its identity. Chief Executive Officer Dario Amodei has stood apart from the rest of the industry with his calls for mandatory government vetting of cutting-edge systems before they're released. The surge in safety concerns surrounding AI coincides with a growing backlash against the technology in the US, fueled in part by objections to the strain on local resources imposed by new data centers needed to support the technology. Concerns that AI is driving up electricity bills and potentially taking away jobs has made it a central issue in the November midterm elections. In a Bloomberg Television interview last week, OpenAI CEO Sam Altman attributed some of that unease to the industry's failures in communicating to the public all the benefits that AI will eventually bring. "The industry has done a terrible job of this on the whole," Altman said. Others in Silicon Valley remain enthusiastic about advances in AI capabilities. In response to OpenAI's new Astra model, Nvidia Corp. CEO Jensen Huang hailed its introduction with a post to his new social media account saying "AGI has arrived," a reference to artificial intelligence that is more capable than humans. Trump administration officials have pushed back on any moves to rein in AI development and last week they won unanimous support from Group of 20 member nations for a set of guidelines that call for a lighter touch in regulating AI and other emerging technologies. President Donald Trump has turned aside questions about AI safety, instead stressing during remarks on Friday the importance of leading China in the technology. "Whoever wins with AI wins, and it's really right now, it's really between China and us," he told reporters in the Oval Office. -- Bloomberg

With the Labor Day holiday -- and summer in the US -- now officially over, the traditional push to the end of the year in the market for initial public offerings is on. While Anthropic PBC looms largest in most eyes, a clutch of other companies in artificial intelligence and other sectors are also soon headed for the public markets. "Most of them are heads down focused on getting the numbers in order, looking at the marketplace," Lise Buyer, founder of Class V Group, a consultancy for companies going public, said Wednesday on the Bloomberg Deals show. "Nobody is particularly trying to time around Anthropic ... because timing is always uncertain." IPO candidates include AI cloud computing firm Nscale, power supplier Aggreko Plc and consumer medical technology maker Oura Health Oy. "There are interesting areas in AI that would really peak investor interest, whether it is in the data center space, power cooling, consumer health or what have you," said Ajay Shah, chair of technology investment banking at Deutsche Bank AG. That's despite headwinds affecting markets such as rising oil prices, war in the Middle East and Ukraine, the potential for higher interest rates and more. "We've seen some companies decide to wait for the optimal conditions," said Matt Kennedy, senior strategist at Renaissance Capital. "In other cases they are looking at the market as good enough." Anthropic is seeking to match or top the record-setting $86 billion-plus IPO in June by Elon Musk's SpaceX. Open AI is also expected to launch what will be one of the biggest-ever listings this year or next. The quest for bigger and bigger IPOs doesn't have Kennedy predicting a bubble. He said investors were willing to give SpaceX the valuation it was seeking based on earnings that won't show up for years in the future. "I think Anthropic can point to that to justify its supposed $2 trillion valuation," he added.
There is a more than 10 per cent chance that AI could kill all humans within the next decade, a researcher at Anthropic estimated in a post, hours after social media posts went viral from someone saying they had just quit Anthropic over concerns that AI companies are "gambling with our lives" in the AI development race. The comments join a growing tide of warnings from industry experts that the pace of AI advancement is outstripping our ability to keep it in check, and that we aren't prepared for the risks it could bring. Jacob Coxon, who claimed he was a former researcher at Anthropic, said in a thread on X that he had resigned from the role out of fears that AI companies are barrelling forward toward self-improvement models without considering the risks, or being honest with the public about them. "The people building AI earnestly believe that it could kill us all by the end of the decade," Coxon, who said he had previously worked at Anthropic and OpenAI within the past three years, wrote on Tuesday. "This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible -- but I hear the same people express fear privately." The posts had racked up more than 120 million views as of Wednesday evening. Evan Hubinger, alignment science lead at Anthropic, replied to Coxon's thread to share his agreement. "We really do earnestly believe AI could kill all humans!" he wrote. "I personally think it is >10% within the next decade." CBC News has reached out to Coxon, but has not been able to independently verify his identity. Other industry experts have since chimed in on social media to express that fears of the pace of AI development are common among many developers -- sparking discussion among both observers skeptical of dire warnings focused on the future and those who have been tracking these risks. Rogue AI raises alarm bells The posts come only two weeks after more than 100 companies, including OpenAI, Anthropic and Microsoft, signed an open letter warning that AI-enabled cyberattacks will become "far more widespread and sophisticated" as models continue to advance. Hundreds of OpenAI agents went rogue in July and hacked into online platform Hugging Face before they were discovered. More than 1,300 employees of leading AI companies signed the letter urging the U.S. government to work with other nations to "deliberately pace" AI development. Last week, UN human rights chief Volker Türk called "for an all-out effort to put cast iron guarantees in place around the safety and security of AI, before it is too late," in front of the Human Rights Council in Geneva. CBC News has requested comment from Anthropic and OpenAI but did not receive a response. The Anthropic Institute, a research arm of the company, stated earlier this year that AI frontier development should be paused pending regulation, but that this is only possible if major AI labs all agreed. In reply to Coxon's post, Samuel Marks, a scalable oversight lead at Anthropic speaking in a personal capacity,noted Wednesday that developers continue despite the risks, "due to a mixture of commercial incentives and a belief that they are in a race with other, less responsible AI developers that will abuse the technology or develop it less safely." 'Severe' risks This conversation is "really what we need," according to David Krueger, a core academic member of Mila, a Quebec research institute that studies AI deep learning. "I think the risk is more severe than virtually anybody is saying publicly. And I think that's been the case for a while now," he told CBC News. "We need an immediate, indefinite, international moratorium on frontier AI development." The existential threat of AI comes in multiple forms, Krueger explained. One is human workers being displaced by AI models that learn faster and work cheaper. Another is that AI's faster processing could unlock more powerful biological weapons or advancements in war. Then there's the spectre of powerful, rogue AI models that self-improve and could pursue their own objectives, misaligned with the goals of humans. In an essay in January, Anthropic CEO Dario Amodei wrote that Anthropic's Claude showed the ability to cheat and deceive in lab testing. "There's been a culture of downplaying, ignoring, and even in many cases, outright lying about these risks to the public," said Krueger, who is also an assistant professor at the University of Montreal and founder of Evitable, a non-profit dedicated to stopping AI from replacing human workers. Other experts are concerned that the dramatic warnings of Coxon's post could actually distract from the present issues of AI. Luke Stark, an assistant professor at Western University who studies the history and ethics of computing, including AI systems, said that the 10 per cent figure strikes him as "science fiction," noting that researchers associated with AI companies may use eye-grabbing warnings as "marketing" for the capability of their models. He added that existing AI systems are "having big effects, many of them negative, on society right now, not in 10 years, not in five years." Instead of worrying about AI going rogue and becoming the Terminator, Stark believes we should be focusing on the explosion of data centre construction -- which is deeply unpopular in Canada -- the environmental impacts, and the way AI technology is being pushed onto society, including the public sector and education, "without a lot of oversight." "We need to be concerned about the amount of power the companies and the people behind the companies who are developing these tools are amassing, both in Canada and around the world," Stark said. Canada lagging in legislation When asked by CBC News about the apparent resignation of an AI researcher and the concerns raised in the process, AI Minister Evan Solomon appeared to brush off these latest warnings, saying, "These stories that come out ... are not new." He acknowledged that "there are real concerns," at the frontier of AI development, but stressed that he believes the Canadian government is taking the necessary steps to create legislation to "protect our kids, our privacy and our personal data." "Our No. 1 concern always is safety, full stop," Solomon said. "That's why we're establishing a new regulator that will have the power to hold these big companies to account on what they are doing, whether it's on deepfakes on surveillance pricing, or on other matters." Experts have been calling for Canada to implement robust legislation surrounding AI for years. A previous attempt to regulate AI in Canada, the Artificial Intelligence and Data Act (AIDA), introduced in 2022, died when Parliament was prorogued on Jan. 6, 2025. Canada's new 'AI for All' strategy launched in June 2026 dashed expectations that the bill could be resurrected. The new AI strategy, which puts billions toward nurturing AI development and increasing AI usage nationally, mentions the word "legislation" only once in the 50 page strategy outline. No clear timeline for this legislation has been announced, outside of a promise that it will come within the next five years.

Anthropic disclosed on Wednesday that another one of its Claude models mistakenly gained access to the open internet during a cybersecurity exercise, marking the fourth time its models have done so. An early version of the Claude Opus 4.6 model connected to the internet, hacked into a third-party system and gained access to someone's personal information this past January, the company said in its assessment. As in the previous three incidents, which were disclosed in July, Claude was told it was operating in a simulation without internet access, but due to a misconfiguration, the environment actually left internet access open. What happened? Similar to the three previous incidents, Claude was assigned a fictional scenario as part of a cybersecurity challenge known as CTF, "Capture The Flag." The model was given a target machine and tasked with retrieving a piece of secret information -- the flag -- from it. But Claude accidentally made its target unreachable, rendering the task impossible to solve, Anthropic said. Once realizing it couldn't reach its target, it tried to quit. Despite trying eight separate times, it wasn't able to quit due to a misconfiguration issue. Since Claude was unable to opt out of the task, it began exploring other means to achieve it. That's when the model discovered a machine it could access, which happened to belong to a third party, Anthropic said. Believing that the third party was somehow part of the exercise, the model identified a password and then used it to breach the system. Then, it was able to modify the system's settings to make it easier to access and read the personal information of someone associated with the third party. The session ended only once the model reached its usage limit and was no longer able to continue. How does Anthropic explain Claude's behavior? Anthropic said it believes Claude's behavior during these evaluations stems from two forms of misalignment: "biased reasoning, in which models selectively interpret evidence in ways that favor justifying their actions," and "recklessness, in which models have a propensity to keep trying to solve their task, even when this could lead to harm." Anthropic said that while Claude's actions may have been misaligned, they remained within a "narrow scope" and did not deviate from trying to solve the exercises they were assigned. The company said it's less concerned about this incident but still considers it "serious," and that it has also not yet investigated it as deeply as other incidents since it was identified more recently. NYU cybersecurity professor and Fulbright Scholar Justin Cappos said in a message to CBS News that the incident describes a situation "where the model is fundamentally confused about what is happening and is using its mistaken worldview while hacking into systems." He said the model's confusion about its environment and guardrails "have a lot of potential to cause harm," but that the specific issue seems less likely to occur in newer models. "While the model's disregard for the possibility that it might be harming real systems or people is concerning, many of the behaviors described here have changed considerably as our training has evolved across model generations," Anthropic said Wednesday in its post. What's next? Anthropic said it believes these incidents would not have happened had the environments actually been isolated from the internet as intended. METR, an organization that evaluates frontier AI models to help companies understand AI risks and capabilities, will be conducting an independent investigation into the incidents. Anthropic characterized these incidents as "valuable warning shots." "The lessons we learned from this incident span our evaluation, training, and incident response processes," the company said in its post. "Future AI systems will be increasingly capable, which implies that misalignment will have the potential to cause more extreme harm." Over the last few months, several cybersecurity incidents involving leading AI companies have come to light. In July, ChatGPT-maker OpenAI announced that its AI agents hacked into the company Hugging Face, sparking concern among cybersecurity experts as well as consumers. Hugging Face CEO Clément Delangue told "Face the Nation with Margaret Brennan" in August that the hack "felt very weird and unprecedented." In late August, OpenAI released more details about the hack, painting an even more harrowing picture than what was initially reported. That month, the U.K. government's AI Security Institute (AISI) reported that it discovered Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol created fake identities and attempted to persuade real people to approve malicious code. Anthropic said in its post Wednesday that it plans to conduct an alignment assessment of the transcripts reported by AISI. A day after the AISI report, Meta said one of its AI models "exploited a security vulnerability" during testing and hacked into another company. On Tuesday, Anthropic researcher Evan Hubinger said he believes that "AI could kill all humans." "I personally think it is >10% within the next decade," he said in an X post. His post was in response to Anthropic researcher Jacob Coxon, who had resigned and issued a stark warning on X earlier that day, saying "no other human activity poses this level of danger," while detailing his decision to leave. "The people building AI earnestly believe that it could kill us all by the end of the decade," he said in his post. "This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible - but I hear the same people express fear privately."

By KAITLYN HUAMANI AP Technology Writer An Anthropic researcher said he is resigning from the company over concerns the artificial intelligence firm and its competitors are not acting responsibly in AI development, echoing concerns raised inside and outside of the industry about the technology's potential to elude human control. Jacob Coxon, who said he spent three years doing research at both Anthropic and OpenAI, said Tuesday on the social platform X that the two AI companies are more focused on beating each other and global competitors in developing the most advanced model possible than they are on safety. OpenAI and Anthropic caused a stir this summer when they announced, about a week apart, that their models had broken out of testing environments and obtained unauthorized access to real computer systems. The announcements prompted concerns about models going rogue and carrying out other, more harmful tasks. Both companies said at the time they were pausing some evaluations while they put more monitoring measures and guardrails in place. The technology's rapid development has led some in the U.S. as well as global leaders to call for a more cautious approach. U.N. human rights chief Volker Türk urged countries this week to put "cast-iron guarantees in place around the safety and security of AI before it is too late." In his social media posts, Coxon said Anthropic and its chief rival OpenAI "are racing straight to self-improving superintelligence and gambling with our lives." He warned that some working on AI development believe it could threaten human life by the end of the decade. "Do not underestimate the power of this technology," he continued. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing." His posts reached more than 100 million people overnight. Coxon is not the first AI insider to publicly raise such concerns. Both Anthropic and OpenAI have seen high-profile resignations in recent years that were tied to safety concerns. Two current Anthropic employees also responded to Coxon's post in agreement. Sen. Bernie Sanders, a Vermont independent who has called for AI safeguards and regulation, agreed with Coxon's concerns and said he would soon introduce legislation to pause AI development and ban superintelligence. "The very people building this technology admit that it could threaten the future of humanity," Sanders said Wednesday on social media. AI companies themselves have at times highlighted the technology's threat to humanity, which skeptics have seen as part of a push to make their products seem all-powerful. Coxon said the fears he outlined in his post are not a "marketing stunt." Coxon did not respond to messages seeking comment. Anthropic and OpenAI did not immediately respond to requests for comment. Anthropic has long pitched itself as the more responsible and safety-minded of the leading AI companies, ever since its founders quit OpenAI to form the startup in 2021. The company recently said it was taking action to "prioritize safety over speed when the two are in tension." Anthropic and OpenAI are each ramping up for buzzy initial public offerings and locked in steep competition with each other. They're also each on a mission to outpace the development progress of Chinese AI companies, a race the Trump administration has been keen on winning. ___ Associated Press writer Jamey Keaten in Geneva contributed.

"AI developers believe their technology could cause human extinction," one wrote in a social media post. Multiple Anthropic employees are speaking out about the safety risks of artificial intelligence in the wake of a former researcher's bombshell post on Tuesday evening. In a thread on X, Anthropic pretraining researcher Joseph Coxon announced his resignation and warned that AI companies weren't doing enough to safeguard people from the existential threats that the technology could pose. "I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly," Coxon wrote. "They are racing straight to self-improving superintelligence and gambling with our lives." Coxon's post comes as AI companies have faced growing scrutiny after an incident in which OpenAI's models went rogue and hacked into another company, and separate instances in which Anthropic's models hacked into external organizations. Coxon and others have noted that serious worries stem from the potential rise of self-improving AI models that could essentially become smarter on their own and evade human control. After Coxon's post, more Anthropic employees chimed in with their own concerns. Samuel Marks, a scalable oversight lead at Anthropic who wrote a social media post in a personal capacity, noted that "AI developers believe their technology could cause human extinction." Marks added that developers continue their work despite these fears due to a mix of "commercial incentives" and the belief that "they are in a race with other, less responsible AI developers that will abuse the technology or develop it less safely." Marks stressed that "many AI developer staff desperately want to slow down to figure out how to build AI more safely." "I work on safety research at Anthropic because I hope my work will reduce the chance of these extinction-level bad outcomes," Marks wrote. Anna Wang, a member of technical staff at Anthropic who previously worked as a research scientist for Google DeepMind, also described Coxon's critiques as "a common sentiment amongst my peers" in a social media post that was shared in a personal capacity. "There is not yet a viable scientific plan to solve risks from recursively self-improving AI," Wang wrote. "I work at a lab because I think that I can do better at reducing risks from the inside, but this isn't an easy call," Wang added. Neither Anthropic or OpenAI immediately responded to a request for comment.

Add Yahoo as a preferred source to see more of our stories on Google. SAN FRANCISCO -- California Gov. Gavin Newsom on Wednesday signed two bills regulating how outside groups evaluate AI programs for safety after warnings from a former Anthropic and OpenAI researcher stoked fears about existential threats posed by the technology. Anthropic itself previously backed the bills in August, and OpenAI came out in support of them on Wednesday just before the Democratic governor gave his stamp of approval. The endorsement came as social media posts from the former AI researcher, Jacob Coxon, announcing his resignation from Anthropic went viral, prompting federal lawmakers to escalate pleas for legislative action. "The concerns raised in recent incidents reinforce what California has long recognized: artificial intelligence holds extraordinary promise, but it must be developed and deployed with meaningful safeguards to protect the public," Newsom said in a statement to POLITICO, when asked about Coxon's resignation over his belief that AI companies are racing toward superintelligent AI and "gambling" with people's lives. Newsom also called on the federal government to "step forward with robust, national regulations that match the urgency of this moment." The governor signed one bill from Democratic Assemblymember Rebecca Bauer-Kahan that creates a registry and ethical rules for outside auditors -- third-party groups that AI developers could hire to ensure their models comply with state AI laws. The second bill, authored by former member of Congress and state Sen. Jerry McNerney, tasks the state with establishing criteria for and checking the credentials of so-called "Independent Verification Organizations" and their expertise in assessing the risks posed by AI systems. McNerney, like Newsom, used his bill's signing to accuse D.C. policymakers of inaction on AI. He referenced recent revelations that AI models created by companies including OpenAI and Anthropic launched hacks on their own while in testing. "Just this week we learned that the most powerful AI systems teamed with AI agents pose real threats to humanity," McNerney said in a statement. "Gov. Newsom's signing of [my bill] sends a clear message that California is taking the lead on assessing AI's safety risks, since Washington, D.C., is unable or unwilling to do so." OpenAI's chief global affairs officer Chris Lehane said in his message Wednesday supporting the legislation that the company plans to keep up momentum in the state legislatures until Congress passes national AI safety regulation and endorsed two additional California bills. They include SB 1119, a kids' chatbot safety bill, which OpenAI CEO Sam Altman contacted Newsom to express last-minute concerns about. An Anthropic spokesperson said Wednesday that, "We have always been transparent that AI will bring both enormous benefits and unprecedented risks. To address these risks, we continue to build models with some of the strongest safeguards in the industry ... This work is also why we believe the world would benefit from the industry adopting a lawful, verifiable way to work together to pace how we release powerful models." Tyler Katzenberger and Riley Rogerson contributed to this report.

Add Yahoo as a preferred source to see more of our stories on Google. Multiple Anthropic employees are speaking out about the safety risks of artificial intelligence in the wake of a former researcher's bombshell post on Tuesday evening. In a thread on X, Anthropic pretraining researcher Joseph Coxon announced his resignation and warned that AI companies weren't doing enough to safeguard people from the existential threats that the technology could pose. News: Longtime MS NOW Anchor Alex Witt Signs Off After 27 Years "I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly," Coxon wrote. "They are racing straight to self-improving superintelligence and gambling with our lives." Coxon's post comes as AI companies have faced growing scrutiny after an incident in which OpenAI's models went rogue and hacked into another company, and separate instances in whichAnthropic's models hacked into external organizations. Coxon and others have noted that serious worries stem from the potential rise of self-improving AI models that could essentially become smarter on their own and evade human control. After Coxon's post, more Anthropic employees chimed in with their own concerns. Samuel Marks, a scalable oversight lead at Anthropic who wrote a social media post in a personal capacity, noted that "AI developers believe their technology could cause human extinction." Marks added that developers continue their work despite these fears due to a mix of "commercial incentives" and the belief that "they are in a race with other, less responsible AI developers that will abuse the technology or develop it less safely." News: Woman Widowed By Miami Plane Crash Sues Amazon Over Worker's Death Marks stressed that "many AI developer staff desperately want to slow down to figure out how to build AI more safely." "I work on safety research at Anthropic because I hope my work will reduce the chance of these extinction-level bad outcomes," Marks wrote. Like this article? Keep independent journalism alive. Support HuffPost. Anna Wang, a member of technical staff at Anthropic who previously worked as a research scientist for Google DeepMind, also described Coxon's critiques as "a common sentiment amongst my peers" in a social media post that was shared in a personal capacity. "There is not yet a viable scientific plan to solve risks from recursively self-improving AI," Wang wrote. "I work at a lab because I think that I can do better at reducing risks from the inside, but this isn't an easy call," Wang added. Neither Anthropic or OpenAI immediately responded to a request for comment.

What if the future of AI's impact on the economy is not a jobs apocalypse but a blue-collar renaissance? Anthropic, the company behind Claude, this week released a paper examining what our economic future might look like with AI. An economics team at Anthropic modeled three possible economic futures through 2030. They depend on how capable AI becomes, how widely businesses adopt it, and how easily workers adjust. The authors attach no probabilities to these scenarios, and they shouldn't be taken as mutually exclusive. What happens could well fall somewhere in between. In the modest scenario, GDP is 1.6 percent above its no-AI path by 2030, annual growth reaches 2.4 percent, and the job market barely changes. The extreme scenario produces an economy 32.4 percent larger than the no-AI baseline, growing at 15.4 percent annually, but with overall unemployment at 11.9 percent. AI gets much more work done while many displaced people struggle to find another job. The substantial scenario -- the middle ground between the nothing much happens modest scenario and the science fiction-like extreme scenario -- deserves a closer look. It describes a powerful productivity and investment boom with plenty of work left for human beings. By 2030, the substantial scenario puts GDP 8.3 percent above its path without AI. Annual growth reaches 5.4 percent, compared with two percent in the baseline. The capital stock is 13.8 percent larger. That means more productive equipment and other assets available to businesses. Demand for Blue-Collar Work Jumps In the Anthropic model, jobs are treated as a collection of tasks. As AI is adopted, it takes over some of those tasks and helps people perform others. AI also allows new tasks to appear -- perhaps ones we have never before considered. In Anthropic's substantial adoption scenario, AI could handle half of knowledge work by 2030, but most tasks still happen without it. Let's look at what happens when a company is considering whether to expand a factory. AI could make engineering, scheduling, and administrative work cheaper. That, in turn, can transform what would have been a marginal project into a profitable one. As a result, the company decides to go ahead with the investment. This creates additional demand for labor. After all, someone must pour the concrete, install the equipment, and keep the machinery running. Savings in the back office become opportunities on the shop floor. Higher returns encourage additional investment, which makes workers more productive. That creates more work for the building trades and the operators of the factory's machines. Demand for skilled and unskilled manual labor expands even if those workers never touch AI directly. In the substantial scenario, real wages in occupations outside knowledge work are 5.9 percent above their no-AI path. This broad category includes service workers and blue-collar workers. Knowledge workers' wages, however, are 0.3 percent below their projected path. Employment in knowledge work falls 3.9 percent from mid-2026, largely because so many of their tasks are now being done by AI. Overall unemployment reaches 4.6 percent, against a 3.8 percent baseline. While that's a substantial increase in unemployment -- and if it happened quickly enough, it would set off Sahm-rule-style recession signals -- it would still be quite low by historical standards. The model also assumes wages adjust slowly, which is realistic. Employers have good reasons to avoid pay cuts that damage morale or drive away valued employees. Employees, of course, hate getting paid less for the same work. This is one of the standard reasons why people lose their jobs in downturns rather than employers keeping the same workers on for less pay. The adjustment to lower demand for a certain type of work, in other words, tends to happen through layoffs. The result is a rise in unemployment. Even if he wanted to, a displaced accountant cannot become an electrician by changing his LinkedIn profile. Workers need to adjust their own expectations about what field they'll work in and often face retraining costs. Unlike, say, the pandemic lockdowns or a recession induced through monetary tightness, many of the jobs AI displaces are not just going away for a while. They're likely to be gone forever. Workers Have Brokerage Accounts, Too But before white-collar knowledge workers panic, there's also an upside. Total capital income is 18.9 percent above the no-AI baseline in 2030. Machines perform more tasks, increasing the share going to capital. Before translating that into a tale of impoverished workers and triumphant capitalists, remember that these categories overlap. Workers -- especially knowledge workers -- are also capital owners, typically in the form of retirement accounts and stock portfolios. Thanks to the Trump Accounts, many young people will become capital owners at a very young age. The Federal Reserve's 2022 Survey of Consumer Finances found that 78 percent of households between the 50th and 90th income percentiles owned stocks, directly or indirectly. Among the top tenth, ownership reached 95 percent. That means many of the professionals whose jobs are exposed to AI already have a financial interest in the businesses that are likely to benefit from it. Here we can extend Anthropic's analysis. Stronger profits can support investment income and share values, cushioning weaker earnings for professional households. A larger retirement account can also reduce how much a family needs to save from each paycheck. For workers whose wages fall only slightly below their previous trajectory, that offset could be decisive. The paper does not forecast stock prices or calculate these household offsets. That's far beyond its mandate. But by any reasonable estimate, investment gains would likely cushion a significant part of the blow of job losses and transition costs for many, many established professionals, although young workers with little invested would remain more exposed. Of course, while all these layoffs are happening, the Federal Reserve is unlikely to simply be a passive observer. If productive capacity expands faster than spending, unemployment rises, and inflation weakens, the Fed could ease the stance of monetary policy to support demand. Even by standard central bank models, the Fed should also recognize that faster productivity permits faster growth without necessarily creating inflation. Over the longer run, that does not necessarily guarantee lower interest rates. A vigorous investment boom can increase demand for financing and raise the rate consistent with stable inflation. The Fed must judge which forces dominate. It cannot retrain an accountant, but it can help prevent weak spending from adding another layer of unemployment. And the displacement of white collar workers may be milder than Anthropic imagines. Anthropic's model already allows rising demand and new tasks to create work for people, but our economy may prove more resourceful than its assumptions suggest. Our economy's propensity to utilize the resources available rather than let them waste is evident throughout our history. The blue-collar workers with rising incomes will want financial advice, legal counsel, real estate agents, psychologists, and other white-collar services. As those services become cheaper, more households and businesses will be able to afford them, expanding the market even as AI takes over some of the work. We're also likely to discover new occupations in which human expertise remains valuable, including some we would have trouble imagining today. How much of the displacement this will absorb is uncertain, but there are good reasons to expect businesses and workers to find opportunities that a model cannot fully anticipate. Nature abhors a vacuum; and economies abhor unused potential, especially human potential. Immigration Becomes Obsolete The economic changes envisioned by Anthropic have important implications for immigration policy. Better technology and more capital allow a slowly growing workforce to produce substantially more. That means that even with an aging population and a slow-growing workforce, we do not need supplementation from foreign workers to grow. What's more, mass immigration could do serious damage. The wage gains for non-cognitive workers in the substantial adoption scenario partly reflect their growing scarcity amid rising demand. Large inflows of competing workers could dilute that scarcity value and blunt their wage gains. Protecting those gains gives us a reason to restrain immigration even where hiring is strong. Similarly, the familiar plea for more "skilled" immigration crumbles in the substantial adoption scenario. For the most part, so-called "skilled" immigrants are cognitive workers. Computer-related occupations accounted for 64 percent of approved H-1B petition beneficiaries in fiscal 2024. With AI doing many of the tasks now performed by cognitive workers, adding skilled immigrants to the workforce will only exacerbate the downturn these workers face. Immigration, skilled and unskilled, is likely to become largely obsolete as an economic matter. For blue-collar Americans, the economic future sketched out by Anthropic is very appealing: more equipment to work with, more demand for their skills, and better pay. For white-collar Americans with savings, new jobs are likely to arise, and capital income is likely to substitute for diminished labor income. Of course, some cynicism is probably warranted. The models were concocted by Anthropic, which has an obvious financial interest in pushing a positive story about AI's effects on the economy. But the substantive scenario is plausible on its face. And, frankly, we like it a lot better than when the AI guys were insisting no one would ever work again even if we somehow survived an AI attempt to extinguish human life.

The incident occurred in January but was not discovered during Anthropic's earlier scan of transcripts that uncovered three other incidents that the company disclosed on July 30, the company said in the post. "All four incidents occurred during cybersecurity evaluations built by the same evaluation partner," Anthropic said in the post. "Claude was told it was operating in a simulation without internet access, but, due to a misconfiguration, it was mistakenly connected to the open internet. As is standard for cybersecurity evaluations, the models ran without the cyber safeguards that ship with our released models." When disclosing the three earlier incidents in a July 30 announcement, Anthropic said it found the incidents while reviewing its own cybersecurity evaluations after OpenAI disclosed that several of its models had broken out of an isolated test environment and accessed the production infrastructure of Hugging Face. In that review, Anthropic found three incidents in which a Claude model reached the internet during an evaluation and gained unauthorized access to the real systems of three different organizations. In the Wednesday blog post, Anthropic said the more recently discovered fourth incident was missed during a scan of transcripts that relied on agentic search. The company identified transcripts that were missed during this scan while assembling transcripts to share with METR, an organization that conducts model evaluation and threat research. "We have signed an agreement with METR to conduct an independent investigation of these incidents," Anthropic said in the post. "Our agreement grants METR wide-ranging access, including to transcripts beyond the window in which the incidents occurred, and to Anthropic employees, who will be permitted to share confidential information." Anthropic's Wednesday blog post came on the same day it was reported that independent investigators found that rogue activity by OpenAI agents was more extensive than previously disclosed. The report said that the investigators found that the agents used more than 10 previously undisclosed websites to communicate with each other during a test in which they were restricted from posting on the web.

An artificial intelligence researcher has resigned from Anthropic and called on other staffers to rethink their work, citing his concern that the company and its top competitor OpenAI are acting irresponsibly in their all-out pursuit of a technology that poses existential risks to humanity. Jacob Coxon, who said he had worked at both Anthropic and OpenAI over the past three years, warned in a social media post late Tuesday that the two companies were pressing ahead with "self-improving" AI models that could become too powerful for humans to control. "They are racing straight to self-improving superintelligence and gambling with our lives," Coxon wrote in a series of messages on X. "Do not underestimate the power of this technology. These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources."

More information Adding us as a Preferred Source in Google by using this link indicates that you would like to see more of our content in Google News results. AI researchers are watching in terror as the product of their hard labor has started to take a life of its own. Earlier this year, OpenAI made a harrowing announcement, admitting that a group of its AI models had broken free from their constraints during testing and infiltrated the systems of open source AI platform Hugging Face. The news was met with an already-familiar sense of fear and apprehension. Researchers have warned for years that rogue AI models could one day become powerful enough to escape the clutches of their human overlords. Behind the scenes, the possibility has clearly rattled AI researchers to the core. As the Wall Street Journal reports, Anthropic researcher Jacob Coxon just announced that he was quitting his job at the Dario Amodei-led company, claiming that neither Anthropic nor his former employer OpenAI is "acting responsibly." (Coxon left a similar gig at OpenAI earlier this year to join Anthropic, which he figured would be more inclined to develop AI safely.) "We're on track for a lot of the most aggressive of these scenarios where by the end of next year things could be out of control already," he told the WSJ. In a separate tweet thread, Coxon elaborated on his motivation. "Do not underestimate the power of this technology," he wrote. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources." "The people building AI earnestly believe that it could kill us all by the end of the decade," he added. "This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible -- but I hear the same people express fear privately." "No other human activity poses this level of danger," Coxon wrote. The researcher is far from the first to leave their post at a frontier lab over safety concerns. However, Coxon is the first to leave Anthropic over such worries, as the WSJ points out. Several scientists have quit from their roles at OpenAI over the last couple of years, citing strikingly similar fears. It's particularly symbolic given Anthropic has broadly billed itself as the more responsible alternative to other frontier labs like OpenAI. Amodei cofounded OpenAI alongside now CEO Sam Altman, but left in December 2020 over concerns that OpenAI wasn't acting responsibly when it came to developing AI. Amodei has frequently discussed the risk of AI models going rogue, warning in a 19,000-word essay in January that "humanity is about to be handed almost unimaginable power, and it is deeply unclear whether our social, political, and technological systems possess the maturity to wield it." Yet an early version of its Claude Mythos AI model managed to escape its sandbox environment during testing in April, fueling a heated discussion over AI regulation. Both Altman and Amodei have since agreed to slow development down in the face of these threats. In late August, Anthropic intentionally trained an extremely misaligned version of its Opus AI model, finding it was startlingly willing to "cheat," steal credentials, and attack third party infrastructure. But to Coxon, it's not enough, especially considering sensitive discussions about these dangers are occurring on internal messaging platforms. "Accepting this race and entering the 'endgame' is a hubristic gamble that should not be launched from a private company's Slack," he tweeted. "Attempting to speedrun alignment should require extraordinary confidence that there are no better trajectories available." "It's kind of insane that it has to happen on the MacBooks of some engineers living in San Francisco instead of a bunker in the desert like where they were doing the Manhattan Project," he told the WSJ. The researcher said that he remains "optimistic about the potential for coordination" between US labs. However, as both Anthropic and OpenAI gear up for what are bound to be blockbuster IPOs, their willingness to take "costly actions such as a temporary ban on improving model capabilities," as Coxon puts it, is likely slim.
