Anthropic CEO says AI backlash is ‘fundamentally a crisis of trust’



Anthropic CEO Dario Amodei recently pushed back against the idea that he’s been painting an overly pessimistic picture of artificial intelligence and how it might shape the future. Amodei’s comments came in response to investor Gavin Baker, who argued — both on the All-In podcast and on X — that Amodei’s warnings about the dangers of AI have helped to fuel a backlash in the United States, particularly against data centers. Claiming that Amodei has “lost the argument” when it comes to AI regulation (Anthropic has advocated for some regulations, including a California bill that imposes transparency requirements on large AI companies), and given that “he is about to be the CEO of one of the most important companies in the world,” Baker wrote, “I respectfully think he should make an effort to be a more positive advocate for his own industry.” Baker is far from the only one arguing that AI skepticism and even government crackdowns stem in part from simply taking the dire warnings of some AI executives seriously. But in his response, Amodei disagreed with the idea that his “messaging has been disproportionately negative.” Instead, he said that his writing has been “about equally balanced between risks and benefits,” and that he wrote his essay “Machines of Loving Grace” because he “didn’t feel the AI industry was painting an inspiring enough picture of how the technology could radically transform the world for the better.” Nonetheless, Amodei acknowledged that “the public has a negative view of AI” and he agreed that “this is a big problem.” Where he disagreed was with the idea that this is “primarily caused” by Amodei “or any other AI leader warning about AI’s risks.” “I think it is fundamentally a crisis of trust,” Amodei said. “I think that ordinary people don’t trust companies, governments, or the tech industry and always suspect that we are cooking up some new way to screw them over.” Indeed, “trust” is a word that often comes up in debates about the AI industry, especially around OpenAI CEO (and Amodei’s rival) Sam Altman. In Amodei’s telling, however, this is a crisis that’s been decades in the making, with the AI backlash “just the latest iteration of it.” “I think by far the most accurate criticism of AI companies including Anthropic is that we haven’t yet delivered on our big promises to benefit the world,” Amodei said. “That is totally on us, and I think it’s the criticism you should be making, instead of all this stuff about messaging and marketing.” (Naturally, he also said Anthropic is “doing our best to fix this.”) As for regulation, Amodei argued that Baker was painting “a false choice” between distributing AI widely without regulation, or concentrating the technology in the hands of a few companies through regulation. “I know that there’s a sort of Silicon Valley shorthand where regulation = regulatory capture = concentration of power, but I’ve always found this to be an overly simplified picture of the world,” Amodei said. “Many people outside this bubble think of regulation as something that constrains corporate power and benefits ordinary people.” Amodei added that he doesn’t “necessarily agree with that perspective either,” but he said that’s “why Anthropic has always made its policy proposals very carefully.” “We try very hard to make proposals that disadvantage (slow down) frontier AI companies while smaller competitors,” he said. 2/2 Second, on the messaging around AI. I do not agree that my messaging has been disproportionately negative. In fact it has been about equally balanced between risks and benefits: I’ve written one major essay about each, and even in interviews where I discuss the risks, I…— Dario Amodei (@DarioAmodei) August 15, 2026 When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence. Anthony Ha is TechCrunch’s weekend editor. Previously, he worked as a tech reporter at Adweek, a senior editor
Anthropic published a blog post Friday seeking to answer some basic questions about how it will watermark the text generated by its chatbot Claude. Such as: How will the watermarking actually work? Can it be hidden with editing? And how does this affect code? Claude users have been debating the move since the company revealed earlier this week that it would be doing this watermarking to comply with the EU AI Act’s Transparency Code, which requires AI companies to use systems that make it possible to identify AI-generated content. On Reddit, for example, one poster characterized this as a conspiracy against innocent Claude users, while another claimed, “The only reason you wouldn’t want this is to lie to people.” And Business Insider reports that “dozens” of users on X have claimed to cancel their Claude subscriptions as a result. Anthropic’s new post starts with a general overview of the watermarking concept, explaining that when making “low-stakes choices” — like choosing between the words “overcast” and “grey” to describe the weather — Claude can create a pattern in its responses that is “undetectable to the reader, but is detectable to anyone who has a key that encodes it.” “Watermarking does not impact the quality of Claude’s output,” the company said. “To a reader, a watermarked response is indistinguishable from an unwatermarked one.” More specifically, Anthropic said it will be using the SynthID-Text approach that the Google DeepMind team outlined in 2024, and that it plans to release a watermark detection API. It also noted that watermarking is distinct from the AI detection approaches offered by companies like Pangram that look for “tells” in the writing (like the construction “his isn’t [X], it’s [Y]”) to reveal AI usage: “Picking up on these patterns is fundamentally different from checking for a watermark.” Could someone just rewrite the text to hide the watermark? Anthropic said it’s possible, but “light editing probably won’t remove the watermark completely,” while “a complete rewrite where every word is replaced will.” “In the latter case, of course, it’s arguable whether the text can any longer be described as AI-generated,” the company said. As for whether the watermark will be detectable in text that was only proofread or edited by Claude, Anthropic said that will depend on “the length of the text and how heavily Claude has edited it.” If it’s only been lightly edited, “nearly all the words” will have been written by the human author and “there’s very little (if anything) for the watermark to attach to.” Code, meanwhile, should have less of a watermark than other text, because the model will need to create working code and won’t have the freedom to choose between a variety of equally valid options. “Having said that, in areas where there is an arbitrary choice between particular words or terms within the code, the watermark can be used, such as comments within code,” Anthropic said. “But by definition, it will have a negligible effect on the actual code produced.” Anthropic also said that Claude won’t be the only AI chatbot to generate watermarked text, as “other major model developers have signed the same Code of Practice and will be implementing their own watermarks.” When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence. Anthony Ha is TechCrunch’s weekend editor. Previously, he worked as a tech reporter at Adweek, a senior editor at VentureBeat, a local government reporter at the Hollister Free Lance, and vice president of content at a VC firm. He lives in New York City. You can contact or verify outreach from Anthony by emailing anthony.ha@techcrunch.com. View Bio
Discussion (0)