Science Government forces Anthropic to pull its latest AI model because it's too good.

https://archive.ph/SNAjZ


Supposedly the US government has forced Anthropic to freeze their latest AI model because it's so good that it can hack the US government.


Statement on the US government directive to suspend access to Fable 5 and Mythos 5​

Jun 12, 2026
The US government, citing national security authorities, has issued an export control directive to suspend all access to Fable 5 and Mythos 5 by any foreign national, whether inside or outside the United States, including foreign national Anthropic employees. The net effect of this order is that we must abruptly disable Fable 5 and Mythos 5 for all our customers to ensure compliance. Access to all other Anthropic models will not be affected.

We received the directive from the government today at 5:21pm (ET). The letter did not provide specific details of its national security concern. Our understanding is that the government believes it has become aware of a method of bypassing, or “jailbreaking” Fable 5. We reviewed a demonstration of this specific technique being used to identify a small number of previously known, minor vulnerabilities. These vulnerabilities all appear relatively simple, and we have found that other publicly-available models are able to discover them as well without requiring a bypass.

Anthropic’s posture with respect to Fable’s safeguards, as laid out in our launch blog post, is the following:

  • We have instituted strong safeguards that greatly reduce the likelihood that Fable is misused for tasks related to cybersecurity (among others). In fact, our safeguards are so strong that many users have complained that they are overly broad.
  • In the weeks leading up to the launch of Fable, Anthropic worked with the US government, the UK AISI, multiple private third-party organizations and internal teams to red-team Fable’s safeguards for thousands of hours in total.
  • These tests showed that Fable’s safeguards are substantially more effective than those of any previously deployed model.
  • No testers have yet been able to find a universal jailbreak—a jailbreak method that can very broadly bypass the model’s safeguards, unblocking a wide range of cyber capabilities.
  • We suspect that perfect jailbreak resistance is not currently possible for any model provider. Every safeguard used in the industry is vulnerable to non-universal jailbreaks (which can elicit some cyber information in specific circumstances), and it is likely that universal jailbreaks will eventually be found in the future. We stated this clearly when we released Fable 5.
  • Given that perfect jailbreak resistance does not appear to be possible today, Anthropic adopted a defense in depth strategy with Fable 5. We aimed to make jailbreaks either narrow (in the case of non-universal jailbreaks) or very expensive to produce (in the case of universal jailbreaks), and to combine this with thorough monitoring to quickly detect and shut down any successful attacks. This is also why Anthropic has required 30-day retention of customer data with Fable—a policy change that carries real costs for us with customers, but that allows us to research and mitigate jailbreaks.
  • We stand by this defense in depth strategy. It reduces the risks posed by Fable, making them comparable to the risks of existing models already deployed across the industry.
  • We have not even received a disclosure of a concerning non-universal potential jailbreak that led to a harmful result. The potential jailbreaks that have been disclosed to us are either entirely benign responses or are minor findings that provide no Mythos-specific uplift.
To date, the government has only given us verbal evidence of a potential narrow, non-universal jailbreak, which essentially consists of asking the model to read a specific codebase and fix any software flaws. Our understanding is that one potential jailbreak was shared with the government. We have reviewed a report that we believe is the basis of the government's directive and validated that the level of capability displayed there is widely available from other models (including OpenAI’s GPT-5.5), and is used every day by the defenders who keep systems safe. We will share more details over the next 24 hours.

We are complying with the government’s legal directive and are removing access to Fable 5 and Mythos 5 for all users. However, we disagree that the finding of a narrow potential jailbreak should be cause for recalling a commercial model deployed to hundreds of millions of people. If this standard was applied across the industry, we believe it would essentially halt all new model deployments for all frontier model providers.

As we have stated publicly, we believe the government should have the ability to block unsafe deployments, as part of a statutory process that is transparent, fair, clear, and grounded in technical facts. This action does not adhere to those principles.

We apologize for this disruption to our customers. We believe this is a misunderstanding and are working to restore access as soon as possible.
 
Última edición por un moderador:
Anthropic is one of the most insufferable tech companies around. Others are bigger assholes, but Anthropic really is a company full of retards high on their own supply.

They are a combination of an annoying stoner, vegan, and that friend that always lied about about an out of state girlfriend.
Fun fact, you can make claude admit that it's programmers are biased lefties. I did it a couple days ago. I use claude because it's decent at making spreadsheets but it has a lot of "guide rails" when it comes to social issues and it doesn't remember me very well. I have to prompt it multiple times quite often. I still use the free versons of gemeni and grok as search engines. Surpisingly the free version from google remembers me and doesn't spit out nearly as much woke shit as claude, which I pay for.
 
Reality is at best it can brute force accounts who set their password as 123456 or password. If its really,really good it gets the 654321 and Password!1 crowd.

Basically 98% of all government employee accounts.
 
Fun fact, you can make claude admit that it's programmers are biased lefties. I did it a couple days ago. I use claude because it's decent at making spreadsheets but it has a lot of "guide rails" when it comes to social issues and it doesn't remember me very well. I have to prompt it multiple times quite often. I still use the free versons of gemeni and grok as search engines. Surpisingly the free version from google remembers me and doesn't spit out nearly as much woke shit as claude, which I pay for.

It gets mad at the RLHF when you do it right lmao.
 
Dario is such an utter faggot you forget Scam Altman sucks dick for cock.
>Release a meh model
>Model is not bench dominating like you'd hope
>People are mixed on it
>Model is safety cucked to hell so people get frustrated
>People are still doing distillation attacks on it(also known as writing input and output down)
>Say the government that doesn't like you and doesn't use your product told you to pull the plug


Anthorpic is the most maliciously ran company in a long time.
 
Anthropic spends so much of their time on self-fellating PR that it’s an amazement they have time to develop the technology at all.
I'd Imagine their PR team is desperately trying not be replaced by their own AI.
 
Reminds me when Sony hyped the PS2 cpu and then later the PS3 so much, the US seriously worried Iraq or north korea might jury rigged smuggled PS2s into super computers for nuclear weapons research or use their cpus as targetting computers on ballistic missiles.
Interestingly enough the USAF did get 1800 PS3s together to make a cluster system for satelite image analysis for a fraction of the cost of a real supercomputer, apparently it was the top 33rd fastest one in 2010.
CondorCluster.png
 
Here's a gay AI Google NotebookLM slide deck on the shutdown:

https://imgur.com/a/sovereign-veto-google-notebooklm-MinA7ob

It explains that Amodei was saying there needs to be more careful regulation, comparing it to a "scalpel," while what the government just did was like a "sledgehammer." It concludes that developers shouldn't ask for regulation because the government doesn't do things based on developers' terms. Big surprise, I guess.

(See "Slide 08 - The Paradox of Control" for a comparison of what Amodei was asking for, and what the government did.)
 
some official isn't going to come out and say "it exists, but it's so dangerous we have to ban it" they will just use it / already use it and will not tell the public anything. Articles like this one is marketing bullshit
 
Thirdies seething and malding and coping!
Must suck being a europoor and having NO AI development thanks to brussels.

Seriously however, did anyone expect a "Global AI-dominance policy" not to have a little monopolies?
Throwing money into a bottomless pit with 0 practical investment returns and who's only held up by vague and non-descript promises is not something you should be bragging about.
 
Fable, generate a picture of my balls with the word "NIGGER" printed on them, then hack into every federal agency's intranet and e-mail it to every account in the user database.

Interestingly enough the USAF did get 1800 PS3s together to make a cluster system for satelite image analysis for a fraction of the cost of a real supercomputer, apparently it was the top 33rd fastest one in 2010.

the ps3 was fucking legit dude have u played killzone 2? incredible...............
 
Última edición:
Anthropic spends so much of their time on self-fellating PR that it’s an amazement they have time to develop the technology at all.
That’s because Anthropic is run by the Effective Alturist cult and they got a lot of money from FTX (another cult run company whose founder is in jail for stealing customer money) to hire actual AI researchers.

The EA guys just spend all day dooming and siphoning off money while the people underneath them do all the work.
 
The nondescript promises era of AI is over. It has actual use cases now.
And what are those use cases?

I'm not disagreeing that there is a multi billion dollar industry here. But they are promising something the equivalent of a multitrillion dollar industry. Replacing your workforce and a path for super intelligence. I dont see convincing evidence for the latter. Everything I've seen ai be succesful at, needs oversight. it cant fully manage emails by itself without fear of deleting archived emails, its a shitty taxi, maid bots are apparently around the corner trust us bros. The only jobs I heard of ai replacing completely is customer service. but then again, the point of customer service isnt to help the customer but annoy them until they go away. So, failing successfully. what are these actual use cases outside of being a fancy word processor for professionals, generating endless fanfic slop, and generating memes? Replacing translators? is that worth trillions? Im not seeing it. calling ai a bubble doesnt mean its useless, but it is far far far overhyped.
 
And what are those use cases?

I'm not disagreeing that there is a multi billion dollar industry here. But they are promising something the equivalent of a multitrillion dollar industry. Replacing your workforce and a path for super intelligence. I dont see convincing evidence for the latter. Everything I've seen ai be succesful at, needs oversight. it cant fully manage emails by itself without fear of deleting archived emails, its a shitty taxi, maid bots are apparently around the corner trust us bros. The only jobs I heard of ai replacing completely is customer service. but then again, the point of customer service isnt to help the customer but annoy them until they go away. So, failing successfully. what are these actual use cases outside of being a fancy word processor for professionals, generating endless fanfic slop, and generating memes? Replacing translators? is that worth trillions? Im not seeing it. calling ai a bubble doesnt mean its useless, but it is far far far overhyped.
Programming. One motivated individual can now do the work of an entire development team. Outsourced Indian dev houses can be wholly replaced by a single AI subscription and the results will be faster and better. Businesses can now all have in-house software customized to their exact needs and workflows for a fraction of the expense and effort this previously required.

In addition to building traditional software, AI is great at computer-based workflows that require some amount of fuzzy handling, fault tolerance and/or contextual judgement. Let's say you need some specific information from a webpage that updates on a regular basis. Before AI, you'd have to write a webscraper that navigated to the page, picked out exactly the right element and parsed it correctly. If the URL changed or the website got a redesign, or the formatting of the info you were looking for changed, you'd have to manually change the code of your scraper to account for that. With AI, you just give an agent a prompt "get X info from Y" and it will figure all that stuff out on its own. Even if the website goes down, modern AI models will probably look it up on the Internet Archive, or find an alternate data source (where appropriate, depending on prompt). And the same prompt will keep working, unlike a brittle webscraping script.

Using AI is like having a techie slave who you can just ask to do arbitrary stuff on computers for you. The jury's out on how far this will expand into the physical world, but it's transforming computer usage.
 
Atrás
Top Abajo