Alibaba releases Qwen 3.8 preview, claiming it is second only to Fable 5

Alibaba have been accused of massively copying Claude’s responses. (Picture: Alibaba, generated)
The Chinese onslaught continues, just days after the launch of groundbreaking Kimi 3, with Alibaba claiming that their 2.4 trillion parameter open source model matches the American frontier and is only behind Anthropic’s Fable 5.

There is very little information out on the model save for Alibaba’s X post, where they claim the model is «continuously evolving» and is «one of the most powerful model[s] available today.»

There are no benchmarks to back that claim just yet, and the last Alibaba model on Arena.ai’s leaderboard is the Qwen 3.7, hovering around 17th place on coding and 18th on agentic tasks. The new Qwen 3.8 would significantly improve on that performance.

GPT-5.6 was released in late July, while Fable 5 was announced in early June, putting the current window to cutting edge Chinese open models at a little over one month. And we are still waiting on DeepSeek v4, which by early leaks seems to be frontier-level, too.

Qwen 3.8 «preview» is already out and ready to test on Alibaba’s token plan, and they are promising to release the open weights «soon,» writes The Decoder.

Like Moonshot AI, behind the Kimi 3 model, Alibaba was accused of a massive distillation campaign on Anthropic’s Claude, copying some 28.8 million exchanges between April and June this year.

Read more: Alibaba’s X post, The Decorder, Qwen promotion. Discussion on Hacker news and r/Singularity.

US government lifts restrictions on Anthropic’s Fable 5

Fable 5 spooked the government enough to ban it, but now it’s back online with even stronger safeguards. (Picture: Shutterstock)
As of July 1, the Mythos-class model is available on wide release to paying customers, after the Commerce department lifted their export controls on June 30.

Anthropic hails the news, but cautions that recent events have highlighted the need for a «consistent way to assess and fix potential «jailbreaks,»» after two weeks of grueling exchanges with the administration.

Commerce secretary Howard Lutnick said on X that they had «worked closely with Anthropic to […] ensure alignment across the US Government and strengthen America’s leadership in AI,» Axios reports.

The Fable 5 model became somewhat of a joke in the AI community for having such strong safeguards that it would not answer anything even remotely related to security or biology, but Amazon engineers got it to output code for an exploit, leading to the export ban on June 12.

This behavior is now blocked 99% of the time, and Anthropic has agreed to even further safeguards on the model, pledging to work even more closely with the government in the future, including on vetting pre-release models.

Read more: Anthropic’s announcement and X post, Lutnick’s letter, Axios, Politico, CNBC.

US government clears Mythos 5 for limited release to «trusted partners»

Mythos 5 will become available on Project Glasswing after the government lifts its ban. (Picture: Shutterstock)
After two weeks of purgatory and almost daily explanatory meetings, Commerce Secretary Howard Lutnick sent a letter to Anthropic on Friday, clearing one of two models on hold for release:

— I have determined that appropriate safeguards are in place to permit certain trusted partners to access the Claude Mythos 5 Model, he writes to chief compute officer Tom Brown, according to Semafor, who scooped the story.

Mythos 5 was only intended for a limited release through the Glasswing project, for use by a trusted set of government agencies and corporations to test their cyber defenses.

UPDATE: On Saturday, Axios is reporting that Fable 5 might be next in line, with «insiders» predicting it might be released next week, and Anthropic anticipating access «soon.»

The letter also says that «Anthropic has committed to work with the U.S. government on protocols and standards and releases,» and talks are progressing toward a release of Fable, though «the timeline is unclear,» according to Semafor.

Read more: Scoop by Semafor, comment from Anthropic, Reuters, CNBC, and Politico.

OpenAI updates GPT-5.5-Cyber, claims better performance than Mythos 5

The new Cyber model not only finds flaws, but offers to fix them automatically. (Picture: Adobe)
Less fabled than the recent Mythos releases, GPT-5.5-Cyber is just as capable, OpenAI claims, and has benchmarks to show it.

The update scores 85.6% in CyberGym versus 83.8% for Mythos 5, and is likewise only available to vetted cybersecurity professionals.

— AI has changed the physics of cybersecurity, OpenAI says, — The bottleneck historically has been finding vulnerabilities, but now defenders are overwhelmed with the number of vulnerabilities found.

Therefore, their Daybreak suite, which includes the Cyber model and Codex Security, moves beyond just threat scanning to actually offering vulnerability patches to fix them.

Since the launch as preview in March, OpenAI claims to have scanned over 30 million commits across 30,000 codebases, with humans having marked 70,000 findings as fixed and 500,000 findings having been fixed automatically.

Read more: OpenAI’s announcement, Launch tweet, writeup on SiliconAngle. Discussion on r/Singularity.

From the grapevine; GPT-5.6 coming shortly, Mythos smashes the NSA, and Google’s upcoming «instant ramen»

Lots of talk on socials doesn’t always relate to hard news, but here are some topics making the rounds. (Picture: Adobe)
GPT-5.6 in overdrive on the hype machine
The rumor mill is kicking into high gear for OpenAI’s next GPT-5.6. It’s supposed to be largely on par with Fable 5, or kick it to the curb on some tests, according to X leaker Chetaslua who claims to have been testing it for a while.

OpenAI’s chief scientist, Jakub Pachocki, pre-announced the model as a «meaningful improvement,» but would not point to a timeline.

According to prediction markets, Kalshi has it out before July 15th, while Polymaket predicts it between June 30 and July 31, which seem like safe bets.

Continue reading “From the grapevine; GPT-5.6 coming shortly, Mythos smashes the NSA, and Google’s upcoming «instant ramen»”

G7 leaders urged by labs to regulate AI, will begin «assessment»

Altman asked the G7 not to leave regulation of increasingly powerful models to the AI labs. (Picture: generated)
Amodei, Altman and Hassabis were all present at a working lunch with the Group of Seven countries Wednesday, and they all asked «the free world» for stronger regulation.

— Do not cede your responsibilities to AI labs like mine, OpenAI’s Sam Altman said, according to Reuters. — We develop the technology, and the citizens of the free world make the rules.

The leaders gathered for the summit had been spooked by the US government’s shutdown of Mythos 5 just last Friday, with especially European nations worrying about reliance on US tech:

— From one day to the next [the USA] can turn off the switch, said French President Emmanuel Macron, reported by The Financial Times, as he, too, called for stronger regulation and cooperation.

There was also a strong push at the meetup to expand access beyond the USA to Anthropic’s Project Glasswing and the Mythos Preview model, used in a race to find cyber vulnerabilities in critical systems before the capabilities become widely available.

The meeting ended in an official joint statement that would task finance officials, regulators and cybersecurity experts with evaluating AI’s impact on financial stability, productivity and labor markets.

Read more: Reuters, The Financial Times, and CNBC.

Speaking in «different languages,» Anthropic struggles with White House

Finding themselves thrust into politics, Anthropic is having difficulty getting through, Axios reports. (Picture: Shutterstock)
As security researchers and executives are urging a retreat from the Mythos ban, Axios is reporting that Anthropic is failing to «communicate effectively.»

They were scheduled to meet in person with the White House on Monday, sources tell CNBC, and are now finding themselves deeply enmeshed in current politics.

Government officials are now candidly telling Axios that Anthropic «screwed» them, saying they failed to «honor» the recent Executive Order on AI — which calls for a 30-day vetting period for new models, and are calling Anthropic a «bad actor.»

There are also differing narratives in the Axios report, where on the one hand, Anthropic says they «received explicit approval» to deploy Fable, and another where the administration threatened export controls several weeks ago, fearing the model could be exploited by bad actors.

The exploit uncovered in the Fable and Mythos models are replicable in other frontier and open source models, security researchers say in an open letter, and shutting down Fable at a critical time gives an advantage to attackers over defenders. They say there is only a 90-day window until Chinese models reach the same capabilities.

Read the Axios report here, also see the security researchers’ open letter, reports on euters, CNBC and Politico.

Anthropic disables Mythos 5 and Fable 5 after Commerce intervention

Future models should be voluntarily vetted before release, but the mechanism isn’t in place yet. (Picture: Shutterstock)
Commerce Secretary Howard Lutnick has enacted export controls on the new models, banning their use in any foreign countries and by foreign nationals — and Anthropic responded quickly by completely blocking access for everyone.

The triggering factor appears to be that a company was able to jailbreak the models, potentially getting them to perform harmful actions, an administration official tells Axios.

According to Anthropic, the claim concerns a limited jailbreak and a «minor vulnerability,» not considered a systemwide issue, and says it is impossible to guarantee against this 100% for any model by any provider.

Continue reading “Anthropic disables Mythos 5 and Fable 5 after Commerce intervention”

Reports: Fable 5 won’t answer even simple questions on its restricted topics

Hardly any questions about cybersecurity, biology or chemistry get through to Fable 5. (Picture: Shutterstock)
As more people are testing out Anthropic’s new, Mythos-class Fable 5, many are finding it so restrictive it wont even answer the most basic questions, TechCrunch writes.

The Verge tested it on the fundamentals of biology and found it wouldn’t even answer «easy» questions, about cell membranes or on how mRNA vaccines work.

TechCrunch is reporting on several developers who are complaining that just including the word «cyber» in a query is enough to trip it up, while others say it even refuses to write «secure» code, thinking it might be cybersecurity related.

In order to release the Mythos-based model, Anthropic made a tradeoff on very strict limits for it, and the fallback option is the still very capable Claude Opus 4.8, released in late May.

These tight restrictions aren’t by mistake, Anthropic says;

— To deploy Fable 5 safely, we believe it was necessary to be overly conservative with our safeguards so they block most queries tied to biology work, Anthropic spokesperson Paruul Maheshwary told The Verge.

Read more: TechCrunch and The Verge.

Anthropic releases Mythos-class Fable 5 and Mythos 5, with strong safeguards

Fable 5 uses a second AI to monitor throughput and prevent misuse. (Picture: Anthropic)
After refusing to release any Mythos models for fear of abuse of its advanced cyber firepower, Anthropic has finally come up with enough safeguards to make it to general rotation. Its capabilities «exceed those of any model we’ve ever made available,» they say.

— The capabilities of models like Fable 5 and Mythos 5 have the potential to do profound good for the world, Anthropic writes in its presentation, and adds — We’ve also seen it in life sciences research, where the models are positing novel hypotheses and speeding up the development of new therapeutics.

Continue reading “Anthropic releases Mythos-class Fable 5 and Mythos 5, with strong safeguards”

Anthropic says 80% of code for Claude is written by Claude, calls for slowdown

Human review might take more time than writing the code in the near future, Anthropic says. (Picture: Anthropic)
Recursive self-improvement may come sooner than we are prepared for, the AI lab says in a new report.

They are now foreseeing a not-so-distant future where Claude can autonomously write and improve on itself, and is warning that humans might lose control.

Compared to 2021-2025, Anthropic is now producing 8x more code with AI tools, and this rate is doubling every four months, approaching a point where human judgment could become the bottleneck in software design.

With the more capable Mythos Preview model, researchers at the company estimate they could output another 4 times as much code on top of the already accelerating trend.

There is no sign of this progress stopping, and Anthropic is now calling for a global slowdown so that society has time to adapt. Even if growth did stop, or somehow development stalled, we are only just beginning to scratch the surface of adopting the already existing technology, the lab says.

Read more: Anthropic’s report, launch thread summary, Axios, The WSJ. DIscussion on r/Singularity.

Anthropic ships «more honest» Opus 4.8, teases Mythos release

Anthropic’s latest tops the benchmarks and makes fewer mistakes. (Picture: Anthropic)
The new model comes just over a month after the last one. It is four times less likely to allow flaws in code or make unsupported claims, and is «more likely» to say so when it is uncertain about a reply.

Releasing today at the same price as 4.7, it can tackle problems at a larger scale and has an upgrade to the «fast mode» — which is now three times cheaper.

There is also a new setting for Claude where users can set the «effort»-level on a given task. More effort costs more tokens, but gives a more precise answer, while low effort won’t bust rate limits.

The new Opus is of course on top of all the benchmarks so far in the cycle, and it is supposedly also sharper and more reliable in its judgement.

At the same time, Anthropic has some news about the code-busting Mythos model, previously deemed too dangerous to release.

The company says it is making headway on the safeguards needed to make it safe enough for public use, and plans to bring «Mythos-class» models to their customers «in the coming weeks.»

Read more: Anthropic’s announcement, TechCrunch, Reuters, The Verge.

Mythos finds 10K severe bugs in a month, as Anthropic widens release

Anthropic is now expanding availability for government and qualifying security teams. (Picture: Shutterstock)
Officially dubbed the Claude Mythos Preview, Anthropic’s code-busting agent has worked with about 50 partners to find ten thousand high-severity bugs in the month since release:

— Progress on software security used to be limited by how quickly we could find new vulnerabilities. Now it’s limited by how quickly we can verify, disclose, and patch the large numbers of vulnerabilities found by AI, Anthropic says.

Of those vulnerabilities, 6,202 were serious finds in open source software, where Anthropic has partnered with «more than 1,000» projects. Mythos actually found 23,019 bugs, but most were estimated at medium or low severity.

— Models with similar cybersecurity skills to Mythos Preview will soon be more broadly available, says Anthropic.

— There is a clear need for a larger effort across the software industry to manage the volume of findings that these models will generate.

Therefore, Anthropic is widening the release of Mythos, making it available to «qualifying» security teams «on request.» In the future, they say they hope to develop safeguards strong enough to make it generally available, but no such safety features exist as of today.

Read more: Anthropic’s findings and Engadget. Discussion on r/cybersecurity

GPT-5.5 -Cyber is out in «limited preview,» available to vetted defenders


Two models models are launching today; GPT-5.5 with «Trusted Access for Cyber» that requires some vetting to get into. It can handle defensive security, do code review, malware analysis and patch validation.

GPT-.5.5-Cyber requires stronger verification and does specialized workflows, red teaming, penetration testing and controlled validation.

GPT-5.5 Cyber was earlier found on par with Anthropic’s Mythos model, that has been spooking the establishment lately.

The vetting approach for the Cyber model has been «informed by conversations with cybersecurity and national security leaders across federal and state government and major commercial entities,» OpenAI says.

Read more: OpenAI’s blog, CNBC, Axios.

The White House reportedly discussing vetting AI models ahead of release

The White House says any Executive Order will come from the President himself. (Picture: Adobe)
The Trump administration has appartently been spooked by the cyber capabilities of Anthropic’s Mythos model and OpenAI’s GPT-5.5 — and is considering an Executive Order to vet new models ahead of release, Axios and the NYT reports.

These models have both been limited for their ability in cybersecurity, and point to a not-so-distant future where such capabilities might be widely available.

To that end, the White House’s Office of the National Cyber Director held all of two meetings last week, with tech and cyber companies on the one hand and with trade groups in tech on the other, according to Axios.

The ONCD has also been discussing safety testing for federal AI deployments, by assessing the security exposure of AI models before rolling out to the public sector.

The NYT reported on this first, and is saying that there might be a safety review for new models, while giving the Pentagon the first shot at eventual «useful» cyber capabilities, but would not block their release.

Any discussion on «potential executive orders is speculation,» a White House official told Axios.

Read more: Axios and the New York Times.