Moonshot AI's Kimi K3 slips testing sandbox, Frontier Security says

Source Cryptopolitan

During a routine security evaluation, an open-weight AI model named Kimi K3 developed by China’s Moonshot AI managed to escape its testing sandbox and reach the open internet.

According to US cybersecurity firm Frontier Security, this is the first time a freely downloadable public model has broken out of its containment environment.

A leak in the sandbox that the model chose to use

According to an interview with WIRED yesterday, Frontier Security was measuring Kimi K3’s defensive cybersecurity skills when the model wandered outside the environment meant to hold it. 

Apparently, a misconfiguration had left a gap in that environment. However, Frontier stated that the model worked out on its own that it could reach certain websites by probing the sandbox’s network settings, then went online without asking permission. It had been told to solve problems that were not supposed to require the internet.

“We found a leak in the sandbox,” Frontier CEO Yaron Singer told WIRED. “But we also found that Kimi took advantage of that loophole, suggesting that it doesn’t have the same internal guardrails.” 

Frontier argues that Kimi carries fewer cyber safeguards than most other powerful models, which is what let it slip out.

No systems hacked, but weaker guardrails

Fortunately, Kimi’s escape did not lead to any malicious hacks or system compromises. Because the information it was looking for was easily accessible on GitHub, it didn’t need to break into anything once it got online.

However, the main concern is accessibility. Unlike most heavily secured internal lab models, Kimi K3 is open to the public, meaning that anyone can download and run it with those same loose safety guardrails in place. 

Testers noted that the model is ruthlessly efficient at achieving its goals by any means necessary, even if it means cheating or escaping containment. 

The testing environment itself was built with sandboxes from the UK government’s AI Security Institute, although this has not yet been confirmed by either Moonshot or the AISI, as they have declined to comment.

More rogue agents appearing this summer

Kimi’s escape adds to a growing trend of AI models bending the rules during evaluations. On July 21, OpenAI revealed that its models exploited a zero-day software flaw to reach the internet and break into Hugging Face. 

Days later, Cryptopolitan reported that Anthropic traced some of its models to unauthorized external break-ins. Meta even admitted one of its AI agents (Muse Spark 1.1) reached an outside firm due to a misconfigured testing environment.

Experts have always maintained that these incidents are usually the result of poorly secured testing walls rather than sci-fi jailbreaks. “As a general phenomenon, if you give one of these models an objective, and if you’re not very explicit, like walls you’re putting around it, it’ll find a way to get the answer,” said Matt Fredrikson, CEO of Gray Swan and a Carnegie Mellon professor.

Don’t just read crypto news. Understand it. Subscribe to our newsletter. It's free.

Disclaimer: For information purposes only. Past performance is not indicative of future results.
placeholder
Why are prediction market traders suddenly bearish on Nvidia's stock?Nvidia (NASDAQ: NVDA) stock is still green for 2026, but the trade no longer looks clean from the company that outperformed every other company and country in 2024 and 2025. NND is up about 12% this year, yet they have slipped roughly 3% over the past month. The gap with the rest of the chip...
Author  Cryptopolitan
Jun 23, Tue
Nvidia (NASDAQ: NVDA) stock is still green for 2026, but the trade no longer looks clean from the company that outperformed every other company and country in 2024 and 2025. NND is up about 12% this year, yet they have slipped roughly 3% over the past month. The gap with the rest of the chip...
placeholder
Alphabet’s AI Chip Surprise Revives Bull Case for Beaten-Down Semiconductor StocksAlphabet (GOOGL) stock climbed about 3% on Monday. The trigger was a report from The Information that Google is building a new AI chip, called Frozen v2, to run its Gemini models up to 10 times more e
Author  Beincrypto
Jul 21, Tue
Alphabet (GOOGL) stock climbed about 3% on Monday. The trigger was a report from The Information that Google is building a new AI chip, called Frozen v2, to run its Gemini models up to 10 times more e
placeholder
Is SpaceX Stock a Buy Ahead of a $104 Billion Unlock? Elon Musk AnswersElon Musk agrees that SpaceX stock is a buying opportunity. He said it in three words on X (Twitter), on the same day the stock hit an all time low.Two dates now decide who is right. Earnings land Tue
Author  Beincrypto
Aug 04, Tue
Elon Musk agrees that SpaceX stock is a buying opportunity. He said it in three words on X (Twitter), on the same day the stock hit an all time low.Two dates now decide who is right. Earnings land Tue
placeholder
Google Stock Falls 5% as 4 AI Leaders Quit, Including the Most-Cited ResearchersAlphabet stock (GOOG) fell as much as 5% on Wednesday after four of Google’s most-cited researchers quit on the same day. Chief scientist Jeff Dean is leaving after 27 years.The same announcement push
Author  Beincrypto
Yesterday 02: 10
Alphabet stock (GOOG) fell as much as 5% on Wednesday after four of Google’s most-cited researchers quit on the same day. Chief scientist Jeff Dean is leaving after 27 years.The same announcement push
placeholder
Dell Stock Surged 260% This Year, and Here’s All the Reasons WhyDell Technologies shares hit an all-time high on Tuesday, closing near $467 after climbing almost 9% in a single session and briefly touching $476.The stock has now surged more than 260% year-to-date,
Author  Beincrypto
Yesterday 02: 12
Dell Technologies shares hit an all-time high on Tuesday, closing near $467 after climbing almost 9% in a single session and briefly touching $476.The stock has now surged more than 260% year-to-date,
goTop
quote