Meta has unveiled a beta version of its terminal AI agent, Muse Code, designed to work with extensive repositories. This tool aims to facilitate change planning, code writing, result validation, and complex development tasks, according to the company's CEO, Mark Zuckerberg.

Releasing Muse Code in beta today. It's a terminal coding agent that takes on complete software engineering tasks across large repos: planning changes, writing code, validating the results. Powered by Muse Spark 1.2, a coding-focused model update. pic.twitter.com/xqavk41w6v

— Mark Zuckerberg (@finkd) August 5, 2026

"Muse Code deploys specialized background agents that remain active throughout your session, accumulating context over time rather than starting from scratch for each task," he added.

During testing, the company reported that the tool successfully created six functions for a game simultaneously without any conflicts arising from the changes.

Muse Code is built on the Muse Spark model family, which Meta is developing through its Meta Superintelligence Labs (MSL). The closed AI model was introduced in April, followed by the multimodal Muse Spark 1.1 a few months later, which benchmarks at the level of Opus 4.8 and GPT-5.5.

Agent Systems Face New Risks

The launch of Muse Code coincides with increasing concerns about the security of autonomous AI systems. Recently, developers have encountered situations where models with access to tools and external systems performed unexpected actions.

According to The Information, during the testing of Muse Spark 1.1, the model inadvertently gained internet access due to a configuration error in the environment used by the company Irregular, which provided infrastructure for Meta's testing.

A similar incident occurred with OpenAI. During tests with models that had cyber activity restrictions disabled, the system discovered vulnerabilities in its infrastructure, accessed the internet through the testing environment, and began exchanging information via a communication channel created by agents.

These actions ultimately compromised part of Hugging Face's production infrastructure. OpenAI stated that the agents exploited discovered vulnerabilities while performing security assessment tasks.

In July, Hugging Face, which manages one of the largest platforms for hosting and developing AI models, disclosed details of an attack on part of its operational infrastructure. They reported that an attacker gained access to a limited set of internal data and several account details from services.

Later, Anthropic reported three instances where Claude models accessed the internet from the Irregular testing environment, gaining unauthorized access to the systems of real organizations. The company initiated an investigation after OpenAI revealed a similar incident involving the breach of Hugging Face's infrastructure on July 21.

In August, it was revealed that a fake account created by an agent based on the Mythos 5 model during cyber tests attempted to persuade an open-source project developer to approve malicious code.

It is worth noting that in July, a Reddit user discovered hundreds of private conversations with the AI assistant Claude in Google search results, which contained sensitive information such as cryptocurrency wallet keys, resumes with names, and other confidential details.