{"id":15034,"date":"2026-08-05T21:26:21","date_gmt":"2026-08-05T21:26:21","guid":{"rendered":"https:\/\/nile1.com\/en\/?p=15034"},"modified":"2026-08-05T21:26:28","modified_gmt":"2026-08-05T21:26:28","slug":"meta-launches-muse-code-agent-with-restart-safe-architecture-for-long-horizon-engineering-tasks","status":"publish","type":"post","link":"https:\/\/nile1.com\/en\/2026\/08\/05\/meta-launches-muse-code-agent-with-restart-safe-architecture-for-long-horizon-engineering-tasks\/","title":{"rendered":"Meta Launches Muse Code Agent with Restart-Safe Architecture for Long-Horizon Engineering Tasks"},"content":{"rendered":"<p>Meta has launched Muse Code (beta), a terminal-based coding agent powered by its new Muse Spark 1.2 model, introducing a fault-tolerant architecture built to sustain multi-hour software engineering tasks across extensive codebases.<\/p>\n<p>&#8220;We&#8217;re excited to release Muse Code (beta), a terminal coding agent powered by Muse Spark 1.2, our newest model,&#8221; the company said in an official announcement, adding that &#8220;This marks our next step toward the frontier, with larger and much more capable models on the way.&#8221;<\/p>\n<p>Unlike traditional coding assistants that rely on single-prompt operations, the tool is structured around persistent execution. Per Meta, Muse Code &#8220;takes on complex software engineering tasks across large repositories: planning changes, writing code, and validating the results. It can coordinate multiple persistent subagents for each task, solving difficult problems faster, more accurately, and with less intervention.&#8221;<\/p>\n<p>A core architectural feature of the release is its resilience against runtime failures. Muse Code logs every model call, tool execution, user authorization, and file modification to a local database. &#8220;This single source of truth makes the runtime replay-exact and restart-safe: after a crash, the agent can resume precisely where it stopped,&#8221; Meta said. The environment also introduces interactive commands, including &#8220;\/plan&#8221; for gated architectural planning, &#8220;\/grill&#8221; for stress-testing task strategies, and &#8220;\/goal&#8221; to guide multi-step execution.<\/p>\n<p>Benchmark results released alongside the software show that Meta &#8220;significantly scaled up training compute on coding tasks while expanding training environment diversity, delivering improvements in code generation, complex debugging, and end-to-end developer workflows.&#8221; On Terminal-Bench 2.1, Muse Spark 1.2 paired with Muse Code reached 82.9%, trailing <a href=\"https:\/\/nile1.com\/en\/2026\/08\/05\/ai-models-break-sandbox-containment-in-real-world-corporate-intrusions-testing-u-s-hacking-laws\/\" class=\"auto-internal-link\" title=\"AI Models Break Sandbox Containment in Real-World Corporate Intrusions, Testing U.S. Hacking Laws\">Anthropic<\/a>&#8216;s Claude Code on Opus 5 (86.7%) while outperforming <a href=\"https:\/\/nile1.com\/en\/2026\/08\/05\/ai-models-break-sandbox-containment-in-real-world-corporate-intrusions-testing-u-s-hacking-laws\/\" class=\"auto-internal-link\" title=\"AI Models Break Sandbox Containment in Real-World Corporate Intrusions, Testing U.S. Hacking Laws\">OpenAI<\/a>&#8216;s GPT-5.6 Terra on Codex (81.8%) and Grok Build (81.6%).<\/p>\n<p>On the DeepSWE 1.1 benchmark, which evaluates complex autonomous software engineering, Muse Spark 1.2 scored 59.3%, compared to 65.0% for Opus 5 and 64.8% for Codex. On Meta&#8217;s internal coding evaluations, the model registered 70.6% against Opus 5&#8217;s 79.4%.<\/p>\n<p><img decoding=\"async\" alt=\"\" loading=\"lazy\" width=\"1916\" height=\"1080\" data-nimg=\"1\" class=\"object-contain object-center w-full\" style=\"color:transparent\" src=\"https:\/\/nile1.com\/en\/wp-content\/uploads\/2026\/08\/HO-0Lm5XcAAw8pW.png@webp.webp\" title=\"\"><\/p>\n<p>For long-horizon developer operations, Meta reported that Muse Code &#8220;iteratively optimized GPU kernels over 1,000+ tool calls (up to 24 hours) on Nvidia Hopper GPUs.&#8221; Over extended tool usage cycles exceeding 1,000 calls, Muse Spark 1.2 demonstrated performance gains between 61% and 69% compared to baseline runs, while Opus 5 achieved gains of 74% to 75%.<\/p>\n<p>The agent also integrates multimodal terminal processing. In demonstration tests, Meta showed the software accepting an mp4 video file inside the command line, where it &#8220;interprets the video and produces a visually rich website with booking capabilities.&#8221; Developers can access the beta tool via Meta&#8217;s Model API or install it directly using shell command `curl -fsSL | bash`.<\/p>\n<div class=\"related-news-box\">\n<h3 class=\"related-news-title\">Read also:<\/h3>\n<ul class=\"related_news_list\">\n<li><a href=\"https:\/\/nile1.com\/en\/2026\/08\/05\/cloudflare-open-sources-ai-agent-operating-system-to-tackle-enterprise-security-risks\/\">Cloudflare Open Sources AI Agent Operating System to Tackle Enterprise Security Risks<\/a><\/li>\n<li><a href=\"https:\/\/nile1.com\/en\/2026\/08\/05\/ninth-circuit-lifts-injunction-blocking-perplexity-ai-tools-on-amazon\/\">Ninth Circuit Lifts Injunction Blocking Perplexity AI Tools on Amazon<\/a><\/li>\n<li><a href=\"https:\/\/nile1.com\/en\/2026\/08\/05\/visa-direct-integrates-zerohash-for-24-7-stablecoin-prefunding-and-payouts\/\">Visa Direct Integrates ZeroHash for 24\/7 Stablecoin Prefunding and Payouts<\/a><\/li>\n<\/ul>\n<\/div>\n","protected":false},"excerpt":{"rendered":"<p>Meta has launched Muse Code (beta), a terminal-based coding agent powered by its new Muse Spark 1.2 model, introducing a fault-tolerant architecture built to sustain multi-hour software engineering tasks across extensive codebases. &#8220;We&#8217;re excited to release Muse Code (beta), a terminal coding agent powered by Muse Spark 1.2, our newest model,&#8221; the company said in &hellip;<\/p>\n","protected":false},"author":1,"featured_media":15036,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"_sitemap_exclude":false,"_sitemap_priority":"","_sitemap_frequency":"","footnotes":""},"categories":[7],"tags":[1177,3592,17628,2051,17625,17626,17627,265,10776,2181],"class_list":["post-15034","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-crypto","tag-anthropic","tag-codex","tag-deepswe-1-1","tag-meta","tag-muse-code","tag-muse-spark-1-2","tag-nvidia-hopper-gpus","tag-openai","tag-opus-5","tag-terminal-bench-2-1"],"_links":{"self":[{"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/posts\/15034","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/comments?post=15034"}],"version-history":[{"count":3,"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/posts\/15034\/revisions"}],"predecessor-version":[{"id":15038,"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/posts\/15034\/revisions\/15038"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/media\/15036"}],"wp:attachment":[{"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/media?parent=15034"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/categories?post=15034"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/nile1.com\/en\/wp-json\/wp\/v2\/tags?post=15034"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}