Muse Spark 1.1 icon

Muse Spark 1.1

Muse Spark 1.1 is a multimodal reasoning model from Meta Superintelligence Labs for agentic tasks, coding, computer use, and multimodal understanding. It is available in public preview through the Meta Model API and in Thinking mode in the Meta AI app and on meta.ai.

Muse Spark 1.1

Overview

Muse Spark 1.1 is a multimodal reasoning model from Meta Superintelligence Labs designed for agentic tasks. Meta positions it as an upgrade over Muse Spark, with stronger tool and computer use, coding, and multimodal understanding.

The model is available through the new Meta Model API in public preview, and it is also available in Thinking mode in the Meta AI app and on meta.ai. The launch emphasizes long-context work, orchestration across tools and subagents, and workflows that require both perception and action.

Features

Agentic task orchestration

The model can gather context, make a plan, and delegate work across parallel subagents for tasks that span multiple apps and services.

Native tool and MCP support

Meta says Muse Spark 1.1 zero-shot generalizes to new native tools, MCP servers, and custom skills.

1 million token context management

The model is trained to actively manage a very large context window, retain earlier actions, and compact work while preserving important steps.

Computer-use workflow handling

It can switch between scripting, clicking, and batching actions depending on whether automation or direct interaction is faster.

Coding on large codebases

Meta says it performs better on real-world coding tasks such as bug fixing, feature implementation, and code migrations.

Multimodal reasoning and grounded outputs

The model can work with visual and audio inputs and produce outputs useful for tasks like image and video captioning or visual-to-code workflows.

Use Cases

  • Agent workflows across apps

    Useful when a task requires planning, context collection, and coordinated execution across multiple tools or services.

  • Computer-assisted desktop work

    Fits workflows that involve unfamiliar interfaces, changing requirements, or long sessions where the model needs to preserve context.

  • Software development

    Applicable to debugging, implementing features, tracing issues across a codebase, and validating changes in agentic coding setups.

  • Multimodal content workflows

    Useful when a workflow starts from images, video, or audio and ends in an action such as captioning, code generation, or browser-based execution.

  • Marketplace and listing tasks

    Meta highlights a workflow where the model uses smartphone video, extracts useful photos, reasons about the product, and helps create a Facebook Marketplace listing.

Pros and Cons

Pros

  • Stronger tool use, coding, and multimodal understanding than the earlier Muse Spark release.
  • Designed for long-horizon work with multi-agent orchestration and large-context management.
  • Available both through the Meta Model API and in Meta AI surfaces for broader access.
  • Meta reports stronger resistance to jailbreaks, prompt injection, and other adversarial attacks in its safety evaluations.

Cons

  • Public preview status means availability and behavior may still change.
  • The source material does not provide pricing, plan limits, or detailed API documentation.

FAQ

What is Muse Spark 1.1?

Muse Spark 1.1 is Meta's multimodal reasoning model for agentic tasks, released as an upgrade over Muse Spark.

Who can use it?

Developers can access it through the Meta Model API in public preview, and end users can use it in Thinking mode in the Meta AI app and on meta.ai.

What kinds of tasks is it built for?

Meta highlights planning and orchestration across tools, computer use, coding, and multimodal understanding.

Does it support long-context workflows?

Yes. Meta says the model can actively manage a 1 million token context window and preserve important steps over long workflows.

Is pricing published?

The provided source material does not include pricing details, and the pricing page was not available in the collected evidence.

Quick Facts

Category
Multimodal reasoning model
Primary users
Developers and Meta AI users
Access
Public preview via the Meta Model API; Thinking mode in Meta AI app and on meta.ai
Source domain
ai.meta.com
Context window
1 million tokens
Notable workflow
Agentic tasks across tools, apps, and subagents

Muse Spark 1.1 Alternativen

CreateOS Sandbox icon

CreateOS Sandbox

CreateOS Sandbox is an isolated compute environment for running code and agent workloads inside Firecracker micro-VMs. It is designed for workflows that need machine-level isolation, private networking between sandboxes, and programmatic control through SDK, CLI, or MCP.

AakarDev AI icon

AakarDev AI

AakarDev AI helps teams manage AI provider access, project-level setups, logs, and analytics from one dashboard. It supports BYOK workflows and lists providers including OpenAI, Google Gemini, Anthropic, Groq, Mistral AI, and Perplexity AI.

ByteAsk icon

ByteAsk

ByteAsk is a terminal-first AI coding agent for C and C++ that edits repositories and verifies changes with the real compiler, debugger, sanitizers, and tests before showing a diff. It offers a free tier plus paid plans, with editor connectors and zero-retention handling described in the source.

Codex Plugins icon

Codex Plugins

Codex Plugins bundle reusable skills, app integrations, and MCP servers into workflows you can install in the Codex app or use from Codex CLI. They help extend Codex with connected-service tasks, reusable instructions, and shared team workflows.

hob icon

hob

hob is an independent workspace for coding agents that keeps agent sessions, terminals, history, and follow-up work organized around the tools and providers you already use. It is aimed at developers who want local control over routing, history, and workspace structure rather than a bundled model stack.

Wallie icon

Wallie

Wallie is an open-source AI streamer that watches your screen, hears chat, and generates live commentary in a configurable persona. It runs locally on your machine with your own keys and is aimed at faceless content, autonomous streams, and real-time reactions.