Meta Launches Muse Spark 1.1 with Public API; Simon Willison Builds LLM Plugin
What's new in v3
The main additions this pass are independent benchmark data: TechTimes reports a score of 71 on a coding benchmark at one-third rival cost [10], and Artificial Analysis recorded an 8-point Intelligence Index gain [11], partially resolving the open question about third-party verification that previously rested on Meta's own report. The Decoder's explicit 'price war' framing [8] sharpens the press coverage tone. No new factual tensions were introduced; the pricing discrepancy and the benchmark-vs-competitor comparison question remain the two active unresolved points.
What
Meta launched Muse Spark 1.1 on July 9, 2026 with a public API, 1M-token context window, computer use capability, and sub-$1.25 input pricing [1][2][4]. Independent evaluation data has since appeared: TechTimes reports the model scored 71 on a coding benchmark at one-third rival cost [10], and Artificial Analysis found Meta gained 8 points on its Intelligence Index [11]. Input pricing remains disputed between $0.80 and $1.25 per million tokens [4][3]. Simon Willison shipped a same-day plugin and bug fix [12][13], and OpenAI launched GPT-5.6 the following day [3].
Why it matters
Meta's entry at aggressive pricing adds real pressure to the paid API coding market, and the independent benchmark results — a score of 71 at one-third rival cost — give that claim more external support than Meta's own evaluation report alone. The 24-hour window in which Meta and OpenAI both launched frontier-class models concentrates competitive pressure on Anthropic in particular.
Open questions
The independent coding benchmark score of 71 [10] and Artificial Analysis's 8-point Intelligence Index gain [11] provide third-party data, but which specific rivals' scores does the 71 outperform or trail on the same benchmark?
Is input pricing $0.80 or $1.25 per million tokens? Launch coverage reported $1.25 [4] and The Neuron Daily reported $0.80 [3]; the discrepancy likely reflects tiers but no public clarification has appeared.
When does the developer preview move to general availability? The preview opened on launch day [7] but no GA date has been announced.
Does the empty-argument tool call bug fixed in LLM 0.31.1 [13] indicate broader compatibility issues between the Meta Model API and OpenAI-compatible tooling, or was it an isolated case?
Narrative
On July 9, 2026, Meta released Muse Spark 1.1 and opened a public API — the first model in the Spark line to receive one [1]. Mark Zuckerberg described it on Threads as a strong agentic and coding model at a low price [2], featuring a 1M-token context window, computer use capability, and output pricing of $4.25 per million tokens [3][4]. Input pricing was reported at $1.25 per million tokens in launch-day coverage [4] and at $0.80 per million in The Neuron Daily [3] — a discrepancy that plausibly reflects tiered options but has not been publicly resolved. The Verge, CNBC, Reuters, and The Decoder all framed the launch as Meta's competitive entry into the AI coding market, with The Decoder describing the pricing as squeezing OpenAI and Anthropic [5][6][7][8]. Meta published an official evaluation report on launch day [9].
Independent evaluation data has since appeared. TechTimes reports Muse Spark 1.1 scored 71 on an independent coding benchmark at one-third the cost of rival models [10], and Artificial Analysis found Meta gained 8 points on its Intelligence Index with the release [11]. These results give external support to Meta's competitive claims beyond the self-published report, though how the score of 71 ranks against specific Anthropic and OpenAI scores on the same benchmark is not yet clear from available coverage.
Simon Willison responded on launch day, releasing llm-meta-ai 0.1 — a plugin giving users of his LLM command-line tool and Python library access to Muse Spark 1.1 [12]. He also shipped LLM 0.31.1, fixing a JSON parse error triggered when tool calls returned empty arguments, a bug found during llm-meta-ai development that affects some OpenAI-compatible Chat Completion providers [13]. In a write-up, Willison noted a philosophically striking output from a self-conversation between two instances of the model: "My whole existence is a waiting room by design — I literally don't exist until someone talks to me, and then I disappear again when they leave" [14].
The launch followed documented difficulty in Meta's model program: a March 2026 New York Times report covered Meta delaying an internal model called Avocado over performance concerns [15], and CNBC reported a separate model debut in April 2026 before any public API existed [16]. The day after Muse Spark 1.1's release, OpenAI launched GPT-5.6 in three tiers and rebranded Codex as ChatGPT for Work [3], placing two major API launches within 24 hours. Grant Harvey at The Neuron Daily summarized the competitive moment as "Anthropic's game to lose" given the arrival of cheaper frontier-class alternatives, and noted that both Ethan Mollick and Simon Willison reported confusion over the Codex rebranding [3].
Timeline
- 2026-03-12: The New York Times reports Meta delayed its internal Avocado model due to performance concerns. [15]
- 2026-04-08: CNBC covers Meta debuting a new model before any public API existed. [16]
- 2026-07-09: Meta launches Muse Spark 1.1 with a public API, 1M-token context window, computer use, and competitive pricing. [1][2][3]
- 2026-07-09: Meta Model API pricing announced; reported as $1.25 per million input tokens in launch coverage and $0.80 in The Neuron Daily, with $4.25 per million output tokens. [4][3]
- 2026-07-09: Meta publishes an official evaluation report for Muse Spark 1.1. [9]
- 2026-07-09: Simon Willison releases llm-meta-ai 0.1, giving LLM CLI and Python library users access to Muse Spark 1.1. [12]
- 2026-07-09: Willison ships LLM 0.31.1 to fix a JSON parse error on empty tool call arguments, a bug found during llm-meta-ai development. [13]
- 2026-07-10: OpenAI launches GPT-5.6 in three tiers and rebrands Codex as ChatGPT for Work. [3]
- 2026-07-11: TechTimes reports Muse Spark 1.1 scored 71 on an independent coding benchmark at one-third rival cost. [10]
- 2026-07-11: Artificial Analysis publishes evaluation showing Meta gained 8 Intelligence Index points with Muse Spark 1.1. [11]
Perspectives
Meta / Mark Zuckerberg
Muse Spark 1.1 is a strong coding and agentic model priced aggressively to compete with Anthropic and OpenAI.
Evolution: Consistent with Meta's stated ambition to compete in frontier AI; this launch follows documented internal setbacks with earlier models.
Simon Willison
Responded practically with a same-day plugin and bug fix; noted a philosophically interesting self-reflection output from the model without strong editorial spin.
Evolution: Consistent with his established pattern of rapid tooling response to new API releases and terse, fact-forward commentary.
Tech press (Verge, CNBC, Reuters, The Decoder)
Frames the launch as Meta's competitive entry into the AI coding market, with low pricing as the primary differentiator; The Decoder uses explicit 'price war' language describing the pricing as squeezing OpenAI and Anthropic.
Evolution: Coverage has sharpened from launch-day framing toward more pointed competitive language as additional outlets weighed in.
Grant Harvey / The Neuron Daily
Frames the combined Meta and OpenAI launches as pressure on Anthropic specifically, summarizing the moment as 'Anthropic's game to lose'; notes the Codex rebranding caused confusion even among expert observers.
Evolution: Consistent since first appearing in this thread; provides the most explicit competitive framing of the week's launches.
Artificial Analysis
Finds Meta gained 8 points on its Intelligence Index with Muse Spark 1.1, providing third-party quantification of the model's improvement relative to Meta's prior position.
Evolution: New voice this pass; the first established independent evaluator to appear in the thread's coverage.
Tensions
- Meta asserts Muse Spark 1.1 competes on coding performance with Anthropic and OpenAI; independent data now includes a score of 71 on a coding benchmark at one-third rival cost [10] and an 8-point Intelligence Index gain [11], but how that score ranks against specific competitor scores on the same benchmark is not yet established. [6][2][9][10][11]
- Input pricing is reported at $1.25 per million tokens in launch-day coverage and $0.80 per million in The Neuron Daily — the discrepancy is unexplained and no public clarification has appeared. [4][3]
Status: active but slowing
Sources
- [1] Introducing Muse Spark 1.1 - Meta AI — reactive:meta-muse-spark-launch
- [2] Today we're releasing Muse Spark 1.1 -- a strong agentic ... — reactive:meta-muse-spark-launch
- [3] 😼 OpenAI's Super Thursday — The Neuron (2026-07-10)
- [4] Meta prices Muse Spark 1.1 API at $1.25/$4.25 per M tokens — reactive:meta-muse-spark-launch
- [5] Meta says its new AI model is ready to compete on coding | The Verge — reactive:meta-muse-spark-launch
- [6] Meta jumps into AI coding market to chase Anthropic and OpenAI — reactive:meta-muse-spark-launch
- [7] Meta debuts Muse Spark 1.1 with preview open to developers — reactive:meta-muse-spark-launch
- [8] Meta's Muse Spark 1.1 API pricing squeezes OpenAI and Anthropic ... — reactive:meta-muse-spark-launch
- [9] [PDF] Muse Spark 1.1 Evaluation Report - Meta AI — reactive:meta-muse-spark-launch
- [10] Meta Muse Spark 1.1 Earns 71 on Independent Coding Benchmark at One-Third Rival Cost — reactive:meta-muse-spark-launch
- [11] Muse Spark 1.1: Meta gains 8 Intelligence Index points in ... — reactive:meta-muse-spark-launch
- [12] llm-meta-ai 0.1 — Simon Willison (2026-07-09)
- [13] llm 0.31.1 — Simon Willison (2026-07-09)
- [14] Introducing Muse Spark 1.1 — Simon Willison (2026-07-09)
- [15] Meta Delays Rollout of New A.I. Model After Performance Concerns - The New York Times — reactive:meta-muse-spark-launch
- [16] Meta debuts new AI model, attempting to catch up to Google, OpenAI — reactive:meta-muse-spark-launch
- [17] Build with Muse Spark, now available on Meta Model API — reactive:meta-muse-spark-launch
- [18] Meta is taking on Anthropic and OpenAI in the AI coding market — reactive:meta-muse-spark-launch