
check our model out!
Justin@JustinGorya·META COOOOOOKED!!! Wow. Meta muse spark 1.3 was just released and it seems to be amazing?! used my 2 custom skills and got this in a oneshot. The whole Website was a simple prompt and oneshot. Muse spark 1.3 is creative and good at frontend. approved.
you can try our model on opencode too!
Hamza@thegenioo·Muse Spark 1.3 is available on OpenCode Zen for free x.com/BennettBuhner/…
y'all are gonna love muse spark 1.3!
Tim Jayas@TimJayas·Muse Spark 1.3 >> Fable 5 muse spark just teleported into #3 spot and it's the first ever model to get in between Claude and GPT the pricing compared to fable is a joke, what a surprise launch! x.com/alexandr_wang/…
i really hate to say it, but… gemini who? 🏎️💨
Artificial Analysis@ArtificialAnlys·Meta has released Muse Spark 1.3, their fourth Muse Spark model release in five months. Muse Spark 1.3 (max), which is in limited preview for Meta’s partners, scores 62 on the Artificial Analysis Intelligence Index, behind only Claude Fable 5.1 and Claude Opus 5. The variant available now, Muse Spark 1.3 (xhigh), scores 61 and ties with GPT-5.6 Sol (max) and Grok 4.6 (high). Both variants’ gains come primarily from improvements in agentic work and scientific capabilities Muse Spark 1.3 (xhigh) enters the Artificial Analysis Intelligence Index at 61, up 4 points from Muse Spark 1.2 (57, August) and 8 points from Muse Spark 1.1 (53, July). It enters tied with GPT-5.6 Sol (max), Grok 4.6 (high), and Claude Opus 5 (high), and behind Claude Fable 5.1 (max, 66), Claude Opus 5 (max, 63), and Claude Fable 5 (max, 62) Muse Spark 1.3 (max), which is in a limited preview stage, lands at 62. This higher index score is enabled by gains vs. Muse Spark 1.3 (xhigh) in Tau3-Bench Banking (52% vs. 47%) and GDPval-AA v2 (1,754 Elo vs. 1,709). Muse Spark 1.3 (max) is second only to Claude’s Fable and Opus variants in total score Congratulations to @AIatMeta, @finkd, and @alexandr_wang on the release! Key Takeaways: ➤ Continued improvement on agentic knowledge work tasks. At the launch of Muse Spark 1.2, we noted its significant gains in agentic knowledge work performance vs. Muse Spark 1.1. The latest iteration continues this trend, with Muse Spark 1.3 (xhigh) demonstrating a notable 12-point gain vs. Muse Spark 1.2 in Tau3-Bench Banking (35% to 47%), a 5-point gain in Terminal-Bench 2.1 (80% to 85%), and a new GDPval-AA v2 Elo of 1709 against its predecessor’s 1615. Muse Spark 1.3 (max) improves further on Tau3-Bench Banking (52%) and GDPval-AA v2 (1,754 Elo). This Tau3-Bench Banking score is #1 among all models. Muse Spark 1.3 (max) achieves these higher agentic work scores by using more turns and total reasoning tokens, reasoning 62% more on GDPval-AA v2 and 28% more on Tau3-Bench Banking compared to Muse Spark 1.3 (xhigh) ➤ The lowest cost per task for any model at 59+ on the Artificial Analysis Intelligence Index. Muse Spark 1.3 (xhigh) costs $0.55 per Intelligence Index task at Meta's unchanged $1.25/$4.25 per 1M token pricing ($0.15 for cached input), with its peers GPT-5.6 Sol (max) and Grok 4.6 (high) costing $0.95 and $0.94 respectively, a 70%+ premium. This places Muse Spark 1.3 (xhigh) on the Pareto frontier for Intelligence vs. Cost per Task. Its cost per task is higher than Muse Spark 1.2 ($0.40 per task), driven by ~57% more input tokens per task on agentic evaluations, with output tokens up only ~8%. Pricing for Muse Spark 1.3 (max) is not yet publicly available ➤ Scientific Reasoning results rose across the board, led by CritPt. CritPt was the standout non-agentic score gain vs. Muse Spark 1.2, with a material +8 points for the xhigh variant (18% to 26%), and GPQA Diamond achieved +4 points (90% to 94%), while Humanity’s Last Exam and SciCode each gained a more modest 2-3 points (45% to 47% and 56% to 59%, respectively). Muse Spark 1.3 (max) achieved roughly similar scores to the xhigh variant, gaining 2 points in Humanity’s Last Exam, tying on GPQA Diamond, and losing a point on CritPt vs. Muse Spark 1.3 (xhigh) ➤ Minor regressions in only two evaluations. Both Muse Spark 1.3 (xhigh) and Muse Spark 1.3 (max) dropped 4 points in AA-LCR (83% to 79%) when compared to Muse Spark 1.2, and AA-Omniscience (Accuracy) fell 3 points for xhigh and 1 point for max. The drops in AA-Omniscience (Accuracy) are due to a higher abstention rate (not answering questions when unsure), which also lowered the hallucination rate for Muse Spark 1.3 (xhigh) Other model details (xhigh variant): ➤ Context window: 1M tokens, unchanged from Muse Spark 1.2 ➤ Pricing: unchanged from Muse Spark 1.2: $1.25/$4.25 per 1M input/output tokens, with cache hits discounted to $0.15 per 1M ➤ Input modalities: text, image, video ➤ Availability: Meta's first-party API and Muse Code
big improvements
AI/ML API@aimlapi·Muse Spark 1.3 vs 1.2: sculpt viking figurines @AIatMeta dropped Muse Spark 1.3 today — we ran it against Muse Spark 1.2 both models got the same brief: three collectible 3D figurines — a viking helmet, a diamond-studded axe, a longship with a crew — one self-contained HTML file each, one shot. Muse Spark 1.3: 22,474 tokens, $0.10 Muse Spark 1.2: 19,324 tokens, $0.08
some early eval results from Stata benchmark :)
khaled@eltokh7·Very impressive as usual from @Meta and @alexandr_wang! Muse Spark 1.3 debuts as the 3rd top performing model in Stata Benchmark. It beats EVERY model except Claude Fable x.com/alexandr_wang/…
check out our new muse spark 1.3 model!
Mark Zuckerberg@finkd·Muse Spark 1.3 is rolling out today with frontier performance almost too cheap to meter. This is the biggest jump we've made so far on coding and agentic work. Try it in Muse Code and our API. Next up 🍉 and Muse Spark open weights releases coming soon.
1/ today we’re releasing muse spark 1.3—available in muse code & the meta model api. this is our most capable model yet—frontier performance almost too cheap to meter. much stronger at agentic and coding with better usability. we think users will really notice the jump.
2/It's a step up on reasoning and coding. compared to Muse Spark 1.2, Muse Spark 1.3 wastes fewer turns, uses ~20% fewer tool calls, ~25% fewer tokens, and holds onto requirements well during long-horizon tasks.
try the model out!
Artificial Analysis@ArtificialAnlys·Meta has released Muse Voice Transcribe, taking the #1 spot for Final Transcript accuracy on AA-WER Streaming with 3.1% WER at 0.16s after end of speech Muse Voice Transcribe is the first streaming Speech to Text model developed by Meta Superintelligence Labs. Meta states that the model was trained on more than 70 languages, with 25 extensively verified, and supports audio inputs exceeding one hour without required post-processing. It processes audio in 80ms chunks and is available through the Meta Model API, Meta AI for Mac and Muse Code. Key takeaways ➤ Final Transcript: Muse Voice Transcribe achieves 3.1% WER at 0.16s after end of speech. It is more accurate and faster than Cartesia Ink-2 (semantic endpoints) at 3.4% and 0.43s, and more accurate but slower than Cartesia Ink-2 (external endpoints) at 4.0% and 0.07s. It is also more accurate, though slightly slower, than ElevenLabs Scribe v2 Realtime at 3.6% and 0.14s. ➤ First Partial Transcript: The model achieves 3.6% WER at 0.13s, just ahead of ElevenLabs Scribe v2 Realtime on accuracy and latency. It is more accurate and faster than Cartesia Ink-2 (semantic endpoints) at 4.9% and 0.17s, and more accurate but slower than Cartesia Ink-2 (external endpoints) at 4.0% and 0.07s. ➤ Price: Muse Voice Transcribe costs $0.18 per hour, or $3 per 1,000 minutes. This is below Cartesia Ink-2 at $4 and less than half the $6.50 charged for ElevenLabs Scribe v2 Realtime and Deepgram Flux. See more details below ⬇️
1/ today we're rolling out muse voice transcribe, our first real-time audio perception model - SOTA in streaming speech-to-text. also handles speaker diarization and endpointing natively in a single model.
2/ it’s accurate and quick - it uses adaptive delay to decide how much context it needs before transcribing each word. trained on 70+ languages (with 25 validated at launch), handles code-switching mid-sentence, and manages hour+ sessions and 20+ speakers without post-processing.
1/ muse code is out of beta. launching with an sdk in developer preview to build your own agents on top of it and rolling out monthly subscription plans. copy-paste into your terminal: curl -fsSL dev.meta.ai/install.sh | bash
2/ it's built for complex work: agents that share context across sessions, and workflows that split a task across multiple teams of subagents to deliver one result.
top 3 model on opencode! future models will be even better 💪
OpenCode@opencode·meta muse spark has cracked top 3 with 11T tokens for the week
1/ muse image is live on the meta model api $0.01/image - one of the best price-to-quality ratios for production volumes
2/ try it out: developer.meta.com/ai/resources/b…
cool!
Design Arena@DesignArena·BREAKING: Muse Spark 1.2 by @AIatMeta takes 1st for Video-to-Website with an Elo rating of 1279 and impressive scores across all of our multimodal code categories. Muse Spark 1.2 also takes 2nd for Image-to-HTML with an Elo rating of 1252 and 3rd for Image-to-Frontend with Elo rating of 1272. At $1.25/1M input tokens and $4.25/1M output tokens, it lands on the price-preference Pareto frontier for all three categories. Congrats to the @Meta team on the achievement!
1/ muse spark 1.2 is a very strong multimodal model—it can do visual coding, robotics planning, and audio-visual understanding that all come together through agentic tools.
2/ muse spark 1.2 performs quite strongly across a wide variety of multimodal capabilities and evals.
check out the Meta AI desktop app! the dictation has been a game changer for me personally
Spencer Barnett@spencerbarnett·We launched the Meta AI Mac OS app today! 🚀 I particularly love the dictation feature which allows me to dictate anywhere on my computer with ultra high accuracy. Just hold down 'fn' and yap!
great initiative (RSI benchmark) from @ScaleAILabs and @mhrezaeics
MohammadHossein Rezaei@mhrezaeics·Launching t.co/jhIfXxZkmL: The work of AI R&D has always belonged to humans. For the first time, though, it no longer seems certain that it always will. Recursive self-improvement is within a line of sight. It may still be far, but it is close enough that we should start measuring it.
1/ big announcement today: we will be releasing an open weight version of muse spark 1.2 soon. we also are releasing muse glimmer, a 30B agentic model with open weights under apache 2.0. muse glimmer can run on 24GB of VRAM without losing agentic reliability. 🧵
2/ just like much larger models, muse glimmer can operate as a fully capable agent via planning, tool calls, checking its own results, and failure recovery.
excited to be releasing open weights for muse glimmer today, a 30b model that runs on a single consumer gpu, with open weights for a version of muse spark 1.2 coming soon. two very different models, both headed into people's hands, with more to come. x.com/finkd/status/2…
personal superintelligence should be available to everyone, and opening access to our models is abig part of that. read more from mark: meta.com/futureisforeve… x.com/finkd/status/2…
to put ai progress in perspective: 9 months ago: most developers wrote code by hand now: misaligned multi-agent swarm finding and collaborating on 0-days undetected (OpenAI/hugging face) 9 months in the future likely much crazier
muse spark 1.2 on the Pareto frontier x.com/artificialanly…
muse spark 1.2 is SOTA on finance agent v2! x.com/valsai/status/…
concerning that data companies serving the US government (mercor, surge) are also working with Chinese AI labs. serving the US government should not be a commercial convenience, it must be a bedrock principle for startups. x.com/aakashsabharwa…
great progress x.com/shuchaobi/stat…
apparently must spark 1.2 is very good at sidequests x.com/iam_zachi/stat…
our models got gold on a bunch of olympiads. this one hits home! 🏅 Asian Physics Olympiad (APhO): Perfect score, theory exam 🏅 International Physics Olympiad (IPhO): Perfect score, theory exam 🥇 International Mathematical Olympiad (IMO): Gold medal 🥇 International Chemistry Olympiad (IChO): Gold-medal-level performance 🥇 Romanian Masters of Mathematics (RMM): Gold-medal-level performance
🙏 plurality of great AI labs is good for the world! x.com/millionint/sta…
magic 🪄 x.com/goodhunt/statu…
the pelican improves x.com/simonw/status/…
check out muse code! x.com/charlieholtz/s…
update from vals on muse spark 1.2! x.com/ValsAI/status/…
good iteration for muse spark 1.2 onto bigger and better things 🍉 x.com/artificialanly…
in case you missed it—check out our research blog on muse code and muse spark 1.2 research.meta.ai/blog/introduci…
hell yea x.com/haider1/status…
check out muse code! x.com/craigweiss/sta…
check out muse spark 1.2 in muse code! x.com/mweinbach/stat…
muse code in beta is live. first coding agent from msl, built on muse spark 1.2. install: curl -fsS dev.meta.ai/install.sh | bash here's what you should know:
the harness architecture: multiple specialized agents coordinate in parallel on the same task. they persist across your session, building context instead of starting cold each time. for complex work, sub-agents fan out into isolated worktrees so your working copy stays clean.
muse code in beta is here: our first coding agent powered by our latest model, muse spark 1.2. one command to install and start building. get it through Meta Model API. x.com/finkd/status/2…
Welcome @fdesouza to @scale_AI as CEO! I started Scale a decade ago at 19, and it has grown to be the backbone of the AI world, powering frontier labs, Fortune 500 companies, and the US Government. Francis is an exceptional leader and the right steward. After spending a lot of time together, I have no doubts he’ll take Scale to greater heights. Thank you @jdroege for your leadership as interim CEO 🙏
Great letter from Mark articulating the principles guiding MSL: individual empowerment, invention over automation, and balance of power through broad access rather than concentrated control. x.com/finkd/status/2…
this is pretty cool x.com/atomic_chat_hq…
gemini who? 🏎️💨 x.com/firstadopter/s…
muse spark 1.1 is SOTA model on video-to-code 🔥 x.com/designarena/st…
muse spark 1.1 on eyebench x.com/adonis_singh/s…
Muse Spark 1.1 is now available on OpenRouter! This was highly requested by developers! Give it a whirl and let us know! x.com/MetaforDevs/st…
muse spark can refill your fridge for you! full circle, this problem was actual the inspiration for scale ai x.com/tjshibata/stat…
Muse Spark achieved a perfect score on the 2026 Asian Physics Olympiad! Great milestone against our long-term goal of building AI systems to accelerate science. x.com/AIatMeta/statu…
pricemogging x.com/MollySOShea/st…
Muse Spark 1.1 is SOTA on HealthBench Professional! It is the best health model out there :) x.com/MedicalSphereA…
muse spark 1.1 is #3 on Debate Benchmark, only behind Fable 5 and Opus 4.7 and ahead of GPT-5.6 Sol x.com/LechMazur/stat…
Muse Spark 1.1 is SOTA on the Radiologists Last Exam Handover Readiness Index (RadLE-H), nearing human expert performance. x.com/drdatta_aiims/…
muse spark 1.1 outperforms opus, grok 4.5, and gemini on a new challenging finite model theory / theoretical cs eval x.com/s_batzoglou/st…
muse spark can serve all your particle playground needs x.com/justpulket/sta…
okay this looks fun x.com/bijanbowen/sta…
ok the whale from muse spark 1.1 was a total surprise x.com/orisilver/stat…
um can’t we all be friends 🥺 x.com/bigtechalert/s…
muse spark is able to do end-to-end tasks based on short video instructions x.com/aiatmeta/statu…
new benchmark just dropped 🎁
hey everyone listen up x.com/daffaghiffaryk…
muse image explaining schrodinger’s cat! x.com/thelastgrammer…
I’d live in some of these homes designed by muse spark post upload, and they only cost a few cents to build x.com/thehypedotnews…
90% cheaper than Fable and awesome for building whatever you want x.com/_maxblade/stat…
make fun games with muse spark 1.1! x.com/jacksongrove/s…
muse spark 1.1 is really strong at computer use! x.com/yashvarpatel/s…
supersize ur websites with muse spark 1.1 x.com/heyaleksandr/s…
really good at computer use :) x.com/matt_feroz/sta…
if you want to make a minecraft-like planetary simulator, try out muse spark 1.1 x.com/jquave/status/…
muse spark 1.1 is ahead of gpt-5.6 on SciCode x.com/ArtificialAnly…
big jump on coding for muse spark 1.1 x.com/ArtificialAnly…
muse spark 1.1 is pretty good at frontend engineering x.com/tamirspiritt/s…
muse spark 1.1 can be even better than opus 4.8 at 20% of the cost x.com/panda_liyin/st…
make fun games with muse spark 1.1! x.com/jquave/status/…
woah this is actually pretty good x.com/rom1trs/status…
nice 👍 x.com/polymarketmone…
muse spark 1.1 is one of the best models for OpenClaw, HermesAgent, and other harnesses x.com/garrytan/statu…
compute daddy @dylan522p has spoken x.com/semianalysis_/…
check out muse spark 1.1! x.com/kimmonismus/st…