Models·2 min read
By BitsMindsSource: Meta

Meta Launches Muse Spark: Multimodal AI with Parallel Agents, Health Insights, and Shopping Intelligence

Meta Superintelligence Labs unveils Muse Spark, a fast multimodal model that deploys parallel AI subagents, analyzes medical images, generates functional websites, and surfaces social-powered shopping recommendations.

Meta Launches Muse Spark: Multimodal AI with Parallel Agents, Health Insights, and Shopping Intelligence
Share:

Meta has unveiled Muse Spark, its most powerful AI model to date, developed by Meta Superintelligence Labs. The model is built to handle complex reasoning across science, math, and health while introducing a range of practical consumer capabilities — from visual coding and shopping assistance to contextual search powered by real-time public social content.

One of Muse Spark's most distinctive features is its parallel subagent architecture. When given a complex task, the model can simultaneously launch multiple specialized agents working in concert. Asked to plan a trip, for example, Muse Spark might simultaneously draft an itinerary, compare destinations, and research activities — compressing hours of manual research into a single interaction. This multi-agent design positions Muse Spark as a step toward more autonomous AI workflows capable of handling open-ended, multi-step problems.

The model is natively multimodal, capable of analyzing images alongside text. A dedicated health assistance mode allows users to upload medical charts, lab results, or clinical images and receive detailed, physician-informed responses. Muse Spark can also generate functional websites and mini-games from text prompts, putting it in direct competition with OpenAI's GPT-5.4 in the creative coding arena. A shopping mode leverages Meta's vast social platform data to surface product recommendations and styling inspiration from creators and community trends — a capability no other major AI lab can match given Meta's unique social graph.

Muse Spark currently powers the Meta AI app and meta.ai website, with a rolling deployment underway to WhatsApp, Instagram, Facebook, Messenger, and Meta's AI glasses over the coming weeks. Meta described this as the first model in a new generation series, with each iteration validating the scientific advances of its predecessor before scaling further. Select API partners already have private preview access, and open-source versions are planned for future release. Meta Superintelligence Labs framed the launch as a deliberately scientific approach to model scaling — prioritizing depth of capability and reliability over raw parameter count.

Want AI news before everyone else?

The morning's most important AI stories, straight to your inbox. No fluff.

Related Articles

Many kinds of input, one Qwen context On a deep violet field, a filmstrip, an audio waveform, a landscape photograph and a text page curve toward a single luminous sphere bearing the official purple Qwen star. A label below reads “1M context”, representing text, images, audio and video sharing one million-token context window. Aa 1M context BITSMINDS.COM
Models

Qwen3.8-Omni-Flash Undercuts Gemini on Audio and Video

Hill Climb: GPT-6 Astra versus Claude Fable 5.1 A teal desert buggy climbs a sandstone ridge on the left. An orange jeep climbs a green hill on the right, under an arc of gold coins. The two landscapes meet at a diagonal divide. 7 GPT-6 ASTRA CLAUDE FABLE 5.1 VS HILL CLIMB BITSMINDS.COM
Models

GPT-6 Astra vs Claude Fable 5.1: Hill Climb

Gemini 3.8 Live: thinking while the conversation continues An editorial illustration in a dark blue and violet room. A carefully drawn studio microphone and a tilted smartphone flank a translucent speech bubble carrying the multicoloured Gemini star. A continuous luminous audio waveform travels between them. Above the bubble, a separate arc connects small search, reasoning and completion symbols, representing background work continuing during a spoken conversation. The phone screen and visual paths are conceptual, not a reproduction of Google's actual interface or internal reasoning. Original vector illustration for BitsMinds, gemini-3-8-live-extended-thinking-voice. 16 September 2026. GEMINI 3.8 LIVE EXTENDED THINKING Gemini 3.8 LIVE The conversation continues Thinking. Still talking. BITSMINDS.COM
Models

Gemini 3.8 Live Tops Voice AI and Undercuts GPT