Meta's 'Open' Muse Glimmer Model Can Run On a Single Computer (engadget.com) 27
Meta has released Muse Glimmer, a slimmed-down open-weight AI model designed to run locally on a single GPU for agent tasks such as scheduling, file management, coding, and tool use. The release is based on Meta's closed Muse Spark 1.2 model and appears to be aimed at attracting developers who want capable AI agents without relying entirely on cloud-hosted services. Engadget reports: Facebook said that it's making the "weights" that AI systems use to choose responses available to everyone on Hugging Face along with developer documentation. The download is available for free, and users can run the model on their own PCs. The company noted that optimized integrations will land on llama.cpp and other sites, "so you can go from download to working agent in minutes."
The model is powerful for its size, according to Meta, with the "strong success rates" on benchmarks like DeepSearch QA, MCP-Atlas and SWE-Bench (which evaluates its ability write and debug code). It also supports reliable tool use, multi-step reasoning, failure recovery, multimodal input and scaffold compatibility for work with OpenClaw and other agent orchestrators. It was trained on data from over 100 languages, the company added. "Rather than centralizing superintelligence, we should distribute it widely and give every person the ability to direct it," CEO Mark Zuckerberg said in an essay accompanying Muse Glimmer's release. "This has the potential to begin a new era of personal empowerment where individuals can use this powerful new capability to reach their full potential, pursue their interests, and improve their lives and the world more than ever before."
The model is powerful for its size, according to Meta, with the "strong success rates" on benchmarks like DeepSearch QA, MCP-Atlas and SWE-Bench (which evaluates its ability write and debug code). It also supports reliable tool use, multi-step reasoning, failure recovery, multimodal input and scaffold compatibility for work with OpenClaw and other agent orchestrators. It was trained on data from over 100 languages, the company added. "Rather than centralizing superintelligence, we should distribute it widely and give every person the ability to direct it," CEO Mark Zuckerberg said in an essay accompanying Muse Glimmer's release. "This has the potential to begin a new era of personal empowerment where individuals can use this powerful new capability to reach their full potential, pursue their interests, and improve their lives and the world more than ever before."
What card? (Score:2)
So do I finally need to buy an RTX 5090?
Re: (Score:3)
At what cost? (Score:2)
If you want to use it for coding or something else requiring a very large context then yea, 5090 with 32gb.
If only that was cheaper than hiring a human to do it..
Re:What card? (Score:4, Interesting)
And how f-d up is that? Since when does computer hardware increase in value after it was released?
Re: What card? (Score:2)
If you have a decent CPU it might run on that at half speed or so.
Re: (Score:2)
Feels like such an announcement should include a few examples and their relative performance stats. Instead, the main linked article includes nothing of the sort, and Meta's announcement only notes it would fit in 20gb of memory, so you'd only need a 24gb or 32gb card - normal consumer level stuff, right? /s
Re: (Score:2)
MSI Spark. Thank me later.
Re: (Score:3)
16GB plus some offloading for the 4-Bit version I think and people say it quantizes with little quality loss.
Last time I checked the cards recommended were: 5060 Ti for starters (16 GB best value), 3090 if you need 24GB VRAM and 5090 if you need 32 GB. The xx90 are now quite unaffordable and the 5060 Ti now isn't exactly cheap either.
Re: (Score:2)
I'm seeing 3090s go for $1000-1200, which isn't actually much more than they were a few years ago. They haven't sky rocketed in price like the 5090. Solid points on the rest.
Given the scarcity of RAM... (Score:2)
...reported here [slashdot.org], the fact that it runs on your computer does not mean that it will answer your questions....
Re: (Score:2)
appointments (Score:2)
But can/will it hack the scheduling system at the doctor's office and get me the appointment I need? I mean the one I really want.
Zuck Wakes Up (Score:4, Interesting)
Zuckerberg has finally realized his shop doesn't have the AI muscle to compete with the top dogs, so instead he's emulating the Chinese model of releasing open models in an effort to undercut OpenAI, Anthropic, etc.
Re: (Score:2)
Maybe Dario was mean to him.
Re: (Score:2)
- Widespread use will expose the vulnerabilities/inadequacies of the system,
- Widespread use will encourage hackers/tinkerers to solve those vulnerabilities/inadequacies,
- An attempt to gain a large market share that will encourage corporate purchases,
- An attempt to force vendor lock-in to the Meta ecosystem, that can monetize all users
intrests (Score:2)
ok zuck, tell us how they can use the meta pedo glasses to use this powerful new capability to pursue their interests,
Re: Good, we're eventually going to need it (Score:2)
Sovereign systems
https://www.scry.llc/2026/07/2... [scry.llc]
you don't hand your national economy over to Anthropic or OpenAI.
Same applies to State and local governments, etc
'24 GB or 32 GB envelope' (Score:2)
Why is there no AI cube that has a general purpose CPU, plenty of ram and a GPU with 64GB+ that you can just plop on your intranet? Or is that simply not cost competitive to a cloud AI in your neighbor's back yard that has debt pouring out the wazoo and heavily discounted by the local government.
Re: (Score:2)
Strix Halo. That thing has 128 GB of RAM and fast interconnect if you buy two which is enough to run the deepseek flash model. You're currently at 10k for two, though. So if you like to tinker or want privacy that's an option. If you like it cheap you can get a lot of cloud usage for the money.
Re: (Score:2)
There are plenty of those. Mac Minis were hard to get for a while because the shared memory architecture is really good for inferencing, for running a model like this.
nvidia calls theirs spark, AMD calls theirs the "Halo Developer Platform" but both are just a small box with a fair bit of CPU and 128 gigs of shared memory with the GPU.
Re: (Score:2)
There are a bunch of them. People were chasing after Mac minis, which are almost cubes, just a little squashed.
This is a good thing (Score:2)
Local control is good
The cloud is a trap
Hopefully, affordable hardware will be available in the future and the need to use the cloud will diminish
Is this supposed to be new? (Score:1)
What? (Score:2)