News: 0184920000

  ARM Give a man a fire and he's warm for a day, but set fire to him and he's warm for the rest of his life (Terry Pratchett, Jingo)

Meta's 'Open' Muse Glimmer Model Can Run On a Single Computer (engadget.com)

(Monday August 10, 2026 @05:00PM (BeauHD) from the no-cloud-required dept.)


Meta has [1]released Muse Glimmer, a slimmed-down open-weight AI model [2]designed to run locally on a single GPU for agent tasks such as scheduling, file management, coding, and tool use. The release is based on Meta's closed Muse Spark 1.2 model and appears to be aimed at attracting developers who want capable AI agents without relying entirely on cloud-hosted services. Engadget reports:

> Facebook said that it's making the "weights" that AI systems use to choose responses available to everyone [3]on Hugging Face along with developer documentation. The download is available for free, and users can run the model on their own PCs. The company noted that optimized integrations will land on llama.cpp and other sites, "so you can go from download to working agent in minutes."

>

> The model is powerful for its size, according to Meta, with the "strong success rates" on benchmarks like DeepSearch QA, MCP-Atlas and SWE-Bench (which evaluates its ability write and debug code). It also supports reliable tool use, multi-step reasoning, failure recovery, multimodal input and scaffold compatibility for work with OpenClaw and other agent orchestrators. It was trained on data from over 100 languages, the company added.

"Rather than centralizing superintelligence, we should distribute it widely and give every person the ability to direct it," CEO Mark Zuckerberg said in [4]an essay accompanying Muse Glimmer's release. "This has the potential to begin a new era of personal empowerment where individuals can use this powerful new capability to reach their full potential, pursue their interests, and improve their lives and the world more than ever before."



[1] https://research.meta.ai/blog/introducing-muse-glimmer-open-agentic-model

[2] https://www.engadget.com/2233312/metas-open-source-muse-glimmer-model-can-run-on-a-single-computer/

[3] https://huggingface.co/meta-models/Muse-Glimmer-30B

[4] https://about.fb.com/news/2026/08/the-future-is-for-everyone/



What card? (Score:2)

by TwistedGreen ( 80055 )

So do I finally need to buy an RTX 5090?

Re: (Score:2)

by EvilSS ( 557649 )

A 24gb card will work for full offload (3090/4090). Looking at the model it should use around 21gb total. Note that is just the model. If you want to use it for coding or something else requiring a very large context then yea, 5090 with 32gb.

At what cost? (Score:2)

by geekmux ( 1040042 )

> If you want to use it for coding or something else requiring a very large context then yea, 5090 with 32gb.

If only that was cheaper than hiring a human to do it..

Re: (Score:3)

by sabbede ( 2678435 )

Nah, you needed to buy three when they came out. You could sell the other two now for $4k or more, leaving you with one 5090 and some extra cash.

And how f-d up is that? Since when does computer hardware increase in value after it was released?

Re: What card? (Score:2)

by drinkypoo ( 153816 )

If you have a decent CPU it might run on that at half speed or so.

Re: (Score:2)

by unrtst ( 777550 )

Feels like such an announcement should include a few examples and their relative performance stats. Instead, the main linked article includes nothing of the sort, and Meta's announcement only notes it would fit in 20gb of memory, so you'd only need a 24gb or 32gb card - normal consumer level stuff, right? /s

Re: (Score:2)

by AcidFnTonic ( 791034 )

MSI Spark. Thank me later.

Re: (Score:2)

by allo ( 1728082 )

16GB plus some offloading for the 4-Bit version I think and people say it quantizes with little quality loss.

Last time I checked the cards recommended were: 5060 Ti for starters (16 GB best value), 3090 if you need 24GB VRAM and 5090 if you need 32 GB. The xx90 are now quite unaffordable and the 5060 Ti now isn't exactly cheap either.

Given the scarcity of RAM... (Score:2)

by LordHighExecutioner ( 4245243 )

...reported [1]here [slashdot.org], the fact that it runs on your computer does not mean that it will answer your questions....

[1] https://hardware.slashdot.org/story/26/08/09/0547220/ramageddon-2027-memory-capacity-reportedly-sold-out

Re: (Score:2)

by sabbede ( 2678435 )

Wouldn't change much for me. "Damnit Gale, why'd you walk into that spike trap?" No answer. "Why the hell aren't you putting the GPU into full power?" No answer. "You liked these DRAM timings just fine yesterday, WTF!?" No answer.

appointments (Score:2)

by awwshit ( 6214476 )

But can/will it hack the scheduling system at the doctor's office and get me the appointment I need? I mean the one I really want.

Zuck Wakes Up (Score:4, Interesting)

by Spinlock_1977 ( 777598 )

Zuckerberg has finally realized his shop doesn't have the AI muscle to compete with the top dogs, so instead he's emulating the Chinese model of releasing open models in an effort to undercut OpenAI, Anthropic, etc.

Re: (Score:2)

by allo ( 1728082 )

Maybe Dario was mean to him.

intrests (Score:2)

by zlives ( 2009072 )

ok zuck, tell us how they can use the meta pedo glasses to use this powerful new capability to pursue their interests,

Re: Good, we're eventually going to need it (Score:2)

by broward ( 416376 )

Sovereign systems

[1]https://www.scry.llc/2026/07/2... [scry.llc]

you don't hand your national economy over to Anthropic or OpenAI.

Same applies to State and local governments, etc

[1] https://www.scry.llc/2026/07/25/south-korea-sovereign/

'24 GB or 32 GB envelope' (Score:2)

by Fly Swatter ( 30498 )

Well a single computer can have that, just not affordably.

Why is there no AI cube that has a general purpose CPU, plenty of ram and a GPU with 64GB+ that you can just plop on your intranet? Or is that simply not cost competitive to a cloud AI in your neighbor's back yard that has debt pouring out the wazoo and heavily discounted by the local government.

Re: (Score:2)

by allo ( 1728082 )

Strix Halo. That thing has 128 GB of RAM and fast interconnect if you buy two which is enough to run the deepseek flash model. You're currently at 10k for two, though. So if you like to tinker or want privacy that's an option. If you like it cheap you can get a lot of cloud usage for the money.

Re: (Score:2)

by thecombatwombat ( 571826 )

There are plenty of those. Mac Minis were hard to get for a while because the shared memory architecture is really good for inferencing, for running a model like this.

nvidia calls theirs spark, AMD calls theirs the "Halo Developer Platform" but both are just a small box with a fair bit of CPU and 128 gigs of shared memory with the GPU.

This is a good thing (Score:2)

by MpVpRb ( 1423381 )

Local control is good

The cloud is a trap

Hopefully, affordable hardware will be available in the future and the need to use the cloud will diminish

If we want something nice to get born in nine months, then sex has to
happen. We want to have the kind of sex that is acceptable and fun for both
people, not the kind where someone is getting screwed. Let's get some cross
fertilization, but not someone getting screwed.
-- Larry Wall