Muse Glimmer 30B  ·  Released 10 Aug 2026

Meta gave it
away for free.

A 30-billion-parameter AI model that runs on a gaming PC you might already own. What it means if you're a team of one — and the part they buried.

The number in the fine print

28.4%

of prompt-injection attacks succeeded against it. Meta's pitch is an always-on agent with access to your files.

What actually shipped

Muse Glimmer 30B

$0

Apache 2.0

Free forever. Commercial use allowed. No per-token meter, ever.

~20GB

Runs on 24GB

One consumer GPU. People had it working on a 3090 within hours.

131K

Context + vision

Reads text and images. 100+ languages. Distilled from Meta's closed model.

Why a solo operator should care

01

The meter changes what you try

You don't test twenty hooks because twenty feels expensive. The real cost was never the money — it's the ideas you didn't test.

02

Client work stops being a gamble

Contracts, unreleased campaigns, someone's financials. Local means it never leaves the machine. That goes in a proposal.

03

Nobody covers for you

If your pipeline runs on an API and that API has a bad morning, you have a bad morning. When you are the team, it just stops.

What free-and-local actually unlocks

Batch work you keep skipping

200 product descriptions. 48 title variants. Every post reformatted for the newsletter. Not hard — just long.

Overnight jobs, because free

Point it at a folder of transcripts, get every quotable line by morning. On a metered API you'd think twice.

Volume as a taste multiplier

Not one draft — thirty. Most creators aren't short on taste. They're short on options to apply it to.

The honest trade

You're swapping a subscription for setup, troubleshooting, and a weekend of learning what breaks. That isn't free either.

Interactive  ·  pick your hardware

Will it run on your machine?

Pick a card on the left and I'll tell you which build to download, roughly what speed to expect, and whether it's worth your time.

Interactive  ·  Meta's own published numbers

Interactive  ·  fact check

"Meta rigged the benchmarks."

The claim: Meta tested Glimmer at one sampling setting and Qwen at another, then put them in the same chart as if it were a fair fight.

It's circulating right now. It sounds damning. It doesn't survive ten minutes of checking.

Those were each model's own settings.

Qwen's model card specifies top-k 20. Glimmer's specifies 64. Running a model the way its makers tell you to is normal practice, not bias. The analyst who first noted it never called it unfair — that spin got added in the retelling.

What does hold up: for each competitor score, Meta used whichever number was more favorable — the company's self-report, or Meta's own reproduction. No faking required for the chart to drift its way.

What people actually got on day one

The community beat the spec sheet

253

tokens/sec on a 5090

Higher than the 233 Meta published themselves.

14GB

100+ tool calls

A 2-bit build ran a full agent loop well under the stated 24GB floor.

3090

Full context, fits

A card people have owned for years. That's the actual story.

Individual user reports from launch day, not verified benchmarks.

Who this is for

Try it

  • You already own a 24GB+ GPU and it's mostly idle
  • You have repetitive volume work you keep putting off
  • You do client work under NDA and you've been uneasy about it

Skip it

  • You're on integrated graphics — smaller models are coming
  • You need computer-use automation today. Qwen wins right now
  • You were about to buy a GPU for this. Wait for real testing

Verdict: impressive engineering, real limitations, not ready for anything that matters yet. That's a week-one product. Just don't let anyone sell you week one as week fifty.

Twenty minutes, not a weekend

Your actual next step

1

Install LM Studio

Free, normal interface, no terminal. Search Glimmer, grab the 4-bit build.

2

Change nothing else

Text only. No vision file, no drafter, don't touch context. You're taking a temperature, not building a system.

3

Give it one real task

Not a test prompt — something on your list right now. Time it. Compare.

The question was never whether the model is impressive. It's whether it's useful to you — and watching another video won't answer that. Including this one.

Nobody can test everything alone

AI Creators
Roundtable

A private network of creators and builders using AI to automate, create, and scale — comparing what actually worked, what broke, and what wasted a weekend.

Join the Roundtable

Link in the description

01 / 12
Jayy The AI Creative
← → navigate  ·  O overview  ·  F fullscreen