Models
August 10, 2026

NVIDIA brings Meta’s Muse Glimmer to local agentic AI workflows

NVIDIA is supporting Meta’s Muse Glimmer, a 30-billion-parameter open-weight model designed for long-running local AI agents, with a 120K-plus context window and deployment across NVIDIA GPU platforms.

NVIDIA has detailed support for Meta’s Muse Glimmer, a 30-billion-parameter open-weight dense model designed for local, long-running agentic AI workloads.

Featuring a 120K-plus context window, the model activates every parameter per token to provide predictable latency, reliable instruction following, and sustained long-context performance.

Muse Glimmer can run fully on-device across NVIDIA platforms including GeForce RTX 5090, DGX Spark, DGX Station, and Jetson, helping keep sensitive data local. NVIDIA reports throughput exceeding 20,000 tokens per second per GPU on Blackwell Ultra and supports deployment through NVIDIA NIM, SGLang, vLLM, and NemoClaw for agentic workflows.

#
Nvidia

Read Our Content

See All Blogs
LLM Models

GPT-5.6 benchmarks: The full testing breakdown

Deveshi Dabbawala

August 10, 2026
Read more
LLM Models

LLM testing of Claude Opus 5: The first enterprise-ready frontier AI

Sarankumar S

August 10, 2026
Read more