Benchmarks Architecture Text Decoder Perception Encoder Transformers Text-only Inference Prompting the model with images and text Video Inference Multimodal tool calling Object Detection Llama.cpp Speculative Decoding Speculative Decoding with transformers Speculative Decoding with llama.cpp Inference Endpoints Support for Muse Glimmer vLLM with transformers backend Fine-tuning with TRL Demos Connect OpenClaw to Muse Glimmer Hey Muse Glimmer, quantize yourself Hey Muse Glimmer, deploy yourself Hey Muse Glimmer, optimize yourself Hey Muse Glimmer, research the Hub Wrapping Up Great news from the OGs of open source LLMs! Muse Glimmer, released today, is Meta’s new multimodal model, especially designed for local agentic use cases. Distilled from Muse to 30B parameters, and released under the Apache 2.0 license, it’s ideal deploying locally for privacy, reducing costs, or just hacking around. It’s intended for privacy-aware applications such as coding, document analysis, personal assistants, Claw- or Hermes-like setups.