Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

Intel LLM-Scaler Ready With Muse Glimmer Support, Other LLMs & Features

Дата публикации: 12-08-2026 10:12:00

Intel's LLM-Scaler project that was born out of their Project Battlematrix initiative aims to make it easier to run generative AI on Arc (Pro) B-Series graphics cards with the likes of vLLM, ComfyUI, SGLang, and other popular AI software in this Docker-based pre-configured AI stack. This week new LLM-Scaler releases brought same-day support for new models and other enhancements...

Основное содержимое страницы с новостью.

INTEL

Intel's LLM-Scaler project that was born out of their Project Battlematrix initiative aims to make it easier to run generative AI on Arc (Pro) B-Series graphics cards with the likes of vLLM, ComfyUI, SGLang, and other popular AI software in this Docker-based pre-configured AI stack. This week new LLM-Scaler releases brought same-day support for new models and other enhancements.

Most notable with the Intel LLM-Scaler-vLLM beta 0.21.0-b3 release on Monday was delivering same-day support for Meta's new Muse Glimmer 30B model. Muse-Glimmer-30B with FP8 online quantization is supported by LLM-Scaler-vLLM on the likes of the Arc Pro B70.

Intel Arc Pro B70

The new LLM-Scaler-vLLM beta also adds suppport for DFlash for Muse-Glimmer-30B and Qwen3.6-27B. There is also better time-to-first-token performance for Gemma-4-31B and Gemma-4-26B-A4B-it. Plus various bug fixes for this updated vLLM stack for Intel graphics. See this GitHub release for those details.

Released today was LLM-Scaler-Omni beta 0.2.0-b1. This new LLM-Scaler-Omni Docker container upgrades to the ComfyUI 0.31 XPU stack, adds support for MiniMax H3 local video generation, supports Wan Animate 2 on the Arc Pro B70 and B60, and expands optimized model coverage with Wan 2.2 14B T2V Turbo, LTX-2, Z-Image / Lumina, and Krea2. This update also adds managed GGUF Q4_1 support and pinned ComfyUI-GGUF-XPU integration.

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1Lemonade 11.6 Integrates Muse-Glimmer 30B, Experimental TheNoise ROCm Image Generation016.0914-08-2026
2TileRT - Tile-Based Runtime for Ultra-Low-Latency LLM Inference03528-06-2026
3Intel XPU Manager 2.1 Released For Monitoring Arc Pro Graphics On Windows/Linux08.4113-08-2026
4Linux Foundation & Others Launch "Akrites" To Defend Open-Source Software From AI-Enabled Exploits0725-06-2026
5Meta Releases Muse Glimmer, a 30-Billion-Parameter Open-Weight AI Model That Runs on a Single Consumer GPU05.0611-08-2026
6Что такое LiteLLM и для чего его едят06.0729-07-2026
7vLLM vs LMDeploy vs Triton: обзор бэкендов для инференса LLM0718-07-2026
8Hardware-aware framework accelerates large language models without additional training08.5706-08-2026
9Generative AI using Elastic and Amazon SageMaker JumpStart06.2525-07-2023

Классификация: Пресс-релизы. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 14.44. Источник: www.phoronix.com.