aifollow.news Search
Back vLLM 发布记录
vLLM 发布记录· · Date only

v0.22.1

This language is not available yet; showing the source language.

Source excerpt

## Highlights This release features 8 commits from 6 contributors (1 new)! v0.22.1 is a patch release on top of v0.22.0 with targeted bug fixes plus a couple of additions: new model support for JetBrains' Mellum v2, zentorch-accelerated quantized linear inference on AMD Zen CPUs, and fixes for multi-node Ray data-parallel serving, DeepSeek-V4 initialization, and a few model-loading regressions. ### Model Support * New model: JetBrains' **Mellum v2**, an open-weights Mixture-of-Experts code-generation model (#43992). * **DeepSeek-V4**: resolve a CUTLASS `fmin` compatibility issue that broke initialization (0decac0d). * Fix `OlmoHybridForCausal

The complete source text is not yet available.

Read at the original source
Found an error? Send a correction