Published event
ArtificialIntelligence
ModelRelease
1 source(s)
v0.17.1
Summary
v0.17.1 vllm-project / vllm Public Uh oh! There was an error while loading.
Why it matters
This ModelRelease is relevant to the technology intelligence record because it involves Qwen. The source article should remain the factual reference for follow-up coverage.
Key facts
- vllm-project / vllm Public Uh oh!
- There was an error while loading.
- Notifications You must be signed in to change notification settings Fork 22.7k Star 92.7k v0.17.1 khluu released this 11 Mar 10:24 · 7564 commits to main since this release v0.17.1 95c0f92 This is a patch release on top of v0.17.0 to address a few issues: New Model: Nemotron 3 Super Fix passing of activation_type to trtllm fused MoE NVFP4 and FP8 ( #36017 ) Fix/resupport nongated fused moe triton ( #36412 ) Re-enable EP for trtllm MoE FP8 backend ( #36494 ) [Mamba][Qwen3.5] Zero freed SSM cache blocks on GPU ( #35219 ) Fix TRTLLM Block FP8 MoE Monolithic ( #36296 ) [DSV3.2][MTP] Optimize Indexer MTP handling ( #36723 ) Assets 9 Loading Uh oh!
Entities in this story
AI models
qwen→Related events