-
Notifications
You must be signed in to change notification settings - Fork 45
Pull requests: vllm-project/vllm-gguf-plugin
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
[Models] Support DeepSeek-V4 GGUF weights mapping and architecture fa…
#126
opened Sep 7, 2026 by
DjoserKhemSimeu
Loading…
2 tasks done
fix(moe_vec): chunk launches when tokens*top_k exceeds gridDim.z limit
#125
opened Sep 1, 2026 by
BruceLoveDecimal
Loading…
[Models] Serve Qwen3.5/3.6/3.8 text-only GGUF without an mm_proj
#120
opened Aug 24, 2026 by
laureano-arcanio
Loading…
[Models] Support the Muse Glimmer dflash draft model in GGUF
#115
opened Aug 20, 2026 by
WhatGhost
Loading…
feat(weights-adapter): add DeepSeek V2/V3 GGUF adapter
#111
opened Aug 18, 2026 by
Feijia1231
Contributor
Loading…
fix(quantization): handle mixed-precision fused layers in is_layer_skipped_gguf
#107
opened Aug 13, 2026 by
BenWongCityuCS
Loading…
[Bugfix] Preserve GGUF weight path for config source lookup
#70
opened Jun 23, 2026 by
lesj0610
Loading…
4 tasks done
ProTip!
Updated in the last three days: updated:>2026-09-04.