Skip to content

Pull requests: flagos-ai/vllm-plugin-FL

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

fix(worker): account for CUDA graph memory in KV cache sizing core
#381 opened Aug 14, 2026 by CherryLemon Collaborator Loading…
2 of 3 tasks
adapt(metax): support GLM-5.2 W8A8 INT8 core
#378 opened Aug 13, 2026 by liyu030 Loading…
3 tasks
add flaggems-vllm core
#375 opened Aug 13, 2026 by cyber-pioneer Collaborator Loading…
[CICD]: Add Enflame S60 CI support ci docs tests
#373 opened Aug 13, 2026 by HermiaHuan Collaborator Loading…
feat(metax): support DeepSeek V4 W8A8 INT8 inference core
#372 opened Aug 12, 2026 by liyu030 Loading…
3 tasks
[Kunlunxin] Fix garbled output with CUDA Graph core
#368 opened Aug 11, 2026 by 0songHan Loading…
3 tasks
support mla atten in flaggems core
#359 opened Aug 10, 2026 by ZenithJoyH Contributor Loading…
3 tasks
support Qwen3.5 FP8 Model on GCU (Enflame) core
#358 opened Aug 9, 2026 by liuyancong-enflame-tech Contributor Loading…
1 of 3 tasks
gcu: enable enflame GCU300 via vLLM native FLASH_ATTN backend core
#357 opened Aug 9, 2026 by tengqm Contributor Loading…
feat(cambricon): enable MLU590 (device map + graph class) core
#355 opened Aug 8, 2026 by tengqm Contributor Loading…
perf(metax): optimize single-request prefill attention core
#345 opened Aug 5, 2026 by liyu030 Loading…
2 of 3 tasks
support dsv4 int8 version core
#344 opened Aug 5, 2026 by ceci3 Collaborator Loading…
3 tasks
ProTip! Filter pull requests by the default branch with base:main.