-
Notifications
You must be signed in to change notification settings - Fork 16.7k
Pull requests: deepseek-ai/DeepSeek-V3
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
fix(inference): preserve index metadata when converting fp8 to bf16
#1597
opened Aug 23, 2026 by
loulanyue
Loading…
Improve README docs and harden inference/generate.py validation & cleanup
#1593
opened Aug 20, 2026 by
Shyboy0499
Loading…
docs: fix broken TensorRT-LLM example link in README
#1582
opened Aug 16, 2026 by
Kolgrim33
Loading…
docs: add deployment compatibility table and fix broken links in README
#1545
opened Aug 5, 2026 by
usaihack
Loading…
feat(inference): MTP speculative decoding with cache-aware mid-decode mask fix
#1536
opened Jul 30, 2026 by
n-dlms
Loading…
fix(inference): strip whitespace from generate.py CLI paths
#1511
opened Jul 24, 2026 by
Charmve
Loading…
feat(model): add ZazorLayer for hierarchical context memory
#1485
opened Jul 9, 2026 by
AlexShchuka
Loading…
feat: add mobile-compatible requirements (CPU-only)
#1442
opened Jun 20, 2026 by
Bakomebandias
Loading…
fix: add repetition penalty to mitigate multi-turn repetition (fixes #1125)
#1129
opened Mar 1, 2026 by
modimihir07
Loading…
fix: use OrderedDict for proper LRU cache eviction in fp8_cast_bf16.py
#1128
opened Mar 1, 2026 by
modimihir07
Loading…
Add Multi-Token Prediction (MTP) support for speculative decoding
#1122
opened Feb 24, 2026 by
dcol91863
Loading…
fix: fix triton kernel tiling and fp8_gemm swizzle
#1098
opened Jan 29, 2026 by
JackeyLove1
Loading…
feat(inference): Add streaming support imports for high-performance L…
#1080
opened Jan 13, 2026 by
sanjay-aravindh
Loading…
Previous Next
ProTip!
Mix and match filters to narrow down what you’re looking for.