-
-
Notifications
You must be signed in to change notification settings - Fork 745
Pull requests: Blaizzy/mlx-vlm
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
Drop unused DeepSeek-V4 rotating-cache value graphs
#2043
opened Aug 27, 2026 by
byoan
Contributor
Loading…
fix(cache): trim ChunkedKVCache on valid length, not buffer width (#2031)
#2036
opened Aug 26, 2026 by
Anai-Guo
Contributor
Loading…
Fix positioned top-k sampling for speculative generation
#2029
opened Aug 26, 2026 by
tjansn
Loading…
fix(server): stop dropping streamed content after a tool call ends
#2027
opened Aug 26, 2026 by
Anai-Guo
Contributor
Loading…
fix(sam3): let --threshold fall through to the documented per-task defaults
#2025
opened Aug 26, 2026 by
Anai-Guo
Contributor
Loading…
fix(rfdetr): honor --show-boxes instead of ignoring it
#2024
opened Aug 26, 2026 by
Anai-Guo
Contributor
Loading…
Support DeepSeek V4 speculative chunked prefill
#2023
opened Aug 26, 2026 by
byoan
Contributor
Loading…
Report speculative usage from direct generation
#1997
opened Aug 21, 2026 by
byoan
Contributor
Loading…
Apply max_kv_size to models that build their own cache
#1995
opened Aug 21, 2026 by
Lazarus-931
Collaborator
Loading…
reformatting docs: move the README into a docs/ site
#1964
opened Aug 19, 2026 by
Lazarus-931
Collaborator
•
Draft
Reformat trainer for all modalities
#1939
opened Aug 17, 2026 by
Goekdeniz-Guelmez
Contributor
•
Draft
Fix Qwen speculative decoding with quantized batch cache
#1938
opened Aug 17, 2026 by
Ptico
Loading…
fix(server): remove the
thought prefix by prefix, not by character set
#1934
opened Aug 17, 2026 by
Anai-Guo
Contributor
Loading…
Previous Next
ProTip!
Add no:assignee to see everything that’s not assigned.