Skip to content

Pull requests: ml-explore/mlx-lm

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Context sharding
#1905 opened Sep 20, 2026 by danxn Loading…
Fix incorrect assumption when parsing tool_calls[].function.arguments
#1904 opened Sep 20, 2026 by LxYuan0420 Contributor Loading…
Fix Gemma2 attention bug
#1903 opened Sep 20, 2026 by ursk Loading…
Add xing4_0 model support
#1901 opened Sep 18, 2026 by TokenAI-zer Loading…
Add tool call parser for MiniCPM5
#1899 opened Sep 17, 2026 by dasashreeya Contributor Loading…
Add DeepSeek-V4.1 model new_model
#1895 opened Sep 16, 2026 by Goekdeniz-Guelmez Contributor Loading…
Tolerate a null rope_scaling config when constructing PhiMoE models
#1889 opened Sep 15, 2026 by gyanu2507 Contributor Loading…
server: make exact prompt-cache hits generation-safe
#1887 opened Sep 14, 2026 by deyi2026 Loading…
Chunkwise-parallel gated delta rule for training
#1870 opened Sep 10, 2026 by adityak74 Contributor Loading…
Report the matched stop sequence on GenerationBatch.Response
#1869 opened Sep 9, 2026 by mloiterman Contributor Loading…
Support uneven tensor-parallel sharding in shard() (GQA/MLA-aware)
#1863 opened Sep 8, 2026 by twallgren Contributor Loading…
Add global scale for nvfp4 quantized MoEs
#1844 opened Sep 4, 2026 by nastya236 Collaborator Loading…
Add G9v3 (G9v3-39A5B) support new_model
#1831 opened Sep 3, 2026 by AlphaKure Contributor Loading…
Trimmable ArraysCache alternative
#1821 opened Sep 2, 2026 by zcbenz Member Draft
Add DeepSeek-V4-Flash without MTP (yet) new_model
#1797 opened Aug 28, 2026 by nh13 Contributor Loading…
Add Qwen3.8-Flash-Next (qwen4_exp) model support await verification This pull request is non-trivial and requires a human expert to verify its correctness.
#1788 opened Aug 26, 2026 by eauchs Loading…
Fix deepseek_v32 Indexer evicting attention sinks from sparse top-k bug
#1552 opened Jul 11, 2026 by robertlangdonn Contributor Loading…
5 tasks done
Add Nemotron-H Puzzle model support await response This pull request is waiting for response from the author. new_model
#1535 opened Jul 10, 2026 by sxuff Loading…
ProTip! Type g p on any issue or pull request to go back to the pull request listing page.