-
Notifications
You must be signed in to change notification settings - Fork 1k
Pull requests: ml-explore/mlx-lm
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
Fix/mlx_lm.evaluate ignoring max_gen_toks and --batch-size during generation
#1812
opened Sep 1, 2026 by
dasashreeya
Loading…
fix: plumb evaluate batch-size to generate_until, accept max_gen_toks
#1810
opened Aug 31, 2026 by
axiom-of-choice
Contributor
•
Draft
Add pipeline parallelism for Qwen3 MoE
await verification
This pull request is non-trivial and requires a human expert to verify its correctness.
#1801
opened Aug 29, 2026 by
azamamirza
Contributor
Loading…
fix: apply 1/sqrt(head_v_dim) readout scale in gated_delta_update
#1799
opened Aug 28, 2026 by
axiom-of-choice
Contributor
Loading…
Return 503 from /health when the generation thread exits
await response
This pull request is waiting for response from the author.
#1791
opened Aug 27, 2026 by
alexis-ag
Loading…
Add Qwen3.8-Flash-Next (qwen4_exp) model support
await verification
This pull request is non-trivial and requires a human expert to verify its correctness.
#1788
opened Aug 26, 2026 by
eauchs
Loading…
Add Intern-S2-Mobius (interns2_mobius) model
new_model
#1771
opened Aug 21, 2026 by
nightscape
•
Draft
Add M-RoPE support so vision embeddings keep their grid positions
enhancement
#1768
opened Aug 21, 2026 by
lpalbou
Contributor
Loading…
fix(granitemoehybrid): accept nested rope_parameters and tied lm_head weight
await response
This pull request is waiting for response from the author.
bug
#1605
opened Jul 22, 2026 by
lkrapf
Loading…
Fix deepseek_v32 Indexer evicting attention sinks from sparse top-k
bug
#1552
opened Jul 11, 2026 by
robertlangdonn
Contributor
Loading…
5 tasks done
Fail fast when the generation thread stops
bug
#1514
opened Jul 10, 2026 by
dogukanveziroglu
Contributor
Loading…
Keep the server generation loop alive when a batched request fails
bug
#1513
opened Jul 10, 2026 by
rajanshxrma
Loading…
Fix float32 promotion in BatchKVCache/BatchRotatingKVCache extend()
bug
#1491
opened Jul 7, 2026 by
ethanelasky
Loading…
Add support for ESMC and ESMFold2 models
new_model
#1484
opened Jul 6, 2026 by
faustomilletari
Loading…
Fix mixed-bit quantized load for sanitize()-derived MLA projections
bug
#1482
opened Jul 6, 2026 by
rajanshxrma
Loading…
fix(vl): strip model.visual.* in qwen2_vl/qwen3_vl/qwen3_vl_moe sanitize
bug
#1473
opened Jul 4, 2026 by
Jonathangadeaharder
Loading…
DeepSeek-V3.2/GLM DSA: fix silent >128k top-k corruption + sparse-gather prefill
bug
#1454
opened Jul 2, 2026 by
aidiffuser
Loading…
Previous Next
ProTip!
Follow long discussions with comments:>50.