Skip to content

Pull requests: Layr-Labs/mlx-swift-lm

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Preserve optional function strict metadata in OpenAI tools
#147 opened Sep 12, 2026 by jonathan308 Loading…
4 tasks done
perf: port Qwen 3.8 DFlash2 runner
#121 opened Aug 25, 2026 by davidtai Loading…
Add a vLLM-style zero-copy paged prefix cache
#116 opened Aug 23, 2026 by anupsv Loading…
perf(cbv2): gate Qwen D256 attention execution
#109 opened Aug 17, 2026 by Gajesh2007 Member Loading…
fix: skip batch decode masks when unpadded
#40 opened Jun 16, 2026 by anupsv Loading…
Add LlamaModelTP: tensor-parallel variant of LlamaModel
#25 opened May 21, 2026 by anupsv Loading…
3 of 4 tasks
Add Llama callPartial for pipeline-parallel inference
#24 opened May 21, 2026 by anupsv Loading…
2 of 3 tasks
Sync to ml-explore/mlx-swift-lm
#22 opened May 17, 2026 by Gajesh2007 Member Loading…
perf: MLP fusion + Gemma/GPT-OSS inference optimizations
#17 opened May 11, 2026 by 0xClandestine Member Loading…
6 tasks
feat: TurboQuant+ KV cache compression
#16 opened May 11, 2026 by 0xClandestine Member Loading…
6 tasks
Add opt-in Wired Memory
#6 opened May 3, 2026 by ronaldmannak Loading…
Code cleanup
#5 opened May 3, 2026 by ronaldmannak Loading…
Validate BatchGenerator inputs
#4 opened May 3, 2026 by ronaldmannak Loading…
ProTip! Add no:assignee to see everything that’s not assigned.