Skip to content

chore: bump mlx-swift-lm to 4f1fdaa (preload non-expert weights when streaming) - #218

Merged
solderzzc merged 1 commit into
mainfrom
chore/bump-mlx-swift-lm-80
Oct 9, 2026
Merged

solderzzc merged 1 commit into
mainfrom
chore/bump-mlx-swift-lm-80

Conversation

@solderzzc

Copy link
Copy Markdown
Member

Summary

Bumps the mlx-swift-lm submodule 9da07be → 4f1fdaa (mlx-swift-lm#80).

#80 materializes the non-expert weights right after load when --stream-experts is on (lazyLoad leaves them as pending safetensors reads, so a cold page cache could stall a GPU command buffer past the watchdog).

Notes

🤖 Generated with Claude Code

…streaming)

Picks up mlx-swift-lm#80: with --stream-experts, materialize non-expert
weights after load so the GPU command buffer does not wait on cold
safetensors reads (kIOGPUCommandBufferCallbackErrorTimeout, see #216).

Co-Authored-By: Claude Sonnet 5.5 <noreply@anthropic.com>
@solderzzc
solderzzc merged commit 1b6befd into main Oct 9, 2026
15 checks passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant