mirror of
https://github.com/ollama/ollama.git
synced 2026-07-23 09:10:53 -05:00
Bump MLX to the latest selected upstream ref and update the MLX/imagegen wrappers and tests for the new API behavior. Fix the CUDA MLX archive so runtime NVRTC kernels work after deployment: package CUTE/CUTLASS headers, include the CUDA runtime header closure, and stage a coherent CUDA-toolkit-matched CCCL tree instead of MLX's fetched CCCL for CUDA payloads. The previous archive could build successfully but crash at runtime due to missing or incompatible JIT headers.
2 lines
41 B
Plaintext
2 lines
41 B
Plaintext
51b2768da7e1897d3c4258f7ddbb47083d1eef01
|