This website requires JavaScript.
Explore
Help
Register
Sign In
mirror
/
ollama
Watch
1
Star
0
Fork
0
You've already forked ollama
mirror of
https://github.com/ollama/ollama.git
synced
2026-07-30 17:38:02 -04:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
parth-agent-loop
Add File
New File
Upload File
Apply Patch
ollama
/
x
/
mlxrunner
/
model
History
Patrick Devine
3ef69ef784
mlx: allow the embedding layer to use the nvfp4 global scale (
#16527
)
2026-06-04 17:40:01 -07:00
..
base
Revert "mlxrunner: add DFlash speculative decoding (
#16134
)"
2026-05-22 09:32:09 -07:00
embedding_test.go
mlx: allow the embedding layer to use the nvfp4 global scale (
#16527
)
2026-06-04 17:40:01 -07:00
embedding.go
mlx: allow the embedding layer to use the nvfp4 global scale (
#16527
)
2026-06-04 17:40:01 -07:00
linear.go
mlx: Support NVIDIA TensorRT Model Optimizer import (
#15566
)
2026-04-27 18:28:10 -07:00
quant.go
mlx: add mxfp4/mxfp8/nvfp4 importing (
#15015
)
2026-03-24 13:45:44 -07:00
root.go
mlx: Gemma4 MTP speculative decoding (
#15980
)
2026-05-05 08:55:04 -07:00