mirror of
https://github.com/ollama/ollama.git
synced 2026-09-16 16:11:52 -04:00
* fix(docs): correct typos found during code review Non-functional changes only: - Fixed minor spelling mistakes in comments - Corrected typos in user-facing strings - No variables, logic, or functional code was modified. Signed-off-by: Marcel Petrick <mail@marcelpetrick.it> * fix additional typos and shell-unsafe example in docs --------- Co-authored-by: Patrick Devine <patrick@ollama.com>
runner
Note: this is a work in progress
A minimal runner for loading a model and running inference via a http web server.
./runner -model <model binary>
Completion
curl -X POST -H "Content-Type: application/json" -d '{"prompt": "hi"}' http://localhost:8080/completion
Embeddings
curl -X POST -H "Content-Type: application/json" -d '{"prompt": "turn me into an embedding"}' http://localhost:8080/embedding