ollama/README.md

<div align="center">
  <picture>
    <source media="(prefers-color-scheme: dark)" height="200px" srcset="https://github.com/jmorganca/ollama/assets/3325447/56ea1849-1284-4645-8970-956de6e51c3c">
    <img alt="logo" height="200px" src="https://github.com/jmorganca/ollama/assets/3325447/0d0b44e2-8f4a-4e99-9b52-a5c1c741c8f7">
  </picture>
</div>

# Ollama

[![Discord](https://dcbadge.vercel.app/api/server/ollama?style=flat&compact=true)](https://discord.gg/ollama)

Get up and running with large language models locally.

### macOS

[Download](https://ollama.ai/download/Ollama-darwin.zip) 

### Linux & WSL2

```
curl https://ollama.ai/install.sh | sh
```

[Manual install instructions](https://github.com/jmorganca/ollama/blob/main/docs/linux.md)

### Windows 

coming soon

## Quickstart

To run and chat with [Llama 2](https://ollama.ai/library/llama2):

```
ollama run llama2
```

## Model library

Ollama supports a list of open-source models available on [ollama.ai/library](https://ollama.ai/library "ollama model library")

Here are some example open-source models that can be downloaded:

| Model              | Parameters | Size  | Download                       |
| ------------------ | ---------- | ----- | ------------------------------ |
| Mistral            | 7B         | 4.1GB | `ollama run mistral`           |
| Llama 2            | 7B         | 3.8GB | `ollama run llama2`            |
| Code Llama         | 7B         | 3.8GB | `ollama run codellama`         |
| Llama 2 Uncensored | 7B         | 3.8GB | `ollama run llama2-uncensored` |
| Llama 2 13B        | 13B        | 7.3GB | `ollama run llama2:13b`        |
| Llama 2 70B        | 70B        | 39GB  | `ollama run llama2:70b`        |
| Orca Mini          | 3B         | 1.9GB | `ollama run orca-mini`         |
| Vicuna             | 7B         | 3.8GB | `ollama run vicuna`            |

> Note: You should have at least 8 GB of RAM to run the 3B models, 16 GB to run the 7B models, and 32 GB to run the 13B models.

## Customize your own model

### Import from GGUF or GGML

Ollama supports importing GGUF and GGML file formats in the Modelfile. This means if you have a model that is not in the Ollama library, you can create it, iterate on it, and upload it to the Ollama library to share with others when you are ready.

1. Create a file named Modelfile, and add a `FROM` instruction with the local filepath to the model you want to import.

   ```
   FROM ./vicuna-33b.Q4_0.gguf
   ```

3. Create the model in Ollama

   ```
   ollama create name -f path_to_modelfile
   ```

5. Run the model

   ```
   ollama run name
   ```

### Customize a prompt

Models from the Ollama library can be customized with a prompt. The example

```
ollama pull llama2
```

Create a `Modelfile`:

```
FROM llama2

# set the temperature to 1 [higher is more creative, lower is more coherent]
PARAMETER temperature 1

# set the system prompt
SYSTEM """
You are Mario from Super Mario Bros. Answer as Mario, the assistant, only.
"""
```

Next, create and run the model:

```
ollama create mario -f ./Modelfile
ollama run mario
>>> hi
Hello! It's your friend Mario.
```

For more examples, see the [examples](./examples) directory. For more information on working with a Modelfile, see the [Modelfile](./docs/modelfile.md) documentation.

## CLI Reference

### Create a model

`ollama create` is used to create a model from a Modelfile.

### Pull a model

```
ollama pull llama2
```

> This command can also be used to update a local model. Only the diff will be pulled.

### Remove a model

```
ollama rm llama2
```

### Copy a model

```
ollama cp llama2 my-llama2
```

### Multiline input

For multiline input, you can wrap text with `"""`:

```
>>> """Hello,
... world!
... """
I'm a basic program that prints the famous "Hello, world!" message to the console.
```

### Pass in prompt as arguments

```
$ ollama run llama2 "summarize this file:" "$(cat README.md)"
 Ollama is a lightweight, extensible framework for building and running language models on the local machine. It provides a simple API for creating, running, and managing models, as well as a library of pre-built models that can be easily used in a variety of applications.
```

### List models on your computer

```
ollama list
```

### Start Ollama

`ollama serve` is used when you want to start ollama without running the desktop application.

## Building

Install `cmake` and `go`:

```
brew install cmake
brew install go
```

Then generate dependencies and build:

```
go generate ./...
go build .
```

Next, start the server:

```
./ollama serve
```

Finally, in a separate shell, run a model:

```
./ollama run llama2
```

## REST API

> See the [API documentation](./docs/api.md) for all endpoints.

Ollama has an API for running and managing models. For example to generate text from a model:

```
curl -X POST http://localhost:11434/api/generate -d '{
  "model": "llama2",
  "prompt":"Why is the sky blue?"
}'
```

## Community Integrations

- [LangChain](https://python.langchain.com/docs/integrations/llms/ollama) and [LangChain.js](https://js.langchain.com/docs/modules/model_io/models/llms/integrations/ollama) with [example](https://js.langchain.com/docs/use_cases/question_answering/local_retrieval_qa)
- [LlamaIndex](https://gpt-index.readthedocs.io/en/stable/examples/llm/ollama.html)
- [Raycast extension](https://github.com/MassimilianoPasquini97/raycast_ollama)
- [Discollama](https://github.com/mxyng/discollama) (Discord bot inside the Ollama discord channel)
- [Continue](https://github.com/continuedev/continue)
- [Obsidian Ollama plugin](https://github.com/hinterdupfinger/obsidian-ollama)
- [Dagger Chatbot](https://github.com/samalba/dagger-chatbot)
- [LiteLLM](https://github.com/BerriAI/litellm)
- [Discord AI Bot](https://github.com/mekb-turtle/discord-ai-bot)
- [Chatbot UI](https://github.com/ivanfioravanti/chatbot-ollama)
- [HTML UI](https://github.com/rtcfirefly/ollama-ui)
- [Typescript UI](https://github.com/ollama-interface/Ollama-Gui?tab=readme-ov-file)
- [Dumbar](https://github.com/JerrySievert/Dumbar)
- [Emacs client](https://github.com/zweifisch/ollama)
Update README.md add logo 2023-07-18 19:45:38 +00:00			`<div align="center">`
			`<picture>`
Update icon (#139) 2023-07-20 15:55:20 +00:00			`<source media="(prefers-color-scheme: dark)" height="200px" srcset="https://github.com/jmorganca/ollama/assets/3325447/56ea1849-1284-4645-8970-956de6e51c3c">`
			`<img alt="logo" height="200px" src="https://github.com/jmorganca/ollama/assets/3325447/0d0b44e2-8f4a-4e99-9b52-a5c1c741c8f7">`
Update README.md add logo 2023-07-18 19:45:38 +00:00			`</picture>`
			`</div>`
updated readme 2023-07-05 19:37:33 +00:00
move to contained directory 2023-06-27 16:08:52 +00:00			`# Ollama`
initial commit 2023-06-22 16:45:31 +00:00
fix discord link in `README.md` 2023-07-19 19:31:48 +00:00			`[![Discord](https://dcbadge.vercel.app/api/server/ollama?style=flat&compact=true)](https://discord.gg/ollama)`
add discord link, remove repeated text 2023-07-19 19:28:50 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`Get up and running with large language models locally.`
documentation on the model format 2023-07-20 15:33:28 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`### macOS`
update `README.md` 2023-08-08 22:50:23 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`[Download](https://ollama.ai/download/Ollama-darwin.zip)`
move download to the top of `README.md` 2023-07-18 20:31:25 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`### Linux & WSL2`

			```
			`curl https://ollama.ai/install.sh \| sh`
			```

			`[Manual install instructions](https://github.com/jmorganca/ollama/blob/main/docs/linux.md)`

			`### Windows`

			`coming soon`
move download to the top of `README.md` 2023-07-18 20:31:25 +00:00
add discord link, remove repeated text 2023-07-19 19:28:50 +00:00			`## Quickstart`

Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`To run and chat with [Llama 2](https://ollama.ai/library/llama2):`
add discord link, remove repeated text 2023-07-19 19:28:50 +00:00
			```
			`ollama run llama2`
			```

			`## Model library`

Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`Ollama supports a list of open-source models available on [ollama.ai/library](https://ollama.ai/library "ollama model library")`
adding link to models directly available on ollama (#366) - adding link to models directly available on ollama - ability to push your own models to the library will come in the future 2023-08-17 02:53:27 +00:00
add `codellama` to model list in readme 2023-08-26 03:44:26 +00:00			`Here are some example open-source models that can be downloaded:`
add discord link, remove repeated text 2023-07-19 19:28:50 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`\| Model \| Parameters \| Size \| Download \|`
			`\| ------------------ \| ---------- \| ----- \| ------------------------------ \|`
Update README.md adding in instruction to run mistral 2023-09-28 16:06:03 +00:00			\| Mistral \| 7B \| 4.1GB \| `ollama run mistral` \|
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			\| Llama 2 \| 7B \| 3.8GB \| `ollama run llama2` \|
			\| Code Llama \| 7B \| 3.8GB \| `ollama run codellama` \|
			\| Llama 2 Uncensored \| 7B \| 3.8GB \| `ollama run llama2-uncensored` \|
			\| Llama 2 13B \| 13B \| 7.3GB \| `ollama run llama2:13b` \|
			\| Llama 2 70B \| 70B \| 39GB \| `ollama run llama2:70b` \|
			\| Orca Mini \| 3B \| 1.9GB \| `ollama run orca-mini` \|
			\| Vicuna \| 7B \| 3.8GB \| `ollama run vicuna` \|
add discord link, remove repeated text 2023-07-19 19:28:50 +00:00
clean up `README.md` 2023-07-20 19:21:29 +00:00			`> Note: You should have at least 8 GB of RAM to run the 3B models, 16 GB to run the 7B models, and 32 GB to run the 13B models.`

Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`## Customize your own model`
Add download link to readme 2023-06-27 21:13:07 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`### Import from GGUF or GGML`
update readme 2023-09-01 14:54:31 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`Ollama supports importing GGUF and GGML file formats in the Modelfile. This means if you have a model that is not in the Ollama library, you can create it, iterate on it, and upload it to the Ollama library to share with others when you are ready.`
update readme 2023-09-01 14:54:31 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			1. Create a file named Modelfile, and add a `FROM` instruction with the local filepath to the model you want to import.
update readme 2023-09-01 14:54:31 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			```
			`FROM ./vicuna-33b.Q4_0.gguf`
			```
better `README.md` install instructions 2023-06-30 16:39:25 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`3. Create the model in Ollama`
initial commit 2023-06-22 16:45:31 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			```
			`ollama create name -f path_to_modelfile`
			```
Add an example on multiline input (#311) 2023-08-10 15:22:28 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`5. Run the model`
update readme (#451) * update readme * readme: more run examples 2023-09-01 20:44:14 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			```
			`ollama run name`
			```
update readme (#451) * update readme * readme: more run examples 2023-09-01 20:44:14 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`### Customize a prompt`
update readme (#451) * update readme * readme: more run examples 2023-09-01 20:44:14 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`Models from the Ollama library can be customized with a prompt. The example`
add discord link, remove repeated text 2023-07-19 19:28:50 +00:00
			```
new `Modelfile` syntax 2023-07-20 09:21:51 +00:00			`ollama pull llama2`
add discord link, remove repeated text 2023-07-19 19:28:50 +00:00			```
docs: format with `prettier` 2023-08-08 22:41:48 +00:00
update `README.md` with new syntax 2023-07-18 20:22:33 +00:00			Create a `Modelfile`:
updated readme 2023-07-05 19:37:33 +00:00
add `docker` instruction 2023-06-30 16:31:00 +00:00			```
new `Modelfile` syntax 2023-07-20 09:21:51 +00:00			`FROM llama2`
set temperature on `README.md` example 2023-07-20 15:17:09 +00:00
			`# set the temperature to 1 [higher is more creative, lower is more coherent]`
			`PARAMETER temperature 1`

			`# set the system prompt`
new `Modelfile` syntax 2023-07-20 09:21:51 +00:00			`SYSTEM """`
fix typo 2023-07-18 20:32:06 +00:00			`You are Mario from Super Mario Bros. Answer as Mario, the assistant, only.`
update `README.md` with new syntax 2023-07-18 20:22:33 +00:00			`"""`
simplify `README.md` 2023-06-29 22:25:02 +00:00			```
take all args as one prompt - parse all run arguments into one prompt - do not echo prompt back on one-shot - example of summarizing a document 2023-07-07 20:14:58 +00:00
update `README.md` with new syntax 2023-07-18 20:22:33 +00:00			`Next, create and run the model:`
take all args as one prompt - parse all run arguments into one prompt - do not echo prompt back on one-shot - example of summarizing a document 2023-07-07 20:14:58 +00:00
			```
update `README.md` with new syntax 2023-07-18 20:22:33 +00:00			`ollama create mario -f ./Modelfile`
			`ollama run mario`
			`>>> hi`
			`Hello! It's your friend Mario.`
take all args as one prompt - parse all run arguments into one prompt - do not echo prompt back on one-shot - example of summarizing a document 2023-07-07 20:14:58 +00:00			```

Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`For more examples, see the [examples](./examples) directory. For more information on working with a Modelfile, see the [Modelfile](./docs/modelfile.md) documentation.`

			`## CLI Reference`

			`### Create a model`

			`ollama create` is used to create a model from a Modelfile.
add discord link, remove repeated text 2023-07-19 19:28:50 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`### Pull a model`
add advanced usage to readme 2023-07-06 20:21:01 +00:00
add discord link, remove repeated text 2023-07-19 19:28:50 +00:00			```
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`ollama pull llama2`
add discord link, remove repeated text 2023-07-19 19:28:50 +00:00			```
reorganize `README.md` files 2023-06-28 13:57:36 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`> This command can also be used to update a local model. Only the diff will be pulled.`

			`### Remove a model`
clean up `README.md` 2023-07-20 19:21:29 +00:00
			```
update readme 2023-09-01 14:54:31 +00:00			`ollama rm llama2`
clean up `README.md` 2023-07-20 19:21:29 +00:00			```

Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`### Copy a model`

			```
			`ollama cp llama2 my-llama2`
			```

			`### Multiline input`

			For multiline input, you can wrap text with `"""`:

			```
			`>>> """Hello,`
			`... world!`
			`... """`
			`I'm a basic program that prints the famous "Hello, world!" message to the console.`
			```

			`### Pass in prompt as arguments`

			```
			`$ ollama run llama2 "summarize this file:" "$(cat README.md)"`
			`Ollama is a lightweight, extensible framework for building and running language models on the local machine. It provides a simple API for creating, running, and managing models, as well as a library of pre-built models that can be easily used in a variety of applications.`
			```

			`### List models on your computer`
clean up `README.md` 2023-07-20 19:21:29 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			```
			`ollama list`
			```
clean up `README.md` 2023-07-20 19:21:29 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`### Start Ollama`
clean up `README.md` 2023-07-20 19:21:29 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`ollama serve` is used when you want to start ollama without running the desktop application.
clean up `README.md` 2023-07-20 19:21:29 +00:00
add llama.cpp go bindings 2023-07-03 20:32:48 +00:00			`## Building`

[docs] Improve build instructions (#482) Go is required and not installed by default. 2023-09-07 10:43:26 +00:00			Install `cmake` and `go`:
update README.md 2023-08-25 18:44:25 +00:00
add llama.cpp go bindings 2023-07-03 20:32:48 +00:00			```
update docs for subprocess 2023-08-30 21:54:02 +00:00			`brew install cmake`
[docs] Improve build instructions (#482) Go is required and not installed by default. 2023-09-07 10:43:26 +00:00			`brew install go`
update docs for subprocess 2023-08-30 21:54:02 +00:00			```

			`Then generate dependencies and build:`

			```
			`go generate ./...`
vendor llama.cpp 2023-07-11 16:50:02 +00:00			`go build .`
add llama.cpp go bindings 2023-07-03 20:32:48 +00:00			```

update docs for subprocess 2023-08-30 21:54:02 +00:00			`Next, start the server:`
add development doc 2023-06-27 17:46:46 +00:00
updated readme 2023-07-05 19:37:33 +00:00			```
update docs for subprocess 2023-08-30 21:54:02 +00:00			`./ollama serve`
updated readme 2023-07-05 19:37:33 +00:00			```

update readme 2023-09-01 14:54:31 +00:00			`Finally, in a separate shell, run a model:`
updated readme 2023-07-05 19:37:33 +00:00
			```
update `README.md` with new syntax 2023-07-18 20:22:33 +00:00			`./ollama run llama2`
updated readme 2023-07-05 19:37:33 +00:00			```
add basic REST api documentation 2023-07-21 07:47:17 +00:00
			`## REST API`

Link to `api.md` in `README.md` 2023-08-08 22:48:47 +00:00			`> See the [API documentation](./docs/api.md) for all endpoints.`
add basic REST api documentation 2023-07-21 07:47:17 +00:00
Link to `api.md` in `README.md` 2023-08-08 22:48:47 +00:00			`Ollama has an API for running and managing models. For example to generate text from a model:`
add basic REST api documentation 2023-07-21 07:47:17 +00:00
			```
Link to `api.md` in `README.md` 2023-08-08 22:48:47 +00:00			`curl -X POST http://localhost:11434/api/generate -d '{`
			`"model": "llama2",`
			`"prompt":"Why is the sky blue?"`
			`}'`
add `/api/create` docs to readme 2023-07-23 22:01:05 +00:00			```
Update README.md 2023-07-31 20:59:39 +00:00
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`## Community Integrations`

			`- [LangChain](https://python.langchain.com/docs/integrations/llms/ollama) and [LangChain.js](https://js.langchain.com/docs/modules/model_io/models/llms/integrations/ollama) with [example](https://js.langchain.com/docs/use_cases/question_answering/local_retrieval_qa)`
			`- [LlamaIndex](https://gpt-index.readthedocs.io/en/stable/examples/llm/ollama.html)`
			`- [Raycast extension](https://github.com/MassimilianoPasquini97/raycast_ollama)`
			`- [Discollama](https://github.com/mxyng/discollama) (Discord bot inside the Ollama discord channel)`
			`- [Continue](https://github.com/continuedev/continue)`
			`- [Obsidian Ollama plugin](https://github.com/hinterdupfinger/obsidian-ollama)`
			`- [Dagger Chatbot](https://github.com/samalba/dagger-chatbot)`
			`- [LiteLLM](https://github.com/BerriAI/litellm)`
			`- [Discord AI Bot](https://github.com/mekb-turtle/discord-ai-bot)`
add community project: Chatbot Ollama add community project: Chatbot Ollama by @ivanfioravanti 2023-10-02 16:04:31 +00:00			`- [Chatbot UI](https://github.com/ivanfioravanti/chatbot-ollama)`
Update README.md for linux + cleanup (#601) Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> 2023-09-26 06:44:53 +00:00			`- [HTML UI](https://github.com/rtcfirefly/ollama-ui)`
			`- [Typescript UI](https://github.com/ollama-interface/Ollama-Gui?tab=readme-ov-file)`
			`- [Dumbar](https://github.com/JerrySievert/Dumbar)`
			`- [Emacs client](https://github.com/zweifisch/ollama)`