ollama

Author	SHA1	Message	Date
Daniel Hiltgen	f2ea8470e5	Local unicode test case	2024-04-22 19:29:12 -07:00
Daniel Hiltgen	34b9db5afc	Request and model concurrency This change adds support for multiple concurrent requests, as well as loading multiple models by spawning multiple runners. The default settings are currently set at 1 concurrent request per model and only 1 loaded model at a time, but these can be adjusted by setting OLLAMA_NUM_PARALLEL and OLLAMA_MAX_LOADED_MODELS.	2024-04-22 19:29:12 -07:00
Daniel Hiltgen	ee448deaba	Merge pull request #3835 from dhiltgen/harden_llm_override Trim spaces and quotes from llm lib override	2024-04-22 19:06:54 -07:00
Bruce MacDonald	6e8db04716	tidy community integrations - move some popular integrations to the top of the lists	2024-04-22 17:29:08 -07:00
Bruce MacDonald	658e60cf73	Revert "stop running model on interactive exit" This reverts commit `fad00a85e5`.	2024-04-22 17:23:11 -07:00
Bruce MacDonald	4c78f028f8	Merge branch 'main' of https://github.com/ollama/ollama	2024-04-22 17:22:28 -07:00
Hao Wu	c7d3a558f6	docs: update README to add chat (web UI) for LLM (#3810 ) * add chat (web UI) for LLM I have used chat with llama3 in local successfully and the code is MIT licensed. * Update README.md --------- Co-authored-by: Bruce MacDonald <brucewmacdonald@gmail.com>	2024-04-22 20:19:39 -04:00
Maple Gao	089cdb2877	docs: Update README for Lobe-chat integration. (#3817 ) Co-authored-by: Bruce MacDonald <brucewmacdonald@gmail.com>	2024-04-22 20:18:15 -04:00
Võ Đình Đạt	ea1e9aa36b	Update README.md (#3655 )	2024-04-22 20:16:55 -04:00
Jonathan Smoley	d0d28ef90d	Update README.md with Discord-Ollama project (#3633 ) Co-authored-by: Bruce MacDonald <brucewmacdonald@gmail.com>	2024-04-22 20:14:20 -04:00
Eric Curtin	6654186a7c	Add podman-ollama to terminal apps (#3626 ) The goal of podman-ollama is to make AI even more boring. Signed-off-by: Eric Curtin <ecurtin@redhat.com>	2024-04-22 20:13:23 -04:00
Daniel Hiltgen	aa72281eae	Trim spaces and quotes from llm lib override	2024-04-22 17:11:14 -07:00
reid41	74bcbf828f	add qa-pilot link (#3612 ) * add qa-pilot link * format the link * add shell-pilot	2024-04-22 20:10:34 -04:00
Christian Neff	fe39147e64	Add Chatbot UI v2 to Community Integrations (#3503 )	2024-04-22 20:09:55 -04:00
Bruce MacDonald	fad00a85e5	stop running model on interactive exit	2024-04-22 16:22:14 -07:00
Cheng	62be2050dd	chore: use errors.New to replace fmt.Errorf will much better (#3789 )	2024-04-20 22:11:06 -04:00
Blake Mizerany	56f8aa6912	types/model: export IsValidNamePart (#3788 )	2024-04-20 18:26:34 -07:00
Sri Siddhaarth	e6f9bfc0e8	Update api.md (#3705 )	2024-04-20 15:17:03 -04:00
Daniel Hiltgen	8d1995c625	Merge pull request #3708 from remy415/arm64static move Ollama static build to its own flag	2024-04-18 16:04:12 -07:00
Daniel Hiltgen	fd01fbf038	Merge pull request #3710 from remy415/update-jetson-docs update jetson tutorial	2024-04-18 16:02:08 -07:00
Blake Mizerany	0408205c1c	types/model: accept former `:` as a separator in digest (#3724 ) This also converges the old sep `:` to the new sep `-`.	2024-04-18 14:17:46 -07:00
Jeffrey Morgan	63a7edd771	Update README.md	2024-04-18 16:09:38 -04:00
Michael	554ffdcce3	add llama3 to readme add llama3 to readme	2024-04-18 15:18:48 -04:00
Jeremy	9850a4ce08	Merge branch 'ollama:main' into update-jetson-docs	2024-04-18 09:55:17 -04:00
Jeremy	fd048f1367	Merge branch 'ollama:main' into arm64static	2024-04-18 09:55:04 -04:00
Michael Yang	8645076a71	Merge pull request #3712 from ollama/mxyng/mem add stablelm graph calculation	2024-04-17 15:57:51 -07:00
Michael Yang	05e9424824	Merge pull request #3664 from ollama/mxyng/fix-padding-2 fix padding to only return padding	2024-04-17 15:57:40 -07:00
Michael Yang	52ebe67a98	Merge pull request #3714 from ollama/mxyng/model-name-host types/model: support : in PartHost for host:port	2024-04-17 15:34:03 -07:00
Michael Yang	889b31ab78	types/model: support : in PartHost for host:port	2024-04-17 15:16:07 -07:00
Michael Yang	3cf483fe48	add stablelm graph calculation	2024-04-17 13:57:19 -07:00
Jeremy	8dca03173d	Merge remote-tracking branch 'upstream/main' into update-jetson-docs	2024-04-17 16:18:50 -04:00
Jeremy	85bdf14b56	update jetson tutorial	2024-04-17 16:17:42 -04:00
Jeremy	da8a0c7657	Merge branch 'ollama:main' into arm64static	2024-04-17 15:22:34 -04:00
jmorganca	c8afe7168c	use correct extension for feature and model request issue templates	2024-04-17 15:18:40 -04:00
jmorganca	28d3cd0148	simpler feature and model request forms	2024-04-17 15:17:08 -04:00
jmorganca	eb5554232a	simpler feature and model request forms	2024-04-17 15:14:49 -04:00
Jeremy	ea4c284a48	Merge branch 'ollama:main' into arm64static	2024-04-17 15:11:38 -04:00
jmorganca	2bdc320216	add descriptions to issue templates	2024-04-17 15:08:36 -04:00
jmorganca	32561aed09	simplify github issue templates a bit	2024-04-17 15:07:03 -04:00
Michael Yang	71548d9829	Merge pull request #3706 from ollama/mxyng/mem account for all non-repeating layers	2024-04-17 11:58:20 -07:00
Jeremy	8aec92fa6d	rearranged conditional logic for static build, dockerfile updated	2024-04-17 14:43:28 -04:00
Michael Yang	a8b9b930b4	account for all non-repeating layers	2024-04-17 11:21:21 -07:00
Michael	9755cf9173	acknowledge the amazing work done by Georgi and team!	2024-04-17 13:48:14 -04:00
Jeremy	70261b9bb6	move static build to its own flag	2024-04-17 13:04:28 -04:00
Blake Mizerany	9df6c85c3a	types/model: add FilepathNoBuild (#3680 ) Also, add test for DisplayLongest. Also, plumb fill param to ParseName in MustParseName	2024-04-16 18:35:43 -07:00
Michael Yang	e74163af4c	fix padding to only return padding	2024-04-16 15:43:26 -07:00
Michael Yang	fb9580df85	Merge pull request #3684 from ollama/mxyng/scale-graph scale graph based on gpu count	2024-04-16 14:57:09 -07:00
Michael Yang	26df674785	scale graph based on gpu count	2024-04-16 14:44:13 -07:00
Jeffrey Morgan	7c9792a6e0	Support unicode characters in model path (#3681 ) * parse wide argv characters on windows * cleanup * move cleanup to end of `main`	2024-04-16 17:00:12 -04:00
Michael Yang	7afb2e125a	Merge pull request #3678 from ollama/mxyng/fix-darwin-partial-offloading darwin: no partial offloading if required memory greater than system	2024-04-16 12:05:56 -07:00

1 2 3 4 5 ...

2422 commits