ollama

Commit Graph

Author	SHA1	Message	Date
Jeffrey Morgan	455e61170d	Update openai.md	2024-07-25 18:34:47 -04:00
royjhan	4de1370a9d	openai tools doc (#5617 )	2024-07-25 18:34:06 -04:00
Daniel Hiltgen	6c2129d5d0	Explain font problems on windows 10	2024-07-24 15:22:00 -07:00
Daniel Hiltgen	830fdd2715	Better explain multi-gpu behavior	2024-07-23 15:16:38 -07:00
Michael Yang	9b60a038e5	update api.md	2024-07-22 13:49:51 -07:00
Michael Yang	83a0cb8d88	docs	2024-07-22 13:38:09 -07:00
royjhan	c0648233f2	api embed docs (#5282 )	2024-07-22 13:37:08 -07:00
Daniel Hiltgen	283948c83b	Adjust windows ROCm discovery The v5 hip library returns unsupported GPUs which wont enumerate at inference time in the runner so this makes sure we align discovery. The gfx906 cards are no longer supported so we shouldn't compile with that GPU type as it wont enumerate at runtime.	2024-07-20 15:17:50 -07:00
royjhan	0d41623b52	OpenAI: Add Suffix to `v1/completions` (#5611 ) * add suffix * remove todo * remove TODO * add to test * rm outdated prompt tokens info md * fix test * fix test	2024-07-16 20:50:14 -07:00
Daniel Hiltgen	1f50356e8e	Bump ROCm on windows to 6.1.2 This also adjusts our algorithm to favor our bundled ROCm. I've confirmed VRAM reporting still doesn't work properly so we can't yet enable concurrency by default.	2024-07-10 11:01:22 -07:00
Jeffrey Morgan	8f8e736b13	update llama.cpp submodule to `d7fd29f` (#5475 )	2024-07-05 13:25:58 -04:00
Daniel Hiltgen	52abc8acb7	Document older win10 terminal problems We haven't found a workaround, so for now recommend updating.	2024-07-03 17:32:14 -07:00
Daniel Hiltgen	ef757da2c9	Better nvidia GPU discovery logging Refine the way we log GPU discovery to improve the non-debug output, and report more actionable log messages when possible to help users troubleshoot on their own.	2024-07-03 10:50:40 -07:00
Daniel Hiltgen	d2f19024d0	Merge pull request #5442 from dhiltgen/concurrency_docs Add windows radeon concurrency note	2024-07-02 12:47:47 -07:00
Daniel Hiltgen	69c04eecc4	Add windows radeon concurreny note	2024-07-02 12:46:14 -07:00
royjhan	996bb1b85e	OpenAI: /v1/models and /v1/models/{model} compatibility (#5007 ) * OpenAI v1 models * Refactor Writers * Add Test Co-Authored-By: Attila Kerekes * Credit Co-Author Co-Authored-By: Attila Kerekes <439392+keriati@users.noreply.github.com> * Empty List Testing * Use Namespace for Ownedby * Update Test * Add back envconfig * v1/models docs * Use ModelName Parser * Test Names * Remove Docs * Clean Up * Test name Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Add Middleware for Chat and List * Testing Cleanup * Test with Fatal * Add functionality to chat test * OpenAI: /v1/models/{model} compatibility (#5028) * Retrieve Model * OpenAI Delete Model * Retrieve Middleware * Remove Delete from Branch * Update Test * Middleware Test File * Function name * Cleanup * Test Update * Test Update --------- Co-authored-by: Attila Kerekes <439392+keriati@users.noreply.github.com> Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com>	2024-07-02 11:50:56 -07:00
Daniel Hiltgen	dfded7e075	Merge pull request #5364 from dhiltgen/concurrency_docs Document concurrent behavior and settings	2024-07-01 09:49:48 -07:00
Eduard	27402cb7a2	Update gpu.md (#5382 ) Runs fine on a NVIDIA GeForce GTX 1050 Ti	2024-06-30 21:48:51 -04:00
Jeffrey Morgan	c1218199cf	Update api.md	2024-06-29 16:22:49 -07:00
Daniel Hiltgen	aae56abb7c	Document concurrent behavior and settings	2024-06-28 13:15:57 -07:00
royjhan	6d4219083c	Update docs (#5312 )	2024-06-28 09:58:14 -07:00
royjhan	fedf71635e	Extend api/show and ollama show to return more model info (#4881 ) * API Show Extended * Initial Draft of Information Co-Authored-By: Patrick Devine <pdevine@sonic.net> * Clean Up * Descriptive arg error messages and other fixes * Second Draft of Show with Projectors Included * Remove Chat Template * Touches * Prevent wrapping from files * Verbose functionality * Docs * Address Feedback * Lint * Resolve Conflicts * Function Name * Tests for api/show model info * Show Test File * Add Projector Test * Clean routes * Projector Check * Move Show Test * Touches * Doc update --------- Co-authored-by: Patrick Devine <pdevine@sonic.net>	2024-06-19 14:19:02 -07:00
Daniel Hiltgen	9d8a4988e8	Implement log rotation for tray app	2024-06-19 12:53:34 -07:00
Jeffrey Morgan	176d0f7075	Update import.md	2024-06-17 19:44:14 -04:00
Jeffrey Morgan	c7b77004e3	docs: add missing powershell package to windows development instructions (#5075 ) * docs: add missing instruction for powershell build The powershell script for building Ollama on Windows now requires the `ThreadJob` module. Add this to the instructions and dependency list. * Update development.md	2024-06-15 23:08:09 -04:00
Jeffrey Morgan	6b800aa7b7	openai: do not set temperature to 0 when setting seed (#5045 )	2024-06-14 13:43:56 -07:00
Patrick Devine	4dc7fb9525	update 40xx gpu compat matrix (#5036 )	2024-06-13 17:10:33 -07:00
Jeffrey Morgan	ead259d877	llm: fix seed value not being applied to requests (#4986 )	2024-06-11 14:24:41 -07:00
Michael Yang	5bc029c529	Merge pull request #4921 from ollama/mxyng/import-md update import.md	2024-06-10 11:41:09 -07:00
Napuh	896495de7b	Add instructions to easily install specific versions on faq.md (#4084 ) * Added instructions to easily install specific versions on faq.md * Small typo * Moved instructions on how to install specific version to linux.md * Update docs/linux.md * Update docs/linux.md --------- Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com>	2024-06-09 10:49:03 -07:00
Jeffrey Morgan	943172cbf4	Update api.md	2024-06-08 23:04:32 -07:00
Michael Yang	b9ce7bf75e	update import.md	2024-06-07 16:45:15 -07:00
royjhan	28c7813ac4	API PS Documentation (#4822 ) * API PS Documentation	2024-06-05 11:06:53 -07:00
Shubham	60323e0805	add embed model command and fix question invoke (#4766 ) * add embed model command and fix question invoke * Update docs/tutorials/langchainpy.md Co-authored-by: Kim Hallberg <hallberg.kim@gmail.com> * Update docs/tutorials/langchainpy.md --------- Co-authored-by: Kim Hallberg <hallberg.kim@gmail.com> Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com>	2024-06-03 22:20:48 -07:00
Daniel Hiltgen	0fc0cfc6d2	Merge pull request #4594 from dhiltgen/doc_container_workarounds Add isolated gpu test to troubleshooting	2024-05-30 13:10:54 -07:00
Daniel Hiltgen	1b2d156094	Tidy up developer guide a little	2024-05-23 15:14:05 -07:00
Daniel Hiltgen	f77713bf1f	Add isolated gpu test to troubleshooting	2024-05-23 09:33:25 -07:00
Patrick Devine	3bade04e10	doc updates for the faq/troubleshooting (#4565 )	2024-05-21 15:30:09 -07:00
alwqx	8800c8a59b	chore: fix typo in docs (#4536 )	2024-05-20 14:19:03 -07:00
Patrick Devine	f1548ef62d	update the FAQ to be more clear about windows env variables (#4415 )	2024-05-13 18:01:13 -07:00
睡觉型学渣	9c76b30d72	Correct typos. (#4387 ) * Correct typos. * Correct typos.	2024-05-12 18:21:11 -07:00
Daniel Hiltgen	8cc0ee2efe	Doc container usage and workaround for nvidia errors	2024-05-09 09:26:45 -07:00
Jeffrey Morgan	d5eec16d23	use model defaults for `num_gqa`, `rope_frequency_base ` and `rope_frequency_scale` (#1983 )	2024-05-09 09:06:13 -07:00
Carlos Gamez	daa1a032f7	Update langchainjs.md (#2027 ) Updated sample code as per warning notification from the package maintainers	2024-05-08 20:21:03 -07:00
boessu	5d3f7fff26	Update langchainpy.md (#4236 ) fixing pip code.	2024-05-07 16:36:34 -07:00
CrispStrobe	7c5330413b	note on naming restrictions (#2625 ) * note on naming restrictions else push would fail with cryptic retrieving manifest Error: file does not exist ==> maybe change that in code too * Update docs/import.md --------- Co-authored-by: C-4-5-3 <154636388+C-4-5-3@users.noreply.github.com> Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com>	2024-05-06 16:03:21 -07:00
Jeffrey Chen	d091fe3c21	Windows automatically recognizes username (#3214 )	2024-05-06 15:03:14 -07:00
Mohamed A. Fouad	ee02f548c8	Update linux.md (#3847 ) Add -e to viewing logs in order to show end of ollama logs	2024-05-06 15:02:25 -07:00
Darinka	3ecae420ac	Update api.md (#3945 ) * Update api.md Changed the calculation of tps (token/s) in the documentation * Update docs/api.md --------- Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com>	2024-05-06 14:39:58 -07:00
Adrien Brault	aa93423fbf	docs: pbcopy on mac (#3129 )	2024-05-06 13:47:00 -07:00
Hyden Liu	fb8ddc564e	chore: delete `HEAD` (#4194 )	2024-05-06 10:32:30 -07:00
Daniel Hiltgen	20f6c06569	Make maximum pending request configurable This also bumps up the default to be 50 queued requests instead of 10.	2024-05-04 21:00:52 -07:00
Daniel Hiltgen	e006480e49	Explain the 2 different windows download options	2024-05-04 12:50:05 -07:00
Dr Nic Williams	e8aaea030e	Update 'llama2' -> 'llama3' in most places (#4116 ) * Update 'llama2' -> 'llama3' in most places --------- Co-authored-by: Patrick Devine <patrick@infrahq.com>	2024-05-03 15:25:04 -04:00
Michael Yang	94c369095f	fix line ending replace CRLF with LF	2024-05-02 14:53:13 -07:00
alwqx	68755f1f5e	chore: fix typo in docs/development.md (#4073 )	2024-05-01 15:39:11 -04:00
Christian Frantzen	5950c176ca	Update langchainpy.md (#4037 ) Updated the code a bit	2024-04-29 23:19:06 -04:00
Quinten van Buul	2a80f55e2a	Update windows.md (#3855 ) Fixed a typo	2024-04-26 16:04:15 -04:00
Patrick Devine	74d2a9ef9a	add OLLAMA_KEEP_ALIVE env variable to FAQ (#3865 )	2024-04-23 21:06:51 -07:00
Sri Siddhaarth	e6f9bfc0e8	Update api.md (#3705 )	2024-04-20 15:17:03 -04:00
Jeremy	85bdf14b56	update jetson tutorial	2024-04-17 16:17:42 -04:00
Carlos Gamez	a27e419b47	Update langchainjs.md (#2030 ) Changed ollama.call() for ollama.invoke() as per deprecated documentation from langchain	2024-04-15 18:37:30 -04:00
Jeffrey Morgan	e54a3c7fcd	Update modelfile.md Remove Modelfile parameters that are decided at runtime	2024-04-15 15:35:44 -04:00
Blake Mizerany	1524f323a3	Revert "build.go: introduce a friendlier way to build Ollama (#3548 )" (#3564 )	2024-04-09 15:57:45 -07:00
Blake Mizerany	fccf3eecaa	build.go: introduce a friendlier way to build Ollama (#3548 ) This commit introduces a more friendly way to build Ollama dependencies and the binary without abusing `go generate` and removing the unnecessary extra steps it brings with it. This script also provides nicer feedback to the user about what is happening during the build process. At the end, it prints a helpful message to the user about what to do next (e.g. run the new local Ollama).	2024-04-09 14:18:47 -07:00
Thomas Vitale	cb03fc9571	Docs: Remove wrong parameter for Chat Completion (#3515 ) Fixes gh-3514 Signed-off-by: Thomas Vitale <ThomasVitale@users.noreply.github.com>	2024-04-06 09:08:35 -07:00
Daniel Hiltgen	0a74cb31d5	Safeguard for noexec We may have users that run into problems with our current payload model, so this gives us an escape valve.	2024-04-01 16:48:33 -07:00
Jeffrey Morgan	856b8ec131	remove need for `$VSINSTALLDIR` since build will fail if `ninja` cannot be found (#3350 )	2024-03-26 16:23:16 -04:00
Patrick Devine	1b272d5bcd	change `github.com/jmorganca/ollama` to `github.com/ollama/ollama` (#3347 )	2024-03-26 13:04:17 -07:00
Jeffrey Morgan	f38b705dc7	Fix ROCm link in `development.md`	2024-03-25 16:32:44 -04:00
Blake Mizerany	22921a3969	doc: specify ADAPTER is optional (#3333 )	2024-03-25 09:43:19 -07:00
Daniel Hiltgen	d8fdbfd8da	Add docs for GPU selection and nvidia uvm workaround	2024-03-21 11:52:54 +01:00
Bruce MacDonald	a5ba0fcf78	doc: faq gpu compatibility (#3142 )	2024-03-21 05:21:34 -04:00
Jeffrey Morgan	3a30bf56dc	Update faq.md	2024-03-20 17:48:39 +01:00
Jeffrey Morgan	7ed3e94105	Update faq.md	2024-03-18 10:24:39 +01:00
jmorganca	2297ad39da	update `faq.md`	2024-03-18 10:17:59 +01:00
Daniel Hiltgen	6459377ae0	Add ROCm support to linux install script (#2966 )	2024-03-14 18:00:16 -07:00
Jeffrey Morgan	5ce997a7b9	Update README.md	2024-03-13 21:12:17 -07:00
Patrick Devine	ba7cf7fb66	add more docs on for the modelfile message command (#3087 )	2024-03-12 16:41:41 -07:00
Daniel Hiltgen	b53229a2ed	Add docs explaining GPU selection env vars	2024-03-12 11:33:06 -07:00
Jeffrey Morgan	6d3adfbea2	Update troubleshooting.md	2024-03-11 13:22:28 -07:00
Daniel Hiltgen	0fdebb34a9	Doc how to set up ROCm builds on windows	2024-03-09 11:29:45 -08:00
Daniel Hiltgen	4a5c9b8035	Finish unwinding idempotent payload logic The recent ROCm change partially removed idempotent payloads, but the ggml-metal.metal file for mac was still idempotent. This finishes switching to always extract the payloads, and now that idempotentcy is gone, the version directory is no longer useful.	2024-03-09 08:34:39 -08:00
Jeffrey Morgan	6c0af2599e	Update docs `README.md` and table of contents	2024-03-08 22:45:11 -08:00
Daniel Hiltgen	280da44522	Merge pull request #2988 from dhiltgen/rocm_docs Refined ROCm troubleshooting docs	2024-03-08 13:33:30 -08:00
Jeffrey Morgan	b886bec3f9	Update api.md	2024-03-07 23:27:51 -08:00
Daniel Hiltgen	69f0227813	Refined ROCm troubleshooting docs	2024-03-07 11:22:37 -08:00
Daniel Hiltgen	6c5ccb11f9	Revamp ROCm support This refines where we extract the LLM libraries to by adding a new OLLAMA_HOME env var, that defaults to `~/.ollama` The logic was already idempotenent, so this should speed up startups after the first time a new release is deployed. It also cleans up after itself. We now build only a single ROCm version (latest major) on both windows and linux. Given the large size of ROCms tensor files, we split the dependency out. It's bundled into the installer on windows, and a separate download on windows. The linux install script is now smart and detects the presence of AMD GPUs and looks to see if rocm v6 is already present, and if not, then downloads our dependency tar file. For Linux discovery, we now use sysfs and check each GPU against what ROCm supports so we can degrade to CPU gracefully instead of having llama.cpp+rocm assert/crash on us. For Windows, we now use go's windows dynamic library loading logic to access the amdhip64.dll APIs to query the GPU information.	2024-03-07 10:36:50 -08:00
Jeffrey Morgan	d481fb3cc8	update go to 1.22 in other places (#2975 )	2024-03-07 07:39:49 -08:00
John	23ebe8fe11	fix some typos (#2973 ) Signed-off-by: hishope <csqiye@126.com>	2024-03-06 22:50:11 -08:00
Jeffrey Morgan	ce9f7c4674	Update api.md	2024-03-05 13:13:23 -08:00
Jeffrey Morgan	3b4bab3dc5	Fix embeddings load model behavior (#2848 )	2024-02-29 17:40:56 -08:00
elthommy	1f087c4d26	Update langchain python tutorial (#2737 ) Remove unused GPT4all Use nomic-embed-text as embedded model Fix a deprecation warning (__call__)	2024-02-25 00:31:36 -05:00
Jeffrey Morgan	bdc0ea1ba5	Update import.md	2024-02-22 02:08:03 -05:00
Jeffrey Morgan	7fab7918cc	Update import.md	2024-02-22 02:06:24 -05:00
Jeffrey Morgan	f0425d3de9	Update faq.md	2024-02-20 20:44:45 -05:00
Jeffrey Morgan	8125ce4cb6	Update import.md Add instructions to get public key on windows	2024-02-19 22:48:24 -05:00
Jeffrey Morgan	df56f1ee5e	Update faq.md	2024-02-19 22:16:42 -05:00
Jeffrey Morgan	41aca5c2d0	Update faq.md	2024-02-19 21:11:01 -05:00
Jeffrey Morgan	753724d867	Update api.md to include examples for reproducible outputs	2024-02-19 20:36:16 -05:00
Patrick Devine	9a7a4b9533	add faqs for memory pre-loading and the keep_alive setting (#2601 )	2024-02-19 14:45:25 -08:00
Daniel Hiltgen	b338c0635f	Document setting server vars for windows	2024-02-19 13:30:46 -08:00
Tristan Rhodes	9774663013	Update faq.md with the location of models on Windows (#2545 )	2024-02-16 11:04:19 -08:00
Daniel Hiltgen	1ba734de67	typo	2024-02-15 14:56:55 -08:00
Daniel Hiltgen	29e90cc13b	Implement new Go based Desktop app This focuses on Windows first, but coudl be used for Mac and possibly linux in the future.	2024-02-15 05:56:45 +00:00
Jeffrey Morgan	48a273f80b	Fix issues with templating prompt in chat mode (#2460 )	2024-02-12 15:06:57 -08:00
Jeffrey Morgan	1c8435ffa9	Update domain name references in docs and install script (#2435 )	2024-02-09 15:19:30 -08:00
Jeffrey Morgan	42b797ed9c	Update openai.md	2024-02-08 15:03:23 -05:00
Jeffrey Morgan	336aa43f3c	Update openai.md	2024-02-08 12:48:28 -05:00
Jeffrey Morgan	ab0d37fde4	Update openai.md	2024-02-07 17:25:33 -05:00
Jeffrey Morgan	14e71350c8	Update openai.md	2024-02-07 17:25:24 -05:00
Jeffrey Morgan	453f572f83	Initial OpenAI `/v1/chat/completions` API compatibility (#2376 )	2024-02-07 17:24:29 -05:00
Bruce MacDonald	128fce5495	docs: keep_alive (#2258 )	2024-02-06 11:00:05 -05:00
Jeffrey Morgan	b9f91a0b36	Update import instructions to use convert and quantize tooling from llama.cpp submodule (#2247 )	2024-02-05 00:50:44 -05:00
Jeffrey Morgan	f0e9496c85	Update api.md	2024-02-02 12:17:24 -08:00
Daniel Hiltgen	e7dbb00331	Add container hints for troubleshooting Some users are new to containers and unsure where the server logs go	2024-01-29 08:53:41 -08:00
Daniel Hiltgen	e02ecfb6c8	Merge pull request #2116 from dhiltgen/cc_50_80 Add support for CUDA 5.0 cards	2024-01-27 10:28:38 -08:00
Jeffrey Morgan	5be9bdd444	Update modelfile.md	2024-01-25 16:29:48 -08:00
Jeffrey Morgan	b706794905	Update modelfile.md to include `MESSAGE`	2024-01-25 16:29:32 -08:00
Michael Yang	93a756266c	faq: update to use launchctl setenv	2024-01-22 13:10:13 -08:00
Daniel Hiltgen	df54c723ae	Make CPU builds parallel and customizable AMD GPUs The linux build now support parallel CPU builds to speed things up. This also exposes AMD GPU targets as an optional setting for advaced users who want to alter our default set.	2024-01-21 15:12:21 -08:00
Daniel Hiltgen	a447a083f2	Add compute capability 5.0, 7.5, and 8.0	2024-01-20 14:24:05 -08:00
Daniel Hiltgen	abec7f06e5	Merge pull request #2056 from dhiltgen/slog Mechanical switch from log to slog	2024-01-18 14:27:24 -08:00
Daniel Hiltgen	ecbfc0182f	Go bump to v1.21 to pick up slog	2024-01-18 14:12:57 -08:00
Daniel Hiltgen	fedd705aea	Mechanical switch from log to slog A few obvious levels were adjusted, but generally everything mapped to "info" level.	2024-01-18 14:12:57 -08:00
Daniel Hiltgen	9cd20b0ec8	Refine the linux cuda/rocm developer docs	2024-01-18 09:44:44 -08:00
Tristram Oaten	40a0a90a88	Add group delete to uninstall instructions (#1924 ) After executing the `userdel ollama` command, I saw this message: ```sh $ sudo userdel ollama userdel: group ollama not removed because it has other members. ``` Which reminded me that I had to remove the dangling group too. For completeness, the uninstall instructions should do this too. Thanks!	2024-01-12 00:07:00 -05:00
Daniel Hiltgen	d88c527be3	Build multiple CPU variants and pick the best This reduces the built-in linux version to not use any vector extensions which enables the resulting builds to run under Rosetta on MacOS in Docker. Then at runtime it checks for the actual CPU vector extensions and loads the best CPU library available	2024-01-11 08:42:47 -08:00
Robin Glauser	e868c8a5c7	Update api.md (#1878 ) Fixed assistant in the example response.	2024-01-09 16:21:17 -05:00
Bruce MacDonald	3f3eb19a3b	document response in modelfile template variables (#1428 )	2024-01-08 14:38:51 -05:00
Daniel Hiltgen	2d9dd14f27	Merge pull request #1697 from dhiltgen/win_docs Add windows native build instructions	2024-01-05 19:34:20 -08:00
Matt Williams	df086d3c8c	fix docker doc to point to hub Signed-off-by: Matt Williams <m@technovangelist.com>	2024-01-04 18:42:23 -08:00
Bruce MacDonald	b846eb64d0	Fix `template` api doc description (#1661 )	2024-01-03 11:00:59 -05:00
Cole Gillespie	3c5dd9ed1d	Update README.md (#1766 )	2024-01-03 10:44:22 -05:00
Jeffrey Morgan	b17ccd0542	Update import.md	2024-01-02 22:28:18 -05:00
Jeffrey Morgan	2a2fa3c329	`api.md` cleanup & formatting	2023-12-27 14:32:35 -05:00
Daniel Hiltgen	e201efa14b	Add windows native build instructions	2023-12-25 08:31:34 -08:00
K0IN	10da41d677	Add Cache flag to api (#1642 )	2023-12-22 17:16:20 -05:00
Matt Williams	511069a2a5	update where are models stored q Signed-off-by: Matt Williams <m@technovangelist.com>	2023-12-22 09:48:44 -08:00
Matt Williams	291700c92d	Clean up documentation (#1506 ) * Clean up documentation Will probably need to update with PRs for new release. Signed-off-by: Matt Williams <m@technovangelist.com> * Correcting to fit in 0.1.15 changes Signed-off-by: Matt Williams <m@technovangelist.com> * Update README.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * addressing comments Signed-off-by: Matt Williams <m@technovangelist.com> * more api cleanup Signed-off-by: Matt Williams <m@technovangelist.com> * its llava not llama Signed-off-by: Matt Williams <m@technovangelist.com> * Update docs/troubleshooting.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Updated hosting to server and documented all env vars Signed-off-by: Matt Williams <m@technovangelist.com> * remove last of the cli descriptions Signed-off-by: Matt Williams <m@technovangelist.com> * Update README.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * update further per conversation with jeff earlier today Signed-off-by: Matt Williams <m@technovangelist.com> * cleanup the doc readme Signed-off-by: Matt Williams <m@technovangelist.com> * move upgrade to faq Signed-off-by: Matt Williams <m@technovangelist.com> * first change Signed-off-by: Matt Williams <m@technovangelist.com> * updated Signed-off-by: Matt Williams <m@technovangelist.com> * Update docs/faq.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/README.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update README.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/README.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/README.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/README.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/README.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * examples in parent Signed-off-by: Matt Williams <m@technovangelist.com> * add exapmle for create model. Signed-off-by: Matt Williams <m@technovangelist.com> * update faq Signed-off-by: Matt Williams <m@technovangelist.com> * update create model api Signed-off-by: Matt Williams <m@technovangelist.com> * Update docs/api.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/faq.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/troubleshooting.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * update the readme in docs Signed-off-by: Matt Williams <m@technovangelist.com> * update a few more things Signed-off-by: Matt Williams <m@technovangelist.com> * Update docs/troubleshooting.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/faq.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update README.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/modelfile.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> * Update docs/troubleshooting.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com> --------- Signed-off-by: Matt Williams <m@technovangelist.com> Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com>	2023-12-22 09:10:01 -08:00
Daniel Hiltgen	e5202eb687	Quiet down llama.cpp logging by default By default builds will now produce non-debug and non-verbose binaries. To enable verbose logs in llama.cpp and debug symbols in the native code, set `CGO_CFLAGS=-g`	2023-12-22 08:47:18 -08:00
Daniel Hiltgen	96fb441abd	Merge pull request #1146 from dhiltgen/ext_server_cgo Add cgo implementation for llama.cpp	2023-12-22 08:16:31 -08:00
Daniel Hiltgen	495c06e4a6	Fix doc glitch	2023-12-21 18:21:31 -08:00
Patrick Devine	a607d922f0	add FAQ for slow networking in WSL2 (#1646 )	2023-12-20 16:27:24 -08:00
Jeffrey Morgan	df06812494	Update api.md	2023-12-20 08:47:53 -05:00
Daniel Hiltgen	1b991d0ba9	Refine build to support CPU only If someone checks out the ollama repo and doesn't install the CUDA library, this will ensure they can build a CPU only version	2023-12-19 09:05:46 -08:00
Bruce MacDonald	811b1f03c8	deprecate ggml - remove ggml runner - automatically pull gguf models when ggml detected - tell users to update to gguf in the case automatic pull fails Co-Authored-By: Jeffrey Morgan <jmorganca@gmail.com>	2023-12-19 09:05:46 -08:00
Bruce MacDonald	6e16098a60	remove sample_count from docs (#1527 ) this info has not been returned from these endpoints in some time	2023-12-14 17:49:00 -05:00
Jeffrey Morgan	fedba24a63	Docs for multimodal support (#1485 ) * add multimodal docs * add chat api docs * consistency between `/api/generate` and `/api/chat` * simplify docs	2023-12-13 13:59:33 -05:00
pepperoni21	e3b090dbc5	Added message format for chat api (#1488 )	2023-12-13 11:21:23 -05:00
Jeffrey Morgan	0a9d348023	Fix issues with `/set template` and `/set system` (#1486 )	2023-12-12 14:43:19 -05:00
Patrick Devine	910e9401d0	Multimodal support (#1216 ) --------- Co-authored-by: Matt Apperson <mattapperson@Matts-MacBook-Pro.local>	2023-12-11 13:56:22 -08:00
Jeffrey Morgan	5d4d2e2c60	update docs with chat completion api	2023-12-10 13:53:36 -05:00
Jeffrey Morgan	32064a0646	fix empty response when receiving runner error	2023-12-10 10:53:38 -05:00
Jeffrey Morgan	b74580c913	Update api.md	2023-12-08 16:02:07 -08:00
Jeffrey Morgan	2a2289fb6b	Update api.md	2023-12-08 09:36:45 -08:00
Jeffrey Morgan	ba264e9da8	add future version note to chat api docs	2023-12-07 09:42:15 -08:00
Xe Iaso	f9b7d65e2b	docs/tutorials: add bit on how to use Fly GPUs on-demand with Ollama (#1406 ) Signed-off-by: Xe Iaso <xe@camellia.finch-kitefin.ts.net>	2023-12-06 14:14:02 -08:00
Samuel Calderon	13524b5e72	List "Send chat messages" in table of contents (#1399 ) Thank you @calderonsamuel	2023-12-06 12:34:27 -08:00
Jeffrey Morgan	97c5696945	fix base urls in chat examples	2023-12-06 12:10:20 -08:00
Bruce MacDonald	195e3d9dbd	chat api endpoint (#1392 )	2023-12-05 14:57:33 -05:00
Jeffrey Morgan	00d06619a1	Revert "chat api (#991 )" while context variable is fixed This reverts commit `7a0899d62d`.	2023-12-04 21:16:27 -08:00
Matt Williams	f1ef3f9947	remove mention of gpt-neox in import (#1381 ) Signed-off-by: Matt Williams <m@technovangelist.com>	2023-12-04 20:58:10 -08:00
Bruce MacDonald	7a0899d62d	chat api (#991 ) - update chat docs - add messages chat endpoint - remove deprecated context and template generate parameters from docs - context and template are still supported for the time being and will continue to work as expected - add partial response to chat history	2023-12-04 18:01:06 -05:00
James Radtke	7eda3d0c55	Corrected transposed 129 to 192 for OLLAMA_ORIGINS example (#1325 )	2023-11-29 22:44:17 -05:00
Alec Hammond	91897a606f	Add OllamaEmbeddings to python LangChain example (#994 ) * Add OllamaEmbeddings to python LangChain example * typo --------- Co-authored-by: Alec Hammond <alechammond@fb.com>	2023-11-29 16:25:39 -05:00
ToasterUwU	63097607b2	Correct MacOS Host port example (#1301 )	2023-11-29 11:44:03 -05:00
ftorto	e1a69d44c9	Update faq.md (#1299 ) Fix a typo in the CA update command	2023-11-28 09:54:42 -05:00
Jeffrey Morgan	2eaa95b417	Update api.md	2023-11-21 15:32:05 -05:00
James Braza	f24741ff39	Documenting how to view `Modelfile`s (#723 ) * Documented viewing Modelfiles in ollama.ai/library * Moved Modelfile in ollama.ai down per request	2023-11-20 15:24:29 -05:00
Jeffrey Morgan	1657c6abc7	add note to specify JSON in the prompt when using JSON mode	2023-11-18 22:59:26 -05:00
Michael Yang	c82ead4d01	faq: fix heading and add more details	2023-11-17 09:02:17 -08:00
Michael Yang	90860b6a7e	update faq (#1176 )	2023-11-17 11:42:58 -05:00
Jeffrey Morgan	81092147c4	remove unnecessary `-X POST` from example `curl` commands	2023-11-17 09:50:38 -05:00
Jeffrey Morgan	92656a74b7	Use `llama2` as the model in `api.md`	2023-11-17 07:17:51 -05:00
Michael Yang	d8842b4d4b	update faq	2023-11-16 17:07:36 -08:00
Michael Yang	c13bde962d	Update docs/faq.md Co-authored-by: Jeffrey Morgan <jmorganca@gmail.com>	2023-11-16 16:48:38 -08:00
Michael Yang	ee307937fd	update faq	2023-11-16 16:46:43 -08:00
Michael Yang	b5f158f046	add faq for proxies (#1147 )	2023-11-16 11:43:37 -05:00
Michael Yang	77954bea0e	Merge pull request #898 from jmorganca/mxyng/build-context create remote models	2023-11-15 16:41:12 -08:00
Michael Yang	54f92f01cb	update docs	2023-11-15 15:28:15 -08:00
Jeffrey Morgan	ecd71347ab	Update faq.md	2023-11-15 18:17:13 -05:00
Jeffrey Morgan	8ee4cbea0f	Remove table of contents in `faq.md`	2023-11-15 18:16:27 -05:00
Michael Yang	71d71d0988	update docs	2023-11-15 15:16:23 -08:00
Michael Yang	cac11c9137	update api docs	2023-11-15 15:16:23 -08:00
Matt Williams	f61f340279	FAQ: answer a few faq questions (#1128 ) * faq: does ollama share my prompts Signed-off-by: Matt Williams <m@technovangelist.com> * faq: ollama and openai Signed-off-by: Matt Williams <m@technovangelist.com> * faq: vscode plugins Signed-off-by: Matt Williams <m@technovangelist.com> * faq: send a doc to Ollama Signed-off-by: Matt Williams <m@technovangelist.com> * extra spacing Signed-off-by: Matt Williams <m@technovangelist.com> * Update faq.md * Update faq.md --------- Signed-off-by: Matt Williams <m@technovangelist.com> Co-authored-by: Michael <mchiang0610@users.noreply.github.com>	2023-11-15 18:05:13 -05:00
bnodnarb	85951d25ef	Created tutorial for running Ollama on NVIDIA Jetson devices (#1098 )	2023-11-15 12:32:37 -05:00
Bruce MacDonald	df18486c35	Move /generate format to optional parameters (#1127 ) This field is optional and should be under the `Advanced parameters` header	2023-11-14 16:12:30 -05:00
Jeffrey Morgan	5cba29b9d6	JSON mode: add `"format" as an api parameter (#1051 ) * add `"format": "json"` as an API parameter --------- Co-authored-by: Bruce MacDonald <brucewmacdonald@gmail.com>	2023-11-09 16:44:02 -08:00
Bruce MacDonald	5b39503bcd	document specifying multiple stop params (#1061 )	2023-11-09 13:16:26 -08:00
Matt Williams	dd3dc47ddb	Merge pull request #992 from aashish2057/aashish2057/langchainjs_doc_update	2023-11-09 05:08:31 -08:00
Bruce MacDonald	a49d6acc1e	add a complete /generate options example (#1035 )	2023-11-08 16:44:36 -08:00
Bruce MacDonald	ec2a31e9b3	support raw generation requests (#952 ) - add the optional `raw` generate request parameter to bypass prompt formatting and response context -add raw request to docs	2023-11-08 14:05:02 -08:00
Matt Williams	1d155caba3	docs: clarify where the models are stored in the faq Signed-off-by: Matt Williams <m@technovangelist.com>	2023-11-06 14:38:49 -08:00
aashish2057	b13586cc72	update langchainjs doc	2023-11-03 18:45:19 -05:00
Bruce MacDonald	6109bebba6	reformat api docs for more examples (#972 )	2023-11-03 10:57:00 -04:00
Matt Williams	f21bd6210d	docs: clarify and clean up API docs Signed-off-by: Matt Williams <m@technovangelist.com>	2023-10-31 13:11:33 -07:00
Dirk Loss	874bb31986	Fix conversion command for gptneox (#948 )	2023-10-30 14:34:29 -04:00
Jeffrey Morgan	c0dcea1398	Update faq.md	2023-10-27 18:29:00 -07:00
Bruce MacDonald	5c3491f425	allow for a configurable ollama model storage directory (#897 ) * allow for a configurable ollama models directory - set OLLAMA_MODELS in the environment that ollama is running in to change where model files are stored - update docs Co-Authored-By: Jeffrey Morgan <jmorganca@gmail.com> Co-Authored-By: Jay Nakrani <dhananjaynakrani@gmail.com> Co-Authored-By: Akhil Acharya <akhilcacharya@gmail.com> Co-Authored-By: Sasha Devol <sasha.devol@protonmail.com>	2023-10-27 10:19:59 -04:00

... 2 3 4 5 6 ...

473 Commits