Commit Graph

  • c16f8af911 fix: multiple templates when creating from model Michael Yang 2024-06-12 13:30:08 -07:00
  • 217f60c3d9 Merge pull request #4987 from ollama/mxyng/revert-byte-order v0.1.43 Michael Yang 2024-06-11 16:04:20 -07:00
  • 7bdcd1da94 Revert "Merge pull request #4938 from ollama/mxyng/fix-byte-order" Michael Yang 2024-06-11 15:55:44 -07:00
  • ead259d877 llm: fix seed value not being applied to requests (#4986) Jeffrey Morgan 2024-06-11 14:24:41 -07:00
  • 2ff45d571d Add Ollama-hpp to Community Libraries in README. (#4983) James Montgomery 2024-06-11 14:15:05 -04:00
  • 157f09acdf fix: "Skip searching for network devices" jayson-cloude 2024-06-11 16:11:35 +08:00
  • 0f3cf1d42e Merge pull request #4715 from ollama/mxyng/utf16-parser Michael Yang 2024-06-10 11:41:29 -07:00
  • 5bc029c529 Merge pull request #4921 from ollama/mxyng/import-md Michael Yang 2024-06-10 11:41:09 -07:00
  • e9a9c6a8e8 Merge pull request #4965 from ollama/mxyng/skip-layer-remove Michael Yang 2024-06-10 11:40:03 -07:00
  • 515f497e6d fix: skip removing layers that no longer exist Michael Yang 2024-06-10 11:15:03 -07:00
  • b27268aaef add test Michael Yang 2024-06-10 11:31:34 -07:00
  • f5f245cc15 Merge pull request #4938 from ollama/mxyng/fix-byte-order Michael Yang 2024-06-10 09:38:12 -07:00
  • 94d37fdcae fix: examples/langchain-python-rag-privategpt/requirements.txt (#3382) Jim Scardelis 2024-06-09 10:58:09 -07:00
  • b84aea1685 Critical fix from llama.cpp JSON grammar to forbid un-escaped escape characters inside strings, which breaks parsing. (#3782) Craig Hughes 2024-06-09 13:57:09 -04:00
  • 896495de7b Add instructions to easily install specific versions on faq.md (#4084) Napuh 2024-06-09 19:49:03 +02:00
  • 5528dd9d11 Error handling load_single_document() in ingest.py (#4852) dcasota 2024-06-09 19:41:07 +02:00
  • 943172cbf4 Update api.md Jeffrey Morgan 2024-06-08 23:04:32 -07:00
  • 85169e8d6f Added headless-ollama (#4612) Nischal Jain 2024-06-09 07:21:16 +05:30
  • 34f142797a llm: always add bos token to prompt (#4941) Jeffrey Morgan 2024-06-08 18:47:10 -07:00
  • 46a7f1e74a Update README.md with LangChainRust (#4854) Erhan 2024-06-09 03:29:36 +03:00
  • 620d5c569e fix parsing big endian gguf Michael Yang 2024-06-08 12:32:02 -07:00
  • b9ce7bf75e update import.md Michael Yang 2024-06-07 16:45:15 -07:00
  • cddc63381c Merge pull request #4909 from dhiltgen/oneapi_disable Daniel Hiltgen 2024-06-07 14:07:15 -07:00
  • 385a32ecb5 Merge pull request #4910 from ollama/mxyng/detect-chat-template v0.1.42 Michael Yang 2024-06-07 11:07:39 -07:00
  • 030e765e76 fix create model when template detection errors Michael Yang 2024-06-07 08:55:46 -07:00
  • ab8c929e20 Add ability to skip oneapi generate Daniel Hiltgen 2024-06-07 08:32:49 -07:00
  • ce0dc33cb8 llm: patch to fix qwen 2 temporarily on nvidia (#4897) Jeffrey Morgan 2024-06-06 23:14:33 -07:00
  • 78f81fc0e5 Merge pull request #4800 from ollama/mxyng/detect-chat-template Michael Yang 2024-06-06 16:17:18 -07:00
  • 9b6c2e6eb6 detect chat template from KV Michael Yang 2024-06-03 11:06:29 -07:00
  • 1a29e9a879 API app/browser access (#4879) royjhan 2024-06-06 15:19:03 -07:00
  • 4bf1da4944 Separate ListResponse and ModelResponse for api/tags vs api/ps (#4842) royjhan 2024-06-06 10:11:45 -07:00
  • de5beb06b3 server: skip blob verification for already verified blobs Blake Mizerany 2024-05-24 08:40:40 -07:00
  • 98e65929dc docs(tools): add gollama (#4829) Sam 2024-06-06 09:13:39 +12:00
  • 66ab48772f proper utf16 support Michael Yang 2024-05-29 21:37:07 -07:00
  • 22fcf8f7de Merge pull request #3737 from ollama/mxyng/modelname-4 Michael Yang 2024-06-05 12:05:05 -07:00
  • 28c7813ac4 API PS Documentation (#4822) royjhan 2024-06-05 11:06:53 -07:00
  • 1d8616d30f docs: update to add LLocal.in to web & desktop integrations (#4719) Kartikeya Mishra 2024-06-05 03:13:59 +05:30
  • d61ef8b954 update create handler to use model.Name Michael Yang 2024-05-08 14:36:08 -07:00
  • 89d9900152 Merge pull request #4570 from ollama/mxyng/slices Michael Yang 2024-06-04 13:27:05 -07:00
  • 4a048715b6 local wording was confusing people Michael 2024-06-04 13:25:25 -07:00
  • 6297f85606 gofmt, goimports Michael Yang 2024-06-04 11:53:23 -07:00
  • ed56428dd7 warn on intrange, usestdlibvars Michael Yang 2024-06-04 11:51:39 -07:00
  • ad40b92b6a disable intrange Michael Yang 2024-06-04 11:35:30 -07:00
  • 8ce4032e72 more lint Michael Yang 2024-05-29 18:22:03 -07:00
  • 42660466f8 no usestdlibvars Michael Yang 2024-05-23 11:04:46 -07:00
  • e919f6811f lint windows Michael Yang 2024-05-22 09:26:45 -07:00
  • bf7edb0d5d lint linux Michael Yang 2024-05-22 09:08:01 -07:00
  • f38353d6b9 stdin.fd Michael Yang 2024-05-22 09:00:38 -07:00
  • 201d853fdf nolintlint Michael Yang 2024-05-22 08:52:00 -07:00
  • e40145a39d lint Michael Yang 2024-05-21 22:21:04 -07:00
  • c895a7d13f some gocritic Michael Yang 2024-05-21 22:07:57 -07:00
  • dad7a987ae nosprintfhostport Michael Yang 2024-05-21 21:53:44 -07:00
  • 8ffb51749f nolintlint Michael Yang 2024-05-21 21:52:20 -07:00
  • 55f6eba049 gofmt Michael Yang 2024-05-21 21:32:43 -07:00
  • 04f3c12bb7 replace x/exp/slices with slices Michael Yang 2024-05-21 21:30:52 -07:00
  • 60323e0805 add embed model command and fix question invoke (#4766) Shubham 2024-06-04 10:50:48 +05:30
  • d4a86102fd update welcome prompt in windows to llama3 (#4779) Jeffrey Morgan 2024-06-01 21:05:51 -07:00
  • 476fb8e892 Limit GPU lib search for now (#4777) v0.1.41 Jeffrey Morgan 2024-06-01 19:24:33 -07:00
  • 829ff87bd1 revert tokenize ffi (#4761) v0.1.40 Michael Yang 2024-05-31 18:54:21 -07:00
  • f6b622c4b3 Merge pull request #4733 from ollama/jyan/isvalidname Josh 2024-05-31 14:08:45 -07:00
  • 2e4da8eec2 added tests for IsValidNamespace Josh Yan 2024-05-31 11:48:07 -07:00
  • 763bb65dbb use int32_t for call to tokenize (#4738) v0.1.40-rc1 Jeffrey Morgan 2024-05-30 21:43:30 -07:00
  • 7ca9605f54 speed up tests by only building static lib (#4740) Jeffrey Morgan 2024-05-30 21:43:15 -07:00
  • eb2c443a79 Merge pull request #4736 from ollama/mxyng/vocab-only Michael Yang 2024-05-30 17:21:00 -07:00
  • 278e25ea44 Merge pull request #4737 from ollama/mxyng/less-generate Michael Yang 2024-05-30 17:17:50 -07:00
  • a50a87a7b8 partial offloading: allow flash attention and disable mmap (#4734) Jeffrey Morgan 2024-05-30 16:58:01 -07:00
  • 98085015d5 only generate on relevant changes Michael Yang 2024-05-22 09:58:26 -07:00
  • bf54c845e9 vocab only Michael Yang 2024-05-30 16:49:28 -07:00
  • c365f195a8 directly use isvalidpart Josh Yan 2024-05-30 16:40:04 -07:00
  • e91d0ef737 Merge pull request #4728 from ollama/jyan/japanese Josh 2024-05-30 16:25:12 -07:00
  • 22f5c12ced Update llama.cpp submodule to 5921b8f0 (#4731) Jeffrey Morgan 2024-05-30 16:20:22 -07:00
  • 298c996e54 added IsValidNamespace function Josh Yan 2024-05-30 16:02:07 -07:00
  • 0fc0cfc6d2 Merge pull request #4594 from dhiltgen/doc_container_workarounds Daniel Hiltgen 2024-05-30 13:10:54 -07:00
  • 914f68f021 replaced duplicate call with variable Josh Yan 2024-05-30 10:38:07 -07:00
  • bd1d119ba9 fixed japanese characters deleted at end of line Josh Yan 2024-05-30 10:24:21 -07:00
  • a03be18189 Fix OLLAMA_LLM_LIBRARY with wrong map name and add more env vars to help message (#4663) Lei Jitang 2024-05-31 00:36:51 +08:00
  • 96bc232b43 Merge pull request #4413 from ollama/mxyng/name-check Michael Yang 2024-05-29 12:06:58 -07:00
  • bca7b12284 Merge pull request #3718 from ollama/mxyng/modelname-3 Michael Yang 2024-05-29 12:02:07 -07:00
  • 32cb1960c1 Merge pull request #4380 from ollama/mxyng/tokenize Michael Yang 2024-05-29 12:00:59 -07:00
  • de781b37c8 rm unused infill Michael Yang 2024-05-12 09:21:35 -07:00
  • 3e21799377 rm unused system prompt Michael Yang 2024-05-12 09:20:39 -07:00
  • 26a00a0410 use ffi for tokenizing/detokenizing Michael Yang 2024-05-11 12:49:24 -07:00
  • 646371f56d Merge pull request #3278 from zhewang1-intc/rebase_ollama_main Daniel Hiltgen 2024-05-28 16:30:50 -07:00
  • 1f5008544b Update install.sh Jeffrey Morgan 2024-05-28 15:01:22 -07:00
  • 45cbfc5aee fix wsl2 status check for nvidia cards (#4689) Jeffrey Morgan 2024-05-28 14:49:46 -07:00
  • 6d423b383b Improve install experience on WSL2 and Linux (#4653) Jeffrey Morgan 2024-05-28 14:41:50 -07:00
  • ad897080a2 working on integration of multi-byte and multi-width runes (#4549) v0.1.39 Josh 2024-05-28 12:04:03 -07:00
  • b7d316d98d fix nvidia detection in install script (#4683) Jeffrey Morgan 2024-05-28 09:59:36 -07:00
  • d7339fad52 Merge pull request #4682 from dhiltgen/more_time Daniel Hiltgen 2024-05-28 09:36:02 -07:00
  • 92c81e8117 Give the final model loading more time Daniel Hiltgen 2024-05-28 08:56:18 -07:00
  • 9db0996ed4 Add OllamaSpring Project to Readme (#4672) Tai 2024-05-28 10:58:26 +08:00
  • 6f43898b17 Adds olpaka flutter client (#4647) Orfeo Ciano 2024-05-28 01:22:01 +01:00
  • 7487229c34 llm/server.go: Fix 2 minor typos (#4661) Lei Jitang 2024-05-28 08:21:10 +08:00
  • 8a8e7afa96 small fix on examples/python-simplechat/client.py to actually get a streamed response and get tokens printed as we receive it (#4671) Rayan Mostovoi 2024-05-28 02:19:20 +02:00
  • c79f8c9c39 Ensure nvidia and nvidia_uvm kernel modules are loaded in install.sh script and at startup (#4652) Jeffrey Morgan 2024-05-26 14:57:17 -07:00
  • 485016bfbb Update install.sh Jeffrey Morgan 2024-05-26 11:46:00 -07:00
  • 0165ba1651 Merge pull request #4638 from dhiltgen/better_error Daniel Hiltgen 2024-05-25 14:32:28 -07:00
  • c4209d6d21 Report better warning on client closed abort of load Daniel Hiltgen 2024-05-25 09:23:28 -07:00
  • 6adca97f37 Merge pull request #4619 from noxer/patch-1 Michael Yang 2024-05-24 17:21:57 -07:00
  • 9a3c8003c8 Merge pull request #4624 from ollama/mxyng/fix-5 Michael Yang 2024-05-24 16:11:21 -07:00