102 Commits

Author SHA1 Message Date
Concedo 4e1ea2ac61 hopefully fixed the ooms for good 2023-04-23 13:49:50 +08:00
Gustavo Rocha Dias 3f21bd81f3 doc - Better explanation of how to build the libraries at Windows. (#107) 2023-04-23 13:40:09 +08:00
Concedo c454f8b848 Gpt NeoX / Pythia integration completed 2023-04-22 11:23:25 +08:00
Concedo 07bb31b034 wip dont use 2023-04-21 00:35:54 +08:00
Concedo 93761e7baf slightly clarified the library replacement steps - replacing the dll is necessary in addition to replacing the library imports 2023-04-20 12:23:54 +08:00
Gustavo Rocha Dias 5ca2d774cc doc - explanation of how to use a custom version of the windows libraries at the lib folder. (#92)
the dynamic libraries also need to be updated if you replace the import libraries
2023-04-20 12:20:11 +08:00
Concedo f39def81d4 Update readme with more info 2023-04-18 21:44:26 +08:00
Concedo 3614956bc7 update readme 2023-04-18 21:39:05 +08:00
Gustavo Rocha Dias ed5b5c45a9 doc - enhanced readme explaing how to compile at Windows. (#80) 2023-04-18 17:40:04 +08:00
AlpinDale 624dc8809e Added openblas and clblas package names for debian (#63) 2023-04-15 01:08:56 +08:00
Concedo ca69e05d1f update readme and fixed typos 2023-04-11 23:53:21 +08:00
ariez-xyz b48255db19 add more precise instructions for arch 2023-04-08 10:41:57 +02:00
Concedo 1369b46bb7 notice about false positives 2023-04-08 12:20:48 +08:00
Concedo eb5b22dda2 rebrand to koboldcpp 2023-04-03 10:35:18 +08:00
Concedo bb965cc120 Merge branch 'master' into concedo
# Conflicts:
#	README.md
2023-04-02 17:13:28 +08:00
rimoliga d0a7f742e7 readme: replace termux links with homepage, play store is deprecated (#680) 2023-04-01 16:57:30 +02:00
Concedo 801b178f2a still refactoring, but need a checkpoint to prepare build for 1.0.7 2023-04-01 08:55:14 +08:00
Concedo 559a1967f7 Backwards compatibility formats all done
Merge branch 'master' into concedo

# Conflicts:
#	CMakeLists.txt
#	README.md
#	llama.cpp
2023-03-31 19:01:33 +08:00
Concedo 9eab39fe6d prepare legacy functions (+1 squashed commits)
Squashed commits:

[8bc8d0d] prepare for big merge
2023-03-31 17:45:49 +08:00
Concedo 79f9743347 improved console info, fixed utf encoding bugs 2023-03-31 15:38:38 +08:00
Pavol Rusnak 9733104be5 drop quantize.py (now that models are using a single file) 2023-03-31 01:07:32 +02:00
Georgi Gerganov 3df890aef4 readme : update supported models 2023-03-30 22:31:54 +03:00
Concedo d8febc8653 renamed main python script 2023-03-30 00:48:44 +08:00
Concedo 664b277c27 integrated libopenblas for greatly accelerated prompt processing. Windows binaries are included - feel free to build your own or to build for other platforms, but that is beyond the scope of this repo. Will fall back to non-blas if libopenblas is removed. 2023-03-30 00:43:52 +08:00
Georgi Gerganov b467702b87 readme : fix typos 2023-03-29 19:38:31 +03:00
Georgi Gerganov 516d88e75c readme : add GPT4All instructions (close #588) 2023-03-29 19:37:20 +03:00
Stephan Walter b391579db9 Update README and comments for standalone perplexity tool (#525) 2023-03-26 16:14:01 +03:00
Georgi Gerganov 348d6926ee Add logo to README.md 2023-03-26 10:20:49 +03:00
Georgi Gerganov 55ad42af84 Move chat scripts into "./examples" 2023-03-25 20:37:09 +02:00
Georgi Gerganov 4a7129acd2 Remove obsolete information from README 2023-03-25 16:30:32 +02:00
Gary Mulder f4f5362edb Update README.md (#444)
Added explicit **bolded** instructions clarifying that people need to request access to models from Facebook and never through through this repo.
2023-03-24 15:23:09 +00:00
LostRuins 1c78ffb964 Update README.md 2023-03-24 22:45:54 +08:00
Georgi Gerganov b6b268d441 Add link to Roadmap discussion 2023-03-24 09:13:35 +02:00
Stephan Walter a50e39c6fe Revert "Delete SHA256SUMS for now" (#429)
* Revert "Delete SHA256SUMS for now (#416)"

This reverts commit 8eea5ae0e5.

* Remove ggml files until they can be verified
* Remove alpaca json
* Add also model/tokenizer.model to SHA256SUMS + update README

---------

Co-authored-by: Pavol Rusnak <pavol@rusnak.io>
2023-03-23 15:15:48 +01:00
Gary Mulder 8a3e5ef801 Move model section from issue template to README.md (#421)
* Update custom.md

* Removed Model section as it is better placed in README.md

* Updates to README.md model section

* Inserted text that was removed from  issue template about obtaining models from FB and links to papers describing the various models

* Removed IPF down links for the Alpaca 7B models as these look to be in the old data format and probably shouldn't be directly linked to, anyway

* Updated the perplexity section to point at Perplexity scores #406 discussion
2023-03-23 11:30:40 +00:00
Georgi Gerganov 93208cfb92 Adjust repetition penalty .. 2023-03-23 10:46:58 +02:00
LostRuins 47ea33ab59 Update README.md 2023-03-23 16:02:19 +08:00
Georgi Gerganov 03ace14cfd Add link to recent podcast about whisper.cpp and llama.cpp 2023-03-23 09:48:51 +02:00
Gary Linscott 40ea807a97 Add details on perplexity to README.md (#395) 2023-03-22 08:53:54 -07:00
LostRuins c5c1c8d5ce Update README.md 2023-03-22 22:54:27 +08:00
Concedo 4e95e7f87f Updated readme 2023-03-22 16:20:37 +08:00
Georgi Gerganov 56817b1f88 Remove temporary notice and update hot topics 2023-03-22 07:34:02 +02:00
Gary Mulder da0e9fe90c Add SHA256SUMS file and instructions to README how to obtain and verify the downloads
Hashes created using:

sha256sum models/*B/*.pth models/*[7136]B/ggml-model-f16.bin* models/*[7136]B/ggml-model-q4_0.bin* > SHA256SUMS
2023-03-21 23:19:11 +01:00
Georgi Gerganov 3366853e41 Add notice about pending change 2023-03-21 22:57:35 +02:00
Georgi Gerganov 1daf4dd712 Minor style changes 2023-03-21 18:10:32 +02:00
Georgi Gerganov dc6a845b85 Add chat.sh script 2023-03-21 18:09:46 +02:00
Georgi Gerganov 3bfa3b43b7 Fix convert script, warnings alpaca instructions, default params 2023-03-21 17:59:16 +02:00
Kevin Kwok e0ffc861fa Update IPFS links to quantized alpaca with new tokenizer format (#352) 2023-03-21 17:34:49 +02:00
Concedo 8d39365af6 update license, added backwards compatibility with both ggml model formats, fixed context length issues. 2023-03-20 23:43:35 +08:00
Mack Straight 074bea2eb1 sentencepiece bpe compatible tokenizer (#252)
* potential out of bounds read

* fix quantize

* style

* Update convert-pth-to-ggml.py

* mild cleanup

* don't need the space-prefixing here rn since main.cpp already does it

* new file magic + version header field

* readme notice

* missing newlines

Co-authored-by: slaren <2141330+slaren@users.noreply.github.com>
2023-03-20 03:17:23 -07:00