Wojtek Kowaluk
1436c5845a
fix ggml detection regex in model downloader ( #1779 )
2023-05-04 11:48:36 -03:00
Mylo
bd531c2dc2
Make --trust-remote-code work for all models ( #1772 )
2023-05-04 02:01:28 -03:00
oobabooga
0e6d17304a
Clearer syntax for instruction-following characters
2023-05-03 22:50:39 -03:00
oobabooga
9c77ab4fc2
Improve some warnings
2023-05-03 22:06:46 -03:00
oobabooga
057b1b2978
Add credits
2023-05-03 21:49:55 -03:00
oobabooga
95d04d6a8d
Better warning messages
2023-05-03 21:43:17 -03:00
oobabooga
0a48b29cd8
Prevent websocket disconnection on the client side
2023-05-03 20:44:30 -03:00
oobabooga
4bf7253ec5
Fix typing bug in api
2023-05-03 19:27:20 -03:00
oobabooga
d6410a1b36
Bump recommended monkey patch commit
2023-05-03 14:49:25 -03:00
oobabooga
60be76f0fc
Revert gradio bump (gallery is broken)
2023-05-03 11:53:30 -03:00
Thireus ☠
4883e20fa7
Fix openai extension script.py - TypeError: '_Environ' object is not callable ( #1753 )
2023-05-03 09:51:49 -03:00
oobabooga
f54256e348
Rename no_mmap to no-mmap
2023-05-03 09:50:31 -03:00
Roberts Slisans
dec31af910
Create .gitignore ( #43 )
2023-05-02 23:47:19 -03:00
Semih Aslan
24c5ba2b9c
Fixed error when $OS_ARCH returns aarch64 ( #45 )
...
For some machines $OS_ARCH returns aarch64 instead of ARM64,and as i see here this should fix it.
2023-05-02 23:47:03 -03:00
oobabooga
875da16b7b
Minor CSS improvements in chat mode
2023-05-02 23:38:51 -03:00
practicaldreamer
e3968f7dd0
Fix Training Pad Token ( #1678 )
...
Currently padding with 0 the character vs 0 the token id (<unk> in the case of llama)
2023-05-02 23:16:08 -03:00
Wojtab
80c2f25131
LLaVA: small fixes ( #1664 )
...
* change multimodal projector to the correct one
* remove reference to custom stopping strings from readme
* fix stopping strings if tokenizer extension adds/removes tokens
* add API example
* LLaVA 7B just dropped, add to readme that there is no support for it currently
2023-05-02 23:12:22 -03:00
oobabooga
c31b0f15a7
Remove some spaces
2023-05-02 23:07:07 -03:00
oobabooga
320fcfde4e
Style/pep8 improvements
2023-05-02 23:05:38 -03:00
oobabooga
ecd79caa68
Update Extensions.md
2023-05-02 22:52:32 -03:00
matatonic
7ac41b87df
add openai compatible api ( #1475 )
2023-05-02 22:49:53 -03:00
oobabooga
4e09df4034
Only show extension in UI if it has an ui() function
2023-05-02 19:20:02 -03:00
oobabooga
d016c38640
Bump gradio version
2023-05-02 19:19:33 -03:00
oobabooga
88cdf6ed3d
Prevent websocket from disconnecting
2023-05-02 19:03:19 -03:00
Ahmed Said
fbcd32988e
added no_mmap & mlock parameters to llama.cpp and removed llamacpp_model_alternative ( #1649 )
...
---------
Co-authored-by: oobabooga <112222186+oobabooga@users.noreply.github.com>
2023-05-02 18:25:28 -03:00
Blake Wyatt
4babb22f84
Fix/Improve a bunch of things ( #42 )
2023-05-02 12:28:20 -03:00
Carl Kenner
2f1a2846d1
Verbose should always print special tokens in input ( #1707 )
2023-05-02 01:24:56 -03:00
Alex "mcmonkey" Goodwin
0df0b2d0f9
optimize stopping strings processing ( #1625 )
2023-05-02 01:21:54 -03:00
oobabooga
e6a78c00f2
Update Docker.md
2023-05-02 00:51:10 -03:00
Tom Jobbins
3c67fc0362
Allow groupsize 1024, needed for larger models eg 30B to lower VRAM usage ( #1660 )
2023-05-02 00:46:26 -03:00
Lawrence M Stewart
78bd4d3a5c
Update LLaMA-model.md ( #1700 )
...
protobuf needs to be 3.20.x or lower
2023-05-02 00:44:09 -03:00
Dhaladom
f659415170
fixed variable name "context" to "prompt" ( #1716 )
2023-05-02 00:43:40 -03:00
dependabot[bot]
280c2f285f
Bump safetensors from 0.3.0 to 0.3.1 ( #1720 )
2023-05-02 00:42:39 -03:00
oobabooga
56b13d5d48
Bump llama-cpp-python version
2023-05-02 00:41:54 -03:00
Lőrinc Pap
ee68ec9079
Update folder produced by download-model ( #1601 )
2023-04-27 12:03:02 -03:00
oobabooga
91745f63c3
Use Vicuna-v0 by default for Vicuna models
2023-04-26 17:45:38 -03:00
oobabooga
93e5c066ae
Update RWKV Raven template
2023-04-26 17:31:03 -03:00
oobabooga
c83210c460
Move the rstrips
2023-04-26 17:17:22 -03:00
oobabooga
1d8b8222e9
Revert #1579 , apply the proper fix
...
Apparently models dislike trailing spaces.
2023-04-26 16:47:50 -03:00
TiagoGF
a941c19337
Fixing Vicuna text generation ( #1579 )
2023-04-26 16:20:27 -03:00
oobabooga
d87ca8f2af
LLaVA fixes
2023-04-26 03:47:34 -03:00
oobabooga
9c2e7c0fab
Fix path on models.py
2023-04-26 03:29:09 -03:00
oobabooga
a777c058af
Precise prompts for instruct mode
2023-04-26 03:21:53 -03:00
oobabooga
a8409426d7
Fix bug in models.py
2023-04-26 01:55:40 -03:00
oobabooga
4c491aa142
Add Alpaca prompt with Input field
2023-04-25 23:50:32 -03:00
oobabooga
68ed73dd89
Make API extension print its exceptions
2023-04-25 23:23:47 -03:00
oobabooga
f642135517
Make universal tokenizer, xformers, sdp-attention apply to monkey patch
2023-04-25 23:18:11 -03:00
oobabooga
f39c99fa14
Load more than one LoRA with --lora, fix a bug
2023-04-25 22:58:48 -03:00
oobabooga
15940e762e
Fix missing initial space for LlamaTokenizer
2023-04-25 22:47:23 -03:00
Vincent Brouwers
92cdb4f22b
Seq2Seq support (including FLAN-T5) ( #1535 )
...
---------
Co-authored-by: oobabooga <112222186+oobabooga@users.noreply.github.com>
2023-04-25 22:39:04 -03:00